Senior Observability Engineer

RBC Capital Markets, LLC

Minneapolis (MN)

On-site

USD 90,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Bonus opportunities
Stock options where applicable
Flexible benefits

Job summary

RBC Capital Markets, LLC in Minneapolis seeks an experienced observability engineer to design and implement monitoring solutions, establish logging standards, and partner with engineering teams to improve system visibility and reliability.

You will build dashboards, evaluate tools, use AI for anomaly detection, and mentor others while promoting best practices across the organization. This role offers a competitive salary and a collaborative, high-performance environment.

Qualifications

  • 5+ years of experience with observability tools and platforms such as ELK, Dynatrace, Prometheus or Grafana.
  • 3+ years of software engineering or infrastructure experience.
  • Expert-level knowledge of logging requirements and best practices for observability.
  • Demonstrated experience building and optimizing monitoring dashboards.
  • Proven ability to use observability data to proactively identify and resolve system issues.
  • Experience using AI tools for anomaly detection and trend analysis.
  • Linux/Unix knowledge.

Responsibilities

  • Lead observability initiatives and establish standards.
  • Partner with engineering teams to improve system visibility, reliability, and performance.
  • Develop and maintain monitoring dashboards for actionable insights.
  • Evaluate and recommend observability tools and vendors.
  • Analyze logs, metrics, and traces to identify issues and translate data into insights.
  • Collaborate with cross-functional teams to communicate observability concepts to stakeholders.

Skills

Observability tools
Python
Java
Go
Software engineering experience
Logging best practices
Monitoring dashboards
AI anomaly detection
Linux/Unix
Query languages

Tools

ELK Stack
Dynatrace
Prometheus
Grafana
OpenTelemetry
Jaeger
Aternity
OpenShift
Kubernetes

Job description

Job Description

What will you do?

  • Design and implement observability solutions using industry-leading platforms, establishing logging standards that enable comprehensive system visibility across all systems
  • Create and maintain monitoring dashboards that provide actionable insights into system health and performance, partnering with platform and application teams to integrate observability into architecture
  • Evaluate and recommend observability tools and vendors to ensure the organization has access to best-in-class solutions
  • Analyze logs, metrics, and traces to proactively identify system issues, performance bottlenecks, and translate observability data into insights about user patterns and system behavior
  • Develop predictive monitoring strategies using AI tools to detect anomalies, identify emerging trends, and prevent incidents before they occur
  • Conduct "what if" analysis using AI capabilities to model potential scenarios and their impact on system performance
  • Apply machine learning-based anomaly detection to identify issues proactively and use predictive analytics to forecast system behavior and prevent failures
  • Collaborate with cross-functional teams (Engineering, DevOps, Security, production support, product) to communicate complex observability concepts to both technical and non-technical stakeholders
  • Lead observability initiatives and mentor junior team members on best practices while facilitating design and problem-solving discussions across the organization
Responsibilities

Lead and execute observability initiatives, establish standards, and partner with engineering teams to improve system visibility, reliability, and performance.

Qualifications
  • 5+ years of experience with observability tools (ELK Stack, Dynatrace, Prometheus, Grafana, OpenTelemetry, Jaeger, Aternity or similar)
  • 3+ years of software engineering or infrastructure experience
    • Python, Java, Go
    • Query languages
  • Expert-level knowledge of logging requirements and best practices for enhanced observability
  • Demonstrated experience building and optimizing monitoring dashboards
  • Proven ability to use observability data to proactively identify and resolve system issues
  • Experience using AI tools for anomaly detection and trend analysis
  • Linux/Unix knowledge
Nice to have
  • Expertise with Tableau or advanced visualization tools
  • Experience in financial technology and financial services environments
  • Previous experience with AI tools such as Anthropic, OpenAI, Devin, Copilot, etc.
  • Experience with OpenShift, Kubernetes and containerized applications
  • Background in DevOps or Site Reliability Engineering (SRE)
  • Experience conducting "what if" analysis and scenario modeling
  • Identify networking slowness and availability issues
What7s in it for you?
  • A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, and stock where applicable
  • Leaders who support your development through coaching and mentoring opportunities
  • Ability to make a difference and lasting impact on system reliability and user experience
  • Work in a dynamic, collaborative, progressive, and high-performing team

The expected salary range for this position is $90,000-$140,000, depending on your experience, skills, and market conditions.

You have the potential to earn more through discretionary variable compensation based on business performance and individual goals.

Job Details
  • Address: 250 Nicollet Mall, Minneapolis, MN, United States
  • City: Minneapolis
  • Country: United States of America
  • Work hours/week: 40
  • Employment Type: Full time
  • Platform: Technology and Operations
  • Job Type: Regular
  • Pay Type: Salaried

RBC is an equal opportunity employer. We are committed to fostering an inclusive workplace that values diverse perspectives. We provide policies and programs to support belonging and opportunity for all employees.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

RBC • Minneapolis (MN)

On-site
USD 80,000 - 140,000
Senior Observability Engineer — AI-Driven Reliability & Dashboards
Senior Observability Engineer — AI-Driven Reliability & Dashboards

RBC Capital Markets, LLC • Minneapolis (MN)

On-site
USD 90,000 - 140,000
Bonus opportunities
Stock options where applicable
Flexible benefits
Staff Software Engineer, Observability
Staff Software Engineer, Observability

United States Digital Space LLC • Menlo Park (CA)

On-site
USD 180,000 - 250,000
Health insurance
Equity ownership
401(k) matching
+1
Observability Engineer
Observability Engineer

Point72 • New York (NY)

On-site
USD 175,000 - 250,000
Fully-paid health care benefits
Generous parental and family leave
Mental and physical wellness programs
+2
Senior Observability Engineer
Senior Observability Engineer

Tata Consultancy Services • Los Angeles (CA)

On-site
USD 120,000 - 130,000
Senior Systems Engineer
Senior Systems Engineer

Cox Automotive Inc. • Atlanta (GA)

On-site
USD 92,000 - 154,000
Observability Engineer
Observability Engineer

ManpowerGroup Global, Inc. • Denver (CO), Town of Norway (WI)

On-site
USD 117,000 - 143,000
Medical and Prescription Drug Plans
Dental Plan
Vision Plan
+3
Staff Software Engineer, Observability
Staff Software Engineer, Observability

Robinhood • Edison (CA)

Hybrid
USD 180,000 - 240,000
Health insurance
401(k) matching
Equity
+3
Senior Manager, IT Observability
Senior Manager, IT Observability

ViziRecruiter,LLC. • Salisbury (NC)

On-site
USD 120,000 - 182,000
Senior Site Reliability Engineer, Observability New York, NY, United States
Senior Site Reliability Engineer, Observability New York, NY, United States

Ripple • New York (NY)

On-site
USD 160,000 - 200,000