Senior Observability Engineer

FanDuel

Atlanta (GA)

Hybrid

USD 149,000 - 186,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, vision, and dental insurance
401(k) matching program
Paid personal time off

Job summary

FanDuel in Atlanta, Georgia is seeking a Senior Observability Engineer to design and enhance the observability ecosystem that supports our platform. This hands-on role involves collaboration with engineering and product teams to deliver scalable observability capabilities and improve system reliability.

The ideal candidate will have experience in observability engineering and tools like Datadog, with the ability to influence technical decisions and enhance monitoring practices across teams.

Qualifications

  • Solid hands-on experience in observability engineering or related roles.
  • Strong expertise in monitoring and observability with tools like Datadog.
  • Experience with Kubernetes and cloud infrastructure (AWS).

Responsibilities

  • Contribute to the observability strategy and roadmap.
  • Design scalable observability solutions for actionable insights.
  • Collaborate to improve system reliability and incident management.

Skills

Monitoring and observability practices
Kubernetes
AWS
Infrastructure as code (Terraform)
Programming (Go, Java, Python, TypeScript)
Analytical and problem-solving skills
Collaboration with stakeholders

Education

Hands-on experience in observability engineering, SRE, or platform engineering
Experience implementing SLOs and SLIs

Tools

Datadog
PagerDuty
Ansible
Helm

Job description

The Position

Senior Observability Engineer

FanDuel is looking for a Senior Observability Engineer to design, build, and mature the observability ecosystem that underpins our platform and services. You will deliver deep visibility into system behavior by combining system telemetry with user signals to provide a holistic view of performance, reliability, and user experience. You’ll explore how AI and machine learning can enhance observability, from intelligent alerting and anomaly detection to accelerating root cause analysis.

This is a hands‑on role. You’ll partner closely with engineering and product teams to deliver scalable observability capabilities, serve as a subject‑matter expert in monitoring, alerting, and incident management, and equip teams with self‑service insights and tooling. By connecting system behavior to real user impact and leveraging AI‑assisted workflows to surface issues faster, you’ll drive improvements in reliability, performance, and data‑informed decision‑making across the organization.

Responsibilities
  • Contribute to the observability strategy and roadmap, partnering with multiple teams to align with business priorities and engineering goals.
  • Design and enhance scalable observability solutions that provide actionable insights into system health, performance, and user experience.
  • Help establish and promote best practices for monitoring, alerting, incident management, and postmortems across teams.
  • Support operational excellence by improving incident response processes, on‑call practices, and post‑incident reviews, focusing on continuous improvement.
  • Collaborate on cross‑team initiatives to improve system reliability, identify risks, and contribute to their resolution.
  • Apply automation and AI‑assisted workflows to improve root cause analysis and reduce operational toil.
  • Work with engineering and product stakeholders to surface observability insights that inform technical decisions and prioritization.
  • Analyze system and user signals to help detect, prevent, and mitigate reliability issues.
  • Contribute to optimizing observability platforms for performance, scalability, and cost‑efficiency.
  • Mentor peers and contribute to raising observability and reliability standards within the team.
Tech Stack

AWS, Kubernetes, Terraform, Helm, Ansible, Vault, Datadog, and PagerDuty.

Qualifications
  • Solid hands‑on experience in observability engineering, SRE, platform engineering, or related roles, with impact across team‑level systems.
  • Strong expertise in monitoring and observability practices, with hands‑on experience using tools such as Datadog.
  • Experience contributing to observability or reliability initiatives across teams or services.
  • Proficiency with Kubernetes, cloud infrastructure (e.g., AWS), and infrastructure‑as‑code tools such as Terraform.
  • Ability to influence technical decisions within and across teams, collaborating effectively with a range of stakeholders.
  • Good understanding of distributed systems principles (e.g., consistency, availability, partition tolerance) and practical trade‑offs.
  • Experience defining and implementing SLOs, SLIs, and alerting strategies, including an understanding of user‑impacting metrics.
  • Strong software engineering fundamentals, with proficiency in at least one modern programming language (Go, Java, Python, or TypeScript) and experience building tooling, automation, and scalable systems.
  • Experience improving systems through automation, helping reduce operational toil and recurring issues.
  • Strong analytical and problem‑solving skills, with the ability to interpret technical signals and relate them to system performance and reliability.
  • Good communication and collaboration skills, with the ability to work effectively with both technical and non‑technical stakeholders.
  • A sense of ownership and accountability, with a focus on delivering reliable, scalable solutions and continuous improvement.

Don’t check all the boxes? That’s okay! We encourage you to still apply if you feel you possess an adjacent skill set and are interested in learning more about this position.

Benefits & Compensation

We offer a competitive salary range of $149,000 – $186,000 USD, which is dependent on relevant experience, location, business needs, and market demand. Additional benefits include: medical, vision, and dental insurance; life insurance; disability insurance; a 401(k) matching program; short‑term and long‑term incentive compensation; paid personal time off; 14 paid company holidays; and paid sick time in accordance with applicable state and federal laws.

Equal Employment Opportunity

FanDuel is an equal opportunity employer and we believe that our workforce is at our core value of “We are One Team!”. The company is committed to equal employment opportunity regardless of race, color, ethnicity, ancestry, religion, creed, sex, national origin, sexual orientation, age, citizenship status, marital status, disability, gender identity, gender expression, veteran status, or any other characteristic protected by state, local, or federal law.

Reasonable Accommodation for Disabilities

FanDuel is committed to providing reasonable accommodations for qualified individuals with disabilities. If you have a disability and need a workplace accommodation or adjustment during the application or hiring process, including support for the interview or onboarding process, please emailBenefits@fanduel.com.

Legal Notice (Massachusetts)

It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Observability Engineer Atlanta, Georgia, United States
Senior Observability Engineer Atlanta, Georgia, United States

FanDuel • Atlanta (GA)

On-site
USD 149,000 - 186,000
Medical, vision, and dental insurance
401(k) matching program
Paid personal time off
+2
Senior Observability Engineer New York City
Senior Observability Engineer New York City

FanDuel • New York (NY)

Hybrid
USD 149,000 - 186,000
Health plans
Mental health support
401(k) matching
+2
Staff Observability Engineer
Staff Observability Engineer

FanDuel • New York (NY)

On-site
USD 170,000 - 213,000
Health plans
401k with up to a 5% match
Generous paid time off
+1
Senior Observability Engineer
Senior Observability Engineer

Omaze • New Jersey

On-site
USD 149,000 - 186,000
Health insurance
401(k) matching up to 5%
Generous paid time off
+2
Senior Observability Engineer
Senior Observability Engineer

Omaze • New York (NY)

On-site
USD 149,000 - 186,000
Health plans including mental health support
Generous PTO and sick leave
401(k) with up to 5% match
+2
Senior Observability Engineer
Senior Observability Engineer

FanDuel • Jersey City (NJ)

On-site
USD 149,000 - 186,000
Health insurance
401k with up to a 5% match
Generous paid time off
+1
Senior Observability Engineer
Senior Observability Engineer

Omaze • Atlanta (GA)

On-site
USD 149,000 - 186,000
Health plans with low premiums
Generous paid time off
401(k) matching program
+2
Senior AI Software Engineer
Senior AI Software Engineer

FanDuel • New York (NY)

On-site
USD 149,000 - 186,000
Health plans with low costs
Generous paid time off
401(k) match up to 5%
+1
Staff Infrastructure Engineer
Staff Infrastructure Engineer

FanDuel • Atlanta (GA)

On-site
USD 159,000 - 209,000
Medical, vision, and dental insurance
401(k) matching program
Paid time off and holidays
+1
Senior Product Analyst
Senior Product Analyst

Omaze • New York (NY)

On-site
USD 110,000 - 160,000
Health plans
Paid time off
Annual bonus
+4