Observability Engineer

Evolutyz Corp

Brea (CA)

On-site

USD 150,000 - 190,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Evolutyz Corp is seeking a Senior Observability Engineer to design, implement, and optimize modern observability across cloud-native environments. You will work on IaC with Terraform, backend APIs in Python/FastAPI, and observability via OpenTelemetry and Datadog.

The role requires deep experience in cloud infrastructure, distributed systems, and CI/CD, with a focus on reliability and proactive monitoring for critical applications and services.

Qualifications

  • Degree in computer science, engineering, or related field or equivalent practical experience.
  • 10 years of experience in cloud infrastructure, software engineering, or observability engineering.
  • Strong hands-on experience with Terraform and Infrastructure as Code practices.
  • Proficiency in Python and backend API development using FastAPI.
  • Experience implementing observability frameworks using OpenTelemetry.
  • Strong experience with Datadog, including dashboards, monitors, alerting, and SLO management.
  • Solid understanding of distributed systems, cloud-native architectures, and microservices.
  • Experience with CI/CD tools and deployment automation.
  • Strong troubleshooting, analytical, and problem-solving skills.

Responsibilities

  • IaC: design, build, and manage scalable cloud infrastructure using Terraform.
  • Develop reusable infrastructure modules and provisioning standards.
  • Implement and maintain CI/CD pipelines to automate deployments.
  • Improve deployment reliability through automation and best practices.
  • Backend: design, develop, and maintain RESTful APIs using Python and FastAPI.
  • Build services that support observability workflows, telemetry processing, and integrations.
  • Optimize services for performance, scalability, reliability, and maintainability.
  • Observability: design and implement solutions using OpenTelemetry for tracing, metrics, logging.
  • Configure and maintain Datadog dashboards, monitors, alerts, and SLOs.
  • Develop monitoring strategies to improve visibility and reduce incidents.
  • Analyze telemetry to identify performance bottlenecks and reliability issues.
  • Establish alerting thresholds and monitoring standards for production ops.

Skills

Terraform
Python
FastAPI
OpenTelemetry
Datadog
CI/CD
Distributed Systems

Education

Bachelor's degree in CS

Tools

Kubernetes
Prometheus
Grafana
Splunk

Job description

Summary

We are seeking a highly skilled Senior Observability Engineer to design, implement, and optimize modern observability solutions across cloud-native environments. The ideal candidate will have strong expertise in infrastructure automation, backend development, and monitoring platforms, with hands-on experience in Terraform, Python, FastAPI, OpenTelemetry, and Datadog. This role will focus on building scalable observability frameworks, improving system reliability, and enabling proactive monitoring for critical applications and infrastructure.

Responsibilities
  • Infrastructure as Code (IaC):
    • Design, build, and manage scalable cloud infrastructure using Terraform.
    • Develop reusable infrastructure modules and maintain infrastructure provisioning standards.
    • Implement and maintain CI/CD pipelines to automate infrastructure and application deployments.
    • Improve deployment reliability through automation and infrastructure best practices.
  • Backend Development:
    • Design, develop, and maintain RESTful APIs using Python and FastAPI.
    • Build backend services that support observability workflows, telemetry processing, and integrations.
    • Optimize services for performance, scalability, reliability, and maintainability.
    • Troubleshoot application and API performance issues.
  • Observability & Monitoring:
    • Design and implement observability solutions using OpenTelemetry for distributed tracing, metrics, and logging.
    • Configure and maintain Datadog dashboards, monitors, alerts, and Service Level Objectives (SLOs).
    • Develop monitoring strategies to improve system visibility and reduce incident response times.
    • Analyze telemetry data to identify performance bottlenecks and reliability issues.
    • Establish alerting thresholds and monitoring standards to support production operations.
Requirements
  • Degree in Computer Science, Engineering, or related field (or equivalent practical experience).
  • 10 years of experience in cloud infrastructure, software engineering, or observability engineering.
  • Strong hands-on experience with Terraform and Infrastructure as Code practices.
  • Proficiency in Python and backend API development using FastAPI.
  • Experience implementing observability frameworks using OpenTelemetry.
  • Strong experience with Datadog, including dashboards, monitors, alerting, and SLO management.
  • Solid understanding of distributed systems, cloud-native architectures, and microservices.
  • Experience with CI/CD tools and deployment automation.
  • Strong troubleshooting, analytical, and problem-solving skills.
Preferred Skills
  • Experience with cloud platforms such as AWS, Microsoft Azure, or Google Cloud.
  • Knowledge of container orchestration platforms such as Kubernetes.
  • Familiarity with logging and monitoring tools such as Prometheus, Grafana, or Splunk.
  • Experience supporting high-availability production environments.
Benefits

Details about benefits are not provided in the draft. Please specify benefits if applicable.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Observability Engineer: Build Scalable Telemetry
Senior Observability Engineer: Build Scalable Telemetry

Evolutyz Corp • Brea (CA)

On-site
USD 150,000 - 190,000
Cloud Observability Architect
Cloud Observability Architect

Tyler Technologies, Inc. • Plano (TX), Yarmouth (ME)

On-site
USD 131,000 - 192,000
Software Engineer (Observability & Monitoring) - West Des Moines, IA
Software Engineer (Observability & Monitoring) - West Des Moines, IA

AHU Technologies Inc • Washington

On-site
USD 120,000 - 150,000
Sr Observability Engineer
Sr Observability Engineer

IT Associates • Irvine (CA)

On-site
USD 150,000 - 210,000
Senior Observability Engineer – NS2JP00000386
Senior Observability Engineer – NS2JP00000386

Prestige Staffing • Oak Hill (WV)

On-site
USD 100,000 - 130,000
Grafana & Observability Engineer - Dallas, Tampa & Jersey City
Grafana & Observability Engineer - Dallas, Tampa & Jersey City

Stradit LLC • Tampa (FL)

On-site
USD 120,000 - 160,000
Lead Software Engineer – Enterprise Observability
Lead Software Engineer – Enterprise Observability

Humana • Atlanta (GA)

On-site
USD 130,000 - 180,000
Medical
Dental
Vision
+7
Observability Operations Engineer
Observability Operations Engineer

Tata Consultancy Services • Phoenix (AZ)

On-site
USD 100,000 - 120,000
Site Reliability Engineer
Site Reliability Engineer

Infosys • Richardson (TX)

On-site
USD 80,000 - 120,000
Hiring | Dynatrace Observability Lead | Remote | Contract
Hiring | Dynatrace Observability Lead | Remote | Contract

Healthcare Triangle Inc • United States

Remote
USD 170,000 - 210,000