Observability Operations Engineer

Tata Consultancy Services

Phoenix (AZ)

On-site

USD 100,000 - 120,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Tata Consultancy Services is seeking an Observability Operations Engineer in Phoenix to administer and optimize Dynatrace, Splunk, and OpenSearch platforms. You will lead monitoring, logging, tracing, and alerting with Grafana/Prometheus/OpenTelemetry, and drive automation using Python and shell scripts.

Responsibilities include supporting Linux, Kubernetes, cloud environments, and performing RCA on production incidents. Collaboration with SRE, DevOps, and Platform teams is essential.

Qualifications

  • Strong Observability Administration experience with Dynatrace, Splunk, and OpenSearch/Elasticsearch.
  • Hands-on experience with monitoring, logging, tracing, alerting, dashboards, and platform performance tuning.
  • Strong knowledge of Linux, Kubernetes, and cloud environments.
  • Experience with Grafana, Prometheus, OpenTelemetry, and related observability technologies.
  • Automation experience using Python/Shell scripting and REST APIs.
  • Experience supporting enterprise-scale production environments, troubleshooting, and RCA.

Responsibilities

  • Administer and optimize enterprise Dynatrace, Splunk, and OpenSearch/Elasticsearch platforms.
  • Maintain platform availability, scalability, performance, security, and reliability.
  • Build and manage monitoring, logging, tracing, dashboards, alerts, and operational metrics.
  • Troubleshoot production issues and perform root cause analysis using observability tools.
  • Support Linux, Kubernetes, container, and cloud-based environments.
  • Automate operational activities and drive self-healing and AI-assisted operations.
  • Manage upgrades, patching, capacity planning, backups, and operational governance.
  • Collaborate with SRE, DevOps, Platform, Infrastructure, and Application teams.

Skills

Observability Administration
Monitoring and logging
Linux Kubernetes cloud
Grafana Prometheus OpenTelemetry
Python Shell scripting REST APIs
Production troubleshooting RCA
SRE DevOps collaboration
Automation for self-healing

Tools

Dynatrace
Splunk
OpenSearch/Elasticsearch
Grafana
Prometheus
OpenTelemetry

Job description

  • Strong Observability Administration experience with Dynatrace, Splunk, and OpenSearch/Elasticsearch.
  • Hands-on experience with monitoring, logging, tracing, alerting, dashboards, and platform performance tuning.
  • Strong knowledge of Linux, Kubernetes, and cloud environments.
  • Experience with Grafana, Prometheus, OpenTelemetry, and related observability technologies.
  • Automation experience using Python/Shell scripting and REST APIs.
  • Experience supporting enterprise-scale production environments, troubleshooting, and RCA.
  • Administer and optimize enterprise Dynatrace, Splunk, and OpenSearch/Elasticsearch platforms.
  • Maintain platform availability, scalability, performance, security, and reliability.
  • Build and manage monitoring, logging, tracing, dashboards, alerts, and operational metrics.
  • Troubleshoot production issues and perform root cause analysis using observability tools.
  • Support Linux, Kubernetes, container, and cloud-based environments.
  • Automate operational activities and drive self-healing and AI-assisted operations.
  • Manage upgrades, patching, capacity planning, backups, and operational governance.
  • Collaborate with SRE, DevOps, Platform, Infrastructure, and Application teams.
Job Description

Observability Operations Engineer

Must Have Technical/Functional Skills
  • Strong Observability Administration experience with Dynatrace, Splunk, and OpenSearch/Elasticsearch.
  • Hands‑on experience with monitoring, logging, tracing, alerting, dashboards, and platform performance tuning.
  • Strong knowledge of Linux, Kubernetes, and cloud environments.
  • Experience with Grafana, Prometheus, OpenTelemetry, and related observability technologies.
  • Automation experience using Python/Shell scripting and REST APIs.
  • Experience supporting enterprise‑scale production environments, troubleshooting, and RCA.
Roles & Responsibilities
  • Administer and optimize enterprise Dynatrace, Splunk, and OpenSearch/Elasticsearch platforms.
  • Maintain platform availability, scalability, performance, security, and reliability.
  • Build and manage monitoring, logging, tracing, dashboards, alerts, and operational metrics.
  • Troubleshoot production issues and perform root cause analysis using observability tools.
  • Support Linux, Kubernetes, container, and cloud-based environments.
  • Automate operational activities and drive self-healing and AI-assisted operations.
  • Manage upgrades, patching, capacity planning, backups, and operational governance.
  • Collaborate with SRE, DevOps, Platform, Infrastructure, and Application teams.

Generic Managerial Skills, If any

Good Communication and assertiveness, Team Player

Salary Range- $100,000-$120,000 a year

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr Observability Engineer
Sr Observability Engineer

IT Associates • Irvine (CA)

On-site
USD 150,000 - 210,000
Site Reliability Engineer
Site Reliability Engineer

BlueSky Resource Solutions • Duluth (GA)

On-site
USD 120,000 - 180,000
Senior Observability Engineer – NS2JP00000386
Senior Observability Engineer – NS2JP00000386

Prestige Staffing • Oak Hill (WV)

On-site
USD 100,000 - 130,000
Senior/Lead Site Reliability Engineer Observability
Senior/Lead Site Reliability Engineer Observability

Tata Consultancy Services • San Jose (CA)

On-site
USD 94,000 - 130,000
Operational Data & Observability Engineer
Operational Data & Observability Engineer

Nscale • Seattle (WA)

Hybrid
USD 145,000 - 180,000
Medical, dental, vision
Flexible PTO
Parental leave
+1
Dynatrace Observability Consultant
Dynatrace Observability Consultant

Infosat IT Services LLC • United States

Remote
USD 150,000 - 190,000
Observability Engineer
Observability Engineer

VMC Soft Technologies, Inc • Allen (TX)

On-site
USD 90,000 - 120,000
Grafana & Observability Engineer - Dallas, Tampa & Jersey City
Grafana & Observability Engineer - Dallas, Tampa & Jersey City

Stradit LLC • Tampa (FL)

On-site
USD 120,000 - 160,000
SRE Observability Architect
SRE Observability Architect

New York Technology Partners • United States

On-site
USD 103,320 - 110,208
Hiring | Dynatrace Observability Lead | Remote | Contract
Hiring | Dynatrace Observability Lead | Remote | Contract

Healthcare Triangle Inc • United States

Remote
USD 170,000 - 210,000