Observability Operations Engineer

Tata Consultancy Services

Phoenix (AZ)

On-site

USD 100,000 - 120,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Tata Consultancy Services is seeking an Observability Operations Engineer in Phoenix to administer and optimize Dynatrace, Splunk, and OpenSearch platforms. You will lead monitoring, logging, tracing, and alerting with Grafana/Prometheus/OpenTelemetry, and drive automation using Python and shell scripts.

Responsibilities include supporting Linux, Kubernetes, cloud environments, and performing RCA on production incidents. Collaboration with SRE, DevOps, and Platform teams is essential.

Qualifications

  • Strong Observability Administration experience with Dynatrace, Splunk, and OpenSearch/Elasticsearch.
  • Hands-on experience with monitoring, logging, tracing, alerting, dashboards, and platform performance tuning.
  • Strong knowledge of Linux, Kubernetes, and cloud environments.
  • Experience with Grafana, Prometheus, OpenTelemetry, and related observability technologies.
  • Automation experience using Python/Shell scripting and REST APIs.
  • Experience supporting enterprise-scale production environments, troubleshooting, and RCA.

Responsibilities

  • Administer and optimize enterprise Dynatrace, Splunk, and OpenSearch/Elasticsearch platforms.
  • Maintain platform availability, scalability, performance, security, and reliability.
  • Build and manage monitoring, logging, tracing, dashboards, alerts, and operational metrics.
  • Troubleshoot production issues and perform root cause analysis using observability tools.
  • Support Linux, Kubernetes, container, and cloud-based environments.
  • Automate operational activities and drive self-healing and AI-assisted operations.
  • Manage upgrades, patching, capacity planning, backups, and operational governance.
  • Collaborate with SRE, DevOps, Platform, Infrastructure, and Application teams.

Skills

Observability Administration
Monitoring and logging
Linux Kubernetes cloud
Grafana Prometheus OpenTelemetry
Python Shell scripting REST APIs
Production troubleshooting RCA
SRE DevOps collaboration
Automation for self-healing

Tools

Dynatrace
Splunk
OpenSearch/Elasticsearch
Grafana
Prometheus
OpenTelemetry

Job description

  • Strong Observability Administration experience with Dynatrace, Splunk, and OpenSearch/Elasticsearch.
  • Hands-on experience with monitoring, logging, tracing, alerting, dashboards, and platform performance tuning.
  • Strong knowledge of Linux, Kubernetes, and cloud environments.
  • Experience with Grafana, Prometheus, OpenTelemetry, and related observability technologies.
  • Automation experience using Python/Shell scripting and REST APIs.
  • Experience supporting enterprise-scale production environments, troubleshooting, and RCA.
  • Administer and optimize enterprise Dynatrace, Splunk, and OpenSearch/Elasticsearch platforms.
  • Maintain platform availability, scalability, performance, security, and reliability.
  • Build and manage monitoring, logging, tracing, dashboards, alerts, and operational metrics.
  • Troubleshoot production issues and perform root cause analysis using observability tools.
  • Support Linux, Kubernetes, container, and cloud-based environments.
  • Automate operational activities and drive self-healing and AI-assisted operations.
  • Manage upgrades, patching, capacity planning, backups, and operational governance.
  • Collaborate with SRE, DevOps, Platform, Infrastructure, and Application teams.
Job Description

Observability Operations Engineer

Must Have Technical/Functional Skills
  • Strong Observability Administration experience with Dynatrace, Splunk, and OpenSearch/Elasticsearch.
  • Hands‑on experience with monitoring, logging, tracing, alerting, dashboards, and platform performance tuning.
  • Strong knowledge of Linux, Kubernetes, and cloud environments.
  • Experience with Grafana, Prometheus, OpenTelemetry, and related observability technologies.
  • Automation experience using Python/Shell scripting and REST APIs.
  • Experience supporting enterprise‑scale production environments, troubleshooting, and RCA.
Roles & Responsibilities
  • Administer and optimize enterprise Dynatrace, Splunk, and OpenSearch/Elasticsearch platforms.
  • Maintain platform availability, scalability, performance, security, and reliability.
  • Build and manage monitoring, logging, tracing, dashboards, alerts, and operational metrics.
  • Troubleshoot production issues and perform root cause analysis using observability tools.
  • Support Linux, Kubernetes, container, and cloud-based environments.
  • Automate operational activities and drive self-healing and AI-assisted operations.
  • Manage upgrades, patching, capacity planning, backups, and operational governance.
  • Collaborate with SRE, DevOps, Platform, Infrastructure, and Application teams.

Generic Managerial Skills, If any

Good Communication and assertiveness, Team Player

Salary Range- $100,000-$120,000 a year

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Back fill Engineer
Back fill Engineer

JPC TECHNO INC • Phoenix (AZ)

On-site
USD 120,000 - 180,000
Local_Observability Operations Engineer
Local_Observability Operations Engineer

Themesoft Inc • Phoenix (AZ)

On-site
USD 110,000 - 140,000
Sr. Observability Engineer
Sr. Observability Engineer

FreedomPay • Select (KY)

On-site
USD 150,000 - 210,000
Observability Engineer (Splunk & Dynatrace)
Observability Engineer (Splunk & Dynatrace)

Veriipro • Plano (TX)

On-site
USD 110,000 - 160,000
Observability Engineer
Observability Engineer

Insight Global • New York (NY)

On-site
USD 120,000 - 180,000
Senior Observability Engineer
Senior Observability Engineer

Tata Consultancy Services • Los Angeles (CA)

On-site
USD 120,000 - 130,000
Observability Architect
Observability Architect

TechDigital Group • Atlanta (GA)

On-site
USD 120,000 - 150,000
Sr Observability Engineer
Sr Observability Engineer

IT Associates • Irvine (CA)

Hybrid
USD 150,000 - 210,000
Senior Observability Engineer – NS2JP00000386
Senior Observability Engineer – NS2JP00000386

Prestige Staffing • Oak Hill (WV)

Remote
USD 100,000 - 130,000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

TechDigital Group • Tyson (AZ)

On-site
USD 120,000 - 160,000