Back fill Engineer

JPC TECHNO INC

Phoenix (AZ)

On-site

USD 120,000 - 180,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

JPC TECHNO INC is seeking a highly skilled Senior Observability Operations Engineer to manage and enhance our enterprise observability platform across Dynatrace, Splunk, and OpenSearch. You will ensure high availability, scalability, and incident resolution.

You will design monitoring, logging, tracing, and alerting, automate tasks with Python and REST APIs, and collaborate with Platform Engineering, SRE, and DevOps teams to drive improvements.

Qualifications

  • Bachelor's degree in CS/IT/Engineering or equivalent experience.
  • 6-10+ years IT infrastructure or observability operations experience.
  • 4+ years administering Dynatrace, Splunk, OpenSearch, or Elasticsearch.
  • Strong troubleshooting and analytical skills.
  • Excellent communication and stakeholder management skills.

Responsibilities

  • Administer and optimize enterprise observability platforms including Dynatrace, Splunk, and OpenSearch/Elasticsearch.
  • Design, deploy, configure, and maintain monitoring, logging, tracing, and alerting solutions.
  • Manage large-scale OpenSearch/Elasticsearch clusters, including indexing strategies and capacity planning.
  • Configure Dynatrace components such as OneAgent, APM, and RUM.
  • Develop dashboards, alerts, and executive operational metrics.
  • Automate operational tasks using Python, Shell, REST APIs, Terraform, or Ansible.
  • Collaborate with Platform Engineering, SRE, DevOps, and Infra teams.

Skills

Strong troubleshooting
Excellent communication
Stakeholder management
Independent worker
Ownership mindset

Education

Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience

Tools

Dynatrace
Splunk
OpenSearch/Elasticsearch
Kubernetes
Linux
OpenTelemetry
Grafana
Kibana
Docker
OpenShift
Rancher
Terraform
Ansible
REST APIs
Python
AI/ML in Observability

Job description

We are seeking a highly skilled Senior Observability Operations Engineer to manage and enhance our enterprise observability platform. The ideal candidate will have deep expertise in Dynatrace, Splunk, OpenSearch/Elasticsearch, Kuberetes, Linux, and cloud-native observability solutions.

Experience leveraging AI/ML and Generative Al to improve observability, automate operations, and accelerate incident resolution is highly desirable.

The role is responsible for ensuring high availability, scalability, operational excellence, and continuous improvement of enterprise monitoring and logging platforms supporting mission-critical applications.

Key Responsibilities
  • Administer and optimize enterprise observability platforms including Dynatrace, Splunk, and OpenSearch/Elasticsearch.
  • Design, deploy, configure, and maintain monitoring, logging, tracing, and alerting solutions.
  • Manage large-scale OpenSearch/Elasticsearch clusters, including indexing strategies, performance tuning, shard optimization, backups, and capacity planning.
  • Configure Dynatrace OneAgent, ActiveGate, Synthefic Monitoring, Real User Monitoring (RUM), Digital Experience Monitoring (DEM), Davis Al, and Application Performance Monitoring (APM).
  • Administer Splunk Enterprise, Universal Forwarders, Indexers, Search Heads, Cluster Manager, Deployment Server, and Splunk ITSI.
  • Develop dashboards; alerts, reports, and executive operational metrics.
  • Support Linux-based infrastructure and Kubernetes environments (Docker/OpenShift/Rancher preferred).
  • Implement observability best practices using OpenTelemetry, distributed tracing, metrics, logs, and events.
  • Perform root cause analysis for production incidents using observability platforms.
  • Collaborate with Platform Engineering, SRE, DevOps, Infrastructure, and Application teams.
  • Automate operational tasks using Python, Shell scripting, REST APls, Terraform, or Ansible.
  • Participate in incident, problem, change, and release management processes.
  • Drive platform upgrades, patching, security compliance, and operational governance.
  • Improve platform reliability through automation, self-healing, and Al-assisted operations.
Observability Platforms
  • Dynatrace Administration
  • OpenSearch Administration
  • Elasticsearch Administration
  • Grafana
  • Kibana
  • Open Telemetry
  • Kafka (preferred)
Infrastructure
  • Docker
  • OpenShift or Rancher
  • Networking (TCP/IP, DNS, Load Balancers, Firewalls)
  • System Administration
  • Git
  • Terraform
  • Ansible
  • REST APIs
Scripting
  • Python
  • Bash/Shell
  • PowerShell (preferred)
AI & Automation Skills (Preferred)
  • Experience using Generative AI (ChatGPT, GitHub Copilot, Amazon Q, Microsoft Copilot, or similar) to improve operational efficiency.
  • Knowledge of AlOps platforms and Al-driven observability.
  • Experience with Dynatrace Davis Al for anomaly detection and root cause analysis.
  • Understanding of machine learning concepts for predictive monitoring and intelligent alerting.
  • Experience building Al-assisted operational runbooks and troubleshooting workflows.
  • Knowledge of Retrieval-Augmented Generation (RAG), vector databases, embeddings, and Al-powered knowledge search is a plus.
  • Experience integrating Al with observability platforms using APIs.
  • Familiarity with LLMs, prompt engineering, and Al-assisted automation.
  • Experience using Python with Al frameworks (LangChain, LangGraph, OpenAI APls, or similar) is desirable.
  • Exposure to Al-driven incident summarization, log analysis, and automated ticket enrichment.
Required Qualifications
  • Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience.
  • 6-10+ years of IT infrastructure or observability operations experience.
  • 4+ years administering Dynatrace, Splunk, OpenSearch, or Elasticsearch.
  • Strong troubleshooting and analytical skills.
  • Excellent communication and stakeholder management Skills.
Preferred Certifications
  • Dynatrace Associate or Professional Certification
  • Elastic Certified Engineer
  • ITIL Foundation
Soft Skills
  • Strong ownership and accountability
  • Excellent problem-solving and analytical thinking
  • Ability to work independently with minimal supervision
  • Strong collaboration across cross-functional teams
  • Continuous learning mindset
  • Ability to thrive in fast-paced production environments
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Observability Operations Engineer
Observability Operations Engineer

Tata Consultancy Services • Phoenix (AZ)

On-site
USD 100,000 - 120,000
Observability Engineer (Splunk & Dynatrace)
Observability Engineer (Splunk & Dynatrace)

Veriipro • Plano (TX)

On-site
USD 110,000 - 160,000
Sr. Observability Engineer
Sr. Observability Engineer

FreedomPay • Select (KY)

On-site
USD 150,000 - 210,000
Solution Architect / Team Lead - Observability
Solution Architect / Team Lead - Observability

VOLTO Consulting • Irvine (CA)

On-site
USD 120,000 - 160,000
Observability Architect
Observability Architect

TechDigital Group • Atlanta (GA)

On-site
USD 120,000 - 150,000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

TechDigital Group • Tyson (AZ)

On-site
USD 120,000 - 160,000
Local_Observability Operations Engineer
Local_Observability Operations Engineer

Themesoft Inc • Phoenix (AZ)

On-site
USD 110,000 - 140,000
Principal Observability Architect (Splunk & Databricks)
Principal Observability Architect (Splunk & Databricks)

Scicominfra • Atlanta (GA)

On-site
USD 140,000 - 180,000
Health insurance
401(k) retirement plan
Paid time off
Senior Dynatrace Engineer
Senior Dynatrace Engineer

GlobalPoint • Georgia

On-site
USD 90,000 - 120,000
Senior Observability Engineer – NS2JP00000386
Senior Observability Engineer – NS2JP00000386

Prestige Staffing • Oak Hill (WV)

Remote
USD 100,000 - 130,000