Engineering Manager

Nabla Infotech LLC

Phoenix (AZ)

Hybrid

USD 180,000 - 240,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Nabla Infotech LLC is seeking an Engineering Manager to lead a greenfield AI Ops and Autonomous Operations Platform. You will guide the architecture, roadmap, and delivery strategy, directing teams across SRE, AI engineering, platform engineering, and automation to build proactive monitoring, self-healing systems, and enterprise observability.

You will establish standards for reliability and operational intelligence, drive cost optimization, and partner with Infrastructure, Cloud, Security,

Qualifications

  • Experience leading enterprise AI ops platforms and SRE teams.
  • Proven ability to design and implement observability, reliability, and incident management frameworks.
  • Strong background in AI/ML driven operations and automation.

Responsibilities

  • Lead greenfield AI Ops and Autonomous Operations platform development.
  • Define AIOps roadmap, architecture, and delivery strategy.
  • Manage teams of SREs, AI/ML, Platform, and Automation engineers.
  • Implement AI-driven event correlation and root cause analysis.
  • Drive automation of incident triage, diagnosis, and remediation workflows.
  • Develop GenAI/Agentic AI solutions for operational support.
  • Build observability using OpenTelemetry and modern monitoring tools.
  • Deliver executive dashboards tied to business outcomes and customer experience.

Skills

AIOps
SRE
Observability
AI/ML
GenAI
Agentic AI
Kubernetes
Cloud Platforms
Incident Management
Automation
Platform Engineering
Reliability Engineering

Tools

OpenTelemetry
Datadog
Splunk
ServiceNow
Kubernetes
Cloud Platforms

Job description

Location: Phoenix, AZ (Hybrid)
Work Stream - AI Ops
Engineering Manager – AI Ops & Autonomous Operations Platform

We are seeking an Engineering Manager to build and lead an enterprise AI Ops platform that transforms IT operations through AI, automation, and observability.

  • Lead the development of a greenfield AI Ops and Autonomous Operations platform.
  • Build capabilities for proactive monitoring, intelligent alerting, and self-healing systems.
  • Establish enterprise standards for observability, reliability, and operational intelligence.
  • Own the AIOps roadmap, architecture, and delivery strategy.
  • Lead teams of SREs, AI Engineers, Platform Engineers, and Automation Engineers.
  • Implement AI-driven event correlation and root cause analysis.
  • Drive automation of incident triage, diagnosis, and remediation workflows.
  • Develop GenAI and Agentic AI solutions for operational support.
  • Leverage telemetry data from logs, metrics, traces, events, and CMDB.
  • Build enterprise observability solutions using OpenTelemetry and modern monitoring platforms.
  • Define and mature SLI, SLO, SLA, and error-budget frameworks.
  • Improve platform reliability, resilience, and operational efficiency.
  • Reduce alert noise and operational toil through intelligent automation.Partner with Infrastructure, Cloud, Security, Risk, and Application teams.
  • Lead major incident management and post-incident improvement programs.
  • Establish AI-assisted operations practices and governance models.
  • Drive cost optimization, capacity forecasting, and predictive operations.
  • Deliver executive dashboards aligned to business outcomes and customer experience.
  • Build a self-service operational intelligence platform for engineering teams.
  • Create measurable improvements in MTTR, MTTA, availability, and engineering productivity.
Key Skills:

AIOps, SRE, Observability, OpenTelemetry, Datadog, Splunk, ServiceNow, AI/ML, GenAI, Agentic AI, Kubernetes, Cloud Platforms, Incident Management, Automation, Platform Engineering, Reliability Engineering.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Manager, AI Ops & Autonomous Platform
Engineering Manager, AI Ops & Autonomous Platform

Nabla Infotech LLC • Phoenix (AZ)

Hybrid
USD 180,000 - 240,000
Managing Director-Delivery
Managing Director-Delivery

Anblicks • Dallas (TX)

On-site
USD 180,000 - 240,000
AI Automation Engineer II
AI Automation Engineer II

iSpace, Inc. • Denver (CO)

On-site
USD 120,000 - 160,000
Senior AIOps and Incident Management / Site Reliability Engineering C2C jobs
Senior AIOps and Incident Management / Site Reliability Engineering C2C jobs

Tech Mirrors • Fort Mill (SC)

Hybrid
USD 140,000 - 190,000
Senior Site Reliability Engineer – AI & Automation.
Senior Site Reliability Engineer – AI & Automation.

Veriipro • Miami (FL)

On-site
USD 130,000 - 170,000
Senior Forward Deployed Engineer (DevOps/SRE)
Senior Forward Deployed Engineer (DevOps/SRE)

LeoForce • Pleasanton (CA)

On-site
USD 300,000 - 350,000
Medical benefits
401(k) plan
Equity
+1
Senior Software Engineering Manager
Senior Software Engineering Manager

Applied Resource Group • Atlanta (GA)

On-site
USD 180,000 - 240,000
Senior AI DevOps Engineer (AI Ops / Platform Engineering)
Senior AI DevOps Engineer (AI Ops / Platform Engineering)

DeepCamp • Tucker (GA)

On-site
USD 96,000 - 165,000
Director - Cloud Platform Engineering & AI Enablement
Director - Cloud Platform Engineering & AI Enablement

Anblicks • Dallas (TX)

On-site
USD 180,000 - 280,000
Engineering Manager, Data / AI
Engineering Manager, Data / AI

brightside • United States

Hybrid
USD 120,000 - 150,000