Senior Manager, Software Engineering - Agentic IT Operations

NVIDIA Corporation

Santa Clara (CA)

On-site

USD 248,000 - 391,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity

Job summary

NVIDIA is seeking a hands-on Senior Engineering Manager to architect and lead enterprise-scale automation platforms, accelerating AI-driven operations across IT. You will guide a team of IT engineers, set coding and design standards, and contribute to system design for critical components.

The role requires deep production software expertise, proven leadership, and a track record of delivering measurable business outcomes through modernized IT platforms and agentic AI.

Qualifications

  • 10+ years of hands-on software engineering experience
  • 5+ years leading engineering teams
  • Experience building teams in high-growth environments
  • Expertise in production software systems, automation, and data pipelines
  • Experience modernizing IT operations and deploying agentic AI in production
  • Proficiency with IaC, CI/CD, Kubernetes, and cloud platforms (AWS/GCP/Azure)
  • Experience with monitoring/observability tools (Prometheus, Grafana, Datadog, PagerDuty)
  • Fluent in Python, Go, or equivalent languages
  • Executive-level communication to influence technical direction
  • Ability to translate tech capabilities into business value for leadership

Responsibilities

  • Own the technical vision, architecture, and delivery of enterprise-scale automation
  • Build and lead IT engineering teams and contribute to system design
  • Architect and ship AI-enabled autonomous enterprise workflows
  • Define multi-quarter roadmaps tied to measurable business outcomes
  • Recruit, develop, and retain top engineering talent
  • Drive ownership and continuous delivery culture

Skills

Python
Go
Infrastructure
SRE
DevOps

Education

Bachelor's or Master's degree

Tools

Kubernetes
CI/CD
AWS
GCP
Azure

Job description

For over 25 years, NVIDIA has been at the forefront of transforming computer graphics, PC gaming, and accelerated computing, driven by a legacy of continuous innovation and exceptional talent. We are now leveraging the immense potential of AI to usher in the next era of computing, where our GPUs power the "brains" of computers, robots, and autonomous vehicles that can comprehend the world! This pioneering work demands vision, innovation, and the world's best talent. Join our diverse and supportive environment, where NVIDIANs are inspired to excel and make a profound global impact.

We are seeking a hands-on technical leader to build and lead a high-performance engineering organization that architects, delivers, and operates production-grade software systems at global scale.

You will be responsible for transforming enterprise IT operations from manual, reactive workflows into fully automated, AI-driven platforms that scale with NVIDIA’s hyper-growth.

This role demands deep software engineering expertise, systems thinking, and the ability to drive large-scale technical transformation with measurable business outcomes. Exceptional interpersonal, written, and verbal communication skills are vital for success.

What You'll Be Doing:

As a Senior Engineering Manager, you will own the technical vision, architecture, and delivery of enterprise-scale automation platforms that eliminate manual workflows and enable NVIDIA to operate at 10x scale without proportional headcount growth. You will build and lead a team of IT engineers, set the technical bar, and personally contribute to system design and code.

Core responsibilities include:
  • Architect and ship agentic AI systems using LLM-based agents, tool calling, RAG, and orchestration frameworks delivering production-grade AI-assisted operations across enterprise IT domains including employee support, endpoint services, and IT support operations.
  • Design and deploy autonomous AI agents that execute complex, multi-step enterprise workflows end-to-end coordinating approvals, vendor handoffs, cross-system data reconciliation, and exception handling with human-in-the-loop controls delivering measurable improvements in availability, cycle time, cost, and compliance.
  • Engineer robust integration and automation platforms spanning ServiceNow, ERP and procurement systems, endpoint-management platforms, Own the full stack infrastructure, data pipelines, APIs, and user-facing applications.
  • Set the engineering standard through hands-on technical leadership co-authoring production code, conducting rigorous code reviews, and personally driving system design for the most critical components.
  • Recruit, develop, and retain top-tier engineering talent.
  • Build a high-performing team culture grounded in engineering excellence, ownership, and continuous delivery.
  • Define and execute a multi-quarter technical roadmap for automation and agentic operations across enterprise IT, with each initiative tied to quantifiable business outcomes (cost reduction, throughput, SLA improvement, headcount avoidance).
What We Need To See:
  • Bachelor's or Master's degree in a related field, or equivalent experience 10+ overall years of hands-on software engineering experience, with deep expertise in at least one of: Infrastructure, SRE, DevOps, or Production Engineering.
  • 5+ years leading engineering teams, with direct experience hiring, growing, and managing IT engineers.
  • Demonstrated ability to build engineering teams from zero and scale them in a high-growth, high-ambiguity environment.
  • Deep expertise in designing and shipping production software systems—including integrations, automation platforms, and data pipelines—for complex enterprise operations at scale.
  • Track record of modernizing enterprise IT operations platforms (e.g., asset management, endpoint services, IT supply chain, infrastructure operations) and deploying agentic AI into production—including multi-step autonomous execution, human-in-the-loop safeguards, exception handling, and governance frameworks with measurable business outcomes.
  • Production-grade proficiency with infrastructure-as-code, CI/CD, containerization (Kubernetes), and cloud platforms (AWS, GCP, or Azure).
  • Experience with monitoring and observability tools (Prometheus, Grafana, Datadog, PagerDuty, or similar).
  • Fluent in Python, Go, or equivalent languages—able to architect, write, and review production-quality code, not just scripts.
  • Executive-level communication skills with the ability to influence technical direction across engineering, product, and senior leadership.
  • Proven ability to translate complex technical capabilities into quantifiable business value and present to VP/C-level audiences.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 248,000 USD - 391,000 USD. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until September 1, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry. Learn more about NVIDIA.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer – Platform Engineering
Senior Software Engineer – Platform Engineering

NVIDIA • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Equity
Benefits
Healthcare
Senior Staff Forward Deployed Engineer, Enterprise AI and Automation
Senior Staff Forward Deployed Engineer, Enterprise AI and Automation

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits
Principal Software Engineer — Agentic AI Applications and Foundations
Principal Software Engineer — Agentic AI Applications and Foundations

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 272,000 - 432,000
Equity
Benefits
Senior Staff Forward Deployed Engineer, Enterprise AI and Automation
Senior Staff Forward Deployed Engineer, Enterprise AI and Automation

NVIDIA • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Senior Software Engineer - Platform Engineering
Senior Software Engineer - Platform Engineering

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Equity
Benefits
Senior Software Engineer, Agentic AI
Senior Software Engineer, Agentic AI

NVIDIA • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Equity
Benefits
Senior Software Engineer, Agentic AI
Senior Software Engineer, Agentic AI

NVIDIA • Redmond (WA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Director, AI Enablement
Director, AI Enablement

NVIDIA Gruppe • Santa Clara (CA)

Hybrid
USD 292,000 - 443,000
Equity
Benefits
Senior Staff Software Engineer — AI Applications and Platform Foundations
Senior Staff Software Engineer — AI Applications and Platform Foundations

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Inclusive work environment
Benefits package
Solutions Architect, AI Factory Infrastructure DevOps
Solutions Architect, AI Factory Infrastructure DevOps

NVIDIA Corporation • Austin (TX)

On-site
USD 152,000 - 288,000