Senior Manager, AI Deployment

General Motors

Sunnyvale, Northern (CA, KY)

Hybrid

USD 296,000 - 454,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

General Motors is seeking a Senior Manager, AI Deployment to lead strategy and execution of model performance and on-vehicle inference for autonomous driving. You will guide optimization across model architecture, operators, kernels, and hardware-aware execution, building dashboards and collaborating with cross-functional teams.

You will manage a team of engineering leaders, define KPIs for latency, throughput, and memory efficiency, and balance near-term needs with longer-term investments in

Qualifications

  • Bachelor’s degree in Computer Science, Electrical or Computer Engineering, Robotics, Machine Learning, or a related field; advanced degree preferred.
  • 10+ years of experience in machine learning systems, model optimization, inference, GPU systems, robotics, autonomous driving, or a related field.
  • 5+ years of people-leadership experience, including experience leading managers or senior technical leaders.
  • Experience shipping production machine-learning inference systems on GPU, accelerator, robotics, automotive, or other edge hardware.
  • Strong understanding of the factors that determine model performance: architecture, tensor shapes, operators, kernels, memory movement, scheduling, runtime execution, and hardware utilization.
  • Hands-on experience with several of the following: PyTorch, CUDA, C++, Python, TensorRT, GPU profiling, benchmarking, performance analysis, or inference runtimes.
  • Experience with quantization, pruning, distillation, architecture optimization, kernel optimization, or memory optimization.
  • Experience building benchmark automation, performance regression detection, telemetry, dashboards, or profiling workflows.
  • Strong systems thinking, communication, decision-making, and cross-functional leadership skills.

Responsibilities

  • Own the strategy, roadmap, and operating plan for AI model performance and inference quality.
  • Establish performance budgets for latency, throughput, memory, GPU utilization, power, and numerical parity.
  • Lead investigations into performance bottlenecks across model architecture, operators, kernels, memory movement, scheduling, runtime behavior, and hardware utilization.
  • Establish repeatable benchmarking and profiling practices across simulation, hardware-in-the-loop, bench, and vehicle environments.
  • Guide optimization through model architecture changes, operator and kernel improvements, memory optimization, scheduling, and hardware-aware execution.
  • Build performance dashboards, regression detection, benchmark automation, and root-cause diagnostics.
  • Partner with Embodied AI, model development, GPU kernel, runtime, system performance, vehicle integration, simulation, and safety teams.
  • Influence model design by translating profiling results into clear recommendations for model architects and researchers.
  • Represent AI Deployment in architecture reviews, program planning, and senior leadership discussions.
  • Build and lead an inclusive, high-performing organization through hiring, coaching, feedback, and manager development.
  • Establish clear ownership, priorities, staffing plans, and operating rhythms across performance workstreams.
  • Define and manage KPIs for inference latency, latency variability, throughput, memory efficiency, GPU utilization, parity, and regression rate.
  • Balance near-term production needs with longer-term investments in profiling, optimization automation, reduced precision, and performance infrastructure.
  • Resolve cross-functional issues and align stakeholders when performance, quality, or implementation trade-offs are contested.
  • Develop technical leaders and succession plans in GPU performance, model optimization, inference systems, and numerical analysis.

Skills

PyTorch
CUDA
C++
Python
TensorRT
GPU profiling
Benchmarking
Performance analysis
Inference runtimes

Education

Bachelor's degree in CS/EE/CE/Robotics
Advanced degree preferred

Tools

PyTorch
CUDA
C++
Python
TensorRT

Job description

## Senior Manager, AI DeploymentApply: Remote: Remote - United States: Sunnyvale, California, United States of America: Full time: Posted Today: JR-202619803**Job Description****About the Organization**General Motors is developing the software and artificial intelligence capabilities for the next generation of autonomous driving. Within AI Foundations, AI Acceleration makes machine learning models faster, more efficient, and more reliable on production vehicle hardware.The AI Deployment team owns model inference performance across simulation, hardware-in-the-loop, bench, and vehicle environments. The team focuses on latency, memory, GPU utilization, numerical parity, profiling, benchmarking, reduced precision, and production readiness.**About the Role**We are looking for a Senior Manager, AI Deployment to lead the strategy and execution of model performance and on-vehicle inference for autonomous driving. You will lead engineering managers and senior technical leaders working across model optimization, GPU systems, inference runtimes, and vehicle integration. You will set performance goals, guide optimization of complex autonomy models, and establish disciplined methods to measure latency, diagnose regressions, and validate improvements. Success requires strong technical judgment, people leadership, and the ability to make clear trade-offs among latency, memory, throughput, accuracy, power, and numerical parity.**What You’ll Do*** Own the strategy, roadmap, and operating plan for AI model performance and inference quality.* Establish performance budgets for latency, throughput, memory, GPU utilization, power, and numerical parity.* Lead investigations into performance bottlenecks across model architecture, operators, kernels, memory movement, scheduling, runtime behavior, and hardware utilization.* Establish repeatable benchmarking and profiling practices across simulation, hardware-in-the-loop, bench, and vehicle environments.* Guide optimization through model architecture changes, operator and kernel improvements, memory optimization, scheduling, and hardware-aware execution.* Build performance dashboards, regression detection, benchmark automation, and root-cause diagnostics.* Partner with Embodied AI, model development, GPU kernel, runtime, system performance, vehicle integration, simulation, and safety teams.* Influence model design by translating profiling results into clear recommendations for model architects and researchers.* Represent AI Deployment in architecture reviews, program planning, and senior leadership discussions.**Leadership Responsibilities*** Build and lead an inclusive, high-performing organization through hiring, coaching, feedback, and manager development.* Establish clear ownership, priorities, staffing plans, and operating rhythms across performance workstreams.* Define and manage KPIs for inference latency, latency variability, throughput, memory efficiency, GPU utilization, parity, and regression rate.* Balance near-term production needs with longer-term investments in profiling, optimization automation, reduced precision, and performance infrastructure.* Resolve cross-functional issues and align stakeholders when performance, quality, or implementation trade-offs are contested.* Develop technical leaders and succession plans in GPU performance, model optimization, inference systems, and numerical analysis.**Your Skills & Abilities (Required Qualifications)*** Bachelor’s degree in Computer Science, Electrical or Computer Engineering, Robotics, Machine Learning, or a related field; advanced degree preferred, or equivalent experience.* 10+ years of experience in machine learning systems, model optimization, inference, GPU systems, robotics, autonomous driving, or a related field.* 5+ years of people-leadership experience, including experience leading managers or senior technical leaders.* Experience shipping production machine-learning inference systems on GPU, accelerator, robotics, automotive, or other edge hardware.* Strong understanding of the factors that determine model performance: architecture, tensor shapes, operators, kernels, memory movement, scheduling, runtime execution, and hardware utilization.* Hands-on experience with several of the following: PyTorch, CUDA, C++, Python, TensorRT, GPU profiling, benchmarking, performance analysis, or inference runtimes.* Experience with quantization, pruning, distillation, architecture optimization, kernel optimization, or memory optimization.* Experience building benchmark automation, performance regression detection, telemetry, dashboards, or profiling workflows.* Strong systems thinking, communication, decision-making, and cross-functional leadership skills.**What Will Give You a Competitive Edge*** Experience optimizing real-time machine-learning systems for autonomous driving, robotics, embedded systems, or computer vision.* Deep experience with GPU performance, memory bandwidth, occupancy, synchronization, stream scheduling, or device-to-device data movement.* Experience with NVIDIA Nsight Systems, NVIDIA Nsight Compute, PyTorch Profiler, TensorRT profiling tools, or equivalent tools.* Experience deploying reduced-precision models and managing calibration, sensitivity, parity, and model-quality risks.* Experience optimizing transformer, vision, lidar, or multimodal workloads.* Experience measuring performance across simulation, hardware-in-the-loop, bench, and vehicle environments.* Experience with safety-critical or highly reliable systems.**Compensation:** The compensation information is a good faith estimate only. It is based on what a successful applicant might be paid in accordance with applicable state laws. The compensation may not be representative for positions located outside of New York, Colorado, California, or Washington* **Compensation:** The expected base compensation for this role is**: $296,300 - $453,900** Actual base compensation within the identified range will vary based on factors relevant to the position.* **Bonus Potential:** An incentive pay program offers payouts based on company performance, job level, and individual performance.* **Benefits:** GM offers a variety of health and wellbeing benefit programs. Benefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays**#GM-AV-1**This role is categorized as remote. This means the selected candidate may be based anywhere in the country of work and is not expected to report to a GM worksite unless directed by their manager.The selected candidate will be required to travel <25% for this role.This job may be eligible for relocation benefits.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Manager, AI Deployment
Senior Manager, AI Deployment

Relha LLC • Northern (KY)

Hybrid
USD 296,000 - 454,000
Staff AI/ML Engineer - Autonomy CI Platform
Staff AI/ML Engineer - Autonomy CI Platform

General Motors • Mountain View (CA), Northern (KY)

Hybrid
USD 219,000 - 335,000
Company vehicle program
GM vehicle discounts
Relocation benefits
Staff ML Engineer - Embodied AI Scaling Foundations
Staff ML Engineer - Embodied AI Scaling Foundations

General Motors • Sunnyvale (CA)

Hybrid
USD 189,000 - 300,000
Health benefits
Company vehicle program
Relocation assistance
Senior ML Accelerator Engineer - GPU
Senior ML Accelerator Engineer - GPU

General Motors • Sunnyvale (CA), Northern (KY)

Hybrid
USD 170,000 - 258,000
Hybrid work model
Relocation benefits
Comprehensive GM benefits package
Senior ML/AI Engineer - Agentic Developer Experience
Senior ML/AI Engineer - Agentic Developer Experience

General Motors • Mountain View (CA)

Hybrid
USD 178,000 - 231,000
Medical, dental, and vision benefits
Health Savings and Flexible Spending Accounts
401(k) retirement savings plan
+2
Senior State Estimation Engineer
Senior State Estimation Engineer

General Motors • Sunnyvale (CA), Northern (KY)

Hybrid
USD 171,000 - 261,000
Health and wellbeing benefits
GM vehicle discounts
Relocation benefits
Sr Manager, AV Behavior Safety Engineering (GPSSC)
Sr Manager, AV Behavior Safety Engineering (GPSSC)

General Motors • Northern (KY)

Hybrid
USD 251,000 - 385,000
Company vehicle program
Bonus program
Health & wellbeing benefits
Staff Software Engineer, AI for Developer Productivity
Staff Software Engineer, AI for Developer Productivity

General Motors • Warren (MI)

Hybrid
USD 160,000 - 247,000
Medical, dental, and vision insurance
Retirement savings plan
Paid vacation and holidays
Staff AI Engineer – Analytics & Domain Intelligence
Staff AI Engineer – Analytics & Domain Intelligence

General Motors • Michigan

Hybrid
USD 172,000 - 221,000
GM vehicle discounts
Relocation benefits
GM vehicle program eligibility
Staff AI/ML Engineer - Future Sensing, Embodied AI
Staff AI/ML Engineer - Future Sensing, Embodied AI

General Motors • New York (NY)

Hybrid
USD 189,000 - 321,000
Health Savings Account
Flexible Spending Accounts
Tuition assistance
+1