Machine Learning Engineer, AI Inference Solutions (Early in Career)

General Motors

Sunnyvale (CA)

Hybrid

USD 119,000 - 151,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Health benefits
Tuition assistance
Employee assistance program
GM vehicle discounts

Job summary

General Motors in Sunnyvale, CA, is hiring an early-career Engineer to help build and optimize the ML deployment platform for on-vehicle inference, collaborating with senior engineers on real deployments and production workflows.

You will learn from experienced teams, work across kernels, compiler, and parity groups, and contribute to model optimization and reliability in safety-critical software. Pending start date in 2026.

Qualifications

  • Recently completed or completing a Bachelor’s or Master’s degree by Spring 2026 in Computer Science, ECE, or a related technical field.

Responsibilities

  • Contribute production code across the ML deployment platform, model-optimization workflows, and inference benchmarking/profiling infrastructure.
  • Pair with senior engineers on deployment workflows, performance investigations, model-optimization experiments (quantization, pruning, distillation), and platform tooling.
  • Build, test, and maintain platform tools (e.g., validators, performance probes, parity and sensitivity analyzers, agentic specialists) with technical guidance and code review support.
  • Investigate and help root-cause production deployment or performance issues; learn and apply the diagnostic playbook for compiler, kernel, runtime, and parity bugs.
  • Collaborate with cross-functional teams across the AV organization; including kernels, compiler, reduced-precision, parity, and model-development groups—to plan and execute model deployments to the AV stack, working under the guidance of senior engineers.
  • Participate in code reviews, design discussions, and technical documentation to ensure reliability, correctness, and clear abstractions in a large-scale codebase.
  • Learn and follow secure coding, safety, and compliance practices required for on-vehicle autonomous driving software.

Skills

Python
C++
AI/ML experience
Computer architecture
Operating systems
Distributed systems
Compilers
Software engineering
Code assistants
Teamwork / communication

Education

Bachelor’s or Master’s in CS or ECE

Tools

PyTorch
TensorRT
ONNX
Triton
Airflow / Kubeflow

Job description

Job Description

General Motors is a global leader in advanced driver assistance, with Super Cruise hands-free technology in more than 500,000 equipped vehicles on the road and over 700 million hands-free miles driven—demonstrating that automation can be trusted, intuitive, and helpful while reaching everyday drivers at unprecedented scale. Within GM AV, the Model Deployment & Inference Solutions team deploys machine learning models from training frameworks (e.g., PyTorch) onto autonomous-vehicle hardware; our two-fold mission is to build the ML deployment platform that makes model rollouts fast and predictable, and to optimize models so they meet the real-time latency and memory budgets required to run on-vehicle. Our work sits on the critical path for GM’s publicly committed launch of eyes-off (hands-free, eyes-free) autonomous driving in 2028 on the Cadillac Escalade IQ, and we’re hiring engineers to help deliver the next generation of safe, delightful personal autonomous-vehicle experiences.

About the Role

As an early career Engineer on the Model Deployment & Inference Solutions team, you’ll contribute across both sides of our mission: building the ML deployment platform and optimizing models for on-vehicle inference. You’ll work with and learn from senior engineers on real production deployments, platform features, and model-optimization workflows that ship to GM’s Super Cruise fleet at large scale, with structured mentorship and a clear onboarding plan. You’ll also collaborate closely with our sister teams (kernels, compiler, reduced precision, and parity) on the end-to-end path that takes trained models from research frameworks to ultra-efficient, safety-critical inference on the car. This is an early-career / new graduate role designed for candidates who have recently or will be completing their degree by August 2026.

What You’ll Do (Responsibilities)
  • Contribute production code across the ML deployment platform, model-optimization workflows, and inference benchmarking/profiling infrastructure.
  • Pair with senior engineers on deployment workflows, performance investigations, model-optimization experiments (e.g., quantization, pruning, distillation), and platform tooling.
  • Build, test, and maintain platform tools (e.g., validators, performance probes, parity and sensitivity analyzers, agentic specialists) with technical guidance and code review support.
  • Investigate and help root-cause production deployment or performance issues; learn and apply the diagnostic playbook for compiler, kernel, runtime, and parity bugs.
  • Collaborate with cross-functional teams across the AV organization; including kernels, compiler, reduced-precision, parity, and model-development groups—to plan and execute model deployments to the AV stack, working under the guidance of senior engineers.
  • Participate in code reviews, design discussions, and technical documentation to ensure reliability, correctness, and clear abstractions in a large-scale codebase.
  • Learn and follow secure coding, safety, and compliance practices required for on-vehicle autonomous driving software.
Your Skills & Abilities (Required Qualifications)
  • Recently completed or completing a Bachelor’s or Master’s degree by Spring 2026 in Computer Science, ECE, or a related technical field. (Degree must be completed before your start date.)
  • Strong computer science fundamentals (e.g., data structures, algorithms, operating systems, computer architecture) and solid coding skills in Python and/or C++, demonstrated through coursework, internships, or substantial projects.
  • Hands‑on experience in AI/ML (e.g., machine learning, deep learning, computer vision, NLP, or ML systems) via classes, research, internships, or personal projects.
  • Depth in at least one of: computer architecture, operating systems, distributed systems, or compilers.
  • Demonstrated software‑engineering experience (internships, coursework, open‑source, research code, or competitions) showing good judgment around reliability, correctness, and clean abstractions.
  • Experience with—or strong interest in—using coding assistants/agents (e.g., Cursor, Claude Code, GitHub Copilot) as part of your workflow.
  • Ability to work effectively in collaborative, cross-functional teams and communicate clearly—both in writing and verbally—including explaining technical work partners.
What Will Give You a Competitive Edge (Preferred Qualifications)
  • Internship, research, or advanced coursework in ML systems, ML compilers, GPU programming (CUDA, OpenAI Triton), inference optimization, or distributed training/serving infrastructure.
  • Familiarity with PyTorch and modern ML compiler/runtime stacks (e.g., torch.compile, TensorRT, ONNX, Triton Inference Server, vLLM, or equivalent).
  • Exposure to model optimization (quantization, pruning, distillation) or GPU profiling tools (Nsight Systems, Nsight Compute, PyTorch Profiler).
  • Familiarity with workflow/ML platforms such as Airflow, Temporal, Flyte, Ray, or Kubeflow.
  • Experience building agentic or LLM-powered tools or workflows.
  • Open-source contributions related to PyTorch, TensorRT, vLLM, OpenAI Triton, or similar projects.
  • Coursework, projects, or publications touching ML systems (e.g., MLSys, OSDI, ASPLOS, HPCA, NeurIPS systems track).
  • Familiarity with a systems language (e.g., C++) and development in a Linux environment.
Location

Sunnyvale, CA

Hybrid/Remote Arrangements

This role is categorized as hybrid. This means the selected candidate is expected to report to a specific location at least 3 times a week. This job may be eligible for relocation benefits.

Compensation

The compensation information is a good faith estimate only. It is based on what a successful applicant might be paid in accordance with applicable state laws. The compensation may not be representative for positions located outside of New York, Colorado, California, or Washington. The salary range for this role is $119,250 to $150,850. The actual base salary a successful candidate will be offered within this range will vary based on factors relevant to the position.

Bonus Potential: An incentive pay program offers payouts based on company performance, job level, and individual performance.

Benefits
  • medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays, tuition assistance programs, employee assistance program, GM vehicle discounts and more.
About GM

Our vision is a world with Zero Crashes, Zero Emissions and Zero Congestion and we embrace the responsibility to lead the change that will make our world better, safer and more equitable for all.

Why Join Us

We believe we all must make a choice every day – individually and collectively – to drive meaningful change through our words, our deeds and our culture. Every day, we want every employee to feel they belong to one General Motors team.

Benefits Overview

From day one, we’re looking out for your well-being–at work and at home–so you can focus on realizing your ambitions. Learn how GM supports a rewarding career that rewards you personally by visiting Total Rewards resources.

Non-Discrimination and Equal Employment Opportunities (U.S.)

General Motors is committed to being a workplace that is not only free of unlawful discrimination, but one that genuinely fosters inclusion and belonging. We strongly believe that providing an inclusive workplace creates an environment in which our employees can thrive and develop better products for our customers. All employment decisions are made on a non-discriminatory basis without regard to sex, race, color, national origin, citizenship status, religion, age, disability, pregnancy or maternity status, sexual orientation, gender identity, status as a veteran or protected veteran, or any other similarly protected status in accordance with federal, state and local laws.

Accommodations

General Motors offers opportunities to all job seekers including individuals with disabilities. If you need a reasonable accommodation to assist with your job search or application for employment, email us or call us at 1-800-865-7580. In your email, please include a description of the specific accommodation you are requesting as well as the job title and requisition number of the position for which you are applying.

Join Us

We are leading the change to make our world better, safer and more equitable for all through our actions and how we behave. Learn more about: Our Company Our Culture How we hire Our diverse team of employees bring their collective passion for engineering, technology and design to deliver on our vision of a world with Zero Crashes, Zero Emissions and Zero Congestion. We are looking for adventure‑seekers and imaginative thought leaders to help us transform mobility. Explore our global locations We are determined to lead change for the world through technology, ingenuity and harnessing the creativity of our diverse team. Join us to help lead the change that will make our world better, safer and more equitable for all by becoming a member of GM’s Talent Community (beamery.com).

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff ML Engineer - Embodied AI Onboard Autonomy
Staff ML Engineer - Embodied AI Onboard Autonomy

General Motors • Sunnyvale (CA)

Hybrid
USD 180,000 - 280,000
Hybrid work model
Relocation benefits
GM vehicle program
Senior ML Infrastructure Engineer - Embodied AI Scaling Foundations
Senior ML Infrastructure Engineer - Embodied AI Scaling Foundations

General Motors • Sunnyvale (CA)

Hybrid
USD 129,000 - 261,000
medical
dental
vision
+7
Senior Machine Learning Engineer - ML Training Infrastructure
Senior Machine Learning Engineer - ML Training Infrastructure

General Motors • Sunnyvale (CA), Northern (KY)

On-site
USD 170,000 - 241,000
Medical
Dental
Vision
+1
Senior ML Infrastructure Engineer (Compute)
Senior ML Infrastructure Engineer (Compute)

General Motors • Sunnyvale (CA), Northern (KY)

Hybrid
USD 155,000 - 206,000
Health and wellbeing benefits
GM vehicle discounts
Tuition assistance
+2
Senior AI/ML Capacity Engineer
Senior AI/ML Capacity Engineer

General Motors • United States

Hybrid
USD 145,000 - 261,000
Hybrid work arrangement
Health benefits
Employee wellbeing programs
Principal AI/ML Engineer, AV ML Infra
Principal AI/ML Engineer, AV ML Infra

Relha LLC • Sunnyvale (CA)

Hybrid
USD 276,000 - 341,000
Health and wellbeing benefits
Retirement savings plan
Paid vacation & holidays
+1
Senior Software Systems Engineer - Autonomous Vehicles
Senior Software Systems Engineer - Autonomous Vehicles

General Motors • Sunnyvale (CA)

On-site
USD 153,000 - 234,000
Health benefits
Retirement savings plan
GM vehicle discounts
Senior ML Infrastructure Engineer - Embodied AI
Senior ML Infrastructure Engineer - Embodied AI

General Motors • United States

Hybrid
USD 153,000 - 235,000
Medical, dental, and vision insurance
Retirement savings plan
Tuition assistance programs
+2
Staff ML Infrastructure Engineer - Embodied AI
Staff ML Infrastructure Engineer - Embodied AI

General Motors • Sunnyvale (CA)

Hybrid
USD 189,000 - 291,000
Health benefits
GM vehicle discounts
Tuition assistance
Senior AI/ML Engineer – Future Sensing, Embodied AI
Senior AI/ML Engineer – Future Sensing, Embodied AI

General Motors • Washington

Hybrid
USD 182,000 - 251,000
Bonus potential
Relocation benefits
GM benefits package