Senior Systems ML Engineer — High-Performance AI Infra

Meta

Des Moines (IA)

On-site

USD 154,000 - 217,000

Full time

5 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Health insurance

Job summary

Meta is seeking a Software Engineer to join our Systems ML Engineering team in Des Moines, IA. You will design and build high-performance ML infrastructure spanning training and inference pipelines, with a focus on scalability and hardware-aware optimizations.

You will work across the full stack with researchers, platform engineers, and product teams to accelerate workloads and improve the efficiency of AI systems serving billions of users.

Qualifications

  • Design, build, and optimize large-scale ML training and inference systems, including distributed computing frameworks and hardware-accelerated pipelines
  • Develop and maintain high-performance ML infrastructure components in C++ and Python, ensuring reliability, scalability, and low-latency execution
  • Identify and resolve performance bottlenecks across the ML stack using profiling, instrumentation, and benchmarking tools
  • Architect and evaluate trade-offs in ML system design, including memory bandwidth, compute utilization, and I/O throughput
  • Partner with research and product teams to translate ML model requirements into efficient infrastructure solutions
  • Define and track system-level metrics and service level objectives to maintain production reliability of ML serving systems
  • Lead technical design reviews and contribute to engineering standards for ML systems across the organization
  • Mentor other engineers on ML infrastructure best practices, debugging methodologies, and performance optimization techniques
  • Drive adoption of AI-augmented development workflows to expand engineering productivity and broaden the scope of deliverables
  • Contribute to staged rollout strategies using feature flagging and experimentation frameworks to safely deploy ML system changes

Responsibilities

  • Design, build, and optimize large-scale ML training and inference systems.
  • Develop and maintain high-performance ML infrastructure components in C++ and Python.
  • Identify and resolve performance bottlenecks across the ML stack using profiling tools.
  • Architect and evaluate trade-offs in ML system design, including memory bandwidth and I/O throughput.
  • Collaborate with research and product teams to translate ML model requirements into infrastructure solutions.
  • Lead technical design reviews and contribute to engineering standards for ML systems.

Skills

Distributed systems
Profiling and optimization
System design
Cross-team collaboration

Education

Bachelor's degree in Computer Science or related field

Tools

C++
Python
CUDA/ROCm
PyTorch/TensorFlow

Job description

Meta is seeking a Software Engineer to join our Systems ML Engineering team in Des Moines, IA. You will design and build high-performance ML infrastructure spanning training and inference pipelines, with a focus on scalability and hardware-aware optimizations.

You will work across the full stack with researchers, platform engineers, and product teams to accelerate workloads and improve the efficiency of AI systems serving billions of users.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Systems ML Engineer — High-Performance AI Infra
Senior Systems ML Engineer — High-Performance AI Infra

Meta • Annapolis (MD)

On-site
USD 154,000 - 217,000
Equity
Benefits
Systems ML Engineer: High-Performance AI Infra
Systems ML Engineer: High-Performance AI Infra

Meta • Saint Paul (MN)

On-site
USD 154,000 - 217,000
Lead Systems ML Engineer – High-Performance AI Infra
Lead Systems ML Engineer – High-Performance AI Infra

Meta • Sacramento (CA)

On-site
USD 154,000 - 217,000
Bonus
Equity
Benefits
Senior Systems ML Engineer — Scalable AI Infra
Senior Systems ML Engineer — Scalable AI Infra

Meta • Nashville (TN)

On-site
USD 154,000 - 217,000
Senior Systems ML Engineer: Scalable AI Infrastructure
Senior Systems ML Engineer: Scalable AI Infrastructure

Meta • Providence (RI)

On-site
USD 154,000 - 217,000
Senior Systems ML Engineer — Scalable AI Infrastructure
Senior Systems ML Engineer — Scalable AI Infrastructure

Meta • Bismarck (ND)

On-site
USD 154,000 - 217,000
Senior Systems ML Engineer - Scalable AI Infra
Senior Systems ML Engineer - Scalable AI Infra

Meta • Raleigh (NC)

On-site
USD 154,000 - 217,000
Systems ML Engineer - Scalable AI Infrastructure
Systems ML Engineer - Scalable AI Infrastructure

Meta • Pierre (SD)

On-site
USD 154,000 - 217,000
Lead Systems ML Engineer — Scalable AI Infrastructure
Lead Systems ML Engineer — Scalable AI Infrastructure

Meta • Frankfort (KY)

On-site
USD 154,000 - 217,000
Senior Systems ML Engineer - Scalable AI Infra (Equity)
Senior Systems ML Engineer - Scalable AI Infra (Equity)

Meta • Charleston (WV)

On-site
USD 154,000 - 217,000