AI Performance Engineer – HPC, ARM & Distributed Inference

EngineersOfAI

Austin (TX)

On-site

USD 90,000 - 120,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

EngineersOfAI is seeking a candidate to optimize AI workloads across various architectures. The role involves collaboration with hardware and software teams to ensure efficient performance of AI/ML models, as well as benchmarking and troubleshooting at scale.

The ideal applicant has a BS/MS in Computer Science or a related field, with strong programming skills in C++ and Python, and experience with distributed systems. This position offers opportunities to work on cutting-edge technology in Austin, Texas.

Qualifications

  • Strong programming skills in C++ and Python.
  • Experience with distributed systems and communication libraries.
  • Experience profiling and optimizing HPC or AI/ML workloads.

Responsibilities

  • Analyze ML models’ compute and memory requirements.
  • Collaborate across hardware and software teams.
  • Benchmark and troubleshoot system performance.

Skills

C++
Python
Distributed systems
Communication libraries (MPI, NCCL, UCX)
Profiling and optimizing HPC or AI/ML workloads

Education

BS/MS in Computer Science, Electrical Engineering, or related field

Tools

ML frameworks (PyTorch, TensorFlow)
HPC networking technologies (InfiniBand, RoCE)

Job description

EngineersOfAI is seeking a candidate to optimize AI workloads across various architectures. The role involves collaboration with hardware and software teams to ensure efficient performance of AI/ML models, as well as benchmarking and troubleshooting at scale.

The ideal applicant has a BS/MS in Computer Science or a related field, with strong programming skills in C++ and Python, and experience with distributed systems. This position offers opportunities to work on cutting-edge technology in Austin, Texas.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Compute Performance Engineer
AI Compute Performance Engineer

Cisco Systems, Inc. • Milpitas (CA)

On-site
USD 155,000 - 223,000
Medical, dental, vision insurance
401(k) with Cisco matching
Paid parental leave
+2
AI/HPC Infrastructure Architect
AI/HPC Infrastructure Architect

AMD • Austin (TX)

On-site
USD 140,000 - 190,000
AMD benefits at a glance
AI Inference & HPC Engineer: Performance & API
AI Inference & HPC Engineer: Performance & API

Topaz Labs • Dallas (TX)

On-site
USD 110,000 - 150,000
100% covered medical/dental/vision
15 days annual PTO
401k matching
+1
AI Inference & HPC Engineer – Performance & APIs
AI Inference & HPC Engineer – Performance & APIs

Topaz Labs • Emeryville (CA)

On-site
USD 90,000 - 150,000
Full medical/dental/vision coverage
15 days PTO
5 personal days + holidays
+2
AI & HPC GPU Compute Performance Engineer
AI & HPC GPU Compute Performance Engineer

engineeringjobs.net, Inc. • San Jose (CA)

On-site
USD 150,000 - 190,000
Medical Insurance
Dental Insurance
Vision Insurance
+16
Senior AI Field Applications Engineer – GPUs & HPC
Senior AI Field Applications Engineer – GPUs & HPC

AMD • Austin (TX)

On-site
USD 120,000 - 160,000
Remote work options
Travel opportunities
Staff AI Performance Engineer
Staff AI Performance Engineer

EngineersOfAI • Austin (TX)

On-site
USD 90,000 - 120,000
Senior AI Performance Engineer — Edge Inference Expert
Senior AI Performance Engineer — Edge Inference Expert

Arm • San Jose (CA)

Hybrid
USD 263,000 - 355,000
Performance Engineer, Inference Engine - High-Performance AI
Performance Engineer, Inference Engine - High-Performance AI

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
AI Compute Performance Engineer
AI Compute Performance Engineer

020 Cisco Systems, Inc. • Milpitas (CA)

On-site
USD 155,000 - 223,000