Senior AI Performance Engineer - Edge Inference

Arm Limited

San Jose (CA)

On-site

USD 263,000 - 355,000

Full time

10 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Arm Limited in San Jose seeks an engineer to optimize AI workloads on Arm technology, delivering best-in-class inference performance for production AI models. You will work closely with customers across the Bay Area, translating complex challenges into actionable insights and driving kernel-level improvements.

You will develop optimized kernel implementations, collaborate with multiple engineering teams, and help shape Arm's tooling roadmap with real-world use cases.

Qualifications

  • Experience optimizing DNNs in kernel-level environments.
  • Strong programming skills in Python and C++.
  • Experience with modern AI frameworks and execution models.

Responsibilities

  • Develop highly optimized solutions for AI workloads from kernel to system level.
  • Create production-quality reference implementations, documentation, and performance-focused content.
  • Act as a technical bridge between customers and internal teams to resolve performance issues.
  • Influence Arm's IP and software roadmap from real-world customer insights.

Skills

Kernel-level programming
Python
C++
DNN optimization
Profiling tools
Communication

Tools

Triton
CUDA

Job description

Arm Limited in San Jose seeks an engineer to optimize AI workloads on Arm technology, delivering best-in-class inference performance for production AI models. You will work closely with customers across the Bay Area, translating complex challenges into actionable insights and driving kernel-level improvements.

You will develop optimized kernel implementations, collaborate with multiple engineering teams, and help shape Arm's tooling roadmap with real-world use cases.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Performance Engineer — Edge Inference Expert
Senior AI Performance Engineer — Edge Inference Expert

Arm • San Jose (CA)

Hybrid
USD 263,000 - 355,000
Senior AI Inference Runtime Engineer - Distributed
Senior AI Inference Runtime Engineer - Distributed

Arm Limited • Seattle (WA)

Hybrid
USD 209,000 - 283,000
AI Inference Engineer - Kernel & Edge Optimization
AI Inference Engineer - Kernel & Edge Optimization

Jobgether SRL • United States

Remote
USD 82,000 - 177,000
Remote-first team
Exposure to cutting-edge AI research
Collaborative engineering environment
Staff AI Inference Runtime Engineer
Staff AI Inference Runtime Engineer

Arm • Seattle (WA)

Hybrid
USD 209,000 - 283,000
Hybrid working
Recruitment accommodations
Senior AI Inference Runtime Architect
Senior AI Inference Runtime Architect

Arm • Seattle (WA)

Hybrid
USD 263,000 - 355,000
Hybrid working
Accommodations during recruitment
Principal AI Inference Runtime Architect
Principal AI Inference Runtime Architect

Arm Limited • Seattle (WA)

Hybrid
USD 263,000 - 355,000
Hybrid working
Accommodations during recruitment
Senior Cloud AI Inference Engineer
Senior Cloud AI Inference Engineer

Arm Limited • Seattle (WA)

Hybrid
USD 209,000 - 283,000
Relocation package
Visa sponsorship
Senior AI Kernel Engineer - Edge ML Optimization
Senior AI Kernel Engineer - Edge ML Optimization

Quadric Inc. • Burlingame (CA)

Hybrid
USD 110,000 - 270,000
Competitive salary
Equity
Health, dental, and vision
+7
AI Inference Engineer: Kernel & Edge Optimization
AI Inference Engineer: Kernel & Edge Optimization

Lever, Inc. • Town of Italy (NY)

On-site
EUR 90,000 - 130,000
Remote-first environment
International team
Exposure to cutting-edge AI research
+1
Edge AI Kernel Engineer - High-Performance Inference
Edge AI Kernel Engineer - High-Performance Inference

QUADRIC PTY LTD • Burlingame (CA)

On-site
USD 170,000 - 230,000
Medical, dental, and vision insurance
Company-paid life insurance
Voluntary life insurance
+11