Senior ML Systems Engineer - Distributed AI Inference
Amazon Web Services (AWS)
Seattle (WA)
On-site
USD 120,000 - 160,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Job summary
A leading cloud services provider is seeking a Senior Software Engineer to join their Machine Learning Applications team in Seattle. This role involves developing and optimizing AI models at scale, focusing on distributed inference solutions and performance tuning for next-generation AI accelerators. Candidates should have strong programming skills in Python or C++, along with a solid foundation in machine learning. Competitive compensation is available.
Qualifications
3+ years of experience in object-oriented design, data structures, algorithm design.
Experience with AI acceleration techniques including quantization and batching.
Fundamentals of machine learning and deep learning models.
Responsibilities
Drive the evolution of distributed AI at AWS Neuron.
Develop optimization bridges between ML frameworks and AI hardware.
Engineer performance optimizations for AWS Trainium and Inferentia.
Skills
Python
C++
ML framework internals
Distributed systems
Performance tuning
Education
Bachelor's degree in computer science or equivalent
Job description
A leading cloud services provider is seeking a Senior Software Engineer to join their Machine Learning Applications team in Seattle. This role involves developing and optimizing AI models at scale, focusing on distributed inference solutions and performance tuning for next-generation AI accelerators. Candidates should have strong programming skills in Python or C++, along with a solid foundation in machine learning. Competitive compensation is available.