Senior Model Inference Engineer for Production-Scale AI
OpenAI
San Francisco (CA)
On-site
USD 325,000 - 490,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Job summary
A leading AI research company in San Francisco seeks an engineer to optimize their powerful AI models for high-volume production environments. The ideal candidate has over 5 years of software engineering experience, strong familiarity with ML architectures, and experience with distributed systems. This role involves collaboration with researchers and focus on performance optimization. Compensation ranges from $325K to $490K.
Qualifications
At least 5 years of professional software engineering experience.
Self-directed with a humble attitude and eagerness to help colleagues.
Responsibilities
Work alongside machine learning researchers and engineers.
Optimize code and architecture for performance and efficiency.
Skills
Understanding of modern ML architectures
Experience with PyTorch
Problem ownership
Familiarity with NVidia GPUs
Experience with distributed systems
Tools
Azure VMs
NCCL
CUDA
InfiniBand
MPI
Job description
A leading AI research company in San Francisco seeks an engineer to optimize their powerful AI models for high-volume production environments. The ideal candidate has over 5 years of software engineering experience, strong familiarity with ML architectures, and experience with distributed systems. This role involves collaboration with researchers and focus on performance optimization. Compensation ranges from $325K to $490K.