Remote Senior ML Systems Engineer – GPU & Kernel Expert
Inferact
United States
Remote
USD 180,000 - 240,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Benefits offered by this job
Competitive salary and equity
Visa sponsorship
Health coverage where applicable
Job summary
A cutting-edge technology company is looking for exceptional generalist engineers who thrive with autonomy. This fully remote role allows you to work on high-impact projects across the vLLM stack, from optimizing CUDA kernels to designing distributed orchestration systems. Ideal candidates will have a Bachelor's degree and a strong track record in systems programming or ML infrastructure. Competitive compensation and benefits are offered.
Qualifications
Ability to work autonomously and drive projects to completion.
Excellent communication skills across time zones.
Deep expertise in systems programming, GPU programming, or distributed systems.
Responsibilities
Work across the entire vLLM stack from GPU kernels to distributed systems.
Optimize CUDA kernels and design orchestration systems.
Build foundational layers for distributed systems at global scale.
Skills
Autonomy
Asynchronous communication
Systems programming
GPU/accelerator programming
Distributed systems
ML infrastructure
Education
Bachelor's degree or equivalent experience
Tools
CUDA
Rust
Go
C++
Python
Kubernetes
Job description
A cutting-edge technology company is looking for exceptional generalist engineers who thrive with autonomy. This fully remote role allows you to work on high-impact projects across the vLLM stack, from optimizing CUDA kernels to designing distributed orchestration systems. Ideal candidates will have a Bachelor's degree and a strong track record in systems programming or ML infrastructure. Competitive compensation and benefits are offered.