Get more replies from employers
Send a job-specific resume in minutes.
Vast.ai in Los Angeles on-site is seeking an engineering role that happens in public — you’ll be Vast’s resident power user: renting GPUs, deploying and serving open-source models with vLLM, SGLang, PyTorch, and ComfyUI, and keeping live endpoints running and benchmark templates ready. You’ll teach it: create short videos and guides, answer questions on Discord and Reddit, and participate in hackathons.
You won’t build the product yourself, you’ll use it intensely and feed friction and gaps back
Vast.ai runs one of the world's largest GPU marketplaces: 20,000+ GPUs, from RTX 4090s to B300s, serving 25,000+ monthly customers who train, fine-tune, and serve AI models. We're profitable, flat, and we ship fast. Applications go straight to the hiring team.
This is an engineering role that happens in public — not a content-calendar job. You'll be Vast's resident power user: most of your week is spent renting GPUs on the platform and making them do impressive things — deploying and serving open-source models with vLLM, SGLang, PyTorch, and ComfyUI, keeping live endpoints running (including token endpoints on markets like OpenRouter), and building the example repos, templates, and benchmarks that show developers exactly how to do the same.
Then you teach it: short videos, technical guides, Discord and Reddit answers, hackathons. You won't be building the product — you'll be using it harder than any customer, in public, and feeding what breaks straight back to engineering.
$160K-$200K base + equity + bonus. Comprehensive health, dental, vision, and life insurance; 401(k) with match; onsite meals; and a travel/conference budget. On-site in San Francisco or Los Angeles.
Compensation Range: $160K - $200K
"