Get more replies from employers
Send a job-specific resume in minutes.
Vast.ai is seeking a hands-on engineering-relations role in San Francisco or Los Angeles. You will deploy and serve open-source models on rented GPUs, build public examples with Docker templates, and create benchmarks to show developers how to reproduce results.
You will teach through guides, short videos, talks, and live sessions, becoming the credible voice in our community and feeding feedback to engineering.
Vast.ai runs one of the world's largest GPU marketplaces: 20,000+ GPUs, from RTX 4090s to B300s, serving 25,000+ monthly customers who train, fine-tune, and serve AI models. We're profitable, flat, and we ship fast. Applications go straight to the hiring team.
This is an engineering role that happens in public — not a content-calendar job. You'll be Vast's resident power user: most of your week is spent renting GPUs on the platform and making them do impressive things — deploying and serving open-source models with vLLM, SGLang, PyTorch, and ComfyUI, keeping live endpoints running (including token endpoints on markets like OpenRouter), and building the example repos, templates, and benchmarks that show developers exactly how to do the same.
Then you teach it: short videos, technical guides, Discord and Reddit answers, hackathons. You won't be building the product — you'll be using it harder than any customer, in public, and feeding what breaks straight back to engineering.
$160K-$200K base + equity + bonus. Comprehensive health, dental, vision, and life insurance; 401(k) with match; onsite meals; and a travel/conference budget. On-site in San Francisco or Los Angeles.