Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
eBay’s AI Platform team seeks an LLM Inference Engineer to remove compute bottlenecks for production LLMs, ensuring frontier-model inference is fast, reliable, and observable—from GPUs to APIs that products rely on.
You will tackle latency, throughput, cost, and reliability across HPC, GPU systems, and MLOps, applying optimizations and robust monitoring. This role sits at the intersection of research and production systems.
As an LLM Inference Engineer on our AI Platform team, you’ll remove the compute-scaling bottleneck for production LLMs. Your job is to make frontier-model inference fast, efficient, reliable, and observable—the “last mile” from GPUs to APIs that products depend on. This role sits at the intersection of HPC, GPU systems, and MLOps, and requires strong intuition for how model architecture, runtimes, and hardware interact.