Get more replies from employers
Send a job-specific resume in minutes.
Jobgether is seeking a Senior Inference Engineer based in the United Arab Emirates to build and own an inference platform from the ground up. You will work with the CTO to turn large language models into production-grade, scalable systems across GPU environments.
You will deploy model-serving infrastructure using vLLM, SGLang, and TensorRT-LLM, focusing on latency, throughput, and cost. This remote-first startup values ownership, speed, and measurable customer impact.
Jobgether is seeking a Senior Inference Engineer based in the United Arab Emirates to build and own an inference platform from the ground up. You will work with the CTO to turn large language models into production-grade, scalable systems across GPU environments.
You will deploy model-serving infrastructure using vLLM, SGLang, and TensorRT-LLM, focusing on latency, throughput, and cost. This remote-first startup values ownership, speed, and measurable customer impact.