Pragmatike is seeking an AI Infrastructure Engineer for a fully remote position. The ideal candidate will have over 4 years of experience in ML Ops or related infrastructure roles, with expertise in production-grade model serving and GPU-based workloads. Responsibilities include building scalable ML inference platforms and optimizing system performance. Join a fast-scaling cloud infrastructure startup focused on next-generation AI-native cloud services. Candidates must be comfortable working in a remote-first environment and have a strong technical background.
Qualifications
4+ years of experience in ML Ops, Platform Engineering, SRE, or similar infrastructure roles focused on ML systems.
Hands-on experience with model serving frameworks such as vLLM, TGI, Triton, or equivalent.
Strong background in container orchestration and operating GPU-based workloads in production.
Responsibilities
Build and operate production-grade model serving infrastructure using frameworks such as vLLM, TGI, Triton, or equivalent.
Design and implement robust deployment pipelines with blue/green and canary rollout strategies for ML models.
Develop and maintain auto-scaling systems, multi-model serving architectures, and intelligent request routing layers.
Job description
Pragmatike is seeking an AI Infrastructure Engineer for a fully remote position. The ideal candidate will have over 4 years of experience in ML Ops or related infrastructure roles, with expertise in production-grade model serving and GPU-based workloads. Responsibilities include building scalable ML inference platforms and optimizing system performance. Join a fast-scaling cloud infrastructure startup focused on next-generation AI-native cloud services. Candidates must be comfortable working in a remote-first environment and have a strong technical background.