Get more replies from employers
Send a job-specific resume in minutes.
EnCharge AI, based in the United States, is on the lookout for an LLM Inference Deployment Engineer. This role involves optimizing and deploying large language models for high‑performance inference using advanced AI accelerators. The ideal candidate will work on LLM integration and performance tuning while developing efficient inference pipelines.
We seek someone with a Bachelor’s or Master’s in Computer Science or Electrical Engineering, with exceptional skills in model optimization and deployment frameworks like Docker and Kubernetes. Join us in pushing the boundaries of AI technology.
EnCharge AI, based in the United States, is on the lookout for an LLM Inference Deployment Engineer. This role involves optimizing and deploying large language models for high‑performance inference using advanced AI accelerators. The ideal candidate will work on LLM integration and performance tuning while developing efficient inference pipelines.
We seek someone with a Bachelor’s or Master’s in Computer Science or Electrical Engineering, with exceptional skills in model optimization and deployment frameworks like Docker and Kubernetes. Join us in pushing the boundaries of AI technology.