Generative AI Engineer: Build Scalable LLMs & RAG Systems
Photon
New York (NY)
On-site
USD 120,000 - 150,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Job summary
A leading technology firm is seeking a Generative AI Engineer to build and optimize production-ready AI applications. The role involves designing complex multi-agent systems and managing LLM deployments. Candidates should have strong Python skills and experience with GenAI applications, vector search, and monitoring in a production setting. This is an opportunity to work with cutting-edge technology and contribute to advanced AI workflows in a collaborative environment.
Qualifications
Hands-on experience building and deploying GenAI applications in a production setting.
Strong proficiency in Python and the modern AI library ecosystem.
Familiarity with production-grade monitoring, API security, and CI/CD for ML.
Responsibilities
Develop and orchestrate sophisticated AI workflows using LangGraph and multi-agent architectures.
Integrate and swap diverse LLMs based on performance and cost requirements.
Optimize GenAI workflows for latency, cost, and reliability.
Skills
GenAI application deployment
Python proficiency
Experience with vector search
Model fine-tuning techniques
Knowledge of CI/CD for ML
Tools
LangChain
LlamaIndex
PostgreSQL
Docker
FastAPI
Job description
A leading technology firm is seeking a Generative AI Engineer to build and optimize production-ready AI applications. The role involves designing complex multi-agent systems and managing LLM deployments. Candidates should have strong Python skills and experience with GenAI applications, vector search, and monitoring in a production setting. This is an opportunity to work with cutting-edge technology and contribute to advanced AI workflows in a collaborative environment.