StackNexus is seeking a GEN AI ENGINEER to design and optimize AI agents and build RAG pipelines. Candidates should have over 5 years of experience in AI/ML or NLP development, strong Python skills, and familiarity with frameworks like FastAPI and TensorFlow. The role involves API development, collaboration on AI integrations, and a focus on model optimization. This position allows for remote work with a preference for candidates familiar with U.S. working hours.
Qualifications
5+ years of experience in product development and AI/ML or NLP solution development.
Strong coding proficiency in Python with practical experience in relevant libraries.
Experience with LLM frameworks and strong understanding of retrieval systems.
Responsibilities
Design, implement, and optimize AI agents with frameworks.
Build and maintain RAG pipelines with LLMs and vector search.
Collaborate to integrate AI components into business systems.
Skills
Python proficiency
Knowledge of AI/ML frameworks
Experience with RESTful APIs
Familiarity with cloud AI services
Understanding of MLOps
Tools
FastAPI
Docker
Kubernetes
TensorFlow
PyTorch
Job description
Overview
Job Title: GEN AI ENGINEER
Experience: +5 Years
Location: Remote (U.S. Working Hours)
Start: Immediate
Responsibilities
Design, implement, and optimize AI agents leveraging frameworks such as LangChain, LlamaIndex, or Haystack.
Build and maintain RAG (Retrieval-Augmented Generation) pipelines combining LLMs, embeddings, and vector search.
Develop and fine-tune ML and NLP models for internal use cases, work with vector databases, and design efficient retrieval workflows.
Collaborate with other product teams to integrate AI components into the orchestration layer and business systems.
Implement APIs and microservices that expose AI functionalities for broader system use.
Experience with context & prompt engineering, model evaluation, and response optimization.
Evaluate, benchmark, and monitor models using standard metrics and observability tools.
Stay up to date with the latest trends in AI frameworks, LLMs, and orchestration technologies.
Job Requirements
5+ years of experience in product development and AI/ML or NLP solution development.
Strong coding proficiency in Python with practical experience in libraries like FastAPI, Pandas, PyTorch, TensorFlow, or Transformers.
Experience with LLM frameworks and strong understanding of retrieval systems, embeddings, and vector databases.
Hands-on experience with RESTful API development, microservices, and containerized deployments (Docker/Kubernetes).
Familiarity with cloud AI services such as GCP Vertex AI (preferred) or AWS Bedrock, Azure OpenAI, etc.
Exposure to data pipelines, preprocessing, and production-grade model deployment.
Experience building multi-agent systems or orchestrated AI workflows.
Familiarity with MLOps, CI/CD, and model lifecycle management.
Contributions to open-source AI/ML projects or custom toolkits.
Understanding of monitoring, cost optimization, and performance tuning for AI workloads.