Senior Engineer AI

Delta Exchange

India

On-site

INR 4,000,000 - 7,000,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Delta Exchange seeks a Senior FullStack Engineer, AI to own and evolve its suite of AI-powered products. You will work across production AI apps including conversational agents, RAG-powered search, and code-generation copilots.

This is a hands-on role focused on architecture, retrieval pipelines, and cost optimization while collaborating with the product team. You will design end-to-end AI applications, implement tool calling and structured outputs, and drive evaluation frameworks to measure

Qualifications

  • 5+ years shipping production software systems.
  • 2+ years building AI/LLM-powered applications end‑to‑end with real users and volume.
  • Strong experience with RAG architectures: vector databases, embedding models, chunking/indexing strategies.
  • Deep understanding of LLM capabilities and limitations: prompt engineering, function/tool calling, structured outputs, context window management.
  • Experience with LLM provider APIs and abstraction layers (OpenAI, Anthropic, LiteLLM, OpenRouter, or similar).
  • Proficiency in Python (Flask/FastAPI) and/or Node.js/TypeScript (Next.js, Vercel AI SDK).
  • Golang experience is a plus.
  • Hands‑on experience building evals, tracking quality metrics, and debugging non‑deterministic outputs in production.
  • Familiarity with cost optimisation: model routing, caching, token usage monitoring, and prompt compression.
  • Solid fundamentals in data structures, algorithms, and system design.
  • Experience with containerised deployments (Docker, Kubernetes) and cloud platforms (AWS/GCP). Practical understanding of k8s concepts and trade‑offs is a must.
  • Nice to Have Experience with agentic frameworks (Vercel AI SDK, LangChain, Mastra, etc.).
  • Background in fine‑tuning or training open‑source models. Large plus if you can demonstrate matching proprietary model quality.
  • Knowledge of cryptocurrency, derivatives trading, or financial systems.
  • Open‑source contributions or personal projects with real traction.

Responsibilities

  • Design, build, and maintain production AI applications end-to-end: backend, frontend, and inference services.
  • Architect RAG systems using vector databases, embedding models, and chunking strategies optimised for accuracy and latency.
  • Build agentic workflows with tool/function calling, multi-step reasoning, and structured output parsing, with accuracy and control as priority.
  • Write and iterate on system prompts, few-shot examples, and prompt chains to maximise output quality.
  • Implement function calling, tool-use patterns, and structured JSON/XML output handling using frontier and lightweight models from providers like Anthropic and OpenAI.
  • Drive cost optimisation: model selection, caching, token budgeting, and request batching at scale.
  • Build and maintain evaluation frameworks to measure accuracy, relevance, hallucination rates, and regression across prompt and model changes.
  • Experience with observability tools (Sentry, Opik, etc.) is a must.
  • Work with message queues (RabbitMQ), caching layers (Redis), and relational databases (PostgreSQL) powering AI service backends.
  • Deploy and manage AI services on Kubernetes with CI/CD pipelines on AWS/GCP.
  • Integrate AI capabilities with third-party platforms (Telegram bots, chat widgets, etc.).
  • Contribute to architectural decisions: model selection, hosting, and build‑vs‑buy trade-offs.

Job description

Role Summary

We are looking for a Senior FullStack Engineer, AI to own and evolve Delta Exchange's suite of AI-powered products. You will work across multiple production AI applications: conversational support agents, RAG-powered search, code-generation copilots, and trading-strategy assistants. This is a hands‑on role where you architect solutions around our existing stack, build retrieval pipelines, improve frontends with product sense, optimise inference costs, and run evaluations, collaborating closely with the product team.

Key Responsibilities
  • Design, build, and maintain production AI applications end-to-end: backend, frontend, and inference services.
  • Architect RAG systems using vector databases, embedding models, and chunking strategies optimised for accuracy and latency.
  • Build agentic workflows with tool/function calling, multi-step reasoning, and structured output parsing, with accuracy and control as priority.
  • Write and iterate on system prompts, few-shot examples, and prompt chains to maximise output quality.
  • Implement function calling, tool-use patterns, and structured JSON/XML output handling using frontier and lightweight models from providers like Anthropic and OpenAI.
  • Drive cost optimisation: model selection, caching, token budgeting, and request batching at scale.
  • Build and maintain evaluation frameworks to measure accuracy, relevance, hallucination rates, and regression across prompt and model changes.
  • Experience with observability tools (Sentry, Opik, etc.) is a must.
  • Work with message queues (RabbitMQ), caching layers (Redis), and relational databases (PostgreSQL) powering AI service backends.
  • Deploy and manage AI services on Kubernetes with CI/CD pipelines on AWS/GCP.
  • Integrate AI capabilities with third-party platforms (Telegram bots, chat widgets, etc.).
  • Contribute to architectural decisions: model selection, hosting (cloud APIs vs. self-hosted), and build‑vs‑buy trade-offs.
Required Skills & Experience
  • 5+ years shipping production software systems.
  • 2 years building AI/LLM-powered applications end‑to‑end with real users and volume. Not prototypes.
  • Strong experience with RAG architectures: vector databases, embedding models, chunking/indexing strategies, and retrieval evaluation.
  • Deep understanding of LLM capabilities and limitations: prompt engineering, function/tool calling, structured outputs, context window management, and multi‑turn conversations.
  • Experience with LLM provider APIs and abstraction layers (OpenAI, Anthropic, LiteLLM, OpenRouter, or similar).
  • Proficiency in Python (Flask/FastAPI) and/or Node.js/TypeScript (Next.js, Vercel AI SDK).
  • Golang experience is a plus.
  • Hands‑on experience building evals, tracking quality metrics, and debugging non‑deterministic outputs in production.
  • Familiarity with cost optimisation: model routing, caching, token usage monitoring, and prompt compression.
  • Solid fundamentals in data structures, algorithms, and system design.
  • Experience with containerised deployments (Docker, Kubernetes) and cloud platforms (AWS/GCP). Practical understanding of k8s concepts and trade‑offs is a must.
  • Nice to Have Experience with agentic frameworks (Vercel AI SDK, LangChain, Mastra, etc.). Even better if you have built Gen AI apps using raw HTTP calls to provider APIs and designed efficient conversation persistence to a database.
  • Background in fine‑tuning or training open-source models. Huge plus if you can demonstrate matching proprietary model quality (Sonnet 4.6, Haiku 4.5, etc.) with fine‑tuned alternatives.
  • Knowledge of cryptocurrency, derivatives trading, or financial systems. Helps understand how AI can improve product UX.
  • Open‑source contributions or personal projects with real traction. Even a tool you built to solve your own problem counts.
  • Your day is spent mostly in AI coding harnesses (Claude Code, Codex, Droid, etc.), MCP servers, custom skills, and similar tooling. We use AI to build AI, and candidates already living this workflow will ramp up fast.
  • Proven ability to debug production AI systems: diagnosing tool call failures, optimising function calling patterns, refining prompts and tool descriptions under real traffic.
  • Extreme ownership and bias towards action. You treat production AI systems as your own, proactively improving quality, latency, and cost without waiting to be asked. You deliver your best work especially when no one is watching.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Artificial Intelligence Engineer
Artificial Intelligence Engineer

Questhiring • Gurugram District

On-site
INR 2,500,000 - 4,500,000
AI Engineer
AI Engineer

Andpayments • India

On-site
INR 1,800,000 - 3,000,000
Senior Full Stack Developer (Kolkata)
Senior Full Stack Developer (Kolkata)

Experiture Global • North 24 Parganas District

On-site
INR 3,000,000 - 4,500,000
Senior AI Engineer (Agentic AI)
Senior AI Engineer (Agentic AI)

Deutsche Telekom Digital Labs • Gurugram District

On-site
INR 1,200,000 - 2,200,000
AI Engineer
AI Engineer

Cosmofeed • Gurugram District

On-site
INR 4,000,000 - 7,000,000
AI Engineer
AI Engineer

Deutsche Telekom Digital Labs • Bengaluru

On-site
INR 1,000,000 - 2,000,000
Full Stack Agentic AI Engineer
Full Stack Agentic AI Engineer

Deutsche Telekom Digital Labs • Gurugram District

On-site
INR 1,500,000 - 2,500,000
Senior AI Engineer
Senior AI Engineer

Spice Money • Dadri

On-site
INR 1,800,000 - 3,200,000
Full Stack AI Engineer
Full Stack AI Engineer

Daten • India

On-site
INR 1,200,000 - 1,800,000
Full Stack Developer (Python & Angular) Assistant Vice President
Full Stack Developer (Python & Angular) Assistant Vice President

Citi • Bengaluru

On-site
INR 1,600,000 - 2,500,000