Agentic AI Engineer (Xora Portfolio Company)

Xora Innovation

Singapore

Hybrid

SGD 120,000 - 170,000

Full time

4 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

ELEMYNT, built by Xora Innovation, seeks a senior software engineer to design a foundational layer for LLM-powered features. You will create provider abstractions, plan and orchestrate agents, ground prompts, and ensure auditable, safe execution in diverse customer environments.

You will implement tool calls, memory management, and a robust evaluation framework to gate changes and maintain reliability in production.

Qualifications

  • Bachelor's or Master's degree in Computer Science or related engineering field and 5+ years building production software.
  • Strong Python and solid engineering practice: async code, typing, testing, modular design, and code review.
  • Hands-on experience building agentic or LLM systems in production with orchestration loops and tool-calling.
  • Experience with multiple model providers behind a single abstraction, including cost and latency trade-offs.
  • Experience building retrieval systems end to end: embeddings, chunking, hybrid search, reranking, and vector databases.
  • Experience with LLM evaluation, guardrails, and observability: tracing, prompts versioning, and evaluation tooling.

Responsibilities

  • Build the provider abstraction that lets any workflow call, swap, or add a model provider by configuration.
  • Develop the agent orchestration where planning agents dispatch sub-agents with durable checkpoints.
  • Create human-in-the-loop checkpoints for high-stakes steps before proceeding.
  • Wrap existing platform capabilities as typed, registered tools for agents to call.
  • Design end-to-end retrieval from ingestion and embeddings through hybrid search and reranking.
  • Build the prompt layer with versioned prompts and captured reasoning for traceability.
  • Expose guarded model access as tools consumed by backend services, frontend, and notebooks.
  • Instrument model calls and tool invocations with tracing and lineage for debuggability.

Skills

Python
Async programming
Typing
Testing
Modular design
Code review
LLM systems
Memory management

Education

Bachelor's or Master's in Computer Science or related field

Tools

LangGraph
AutoGen
Temporal
MLflow
Langfuse

Job description

About Elemynt

ELEMYNT is an early-stage startup built by Xora Innovation. We develop applied intelligence that brings AI into the real world. Our platform combines advanced machine learning, high-performance simulation, and modern software engineering to accelerate the design, validation, and deployment of new materials. Our work sits at the intersection of AI, physics, and large-scale computation. The problems are hard, the stakes are high, and the impact is tangible.

About Elemynt

ELEMYNT is an early-stage startup built by Xora Innovation. We develop applied intelligence that brings AI into the real world. Our platform combines advanced machine learning, high-performance simulation, and modern software engineering to accelerate the design, validation, and deployment of new materials. Our work sits at the intersection of AI, physics, and large-scale computation. The problems are hard, the stakes are high, and the impact is tangible.

About The Role

This role builds the layer every LLM-powered feature on our platform is built on: one interface over many model providers, the agent orchestration that runs multi-step workflows, the retrieval and prompt systems that ground them, and the tracing and evaluation that keep them honest.
It is foundational and deeply hands‑on.
In this role, you'll design agents that plan, call tools, and reason over the results, then make them dependable in production: traceable, evaluated, and safe to run with a person in the loop where it counts.
Because the platform runs inside customers' own secure environments, on their compute clusters, in their cloud, or across a hybrid of the two, the reasoning it produces has to stay auditable and its guardrails have to travel with the software.
Every LLM feature the team ships stands on this layer: what it can do sets what they can build, and how faithfully it reasons sets how far they can trust it.

What You Will Do
  • Build the provider abstraction that lets any workflow call, swap, or add a model provider by configuration, across commercial APIs and self-hosted endpoints, with structured-output validation, retries, and cost tracking.
  • Build the agent orchestration where a planning agent dispatches specialized sub-agents in parallel on a stateful framework, with durable checkpoints, conditional branching, and the context and memory management that keeps multi‑step workflows coherent across long task horizons.
  • Build human‑in‑the‑loop checkpoints so low‑confidence or high‑stakes steps route to a person before an agent proceeds.
  • Wrap existing platform capabilities as typed, registered tools the agents call, with a clean boundary between the agent layer and the systems it builds on.
  • Design retrieval end to end, from ingestion, embeddings, and chunking through hybrid search and reranking, and assemble the context that grounds each model call.
  • Build the prompt layer: versioned prompts, few‑shot sets, and captured reasoning, so every change is tracked and every call is inspectable.
  • Expose agents and guardrailed model access as tools behind one integration point that backend services, the frontend, and notebooks all consume.
  • Instrument every model call, tool invocation, and agent run as traced spans with prompt, model, and tool lineage, so behavior and cost stay debuggable.
  • Build the evaluation framework, deterministic trace metrics alongside LLM‑as‑judge scoring for faithfulness, that gates changes and catches regressions before they ship.
What We Are Looking For
  • Bachelor's or Master's degree in Computer Science or a related engineering field, and 5+ years building and shipping production software, with real depth building LLM or agent systems in production.
  • Strong Python and solid engineering practice: async code, typing, testing, modular design, and code review, plus a track record of shipping systems others depend on.
  • Hands‑on experience building agentic or LLM systems in production: orchestration loops, tool‑calling, structured outputs, and context and memory management for reliable long‑running workflows.
  • Experience working across multiple model providers behind a single abstraction, with routing, fallback, and a feel for the cost and latency trade‑offs.
  • Experience building retrieval systems end to end: embeddings, chunking, hybrid search, reranking, and vector databases.
  • Experience with LLM evaluation and guardrails: building eval sets and harnesses, LLM‑as‑judge scoring, regression gating, and output‑quality and safety checks.
  • Experience instrumenting LLM systems for observability: tracing model and tool calls, versioning prompts, and using traces to debug and improve real behavior.
  • Comfort owning ambiguous systems end to end in a fast‑moving early‑stage environment.
NICE TO HAVE
  • Stateful agent‑orchestration frameworks such as LangGraph or AutoGen, and durable‑execution engines such as Temporal for long‑running workflows.
  • Experience building MCP tools or servers, or similar tool‑calling integration layers.
  • LLMOps and evaluation tooling such as MLflow or Langfuse for tracing, prompt versioning, and evaluation.
  • Human‑in‑the‑loop and interrupt‑driven agent patterns for review and control.
  • Applying LLMs to scientific or technical workflows, grounding reasoning in tool outputs and structured data.
  • Fluency with modern AI coding assistants, or open‑source contributions to AI or agent tooling.
LOCATION

Singapore or United States. We're hiring in both to reach the right person. Work model is on‑site or hybrid, set per location.

CLOSING NOTE

If you don't tick every box but this is clearly your kind of work, get in touch.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Engineer (Xora Portfolio Company)
Senior AI Engineer (Xora Portfolio Company)

Xora Innovation • Singapore

On-site
SGD 120,000 - 180,000
AI Engineer - Agentic & GenAI Systems
AI Engineer - Agentic & GenAI Systems

JOY CONSULTING PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
#EG AI Engineer
#EG AI Engineer

NCS Group • Singapore

On-site
SGD 80,000 - 120,000
AI Engineer (Agentic AI & LLM)
AI Engineer (Agentic AI & LLM)

We+ Asia • Singapore

On-site
SGD 60,000 - 90,000
Senior Data & ML Infrastructure Engineer (Xora Portfolio Company)
Senior Data & ML Infrastructure Engineer (Xora Portfolio Company)

Xora Innovation • Singapore

On-site
SGD 120,000 - 160,000
LLM Application Engineer
LLM Application Engineer

ActAI • Singapore

On-site
SGD 120,000 - 180,000
Full Stack Software Engineer
Full Stack Software Engineer

Knoveleng • Singapore

On-site
SGD 54,000 - 80,000
Senior AI Engineer (Principal-Level Scope)
Senior AI Engineer (Principal-Level Scope)

CHEMT BIOTECHNOLOGY PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Senior AI Engineer Engineering and Technology Singapore Experienced (Individual Contributor) Se[...]
Senior AI Engineer Engineering and Technology Singapore Experienced (Individual Contributor) Se[...]

SEA Singapore • Singapore

On-site
SGD 80,000 - 120,000
Senior AI Engineer - LLM Agents
Senior AI Engineer - LLM Agents

Patsnap • Singapore

On-site
SGD 150,000 - 230,000