Sr Research Engineer/Scientist

Armis

Ahmedabad District

On-site

INR 3,500,000 - 7,000,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Armis is seeking a senior AI research professional to lead LLM mid-training and post-training work, including data decisions, SFT, and RL. You will prototype novel architectures for planning, memory, and tool use, and design robust evaluation methodologies.

You will drive research from hypothesis to validated prototype with measurable impact, collaborating with engineering to push successful ideas into production and contributing to publications or patents as needed.

Qualifications

  • 5+ years of applied ML/AI research with independent project ownership.
  • Hands-on LLM training and post-training including SFT or RL preferred.
  • Strong Python and PyTorch, with capability to modify pipelines or models.
  • Expertise in agentic AI, planning, memory, tool use, and grounding.
  • Experience designing evaluation methodologies and benchmarks.
  • Ability to drive research from idea to prototype with measurable impact.

Responsibilities

  • Own mid-training and post-training research for LLMs, including data decisions.
  • Prototype architectures spanning planning, reasoning, memory, and tools.
  • Develop evaluation methodologies, benchmarks, and human-in-the-loop approaches.
  • Identify failure patterns and translate insights into improvements.
  • Lead projects from hypothesis to validated prototype and publish or patent results.
  • Collaborate with engineering to productionize successful approaches.

Skills

LLM training
Python
PyTorch
Research leadership
Experimentation
Benchmark design

Tools

Research infrastructure

Job description

What youll do
  1. Own LLM mid-training and post-training research, including continued pretraining, SFT, preference optimization, and RL; make data-mixture and experimental decisions and determine how training changes affect downstream agent behavior.
  2. Research and prototype novel agentic architectures and algorithms across planning, reasoning, memory, skills, tool use, retrieval, and multi-agent collaboration, advancing beyond existing approaches where appropriate.
  3. Design and build research harnesses and experimentation methodologies that enable systematic experimentation, trajectory analysis, reproducibility, and rigorous comparison across models, checkpoints, and agent architectures.
  4. Define evaluation methodologies and develop novel benchmarks for measuring agent reasoning, planning, tool use, reliability, factuality, and safety; establish rigorous approaches for LLM‑as‑a‑Judge, trajectory-based, and human evaluation.
  5. Identify systematic model and agent failure patterns, determine their root causes, and translate those insights into new research directions or improvements in training data, architecture, context, or evaluation.
  6. Independently identify research problems, formulate novel hypotheses, and drive projects from research idea to validated prototype and measurable impact, collaborating with engineering to transition successful approaches into production and communicating results through publications, patents, or open-source work.
What were looking for
  • 5+ years of experience in machine learning, deep learning, AI research, or a related field, with demonstrated applied research experience and a track record of independently driving research projects.
  • Hands‑on experience with LLM training and post‑training, including one or more of continued pretraining, SFT, preference optimization, or RL; experience making training‑data decisions and understanding training dynamics and failure modes at scale.
  • Strong Python and advanced PyTorch expertise, with experience modifying models, training pipelines, or research infrastructure to support novel experimentation.
  • Strong practical depth in agentic AI and context engineering, including planning, reasoning, memory, skills, tool use, retrieval, long‑context processing, and knowledge grounding.
  • Experience designing evaluation methodologies, not just running evaluations, including benchmark design, trajectory‑based evaluation, LLM‑as‑a‑Judge, human evaluation, and the ability to determine which metrics and methodologies are appropriate for a research question.
  • Demonstrated research ownership and impact, with the ability to identify important research questions, develop novel hypotheses, conduct rigorous experiments, and communicate findings and implications effectively to technical researchers, engineers, and executive stakeholders.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Scientist - AI / ML Research
Research Scientist - AI / ML Research

Roadzen, Inc. • Delhi

On-site
INR 1,500,000 - 3,500,000
Senior Research Scientist, Agent Evaluation
Senior Research Scientist, Agent Evaluation

Snow Planet • Hyderabad, Ahmedabad District

On-site
INR 2,400,000 - 4,200,000
Research Engineer — Agent Architectures (Coding & Autonomous Systems)
Research Engineer — Agent Architectures (Coding & Autonomous Systems)

Story Terrace Inc. • Mumbai

Remote
INR 1,800,000 - 3,200,000
Senior LLM Engineer
Senior LLM Engineer

Infocusp Innovations • Pune District

On-site
INR 2,000,000 - 3,400,000
Senior LLM Engineer
Senior LLM Engineer

Zoho • Pune District

Hybrid
INR 1,800,000 - 3,000,000
MTS - Research (India)
MTS - Research (India)

Collinear AI, Inc. • Bengaluru

On-site
INR 4,000,000 - 8,000,000
MTS - Research (India)
MTS - Research (India)

Collinear AI, Inc. • India

On-site
INR 700,000 - 1,100,000
AI Engineer (LLMs, Agentic Systems & Model Training)
AI Engineer (LLMs, Agentic Systems & Model Training)

Kayana | Ordering & Payment Solutions • Mumbai

On-site
INR 1,200,000 - 2,000,000
Competitive salary and benefits
Opportunity to work with cutting-edge AI systems
Collaborative environment
+1
Machine Learning Specialist
Machine Learning Specialist

Recro • Bengaluru

On-site
INR 1,200,000 - 2,000,000
Data Scientist
Data Scientist

GK HR Consulting India Private Limited • Bengaluru

On-site
INR 2,500,000 - 4,000,000