AI Safety Research & Testing Professional

Data Security Council of India

Delhi

On-site

INR 1,800,000 - 2,800,000

Full time

12 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Data Security Council of India invites an AI Safety Research & Testing Professional to advance AI safety research and evaluation capabilities at the Noida NCR campus. You will design structured, code-based evaluations of public AI models and build Python-based evaluation pipelines.

The role involves hands-on testing, adversarial evaluation, failure analysis, and benchmarking, translating findings into practical guidance for industry members and partners.

Qualifications

  • 2-3 years in AI/ML engineering, applied ML research or related technical field.
  • Solid Python proficiency including REST APIs and data tooling.
  • Hands-on experience with an LLM evaluation/red-teaming framework or custom tooling.
  • Understanding of modern LLM development and deployment (pretraining, fine-tuning, RAG).
  • Familiar with AI safety failure categories and attack surfaces.
  • Ability to read and synthesize technical research papers.
  • Familiarity with Git and reproducible development practices.
  • Bachelor's degree in CS/AI/DS/Cybersecurity; Master's a plus.

Responsibilities

  • Design and implement evaluation pipelines for public models using open-source tools.
  • Develop Python tooling, test scripts, and reporting pipelines.
  • Perform structured adversarial testing and red-team activities.
  • Benchmark models and products with quantitative and qualitative analysis.
  • Translate results into practical AI risk guidance for stakeholders.
  • Contribute to safety research landscape briefs and workshops.

Skills

Python proficiency
LLM evaluation
LLM development knowledge
AI safety knowledge
Literature comprehension
Git / version control
Security/testing tooling
LangChain / AutoGen

Education

Bachelor's in Computer Science/AI/Data Science/Cybersecurity

Tools

Git
Python tooling

Job description

AI Safety Research & Testing Professional

NASSCOM Campus, Sector 126, Noida, NCR

Role Summary

The AI Safety Research & Testing Professional will play a key technical role in advancing AI safety research and evaluation capabilities. The role involves designing and conducting structured, code-based evaluations of publicly accessible AI models and AI-enabled products, and developing Python-based evaluation pipelines to support these activities.

The Professional will conduct hands‑on technical testing, adversarial evaluation, failure analysis, and comparative benchmarking, while monitoring emerging AI safety and security research. The role will translate empirical findings, published frameworks, and established standards into practical guidance, readiness assessments, and recommendations for industry members and partners.

This is a hands‑on, technically focused research and testing role covering model APIs, open‑source evaluation frameworks, evaluation datasets, automated testing, RAG, tool use, agents, and other AI‑enabled application components.

Key Responsibilities

Research & Landscape Analysis

  • Conduct structured literature reviews of AI safety and security research, including reports from AI safety institutes, evaluation labs, and research organisations.
  • Track frontier and open-weight model releases and assess their safety and security implications.
  • Monitor emerging research on evaluation, red‑team-changing, robustness, interpretability, alignment, AI security, RAG, and agentic systems.
  • Assess relevant evaluation frameworks, benchmarks, attack methodologies, and testing tools.
  • Contribute to the Global & Regional AI Safety Landscape briefing and internal knowledge base.

Technical Testing & Evaluation

  • Design and implement evaluation pipelines using open‑source frameworks and custom Python tooling to test publicly accessible models and AI‑enabled products.
  • Build and maintain test scripts, model/API integrations, structured test cases, scoring rubrics, automated grading, and reporting pipelines.
  • Test key risk areas including prompt injection, indirect injection, jailbreaks, data leakage/PII exposure, hallucination, adversarial robustness, bias/fairness, over‑and‑under‑refusal, and agent/tool‑use safety.
  • Evaluate application‑level risks across RAG, system prompts, tool/function calling, external APIs, memory/state, and agentic workflows.
  • Conduct structured adversarial testing and red‑team‑ing, including multi‑turn and adaptive attacks where appropriate.
  • Develop reusable attack libraries, evaluation datasets, edge cases, and regression test suites.
  • Run comparative benchmarking across open‑weight and commercial models and products, combining quantitative scorecards with qualitative failure analysis.
  • Design rigorous and reproducible evaluations covering objectives, test‑set construction, sampling, baselines, scoring criteria, assumptions, and limitations.
  • Apply appropriate quantitative and statistical methods, including repeated trials, uncertainty analysis, and error analysis.
  • Assess the reliability and limitations of automated evaluation, including LLM‑as‑judge approaches and human‑evaluation agreement where relevant.
  • Investigate, reproduce, classify, and document failures and identify likely model‑, prompt‑, data‑, retrieval‑, tool‑, or application‑level causes.
  • Assess severity and prioritise findings based on likelihood, exploitability, impact, and potential downstream consequences.
  • Develop automated safety regression testing to detect behavioural changes across model versions, prompts, fine‑tuning, RAG configurations, tools, and application releases.
  • Maintain evaluation pipelines and results repositories for longitudinal tracking and benchmarking.
  • Evaluate proposed safeguards through pre‑ and post‑mitigation testing and verify that fixes do not introduce new failure modes.
  • Apply appropriate version control, logging, configuration management, and reproducibility practices.
  • Support AI safety/security readiness assessments, gap analyses, model risk registers, and incident‑response playbooks.
  • Translate testing results into practical recommendations for AI risk identification, evaluation, monitoring, and management.
  • Contribute technical content to workshops and hands‑on AI safety testing modules.
  • Present methodologies and findings at relevant conferences, working groups, and technical dialogues.
Required Qualifications
  • 2‑3 years of relevant experience in AI/ML engineering, applied ML research, AI safety, AI red‑team‑ing, security research, data science, or a related technical field.
  • Solid Python proficiency, including REST APIs, JSON, pandas/requests, and development of maintainable evaluation scripts and pipelines.
  • Hands‑on experience with an LLM evaluation/red‑team‑ing framework or demonstrable experience building custom evaluation tooling.
  • Working understanding of modern LLM development and deployment, including pretraining, fine‑tuning/RLHF, RAG, system prompting, tool use, and agentic architectures.
  • Working knowledge of major AI safety/security failure categories and ability to reason about attack surfaces and abuse cases.
  • Comfortable reading and synthesising technical research papers and evaluation reports.
  • Familiarity with Git, version control, testing, debugging, and reproducible development practices.
  • Bachelor's degree in Computer Science, AI/ML, Data Science, Cybersecurity, or a related technical field; Master's/research experience is a strong plus.
Preferred / Nice‑to‑Have
  • Experience with AI red‑team‑ing, CTFs, security research challenges, or adversarial evaluation.
  • Experience developing AI safety benchmarks, evaluation datasets, attack libraries, or reusable evaluation harnesses.
  • Experience with automated red‑team‑ing, jailbreak research, prompt‑injection testing, or adaptive attacks.
  • Experience with LLM‑as‑judge evaluation, human evaluation, or evaluation reliability analysis.
  • Experience with agent frameworks or RAG pipelines such as LangChain, AutoGen, or similar.
  • Familiarity with AI/ML security risks such as model extraction, membership inference, memorisation, data poisoning, or RAG/agent security.
  • Experience with Docker, cloud environments, CI/CD, experiment tracking, or evaluation infrastructure.
  • Public/open‑source contributions related to AI safety, ML, security, evaluation tooling, or research.
  • Participation in an AI safety fellowship, programme, research group, or professional community.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Safety Research & Policy Professional
AI Safety Research & Policy Professional

Data Security Council of India • Delhi

On-site
INR 1,200,000 - 1,800,000
AI Safety Researcher
AI Safety Researcher

National E-Governance Services Limited • New Delhi

On-site
INR 1,800,000 - 3,200,000
AI Safety & Evaluation Lead - Noida
AI Safety & Evaluation Lead - Noida

Laksh Human Resource • Dadri

On-site
INR 1,800,000 - 3,000,000
Senior AI Security Engineer
Senior AI Security Engineer

Naukri Assist • Pune District, Bengaluru, Delhi

Hybrid
INR 3,000,000 - 5,000,000
AI Security Lead
AI Security Lead

Infosys • Bengaluru

On-site
INR 1,200,000 - 1,800,000
AI Trust and Governance Architect
AI Trust and Governance Architect

Infosys • Bengaluru

On-site
INR 1,600,000 - 2,400,000
Full Stack AI Assurance Engineer
Full Stack AI Assurance Engineer

LTM • Pune District, Chennai District, Bengaluru

Hybrid
INR 2,500,000 - 3,800,000
QA Engineer
QA Engineer

YO IT Consulting • Mumbai

Hybrid
INR 800,000 - 1,400,000
AI Testing Specialist LLM Evaluation & Qualit
AI Testing Specialist LLM Evaluation & Qualit

Hucon Solutions • Hyderabad

On-site
INR 1,000,000 - 2,000,000
Senior GenAI Engineer (AI Evaluation Engineer)
Senior GenAI Engineer (AI Evaluation Engineer)

FM India • Bengaluru

On-site
INR 1,800,000 - 3,000,000