AI Security Research Engineer: Evaluation & Defense

General Analysis

San Francisco, Northern (CA, KY)

Hybrid

USD 150,000 - 230,000

Full time

42 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

General Analysis is seeking a Research Engineer to build the systems that define how frontier AI models are evaluated and secured. You will iterate on agent frameworks and build the scaffolding that lets models perform at their best, incorporating state-of-the-art tools for task delegation and complex multi-step tasks.

You will design and build evaluation environments, including challenges that measure a model's ability to evade detection.

Qualifications

  • Strong production programming skills and ability to move quickly from prototype to reliable infrastructure.
  • Experience building agentic systems, evaluation harnesses, or environment frameworks or eagerness to learn.
  • Security experience or interest in building security into AI systems.

Responsibilities

  • Build the systems that define how frontier AI models are evaluated and secured.
  • Iterate on agent frameworks and scaffolding for multi-step tasks.
  • Design and build evaluation environments including challenges to test model evasion capabilities.
  • Collaborate with red teamers, pentesters, and contracted security firms to ground experiments.
  • Own the pipeline end-to-end from API design to analysis and visualization of thousands of agent trajectories.
  • Publish results in blog posts and academic venues and deliver findings to customers.

Skills

Production coding
Agentic systems
Evaluation harnesses
Security experience
Analytical skills

Job description

General Analysis is seeking a Research Engineer to build the systems that define how frontier AI models are evaluated and secured. You will iterate on agent frameworks and build the scaffolding that lets models perform at their best, incorporating state-of-the-art tools for task delegation and complex multi-step tasks.

You will design and build evaluation environments, including challenges that measure a model's ability to evade detection.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer, Evaluations
Research Engineer, Evaluations

General Analysis • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 230,000
AI Safety & Evaluation Researcher
AI Safety & Evaluation Researcher

CodeGeniusRecruit • United States

On-site
USD 90,000 - 150,000
Agent Security Engineer for AI Systems
Agent Security Engineer for AI Systems

Coinscapture • Northern (KY)

Hybrid
USD 180,000 - 240,000
AI Threat Researcher & Safety Architect
AI Threat Researcher & Safety Architect

Neura Market • San Francisco (CA)

On-site
USD 293,000 - 405,000
Research Engineer: Adversarial RL for AI Security
Research Engineer: Adversarial RL for AI Security

General Analysis • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
AI-Powered Security Engineer — Frontier AI Defense
AI-Powered Security Engineer — Frontier AI Defense

Meta • Carson City (NV)

On-site
USD 154,000 - 217,000
Research Engineer, Post-Training
Research Engineer, Post-Training

General Analysis • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
AI Security Researcher — Build Safer AI Systems
AI Security Researcher — Build Safer AI Systems

0Labs • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Competitive compensation
Flexible work setup
Remote or hybrid work
+6
Applied AI Security Engineer: AI-Driven Defenses
Applied AI Security Engineer: AI-Driven Defenses

Meta • Salem (OR)

On-site
USD 154,000 - 217,000
Security Engineer, AI-Driven Defense Architect
Security Engineer, AI-Driven Defense Architect

Meta • Juneau (AK)

On-site
USD 154,000 - 217,000