AI Red Team Engineer

Visa Hunt

Northern (KY)

Hybrid

USD 140,000 - 210,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Paid time off
Equity package
All hardware & tools
Subscriptions for AI agents
Off-sites

Job summary

White Circle is seeking an AI Safety-focused Red Team Engineer to break LLM-powered systems responsibly, automate repetitive attacks, and turn findings into clear evidence for customer demos and security reviews.

You will own end-to-end adversarial testing: find the failure, prove it, script it, and write it up, contributing to regression tests and product requirements.

Qualifications

  • Background in QA automation, AppSec, API/security/pen testing, or bug bounty.
  • Strong Python scripting skills.
  • Hands-on with LLMs, prompts, system instructions, RAG, agents, and tool/function calling.
  • Experience testing APIs, web apps, backends, or SaaS products.
  • Understand LLM abuse vectors (prompt injection, data leakage, tool misuse).

Responsibilities

  • Red-team LLM-powered systems: chatbots, copilots, RAG pipelines, AI agents, tool-calling workflows, and API-based AI products.
  • Write lightweight Python to automate attacks, run prompt sets, call model APIs, collect and score responses, and generate repeatable reports.
  • Build and maintain an internal attack library: prompts, scenarios, test cases, regression tests, scoring rubrics, and reusable demo cases.
  • Turn model failures into clear reports: what happened, why it matters, how to reproduce it, how severe it is, and how to fix it.
  • Convert successful attacks into regression tests and product requirements.
  • Track new red-team and safety techniques and fold useful ones into tests.
  • Support GTM by producing strong, credible evidence for customer demos, security reviews, and sales conversations.

Skills

Python scripting
QA automation
AppSec
API security
Bug bounty
Burp Suite
Postman
Playwright
pytest
RAG pipelines
LangChain
LLM red-teaming

Tools

Burp Suite
Postman
Playwright
pytest

Job description

TLDR: We're looking for an AI Red Team Engineer to break LLM-powered systems responsibly, automate the repetitive attacks, and turn their findings into clear evidence that powers customer demos, security reviews, and sales conversations. You'll own hands-on adversarial testing end to end: find the failure, prove it, script it, and write it up.

About us

White Circle is an AI Safety company building the safety, reliability, and optimization layer for AI systems. At the core of our platform are policies – simple natural-language rules that define what an AI model should and shouldn’t do. We automatically test, enforce, and continuously improve these policies at scale.

  • We’ve raised $11M from top funds, founders, and senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, DeepMind, Datadog, Sentry, and others
  • We process over one hundred million API calls every month
  • We fine-tune and train our own LLMs so they run faster and cheaper than any open or proprietary model

We’re a small, highly focused team. If you want to work deeply on hard problems, see your work ship to production quickly, and influence how AI safety is actually built – you’re the one we need.

You will:
  • Red-team LLM-powered systems: chatbots, copilots, RAG pipelines, AI agents, tool-calling workflows, and API-based AI products.
  • Test for jailbreaks, prompt injection, system-prompt and tool leakage, sensitive-data and context leakage, unsafe outputs, policy bypass, tool misuse, excessive agency, resource and token-cost abuse, and business-logic abuse.
  • Write lightweight Python to automate attacks, run prompt sets, call model APIs, collect and score responses, and generate repeatable reports.
  • Build and maintain an internal attack library: prompts, scenarios, test cases, regression tests, scoring rubrics, and reusable demo cases.
  • Turn model failures into clear reports: what happened, why it matters, how to reproduce it, how severe it is, and how to fix it.
  • Convert successful attacks into regression tests and product requirements.
  • Track new red-team and safety techniques and fold the useful ones into our tests.
  • Support GTM by producing strong, credible evidence for customer demos, security reviews, and sales conversations.
You will fit right in if you:
  • Genuinely love breaking things and reasoning adversarially.
  • Have a background in QA automation, AppSec, API/security/pen testing, or bug bounty.
  • Have strong Python scripting skills.
  • Have experience testing APIs, web apps, backends, or SaaS products.
  • Are hands-on with LLMs, prompts, system instructions, RAG, agents, and tool/function calling.
  • Understand LLM-specific abuse vectors (prompt injection, jailbreaks, data leakage, tool misuse, excessive agency, token-cost exhaustion).
  • Can find bypasses, abuse edge cases, chain failures, and reason about real-world impact.
  • Can separate real customer risk from low-impact prompt tricks.
  • Write clear, reproducible bug reports in clear English.
  • Can move fast without perfect requirements.
  • Hold a firm ethical line: you red-team to make systems safer, operate within scope and the law, and don't produce or traffic in genuinely harmful material.
A big plus:
  • Experience with Burp Suite, Postman, Playwright, pytest.
  • Experience with modern LLM red-teaming automated agents and pipelines.
  • Familiarity with LangChain, LangGraph, LlamaIndex, RAG pipelines, AI agents, tool/function calling, and LLM-as-judge evaluation.
  • Familiarity with OWASP LLM Top 10, OWASP Web Top 10, MITRE ATLAS, or other AI security taxonomies.
  • Experience testing RAG systems, AI agents, tool-calling workflows, browser agents, or internal copilots.
  • Experience writing customer-facing security reports.
  • Experience with trust & safety, abuse prevention, fraud, moderation, or platform security.
  • Experience building eval pipelines, regression suites, dashboards, or CI-friendly security tests.
  • A track record in CTFs, red-team competitions, or responsible-disclosure / bounty programs.
Why White Circle
  • Paid time off in line with your local regulations, no matter where you work from
  • Meaningful equity package
  • All the hardware, tools, and services you need
  • Covered subscriptions for AI agents
  • Team off-sites twice a year: we've recently been to the Alps and to Saint-Tropez
How we hire
  1. Intro call with HR (25 min)
  2. Take-home test task
  3. Technical interview (60 min)
  4. Final call with CEO (45 min)

Please submit your application in English

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Recruiter
Recruiter

Visa Hunt • Northern (KY)

Hybrid
USD 70,000 - 100,000
Meaningful equity package
Paid time off in line with local regs
All hardware, tools and services
Red Team Engineer, Safeguards
Red Team Engineer, Safeguards

Doist • San Francisco (CA)

Hybrid
USD 320,000 - 405,000
Equity donation matching
Generous vacation
Parental leave
+2
AI Engineer
AI Engineer

Valsoft Corporation • Northern (KY)

Hybrid
USD 120,000 - 180,000
Head of AI Red Teaming
Head of AI Red Teaming

Trajectory Labs, PBC • Berkeley (CA), Northern (KY)

Hybrid
USD 250,000 - 400,000
Equity
Health coverage
401(k)
+1
AI Red Team Engineer
AI Red Team Engineer

Arcitix Security • United States

On-site
USD 100,000 - 140,000
Strategic Projects Lead, Red Team
Strategic Projects Lead, Red Team

Front Door Defense • New York (NY)

On-site
USD 152,000 - 190,000
Comprehensive health coverage
Retirement benefits
Learning and development stipend
+2
Red Team Engineer, Safeguards
Red Team Engineer, Safeguards

Anthropic • San Francisco (CA)

Hybrid
USD 320,000 - 405,000
Strategic Projects Lead, Red Team
Strategic Projects Lead, Red Team

United States Digital Space LLC • New York (NY), San Francisco (CA)

On-site
USD 151,000 - 189,000
Healthcare
Dental & Vision
Retirement benefits
+3
Founding Events Lead
Founding Events Lead

Slope • New York (NY)

Hybrid
USD 140,000 - 180,000
Performance-based bonus
Meaningful equity package
Team off-sites twice a year
AI Engineer
AI Engineer

Pinpoint Global Communications • United States

On-site
USD 120,000 - 180,000