Member of Technical Staff, AI Engineer San Francisco, CA

Parameter

San Francisco (CA)

Hybrid

USD 180,000 - 240,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Parameter builds AI agents that perform autonomous penetration tests against production applications and cloud environments, identifying IDORs, broken access control, XSS, and misconfigurations that scanners miss. You’ll design the agent systems, reasoning loops, tool interfaces, and memory architecture to enable an LLM to behave like a skilled offensive security researcher rather than a simple tester.

This role is onsite in San Francisco, offering a fast-moving, hands-on environment with

Qualifications

  • 3+ years in software engineering with production LLM experience.
  • Strong TypeScript, async/distributed systems, and long-running workflows.
  • Ability to design context engineering and robust agent tooling.
  • Experience with eval harnesses and understanding offline vs production metrics.

Responsibilities

  • Design and ship agent workflows in LangGraph and Temporal that plan, execute, and verify multi-step exploitation chains.
  • Build the tool layer the agents operate through: HTTP clients, auth handling, browser control, cloud enumeration, payload generation.
  • Own the eval loop. Define what 'good' means for agent runs and optimize metrics.
  • Reduce false positives to improve customer trust and value.
  • Manage inference cost and latency, including prompt caching and model selection across providers.

Skills

TypeScript
LLM systems
Context engineering
Eval harness
Distributed systems
Async workflows

Tools

LangGraph
Temporal
GCP
AWS Bedrock
Anthropic API
Langfuse
Braintrust
GitHub Actions
Linear
Graphite

Job description

Parameter builds AI agents that do offensive security work. Our agents run autonomous penetration tests against production applications and cloud environments, finding IDORs, broken access control, XSS, and infrastructure misconfigurations that scanners miss and that human pentest firms only look for once or twice a year.

We are not a theoretical security company. Our team has responsibly disclosed real, high-severity vulnerabilities to well-known technology companies, and our findings are the front door to most of our customer relationships.

You’ll build the agent systems that do the actual hacking. That means designing the reasoning loops, tool interfaces, and memory architecture that let an LLM behave like a competent offensive security researcher rather than a fuzzer with good manners.

The title points at where you’ll spend most of your time. Expect the rest of the week to go wherever the work is: agent systems one day, the product the next, a customer call or a report that has to go out after that. We are small enough that everyone works across every product, and we hire people who want that.

What you’ll do
  • Design and ship agent workflows in LangGraph and Temporal that plan, execute, and verify multi-step exploitation chains
  • Build the tool layer the agents operate through: HTTP clients, auth handling, browser control, cloud enumeration, payload generation
  • Own the eval loop. We use Braintrust and Langfuse. You’ll define what “good” means for agent runs and make the number go up
  • Reduce false positives, which is the single hardest problem in this space and the thing customers care most about
  • Manage inference cost and latency at scale, including prompt caching, context strategy, and model selection across Claude and Bedrock
What we’re looking for
  • 3+ years of software engineering, with at least one year building production LLM systems (not just prototypes or demos)
  • Strong TypeScript. Comfortable with async, distributed systems, and long-running workflow orchestration
  • Real intuition for context engineering, tool design, and why agents fail in the wild
  • You measure things. You’ve built or maintained an eval harness and know why offline metrics diverge from production behavior
  • Bonus: security background, CTF experience, bug bounty history, or a track record of breaking things you were not supposed to break
Interview process
  • Intro call with a founder (30 minutes)
  • Technical deep dive on your past work (60 minutes)
  • Paid work trial or take-home, scoped to roughly one day
  • Onsite in San Francisco with the team
  • References and offer
Stack

TypeScript, LangGraph, Temporal, GCP, AWS Bedrock, Anthropic API, Langfuse, Braintrust, GitHub Actions, Linear, Graphite

We move quickly. Our target is an offer within two weeks of first contact. Small team, high trust, extremely fast. If you want to see your work in front of customers within days, this is that.

This role is onsite in San Francisco. You must be authorized to work in the US; we are not able to sponsor visas at this time.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Member of Technical Staff, Infrastructure San Francisco, CA
Member of Technical Staff, Infrastructure San Francisco, CA

Parameter • San Francisco (CA)

Hybrid
USD 180,000 - 230,000
Member of Technical Staff, Generalist San Francisco, CA
Member of Technical Staff, Generalist San Francisco, CA

Parameter • San Francisco (CA)

Hybrid
USD 140,000 - 180,000
Offensive Security Engineer San Francisco, CA
Offensive Security Engineer San Francisco, CA

Parameter • San Francisco (CA)

Hybrid
USD 150,000 - 210,000
Growth Engineer Intern San Francisco, CA
Growth Engineer Intern San Francisco, CA

Parameter • San Francisco (CA)

Hybrid
USD 85,000 - 121,000
Furnished housing in San Francisco for
Round-trip travel to SF
Laptop and software credits
+1
Security Engineer - Detection & Response
Security Engineer - Detection & Response

LangChain, Inc. • San Francisco (CA), Northern (KY)

On-site
USD 180,000 - 240,000
Medical, dental, and vision coverage
Flexible vacation
401(k) plan
+1
AI Engineer: Offensive Security Agents
AI Engineer: Offensive Security Agents

Zealot • United States

On-site
USD 100,000 - 140,000
Competitive cash
Meaningful early equity
Solving hard problems
Senior Agentic Security Automation Engineer
Senior Agentic Security Automation Engineer

Hyperproof • United States

On-site
USD 120,000 - 160,000
Health Insurance
401k with company match
Flexible PTO
+2
AI Red Team Engineer
AI Red Team Engineer

White Circle • New York (NY)

On-site
USD 130,000 - 210,000
Equity
Flexible time off
Language lessons
+3
AI Red Team Engineer
AI Red Team Engineer

Visa Hunt • Northern (KY)

Hybrid
USD 140,000 - 210,000
Paid time off
Equity package
All hardware & tools
+2
Founding Engineer - Applied AI
Founding Engineer - Applied AI

Ersilia • San Francisco (CA)

On-site
USD 170,000 - 210,000
Unlimited PTO
Health insurance
Visa sponsorship