LLM Safety & Alignment Engineer

Hyphen Connect Limited

San Francisco (CA)

On-site

USD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Hyphen Connect Limited is seeking an AI Safety Specialist in San Francisco, California, to enhance the security and robustness of language models. This role involves conducting adversarial testing, implementing protective measures, and ensuring AI behavior aligns with ethical principles.

The ideal candidate will have a background in cybersecurity or adversarial ML and experience with automated red-teaming frameworks. Strong analytical skills are essential for identifying edge cases and supporting safe deployment of AI systems.

Qualifications

  • Knowledge in prompt engineering or adversarial ML is essential.
  • Experience in automated red-teaming frameworks is preferable.
  • Ability to identify edge cases through strong analytical skills.

Responsibilities

  • Conduct adversarial testing on LLMs and multimodal agents.
  • Implement guardrails and real-time filtering for autonomous tool use.
  • Develop constitutional AI principles and assist with RLHF alignment pipelines.

Skills

Background in cybersecurity
Experience with jailbreak taxonomies
Strong analytical mindset

Job description

Hyphen Connect Limited is seeking an AI Safety Specialist in San Francisco, California, to enhance the security and robustness of language models. This role involves conducting adversarial testing, implementing protective measures, and ensuring AI behavior aligns with ethical principles.

The ideal candidate will have a background in cybersecurity or adversarial ML and experience with automated red-teaming frameworks. Strong analytical skills are essential for identifying edge cases and supporting safe deployment of AI systems.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLM Safety & Alignment Engineer
LLM Safety & Alignment Engineer

Hyphen Connect Limited • Oregon (WI)

Hybrid
USD 90,000 - 115,000
AI Safety Specialist (AI Engineering)
AI Safety Specialist (AI Engineering)

Hyphen Connect Limited • Oregon (WI)

Hybrid
USD 90,000 - 115,000
AI Safety Specialist (AI Engineering)
AI Safety Specialist (AI Engineering)

Hyphen Connect Limited • San Francisco (CA)

On-site
USD 100,000 - 130,000
Senior ML Engineer: AI Safety & Alignment (RLHF)
Senior ML Engineer: AI Safety & Alignment (RLHF)

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 400,000
Top-tier salary and equity grants
Comprehensive medical, dental, and eye
AI Alignment Research Engineer — Evaluation & Safety
AI Alignment Research Engineer — Evaluation & Safety

W3 Sourcing • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
AI Red Teamer: LLM Generalist for Safety Testing
AI Red Teamer: LLM Generalist for Safety Testing

Handshake • Seattle (WA)

On-site
USD 90,000 - 130,000
AI Alignment Research Scientist — Safety & Reasoning
AI Alignment Research Scientist — Safety & Reasoning

Safetytalent • San Francisco (CA)

On-site
USD 100,000 - 150,000
Agentic AI Engineer — LLMs & Safe Automation
Agentic AI Engineer — LLMs & Safe Automation

TRM Labs • San Francisco (CA)

Hybrid
USD 215,000 - 230,000
Competitive salary
Opportunity to work on meaningful projects
Collaborative work environment
AI/LLM Safety Engineer — Guardrails & Red Teaming
AI/LLM Safety Engineer — Guardrails & Red Teaming

Propio • United States

Remote
USD 90,000 - 130,000
Staff AI Safety Engineer - Red Team & Guardrails
Staff AI Safety Engineer - Red Team & Guardrails

B Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive medical, dental, vision, life, and disability insurance
Fully paid parental leave
+2