GenAI Safety Analyst

Alice (Formerly ActiveFence)

United States

On-site

USD 80,000 - 87,000

Full time

6 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Alice is seeking a driven, detail-focused professional to join as a GenAI Safety Analyst. You will analyze content infringements and collaborate with teams across Hate Speech, Misinformation, IP, and Copyright areas to safeguard Generative AI tools.

Responsibilities include writing adversarial prompts, managing data quality, and working with cross-functional teams to craft proactive safety strategies. Strong English and meticulous attention to detail are required.

Qualifications

  • Background in AI Safety, Responsible AI, or Trust & Safety is required.
  • Familiarity with Generative AI models and agents is essential, though direct technical experience is not mandatory.
  • Excellent written and spoken English.
  • Strong organizational skills and ability to juggle multiple tasks.

Responsibilities

  • Develop adversarial and risky prompt strategies to expose potential vulnerabilities in AI models.
  • Manage projects end-to-end from planning to delivery.
  • Handle large datasets across languages and abuse areas with precision.
  • Investigate new tactics to circumnavigate safety measures of foundational models.
  • Collaborate with engineering, product, policy teams to address challenges.
  • Promote knowledge sharing and continual learning within the team.

Skills

Adversarial thinking
Attention to detail
Project management

Education

Background in AI Safety/Responsible AI/Trust and Safety

Tools

Open Source Intelligence (OSINT)

Job description

Alice is seeking a driven, detail-focused professional to become a vital part of our team as a GenAI Safety Analyst. In this role, you'll dive into the cutting-edge of technology, meticulously analyzing various content infringements to secure the new wave of Generative AI tools. Your duties will include collaborating with experts in diverse fields such as Hate Speech, Misinformation, Intellectual Property and Copyright, among others.



Your tasks will involve writing adversarial prompts to identify weaknesses in various AI models, including Large Language Models (LLMs), Text-to-Image, Text-to-Video, AI Agents and beyond. You'll also oversee data management to guarantee the highest quality of outputs.



Responsibilities



  • Developing adversarial and risky prompt strategies across several areas of abuse to expose potential vulnerabilities in models.

  • Managing projects end-to-end, from initial planning and oversight through quality assurance to final delivery.

  • Handling extensive datasets across multiple languages and areas of abuse, ensuring precision and meticulous attention to detail.

  • Ongoing investigation into new tactics for circumventing foundational models' safety measures.

  • Working alongside diverse teams, engineering, product, policy, to tackle new challenges and craft forward-thinking strategies and resolutions.

  • Promoting a culture of knowledge exchange and continual learning within the team.



Requirements


Must have


  • Background in AI Safety and/or Responsible AI and/or Trust and Safety

  • Familiarity with recent Generative AI models and agents is essential, though direct technical experience is not a prerequisite.

  • Command of English at a near-native level.

  • Attention to detail, organizational capabilities, and the capacity to juggle numerous tasks concurrently.



Additional Wants


  • Experience with various model types (Text-to-Text, Text-to-Image) is desirable.

  • Prior experience with OSINT (Open Source Intelligence) will be considered an asset.

  • A self-starter attitude, with the energy to excel in a fast-moving and variable environment.



The salary range for this role in the US is $80K - $87K - Range may vary based on experience. Salary at the time of offer will be commensurate with experience.



About Alice


Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interact- whether with each other or with machines.



In a world where AI has fundamentally changed the nature of risk, Alice provides end-to-end coverage across the entire AI lifecycle. We support frontier model labs, enterprises, and UGC platforms with a comprehensive suite of solutions: from model hardening evaluations and pre-deployment red-teaming to runtime guardrails and ongoing drift detection.



Alice is widely considered a global leader in online safety and AI security. We have some of the most forward-thinking and passionate minds in the world working to safeguard over 3 billion users across the largest AI and tech platforms.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GenAI Analyst
GenAI Analyst

Alice (Formerly ActiveFence) • United States

On-site
USD 63,000 - 90,000
GenAI Safety & Risk Analyst
GenAI Safety & Risk Analyst

Alice (Formerly ActiveFence) • United States

On-site
USD 80,000 - 87,000
GenAI Biosecurity Expert
GenAI Biosecurity Expert

Alice (Formerly ActiveFence) • United States

On-site
USD 150,000 - 178,000
Generative AI Safety Analyst
Generative AI Safety Analyst

Alice (Formerly ActiveFence) • United States

On-site
USD 63,000 - 90,000
GenAI CBRNE Cyber Expert
GenAI CBRNE Cyber Expert

Alice (Formerly ActiveFence) • United States

On-site
USD 150,000 - 178,000
GenAI Account Executive
GenAI Account Executive

Alice • San Francisco (CA)

On-site
USD 250,000 - 300,000
GenAI Biosecurity Architect | AI Safety & Risk Vetting
GenAI Biosecurity Architect | AI Safety & Risk Vetting

Alice • United States

On-site
USD 150,000 - 178,000
Research Lead, Evaluations and Benchmarks
Research Lead, Evaluations and Benchmarks

Alice (Formerly ActiveFence) • New York (NY)

On-site
USD 190,000 - 240,000
Research Lead, Evaluations and Benchmarks
Research Lead, Evaluations and Benchmarks

Alice • New York (NY)

On-site
USD 180,000 - 250,000
Research Lead, Evaluations and Benchmarks
Research Lead, Evaluations and Benchmarks

Alice (Formerly ActiveFence) • United States

On-site
USD 180,000 - 280,000