AI Alignment Engineer — Research & Experimentation

Zizy Inc.

New York, Northern (NY, KY)

Hybrid

USD 120,000 - 220,000

Full time

19 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Mental health support
Generous time off
Parental leave and family-planning
Annual learning & development stipend

Job summary

Zizy Inc. is seeking research engineers and research scientists to design and implement experiments for alignment research, focusing on scalable training signals for AGI alignment with human intent and safety.

You will design experiments to measure scalable oversight, study generalization, manage large datasets from interpretability experiments, and develop LLM-based alignment methods in a fast-paced, collaborative team environment.

Qualifications

  • Experience designing and running ML experiments.
  • Interest in alignment research and AI safety.
  • Ability to work in fast-paced, collaborative research environments.

Responsibilities

  • Design experiments to measure scalable oversight techniques (AI-assisted feedback and debate).
  • Study generalization to see if models trained on easy problems solve hard ones.
  • Manage large datasets from interpretability experiments and create visualizations.
  • Develop experiments to test chain-of-thought reasoning reflecting model cognition.
  • Investigate how training against a reward signal affects outputs.
  • Explore methods to understand/predict model behaviors, incl. anomalous circuits or catastrophic outputs.
  • Design novel approaches for using LLMs in alignment research.

Skills

ML algorithms
PyTorch

Job description

Zizy Inc. is seeking research engineers and research scientists to design and implement experiments for alignment research, focusing on scalable training signals for AGI alignment with human intent and safety.

You will design experiments to measure scalable oversight, study generalization, manage large datasets from interpretability experiments, and develop LLM-based alignment methods in a fast-paced, collaborative team environment.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Alignment Engineer
AI Alignment Engineer

Zizy Inc. • New York (NY), Northern (KY)

Hybrid
USD 120,000 - 220,000
Mental health support
Generous time off
Parental leave and family-planning
+1
AGI Safety & Alignment Research Engineer
AGI Safety & Alignment Research Engineer

Google • United States

On-site
USD 174,000 - 253,000
AI Alignment Research Engineer — Evaluation & Safety
AI Alignment Research Engineer — Evaluation & Safety

W3 Sourcing • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Scalable ML Training Systems Engineer
Scalable ML Training Systems Engineer

Zizy Inc. • New York (NY), Northern (KY)

Hybrid
USD 110,000 - 180,000
Mental health support
Generous time off
Parental leave
+1
AI Research Scientist: Scalable Systems & Algorithms
AI Research Scientist: Scalable Systems & Algorithms

Zizy Inc. • New York (NY), Northern (KY)

Hybrid
USD 100,000 - 150,000
Mental health support
Generous time off
Paid parental leave
+1
Research Scientist
Research Scientist

Zizy Inc. • New York (NY), Northern (KY)

Hybrid
USD 100,000 - 150,000
Mental health support
Generous time off
Paid parental leave
+1
Distributed Model Training Engineer
Distributed Model Training Engineer

Zizy Inc. • New York (NY), Northern (KY)

Hybrid
USD 110,000 - 180,000
Mental health support
Generous time off
Parental leave
+1
Research Engineer: AI Safety & Alignment
Research Engineer: AI Safety & Alignment

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 500,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
AI Alignment Research Scientist — Safety & Reasoning
AI Alignment Research Scientist — Safety & Reasoning

Safetytalent • San Francisco (CA)

On-site
USD 100,000 - 150,000
AI Safety & Alignment Research Scientist (GenAI)
AI Safety & Alignment Research Scientist (GenAI)

DeepMind Technologies Limited • Mountain View (CA)

Hybrid
USD 207,000 - 300,000