An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Anthropic is seeking researchers and engineers for the Interpretability team to reverse-engineer how trained models work and develop mechanistic understanding. The role focuses on methods to understand LLMs, running robust experiments, and constructing interpretable circuits using features.
Based in San Francisco, exceptional candidates may be considered for remote work. This position offers visa sponsorship and collaboration with Alignment Science and Societal Impacts to enhance model safety.
Anthropic is seeking researchers and engineers for the Interpretability team to reverse-engineer how trained models work and develop mechanistic understanding. The role focuses on methods to understand LLMs, running robust experiments, and constructing interpretable circuits using features.
Based in San Francisco, exceptional candidates may be considered for remote work. This position offers visa sponsorship and collaboration with Alignment Science and Societal Impacts to enhance model safety.