Enable job alerts via email!

Research Scientist / Research Engineer, Pre-training

The Rundown AI, Inc.

San Francisco (CA)

On-site

USD 130,000 - 180,000

Full time

12 days ago

Boost your interview chances

Create a job specific, tailored resume for higher success rate.

Job summary

A leading organization in AI research seeks a Research Engineer for its Pretraining team. In this role, you will conduct innovative research to advance large language models, focusing on safe and ethical AI development. Candidates should have an advanced degree and strong software engineering experience, and be eager to collaborate on high-impact projects. Envision a workplace that values diversity and encourages candidates from all backgrounds to apply.

Qualifications

Advanced degree in a relevant field required.
Proven software engineering skills with complex systems.
Familiarity with large-scale machine learning and language modeling.

Responsibilities

Conduct research and implement solutions in model architecture and algorithms.
Design and analyze experiments to improve language model understanding.
Optimize and scale training infrastructure for efficiency.

Skills

Software Engineering

Problem Solving

Communication

Collaboration

Education

Advanced degree (MS or PhD) in Computer Science, Machine Learning, or related field

Tools

Python

PyTorch

Kubernetes

GPUs

Anthropic is at the forefront of AI research, dedicated to developing safe, ethical, and powerful artificial intelligence. Our mission is to ensure that transformative AI systems are aligned with human interests. We are seeking a Research Engineer to join our Pretraining team, responsible for developing the next generation of large language models. In this role, you will work at the intersection of cutting-edge research and practical engineering, contributing to the development of safe, steerable, and trustworthy AI systems.

Key Responsibilities:

Conduct research and implement solutions in areas such as model architecture, algorithms, data processing, and optimizer development
Independently lead small research projects while collaborating with team members on larger initiatives
Design, run, and analyze scientific experiments to advance our understanding of large language models
Optimize and scale our training infrastructure to improve efficiency and reliability
Develop and improve dev tooling to enhance team productivity
Contribute to the entire stack, from low-level optimizations to high-level model design

Qualifications:

Advanced degree (MS or PhD) in Computer Science, Machine Learning, or a related field
Strong software engineering skills with a proven track record of building complex systems
Expertise in Python and experience with deep learning frameworks (PyTorch preferred)
Familiarity with large-scale machine learning, particularly in the context of language models
Ability to balance research goals with practical engineering constraints
Strong problem-solving skills and a results-oriented mindset
Excellent communication skills and ability to work in a collaborative environment
Care about the societal impacts of your work

Preferred Experience:

Work on high-performance, large-scale ML systems
Familiarity with GPUs, Kubernetes, and OS internals
Experience with language modeling using transformer architectures
Knowledge of reinforcement learning techniques
Background in large-scale ETL processes

You'll thrive in this role if you:

Have significant software engineering experience
Are results-oriented with a bias towards flexibility and impact
Willingly take on tasks outside your job description to support the team
Enjoy pair programming and collaborative work
Are eager to learn more about machine learning research
Are enthusiastic to work at an organization that functions as a single, cohesive team pursuing large-scale AI research projects
Are working to align state of the art models with human values and preferences, understand and interpret deep neural networks, or develop new models to support these areas of research
View research and engineering as two sides of the same coin, and seek to understand all aspects of our research program as well as possible, to maximize the impact of your insights
Have ambitious goals for AI safety and general progress in the next few years, and you’re working to create the best outcomes over the long-term.

Sample Projects:

Optimizing the throughput of novel attention mechanisms
Comparing compute efficiency of different Transformer variants
Preparing large-scale datasets for efficient model consumption
Scaling distributed training jobs to thousands of GPUs
Designing fault tolerance strategies for our training infrastructure
Creating interactive visualizations of model internals, such as attention patterns

At Anthropic, we are committed to fostering a diverse and inclusive workplace. We strongly encourage applications from candidates of all backgrounds, including those from underrepresented groups in tech.

If you're excited about pushing the boundaries of AI while prioritizing safety and ethics, we want to hear from you!

Get your free, confidential resume review.

or drag and drop a PDF, DOC, DOCX, ODT, or PAGES file up to 5MB.

Research Scientist, Ads QUEST

Google

Mountain View

On-site

USD 141,000 - 202,000

10 days ago

Research Scientist / Research Engineer, Pre-training

The Rundown AI, Inc.

San Francisco (CA)

On-site

USD 130,000 - 180,000

Full time

Job summary

Qualifications

Responsibilities

Skills

Education

Tools

Job description

Similar jobs

Principal Applied Research Engineer / Scientist

Cupertino

On-site

USD 120,000 - 160,000

Applied Research Engineer/Scientist

Cupertino

On-site

USD 100,000 - 150,000

Principal Applied Research Engineer / Scientist

Cupertino

On-site

USD 130,000 - 180,000

Machine Learning Researcher, Multimodal Foundation Models

Sunnyvale

On-site

USD 143,000 - 265,000

Member of Technical Staff - Applied Scientist, AGI Autonomy

San Francisco

On-site

USD 120,000 - 350,000

Research Engineer - Palo Alto

Palo Alto

On-site

USD 120,000 - 160,000

Research Engineer, World Models

Palo Alto

On-site

USD 130,000 - 250,000

Hewlett Packard Labs - Research Scientist - Generative AI - Early Career

Milpitas

On-site

USD 101,000 - 235,000

Research Scientist, Ads QUEST

Mountain View

On-site

USD 141,000 - 202,000