Staff AI Research Infra Engineer - GPU Scale & Equity

Databricks Inc.

New York (NY)

On-site

USD 199,000 - 270,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Databricks Inc. seeks a Staff Software Engineer - AI Research Infrastructure in New York, NY. You'll develop and manage the infrastructure for AI research, including designing systems for large-scale experiments and model training. With a BS/MS or PhD in Computer Science and 5+ years of software engineering experience, you should excel in building distributed systems and have experience with systems programming languages. A commitment to operational excellence and effective communication with cross-functional teams is essential. The role offers a competitive salary range of $199,000 – $270,000 USD, depending on experience and skills.

Qualifications

  • 5+ years of software engineering experience, including substantial time working on large-scale distributed systems.
  • Deep experience with building and operating distributed systems, data pipelines, or large-scale backend services.
  • Proficient in one or more programming languages (e.g., C++, Rust, Go, Java, Scala).

Responsibilities

  • Develop and run the research stack that powers Databricks AI Research.
  • Design and implement infrastructure for large-scale experiments and model training.
  • Build tooling that improves research developer productivity.

Skills

Large-scale distributed systems
Software engineering
Systems programming languages
Cluster schedulers
Modern ML training and inference workflows

Education

BS/MS or PhD in Computer Science or related field

Tools

Kubernetes
Slurm
Ray

Job description

Databricks Inc. seeks a Staff Software Engineer - AI Research Infrastructure in New York, NY. You'll develop and manage the infrastructure for AI research, including designing systems for large-scale experiments and model training. With a BS/MS or PhD in Computer Science and 5+ years of software engineering experience, you should excel in building distributed systems and have experience with systems programming languages. A commitment to operational excellence and effective communication with cross-functional teams is essential. The role offers a competitive salary range of $199,000 – $270,000 USD, depending on experience and skills.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Research Infra Engineer - Scalable Pipelines
Senior AI Research Infra Engineer - Scalable Pipelines

Menlo Ventures • San Francisco (CA)

On-site
USD 190,000 - 270,000
Comprehensive benefits
Equity options
Annual performance bonus
Senior AI Research Infra Engineer - Scalable Pipelines
Senior AI Research Infra Engineer - Scalable Pipelines

Menlo Ventures • San Francisco (CA)

On-site
USD 190,000 - 270,000
Comprehensive benefits
Equity options
Annual performance bonus
Staff Research Software Engineer - AI Training Infra
Staff Research Software Engineer - AI Training Infra

Reflection • New York (NY)

On-site
USD 120,000 - 160,000
Top-tier compensation
Health & wellness benefits
Paid parental leave
+2
AI Infrastructure Engineer: Scale GPU ML Pipelines
AI Infrastructure Engineer: Scale GPU ML Pipelines

Akkodis • Town of Florida (NY)

On-site
USD 125,000 - 130,000
Staff Software Engineer, AI Runtime - Scalable GPU Training
Staff Software Engineer, AI Runtime - Scalable GPU Training

Databricks Inc. • Mountain View (CA)

On-site
USD 190,000 - 265,000
Staff Software Engineer, AI Runtime - Scalable GPU Training
Staff Software Engineer, AI Runtime - Scalable GPU Training

Databricks Inc. • Mountain View (CA)

On-site
USD 190,000 - 265,000
Staff Software Engineer - AI Infrastructure & Data Analytics
Staff Software Engineer - AI Infrastructure & Data Analytics

Confido • New York (NY)

On-site
USD 280,000 - 330,000
Staff AI Runtime Architect - Scale GPU Training
Staff AI Runtime Architect - Scale GPU Training

Databricks • Mountain View (CA)

On-site
USD 190,000 - 265,000
Annual performance bonus
Equity options
Comprehensive benefits package
Senior AI Runtime Engineer: Scale GPU Training
Senior AI Runtime Engineer: Scale GPU Training

Cacheflow • San Francisco (CA)

On-site
USD 160,000 - 225,000
Equity
Annual performance bonus
Comprehensive benefits package
Staff AI Systems Engineer - Pre-Training Infra
Staff AI Systems Engineer - Pre-Training Infra

Reflection • New York (NY)

On-site
USD 130,000 - 180,000