ML Benchmarking & Infrastructure Engineer

Immunai

United States

Hybrid

USD 88,000 - 122,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Immunai in Ramat Gan, Israel, is seeking an ML Infra Engineer to own benchmarking and evaluation of foundation models and multimodal systems. You will support model development, iteration, and validation, and build out the ML engineering infrastructure for evaluation and deployment.

This role sits at the intersection of software engineering and applied AI, shaping how models are built and compared across the organization, with a hybrid work model.

Qualifications

  • Bachelor's/Master's/PhD in computer science or related field.
  • Strong software engineering skills with modular systems.
  • Hands-on experience with ML models and evaluation pipelines.
  • Proficiency in Python and modern ML ecosystems.
  • Ability to read, modify, and debug deep learning models.
  • Experience with benchmarks, metrics, or evaluation frameworks (preferred).
  • Familiarity with foundation models or multimodal learning (preferred).
  • Comfort navigating complex datasets and doing targeted exploratory analysis.
  • Experience in biomedical or data-intensive domains (a plus).

Responsibilities

  • Own & evolve benchmarking for foundation models and multimodal AI systems.
  • Define core abstractions for datasets, tasks, models, metrics, and evaluation workflows.
  • Develop metrics capturing predictive performance, biological relevance, and multimodal alignment.
  • Support model development by collaborating with AI scientists and data scientists.
  • Bring in new models and baselines to benchmarks and ensure meaningful comparisons.
  • Explore data and results to debug evaluations and understand model behavior.
  • Ensure evaluations are versioned, consistent, and reproducible over time.

Skills

Python
ML evaluation
Software engineering
Data exploration

Education

BSc/MSc/PhD in CS or related

Job description

Immunai in Ramat Gan, Israel, is seeking an ML Infra Engineer to own benchmarking and evaluation of foundation models and multimodal systems. You will support model development, iteration, and validation, and build out the ML engineering infrastructure for evaluation and deployment.

This role sits at the intersection of software engineering and applied AI, shaping how models are built and compared across the organization, with a hybrid work model.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Infra Engineer
ML Infra Engineer

Immunai • United States

Hybrid
USD 88,000 - 122,000
ML Evaluation Engineer: Benchmark & Model Quality
ML Evaluation Engineer: Benchmark & Model Quality

Reducto • San Francisco (CA)

On-site
USD 100,000 - 130,000
Unlimited PTO
Daily free lunch
Reimbursed transportation
+3
Lead ML Infrastructure & Evaluation
Lead ML Infrastructure & Evaluation

Cursor • New York (NY)

On-site
USD 180,000 - 260,000
Remote AI Benchmark Engineer & Researcher
Remote AI Benchmark Engineer & Researcher

Pathway • Palo Alto (CA)

Remote
USD 120,000 - 180,000
AI Benchmarking Engineer — Evaluation & Failure Analysis
AI Benchmarking Engineer — Evaluation & Failure Analysis

Doist • San Francisco (CA)

On-site
USD 150,000 - 210,000
Bi-annual bonus
Equity grant
Relocation bonus
+6
ML Systems Engineer: AI Infra & GPU Acceleration
ML Systems Engineer: AI Infra & GPU Acceleration

Meta • San Francisco (CA)

On-site
USD 180,000 - 240,000
Bonus
Equity
Software Engineer — AI Benchmarking & Systems Infra
Software Engineer — AI Benchmarking & Systems Infra

LatchBio • San Francisco (CA)

On-site
USD 180,000 - 250,000
Unlimited PTO (truly)
Waterfront office in China Basin, San 
Free lunch and dinner
+1
ML Evaluation & Benchmarking Engineer (Equity)
ML Evaluation & Benchmarking Engineer (Equity)

NVIDIA • Town of Florida (NY)

On-site
USD 152,000 - 242,000
Equity
Benefits
Evaluation Platform Engineer: Build Scalable ML Benchmarks
Evaluation Platform Engineer: Build Scalable ML Benchmarks

Thinking Machines Lab • San Francisco (CA)

On-site
USD 300,000 - 475,000
Health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1
Staff Engineer - AI Inference & Benchmarking
Staff Engineer - AI Inference & Benchmarking

Liquid-Ai • Boston (MA)

Hybrid
USD 140,000 - 210,000
Equity
Health insurance
401(k) match
+1