Machine Learning Infrastructure Engineer

Genesis Molecular AI

City of Utica (NY)

On-site

USD 150,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation with salary +

Job summary

Genesis Molecular AI is seeking an experienced ML infrastructure engineer to lead the engineering efforts on our AI platform, focusing on scalable distributed infrastructure for training, inference, and evaluation across multiple clusters and clouds.

You will optimize GPU performance, accelerate workloads, and collaborate with researchers and scientists to advance our generative and predictive AI models for molecular space.

Qualifications

  • Strong engineer with clean code and deep understanding of the codebases you work in.
  • Deeply experienced with distributed training and inference of large models on GPU clusters and core libraries/frameworks: PyTorch, PyTorch Lightning, PyTorch Geometric, and Ray.
  • Independent thinker with ownership and capability of engineering robust systems from first-principles to state-of-the-art realization.
  • Curious, problem-oriented thinker excited by AI, physics, chemistry, and biology intersections with ML.

Responsibilities

  • Lead engineering efforts focused on continuous improvement of the AI platform and scalable distributed infrastructure for ML training, inference, and evaluation.
  • Support model training and deployment across multiple clusters and clouds, optimizing for throughput and cost.
  • Optimize efficiency of ML models and workloads (latency, throughput, memory) including GPU performance tuning.
  • Contribute to the long-term vision for Genesis’ infra platform.

Skills

Distributed training
GPU performance engineering
Ownership
PyTorch
PyTorch Lightning
PyTorch Geometric
Ray
Kubernetes
Terraform
CUDA
XLA
Triton

Tools

Kubernetes
Terraform
CUDA
XLA
Triton

Job description

About The Team

We’re a tight-knit team of proven drug hunters, deep learning researchers, and software engineers united by a common mission — drive AI innovation in biochemistry, discovering and developing groundbreaking therapies for patients suffering from severe disorders. Genesis AI team is focused on developing foundation models for small molecule drug discovery by conducting fundamental research at the intersection of machine learning, physics, and computational chemistry, as well as engineering robust software systems that enable running large scale simulations and training generative and predictive AI models designed to learn from all kinds of molecular data, leveraging our cluster with 1000s of GPUs and 10,000s of CPUs.

About The Role

We’re seeking experienced ML infrastructure engineers to join the team and lead engineering efforts focused on driving forward our ML research agenda for generative modeling of molecular systems, which is instrumental to our mission. As an engineer at Genesis, you will lead rapid iteration on our AI platform and infrastructure, unlocking the next level of performance, efficiency, and scale that was not previously possible. You will build massively distributed training and inference pipelines, core MLOps tools and frameworks, and optimize GPU operations to speed up ML models. Genesis is a highly-collaborative and cross functional environment, and you will work in close partnership with our exceptional engineers, researchers, and scientists.

You Will
  • Lead engineering efforts focused on continuous improvement of the AI platform, focused on rapid build out and iteration on scalable and robust distributed infrastructure for ML training, inference, and evaluation.
  • Support model training and deployment across multiple clusters and multiple clouds, optimizing for throughput and cost.
  • Optimizing efficiency of ML models and other workloads in terms of latency, throughput, memory consumption, etc. (e.g., via GPU performance engineering), pushing the limits of what’s possible with the current hardware.
  • Contribute to the long-term vision for Genesis’ infra platform.
You are
  • Strong engineer who constantly strives for technical excellence. You can write clean code and have a deep understanding of the codebases you work in.
  • Deeply experienced with distributed training and inference of large models on GPU clusters and some of the core libraries and frameworks we use: Pytorch, Pytorch Lightning, Pytorch Geometric, and Ray.
  • Independent thinker with a strong sense of ownership and capability of engineering robust systems from first-principles-based conceptualization to state-of-the-art realization.
  • Curious, problem-oriented thinker who is excited to dive deep into the emerging field at the intersection of AI, physics, chemistry, and biology and make foundational contributions and discoveries (no previous experience in anything but ML necessary).
Nice to haves
  • Experienced with building, maintaining and debugging low-level cluster infrastructure running on multiple clouds using Kubernetes and Terraform.
  • Experienced GPU engineer who can quickly figure out performance bottlenecks and architect highly performant code for large scale ML workloads.
  • Experience with XLA, Triton, CUDA, or similar accelerator programming languages and/or deep learning compiler stacks.
  • Experience working with some of the following: molecular systems (protein sequences and 3D structures, small molecules, etc.), ML force fields or other physics-informed models and methods, or point cloud data in other application domains, such as 3D graphics.
Compensation, Benefits, And Perks
  • Competitive compensation package that includes salary and equity.
  • Comprehensive health benefits: Medical, Dental, and Vision (covered 100% for the employees).
  • 401(k) plan.
  • Open (unlimited) PTO policy.
  • Free lunches and dinners at our offices.
  • Paid family leave (maternity and paternity).
  • Life and long- and short-term disability insurance.
About Genesis Molecular AI

Genesis Molecular AI is pioneering foundation models for molecular AI to unlock a new era of drug design and development. Our generative and predictive AI platform, GEMS (Genesis Exploration of Molecular Space), integrates AI and physics into industry-leading models to generate and optimize drug molecules, including the breakthrough generative diffusion model Pearl for structure prediction. Genesis is backed by premier AI and life science investors, including a16z, NVIDIA, Rock Springs Capital, Menlo Ventures, T. Rowe Price, Fidelity, and Radical Ventures. Genesis has also signed category-leading AI-pharma deals, the most recent of which was a significant expansion with Incyte (see coverage in Forbes and GEN) with a total potential deal value of several billion dollars. Genesis is headquartered in San Mateo, CA, with a fully integrated laboratory in San Diego. We are proud to be an inclusive workplace and an Equal Opportunity Employer.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Machine Learning Infrastructure Engineer
Machine Learning Infrastructure Engineer

Genesis Therapeutics • San Mateo (CA)

On-site
USD 180,000 - 260,000
Salary + equity
Health benefits including Medical, 401
Open PTO
+1
Machine Learning Infrastructure Engineer
Machine Learning Infrastructure Engineer

Genesis Molecular AI • San Mateo (CA)

On-site
USD 140,000 - 210,000
Health benefits
401(k) plan
Open PTO
+3
ML Research Scientist, Foundation Models (Senior / Staff / Principal)
ML Research Scientist, Foundation Models (Senior / Staff / Principal)

Menlo Ventures • San Mateo (CA)

On-site
USD 180,000 - 240,000
Equity
Health benefits
401(k)
+4
ML Research Scientist, Foundation Models (Senior / Staff / Principal)
ML Research Scientist, Foundation Models (Senior / Staff / Principal)

Genesis Molecular AI • New York (NY)

On-site
USD 120,000 - 180,000
Competitive salary and equity
Comprehensive health benefits
Unlimited PTO policy
ML Research Scientist – Foundation Models, Senior / Staff / Principal
ML Research Scientist – Foundation Models, Senior / Staff / Principal

Genesis Molecular AI • San Mateo (CA)

On-site
USD 120,000 - 180,000
Competitive salary and equity
Comprehensive health benefits
401(k) plan
+4
ML Research Engineer, Foundation Models – Senior / Staff / Principal
ML Research Engineer, Foundation Models – Senior / Staff / Principal

Genesis Molecular AI • San Mateo (CA)

On-site
USD 120,000 - 150,000
Competitive compensation package
Comprehensive health benefits
401(k) plan
+4
Applied ML Scientist, Cheminformatics (Staff / Principal)
Applied ML Scientist, Cheminformatics (Staff / Principal)

Genesis • Burlingame (CA)

On-site
USD 180,000 - 240,000
Salary and equity
Health benefits
401(k) plan
+2
Applied ML Scientist (Staff / Principal)
Applied ML Scientist (Staff / Principal)

Genesis Molecular AI • San Mateo (CA)

On-site
USD 180,000 - 260,000
Competitive compensation package
Comprehensive health benefits
401(k) plan
+4
Software Engineer - Core Infrastructure
Software Engineer - Core Infrastructure

Genesis Molecular AI • San Mateo (CA)

On-site
USD 120,000 - 160,000
Competitive salary
Equity
Medical, dental, and vision insurance
+1
Product Management Lead
Product Management Lead

Genesis Therapeutics • San Diego (CA), New York (NY), San Mateo (CA)

Hybrid
USD 120,000 - 190,000