Senior MLOps Engineer

Deep Genomics

Toronto

On-site

CAD 175,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Stock options
Comprehensive benefits (health/vision)
Flexible work environment

Job summary

Deep Genomics in Toronto is seeking a Senior MLOps Engineer to own and evolve the infrastructure powering our ML pipelines. You will work with ML scientists, bioinformaticians, and software engineers to ensure reliable, reproducible, and scalable platforms.

You will manage cloud environments (GCP), CI/CD (CircleCI, GitHub Actions), and container orchestration (Kubernetes, Docker), while supporting GPU workloads and model deployment. A collaborative, growth-oriented team awaits.

Qualifications

  • 4+ years of experience operating production infrastructure.
  • Proficiency with cloud platforms (GCP preferred; AWS/Azure acceptable) and Terraform.
  • Hands-on Kubernetes and containerization experience (Docker).
  • Solid background in CI/CD systems (CircleCI, GitHub Actions, or similar).
  • Experience managing GPU compute and driver issues.
  • Strong Python programming skills.

Responsibilities

  • Maintain and evolve cloud infrastructure (GCP) using IaC tools such as Terraform.
  • Manage IAM, RBAC, and permission policies across cloud environments.
  • Own and evolve CI/CD pipelines (CircleCI, GitHub Actions) for engineering and ML teams.
  • Administer workflow orchestration platforms (Seqera/Nextflow, Argo, Kubeflow).
  • Operate ML experiment tracking and registry tooling (W&B, MLflow).
  • Build and maintain containerized environments (Docker) and manage Kubernetes clusters.
  • Manage GPU resources – provisioning, scheduling, and debugging driver issues.
  • Write Python tooling and integrations to support ML infrastructure.
  • Assist deployment of ML models to production and monitor performance.

Skills

Cloud platforms
Kubernetes
CI/CD
Python
Collaboration

Tools

Terraform
Docker
CircleCI
GitHub Actions
W&B
MLflow
Seqera/Nextflow
Argo
Kubeflow

Job description

About Us

Deep Genomics is at the forefront of using artificial intelligence to transform drug discovery. Our proprietary AI platform decodes the complexity of RNA biology to identify novel drug targets, mechanisms, and therapeutics inaccessible through traditional methods. With expertise spanning machine learning, bioinformatics, data science, engineering, and drug development, our multidisciplinary team in Toronto and Cambridge, MA is revolutionizing how new medicines are created.

Opportunity

Join us in building the future of AI-driven drug discovery as a Senior MLOps Engineer. You will own and evolve the infrastructure that powers our ML pipelines – from cloud environments and CI/CD systems to workflow orchestration and model deployment. You will work closely with ML scientists, bioinformaticians, and software engineers to keep our platform reliable, reproducible, and scalable.

Ideal Candidate

You are someone who enjoys keeping the infrastructure running smoothly so that scientists can focus on their research. You are comfortable working across cloud platforms, CI/CD systems, containers, and GPUs – and you take pride in making these systems reliable and easy for others to use. You have 4+ years of experience in production infrastructure or MLOps, you write solid Python, and you are curious about the ML and scientific workflows your work supports. Above all, you are a collaborative, kind team member who communicates clearly, adapts to evolving needs, and is happy to help colleagues grow their own infrastructure skills along the way. If this sounds like you, we would love to hear from you.

Key Responsibilities
  • Maintain and improve cloud infrastructure (GCP) using Infrastructure-as-Code tools (Terraform).
  • Manage IAM, RBAC, and permission policies across cloud environments.
  • Own and evolve CI/CD pipelines (CircleCI, GitHub Actions) and ensure best practices are followed across the engineering and ML teams.
  • Administer and support workflow orchestration platforms (e.g., Seqera/Nextflow, Argo, Kubeflow).
  • Operate and configure ML experiment tracking and registry tooling (e.g., W&B, MLflow).
  • Build and maintain containerized environments (Docker) and manage Kubernetes clusters.
  • Manage GPU resources – provisioning, scheduling, and debugging hardware and driver issues.
  • Write and maintain Python tooling, scripts, and integrations that support ML infrastructure.
  • Help deploy ML models to production environments and monitor their performance.
Basic Qualifications

4+ years of experience operating production infrastructure.

  • Proficiency with cloud platforms (GCP preferred; AWS/Azure acceptable) and Infrastructure-as-Code (Terraform).
  • Extensive Hands-on experience with Kubernetes and containerization (Docker).
  • Solid background in CI/CD systems (CircleCI, GitHub Actions, or similar).
  • Experience managing GPU compute (provisioning, debugging, driver management).
  • Familiarity with Python package and environment management (e.g., pip, conda, pixi).
  • Strong Python programming skills.
  • Self-motivated problem solver with excellent communication skills.
Preferred Qualifications
  • Understanding of ML frameworks (e.g., PyTorch, PyTorch Lightning), ML workflows (training, inference, evaluation), and the model lifecycle.
  • Familiarity with MLOps tooling (e.g., W&B, Ray, VertexAI) and distributed compute patterns
    (e.g., DDP, realtime/batch inference, multi-node training).
  • Familiarity with Kubernetes CRDs and batch/gang schedulers (e.g., Volcano, Kueue).
  • Experience working with large-scale datasets (storage, versioning, efficient access patterns).
  • Experience working directly with scientists and researchers in an interdisciplinary setting.
  • Knowledge of biology and/or machine learning science.
  • Familiarity with data compliance and governance frameworks (e.g., HIPAA, SOC 2).
  • Previous startup experience.
What We Offer
  • A collaborative and innovative environment at the frontier of computational biology, machine learning, and drug discovery.
  • Highly competitive compensation, including meaningful stock ownership.
  • Comprehensive benefits - including health, vision, and dental coverage for employees and families, employee and family assistance program.
  • Flexible work environment - including flexible hours, extended long weekends, holiday shutdown, unlimited personal days.
  • Maternity and parental leave top-up coverage, as well as new parent paid time off.
  • Focus on learning and growth for all employees - learning and development budget & lunch and learns.
  • Facilities located in the heart of Toronto - the epicenter of machine learning and AI research and development, and in Kendall Square, Cambridge, Mass. - a global center of biotechnology and life sciences.

If you have a disability or special need, accommodation is available on request for candidates taking part in all aspects of the selection process.

*This posting reflects a current vacancy.

We offer competitive compensation aligned with local market benchmarks. The salary range for this role is $175,000 - $200,000, and reflects Canada-based roles; compensation may differ for U.S.-based candidates.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Research Scientist, Machine Learning (BioFM)
Senior Research Scientist, Machine Learning (BioFM)

Amplitude Venture Capital • Toronto

On-site
CAD 175,000 - 200,000
Flexible work environment
Comprehensive health benefits
Stock ownership
+1
Senior Research Scientist, Machine Learning (BioFM)
Senior Research Scientist, Machine Learning (BioFM)

Deep Genomics Inc. • Toronto

On-site
CAD 175,000 - 200,000
Competitive stock ownership
Flexible work environment
Health, vision, and dental coverage
+2
Senior ML Platform Engineer
Senior ML Platform Engineer

Deep Genomics • Toronto

On-site
CAD 175,000 - 200,000
Stock options
Comprehensive benefits (health/vision)
Flexible work environment
Senior LLMOps Engineer -Cloud / AI Infrastructure
Senior LLMOps Engineer -Cloud / AI Infrastructure

Talent To Hire Inc. • Toronto

On-site
CAD 120,000 - 160,000
Competitive salary
Meaningful equity
Innovative work culture
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Jobgether • Toronto

Hybrid
CAD 185,000 - 225,000
Annual bonus
RSU equity
Health benefits
+3
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Motion Recruitment • Toronto

On-site
CAD 140,000 - 190,000
Bonus eligible
Medical, Dental, Vision Insurance
Vacation Time
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Motion Recruitment Partners LLC • Toronto

On-site
CAD 180,000 - 280,000
Senior ML Systems Engineer, Frameworks & Tooling
Senior ML Systems Engineer, Frameworks & Tooling

Cohere • Montreal

On-site
CAD 100,000 - 140,000
Open and inclusive culture
Weekly lunch stipend and snacks
Full health and dental benefits
+2
Sr. AI Engineer
Sr. AI Engineer

TheAppLabb • Toronto

On-site
CAD 110,000 - 170,000
Competitive salary
Opportunities for career growth
Fitness challenge incentives
Senior DevOps Engineer, for AI-based Systems
Senior DevOps Engineer, for AI-based Systems

Oncoustics • Toronto

On-site
CAD 110,000 - 140,000