Senior / Staff Cloud Infrastructure Engineer

Lasso Informatics Inc.

Montreal (administrative region)

Hybrid

CAD 80,000 - 130,000

Full time

11 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Lasso Informatics is a SaaS start-up delivering a live research data management and analysis platform used by researchers globally. We are seeking an experienced cloud infrastructure engineer to own multi-account AWS and GCP environments, drive observability, incident response, and proactive improvements across automation, security, and reliability.

The role requires hands-on experience with AWS, GCP, and Datadog, plus strong scripting and IaC skills.

Qualifications

  • 5+ years of infrastructure/cloud systems experience
  • Strong hands-on AWS ownership
  • Experience with GCP services (Compute Engine, GKE, IAM)
  • Experience building/evolving an observability stack (Datadog preferred)
  • Ability to operate autonomously and drive outcomes
  • Experience in regulated/compliance-driven environments (HIPAA, GDPR, NIST, SOC 2)
  • Strong scripting/automation skills (Bash, Python)
  • Infrastructure as Code experience (Terraform preferred)
  • Ability to produce production-quality technical documentation
  • Must be legally authorized to work in Canada

Responsibilities

  • Own and operate multi-account AWS and GCP environments
  • Lead capacity planning, patching, and cost optimization
  • Manage infrastructure using Infrastructure as Code (Terraform, Ansible, or equivalent)
  • Identify and remediate security risks and reliability gaps
  • Own and evolve the Datadog observability stack
  • Ensure coverage for logs, metrics, and alerts in production systems
  • Build a unified real-time operational dashboard
  • Lead incident response, root cause analysis, and post-mortems
  • Maintain architecture diagrams, SOPs, and incident runbooks

Skills

Autonomy
Problem solving
Documentation
Communication
Security best practices
Incident response
Root cause analysis
Cost optimization

Tools

AWS
GCP
Datadog
Terraform
Ansible
Kubernetes

Job description

Full Time CA

Salary Range: $80,000.00 To $130,000.00 Annually

Lasso Informatics is a SaaS start-up with a live research data management and analysis platform that brings together multi-modal (imaging, genetics, behavioral, and biosample) data for large-scale studies. Thousands of researchers across the globe rely on our platform today, and we are rapidly iterating and improving to push the boundaries of what is possible in research data management.

We live to innovate, and empower scientists to focus on the science, not the technology, leading to a faster time to science and cure.

Our team is incredibly diverse both by background and expertise, and that is not by accident. We believe that the most creative and powerful solutions come from different ways of thinking about the world. You will be working in an inspiring ecosystem alongside world-renowned professionals in medicine, physics, engineering, imaging, epidemiology, software development, and genetics. We thrive on empowering our colleagues to be thought leaders and innovate fresh new solutions for an exciting and rapidly changing field.

What You Will Do
  • Own and operate multi-account AWS and GCP environments (access, reliability, cost, and security posture)
  • Lead capacity planning, patching, and cost optimization proactively
  • Manage infrastructure using Infrastructure as Code (Terraform, Ansible, or equivalent)
  • Identify and remediate security risks, misconfigurations, and reliability gaps
Observability and Reliability (Datadog)
  • Own and evolve the Datadog observability stack across all environments
  • Ensure strong coverage across logs, metrics, and alerting for production systems
  • Continuously improve signal-to-noise ratio in alerting
  • Build and maintain a unified operational dashboard with real-time system visibility
  • Drive measurable improvements in MTTD and MTTR within the first 6 months
Incident Response and Operational Excellence
  • Lead incident response, root cause analysis, and post-mortems
  • Improve system resilience by addressing root causes and preventing recurrence
  • Identify and close monitoring and operational gaps before they surface as incidents
Documentation and Enablement
  • Maintain production-grade documentation for critical systems, including architecture diagrams, SOPs, and incident runbooks
  • Ensure documentation enables independent execution by other engineers
  • Keep documentation aligned with system reality as environments evolve
Proactive Improvement (Core Expectation)

Deliver consistent, measurable improvements across automation and toil reduction, observability coverage and quality, security posture and compliance readiness, and operational processes and tooling. Approach problems with a systems mindset: fix root causes, not symptoms.

CORE EXPECTATION

About Lasso Informatics

Lasso Informatics is a SaaS start-up with a live research data management and analysis platform that brings together multi-modal (imaging, genetics, behavioral, and biosample) data for large-scale studies. Thousands of researchers across the globe rely on our platform today, and we are rapidly iterating and improving to push the boundaries of what is possible in research data management.

We live to innovate, and empower scientists to focus on the science, not the technology, leading to a faster time to science and cure.

Our team is incredibly diverse both by background and expertise, and that is not by accident. We believe that the most creative and powerful solutions come from different ways of thinking about the world. You will be working in an inspiring ecosystem alongside world-renowned professionals in medicine, physics, engineering, imaging, epidemiology, software development, and genetics. We thrive on empowering our colleagues to be thought leaders and innovate fresh new solutions for an exciting and rapidly changing field.

What You Will Do
Cloud Infrastructure Ownership
  • Own and operate multi-account AWS and GCP environments (access, reliability, cost, and security posture)
  • Lead capacity planning, patching, and cost optimization proactively
  • Manage infrastructure using Infrastructure as Code (Terraform, Ansible, or equivalent)
  • Identify and remediate security risks, misconfigurations, and reliability gaps
Observability and Reliability (Datadog)
  • Own and evolve the Datadog observability stack across all environments
  • Ensure strong coverage across logs, metrics, and alerting for production systems
  • Continuously improve signal-to-noise ratio in alerting
  • Build and maintain a unified operational dashboard with real-time system visibility
  • Drive measurable improvements in MTTD and MTTR within the first 6 months
Incident Response and Operational Excellence
  • Lead incident response, root cause analysis, and post-mortems
  • Improve system resilience by addressing root causes and preventing recurrence
  • Identify and close monitoring and operational gaps before they surface as incidents
Documentation and Enablement
  • Maintain production-grade documentation for critical systems, including architecture diagrams, SOPs, and incident runbooks
  • Ensure documentation enables independent execution by other engineers
  • Keep documentation aligned with system reality as environments evolve
Proactive Improvement (Core Expectation)

Deliver consistent, measurable improvements across automation and toil reduction, observability coverage and quality, security posture and compliance readiness, and operational processes and tooling. Approach problems with a systems mindset: fix root causes, not symptoms.

CORE EXPECTATION

The right candidate has a track record of entering imperfect environments and making them materially better — taking ownership without waiting for direction and raising the standard for how infrastructure is operated.

What We Are Looking For
Required
  • 5+ years of infrastructure / cloud systems experience, with substantial recent hands-on AWS ownership
  • Strong hands-on experience with core AWS services (EC2, S3, RDS, EKS, IAM, VPC, etc.)
  • Working experience with GCP (Compute Engine, GKE, IAM, etc.)
  • Experience building or significantly evolving an observability stack (Datadog preferred)
  • Proven ability to operate with autonomy: identifying problems, designing solutions, and driving outcomes
  • Experience in regulated or compliance-driven environments (HIPAA, GDPR, NIST, SOC 2, or equivalent)
  • Strong scripting and automation skills (Bash, Python, or equivalent)
  • Experience with Infrastructure as Code (Terraform preferred)
  • Ability to produce clear, production-quality technical documentation
  • Must be legally authorized to work in Canada
Strongly Preferred
  • AWS or GCP certifications
  • Kubernetes experience (EKS, GKE)
  • Experience scaling infrastructure in a high-growth SaaS environment
  • Background in healthcare, life sciences, or government-facing systems
  • Experience acting as a senior or leading technical voice within a team
Why Lasso

We work on infrastructure supporting real-world clinical and academic research across North America and Europe with high ownership, direct collaboration with engineering and leadership, and the chance to build a best-in-class SysOps function from a strong foundation.

Compensation and benefits
  • Competitive salary band with performance bonus tied to measurable infrastructure outcomes
  • Group benefits plan (100% employer-paid)
  • RRSP contribution matching (2%)
  • Professional development budget (AWS, GCP, Datadog certifications and training)
  • Hybrid flexibility for Montreal-based team members, 2 days/week in office
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior / Staff Cloud Infrastructure Engineer
Senior / Staff Cloud Infrastructure Engineer

Lasso Informatics • Montreal (administrative region)

Hybrid
CAD 90,000 - 120,000
100% employer-paid group benefits
2% RRSP matching
Professional development budget
+1
Full Stack Developer
Full Stack Developer

Lasso Informatics Inc. • Montreal (administrative region)

Hybrid
CAD 100,000 - 120,000
Team Leader, AWS infrastructure/FinOps
Team Leader, AWS infrastructure/FinOps

Croesus • Laval (administrative region)

On-site
CAD 110,000 - 140,000
Annual salary + profit-sharing plan
Gym at Laval head office
Telemedicine + group insurance
+6
Full Stack Developer
Full Stack Developer

Data Sciences • Montreal (administrative region)

Hybrid
CAD 90,000 - 120,000
Assurance maladie et dentaire
15 jours de congés payés + 5 jours PAE
Budget formation & progression de Car
+2
Senior Backend Engineer (AWS)
Senior Backend Engineer (AWS)

Lumenalta • Victoria

Remote
CAD 76,000 - 139,000
Flexible working hours
Health insurance (medical and dental)
Professional development fund
+2
Senior Cloud Engineer
Senior Cloud Engineer

Jobless • Winnipeg

Hybrid
CAD 120,000 - 160,000
Senior Infrastructure Developer
Senior Infrastructure Developer

Blue J Legal • Toronto

Hybrid
CAD 160,000 - 180,000
Competitive base salary and stock options
Flexible remote work options
Healthy work/life balance
+1
Senior Software Developer, Data & MLOps
Senior Software Developer, Data & MLOps

OSEDEA • Montreal (administrative region)

Hybrid
CAD 85,000 - 115,000
Competitive Salary
Pension plan contribution (RRSP)
Flexible work hours
+5
Senior Data Engineer - ETL & Integrations
Senior Data Engineer - ETL & Integrations

Lasso Informatics • Montreal (administrative region)

On-site
CAD 110,000 - 150,000
In-office presence Tue–Thu
Competitive salary & benefits
Leadership opportunities
+1
Snr Cloud Infrastructure Engineer
Snr Cloud Infrastructure Engineer

Bosa Properties Inc • Vancouver

On-site
CAD 105,000 - 143,000