Senior Site Reliability Engineer

Cerebras

Vancouver

On-site

CAD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading AI technology company is seeking a Senior Site Reliability Engineer/DevOps to design and maintain their scalable and secure infrastructure. The role involves working extensively with AWS, Terraform, and Docker, and requires strong troubleshooting skills. Candidates should have over 6 years of SRE or DevOps experience and knowledge of cloud security practices. The position is based in Vancouver, Canada and offers an opportunity to work with innovative technology in a growing field.

Qualifications

  • 6+ years of experience in SRE, DevOps, or infrastructure engineering.
  • Extensive hands-on experience with AWS and core services.
  • Strong experience with scripting languages like Python and Bash.

Responsibilities

  • Design and maintain scalable, secure, and reliable infrastructure.
  • Architect a unified monitoring and alerting system for engineering teams.
  • Drive infrastructure automation and CI/CD improvements.

Skills

AWS
Terraform
Docker
Python
Bash
SQL
NoSQL
Monitoring tools (New Relic, Prometheus, Grafana)
Production-grade infrastructure design
Troubleshooting

Tools

Terraform
Docker
Monitoring tools (New Relic, Prometheus, Grafana)

Job description

Responsibilities

We’re seeking a senior Site Reliability Engineer/DevOps who is passionate about building the best infrastructure and maintaining the health of the systems.

  • Design and maintain scalable, secure, and reliable infrastructure to support Regie.ai's SaaS platform and AI/data workloads.
  • Architect a unified monitoring and alerting system for engineering teams to continuously monitor and improve system availability, reliability, performance.
  • Drive infrastructure automation and CI/CD improvements to reduce operational overhead and deployment risk.
  • Optimize infrastructure costs, support compliance efforts (e.g., SOC 2), and enforce security best practices.
Required Skills & Qualifications
  • 6+ years of experience in SRE, DevOps, or infrastructure engineering roles.
  • Extensive hands‑on experience with AWS and its core services.
  • Strong experience with Terraform (or similar IaC tools), Docker and containerization, and modern CI/CD systems.
  • Proficient in scripting or programming languages such as Python and Bash.
  • Deep experience with monitoring and alerting tools (e.g., New Relic, Prometheus, Grafana, PagerDuty).
  • Strong hands‑on experience with both SQL and NoSQL databases (e.g., MongoDB, PostgreSQL, MySQL).
  • Proven track record of designing and maintaining production‑grade infrastructure with high availability and low latency.
  • Excellent troubleshooting abilities, along with strong communication and collaboration skills.
  • Solid understanding of cloud security and compliance best practices, including SOC 2 readiness and audit support.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Gemini Solutions Pvt Ltd • Toronto

On-site
CAD 120,000 - 170,000
Site Reliability Engineering Manager
Site Reliability Engineering Manager

TekRek • Vancouver

On-site
CAD 180,000 - 240,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

iManage • Toronto

Hybrid
CAD 90,000 - 120,000
Market-competitive salary
Annual performance-based bonus
Comprehensive Health, Vision, Dental, and Life insurance
+4
Site Reliability Engineer
Site Reliability Engineer

Compunnel, Inc. • Montreal (administrative region)

Hybrid
CAD 90,000 - 130,000
Senior Site Reliability Developer
Senior Site Reliability Developer

United States Digital Space LLC • Toronto

On-site
CAD 107,000 - 157,000
Salary transparency
In-person onboarding
Site Reliability Engineer
Site Reliability Engineer

Mantu • Montreal (administrative region)

On-site
CAD 90,000 - 130,000
Software Engineer-Backend
Software Engineer-Backend

Regie.ai • Vancouver

On-site
CAD 80,000 - 100,000
Site Reliability Engineer
Site Reliability Engineer

ALLTECH CONSULTING SVC INC • Quebec

On-site
CAD 90,000 - 130,000
Sr. DevOps Engineer
Sr. DevOps Engineer

AspiringIT • Mississauga

On-site
CAD 120,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

Future Secure AI • Toronto

On-site
CAD 90,000 - 120,000
Flexible work environment
Competitive salary
Diversity and creativity