Software Engineer - Tools & Infrastructure / DevOps

Cerebras

Toronto

On-site

CAD 120,000 - 180,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Cerebras Systems, a leader in AI hardware, seeks a DevOps/Infrastructure engineer to join the team in Toronto. You will help build and maintain CI/CD pipelines, manage artifacts and cloud resources, and collaborate with development teams to optimize workflows and automation.

Strong focus on reliability, scalability, and incident response is expected. Ideal candidates have 4–8 years in DevOps or related roles, with AWS and Kubernetes experience, scripting skills, and proficiency in Linux.

Qualifications

  • 4-8 years of professional experience in DevOps, infrastructure, or software engineering.
  • Familiarity with CI/CD systems and building or maintaining automated pipelines.
  • Understanding of artifact repository management and software packaging concepts.
  • Experience with cloud computing platforms (AWS preferred) and programmatic provisioning.
  • Proficiency with distributed version control and code review processes.
  • Foundational knowledge of Linux/Unix, networking, and scripting for automation.
  • Experience with containerization and orchestration (Kubernetes preferred).
  • Strong troubleshooting skills and a methodical approach to distributed systems.
  • Interest in scaling infrastructure and on-call incident response within Dev Productivity.

Responsibilities

  • Contribute to the development and maintenance of CICD pipelines across the organization.
  • Manage artifact lifecycle, including versioning, storage, and distribution for reproducible builds.
  • Design and improve code review workflows and automated integration processes.
  • Provision, monitor, and optimize cloud infrastructure to support CI workloads.
  • Troubleshoot build failures and pipeline bottlenecks and implement lasting fixes.
  • Contribute to internal build and test infrastructure and tooling to boost productivity.
  • Support AI tooling efforts to enhance engineering efficiency.

Skills

DevOps
CI/CD
Cloud AWS
Kubernetes
Linux/Unix
Scripting
Artifact management
Git / Repository management
Troubleshooting
On-call / Incident response

Education

BS/MS in Computer Science or related field

Tools

Kubernetes
AWS
CI/CD tools

Job description

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.

This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.

Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.

Responsibilities
  • Contribute to the development and maintenance of CICD pipelines, ensuring reliable and efficient build, test, and release workflows across the organization.
  • Help manage artifact lifecycle systems including versioning, storage, distribution, and dependency management to support reproducible builds at scale.
  • Partner with development teams to design and improve code review workflows, branching strategies, and automated integration processes.
  • Provision, monitor, and optimize cloud infrastructure to support CI workloads, balancing cost efficiency with performance and reliability.
  • Troubleshoot build failures, pipeline bottlenecks, and infrastructure issues, driving root-cause analysis and implementing lasting fixes.
  • Contribute to internal build infrastructure, test infrastructure, tooling, and automation that improves developer velocity and engineering productivity.
  • Contribute to the company’s efforts on AI tooling to boost engineering productivity efficiently.
Skills & Qualifications
  • 4-8 years of professional experience in a DevOps, infrastructure, or software engineering role.
  • Familiarity with CICD systems and hands-on experiences in building or maintaining automated build and deployment pipelines.
  • Understanding of artifact repository management and software packaging concepts.
  • Experience with cloud computing platforms (AWS preferred) and programmatic resource provisioning.
  • Proficiency with distributed version control systems, code review processes, and repository management.
  • Foundational knowledge of operating system concepts (Linux/Unix), networking fundamentals, and scripting for automation.
  • Experience with containerization and container orchestration (Kubernetes preferred).
  • Strong troubleshooting skills and a methodical approach to debugging distributed systems.
  • Curiosity about how large-scale infrastructure is built, operated, and improved.
  • Being part of oncall and incident response task force of Dev Productivity org.
Preferred Skills & Qualifications
  • Familiarity with infrastructure-as-code tools and practices.
  • Scripting proficiency in Python or Shell for build automation and tooling.
  • Exposure to build systems and build graph optimization.
  • Understanding of observability practices including monitoring, logging, and alerting.
  • BS/MS in Computer Science or a related field, or equivalent practical experience.
Why Join Cerebras
  • Build a breakthrough AI platform beyond the constraints of the GPU.
  • Publish and open source their cutting-edge AI research.
  • Work on one of the fastest AI supercomputers in the world.
  • Enjoy job stability with startup vitality.
  • Our simple, non-corporate work culture that respects individual beliefs.

Find out more about what it's like to work at Cerebras here!

About

People who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:

Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.

This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Distributed Software Engineer
Distributed Software Engineer

Cerebras Systems, Inc. • Ottawa

On-site
CAD 90,000 - 120,000
Job stability with startup vitality
Open access to cutting-edge AI research
Senior Software Development Engineer in Test (SDET) - AI Cluster
Senior Software Development Engineer in Test (SDET) - AI Cluster

Cerebras • Toronto

On-site
CAD 120,000 - 190,000
FPGA Engineer
FPGA Engineer

Cerebras • Toronto

On-site
CAD 120,000 - 180,000
Cluster Operations Software Engineer
Cluster Operations Software Engineer

Cerebras Systems • Toronto

On-site
CAD 120,000 - 160,000
ML Systems Integration Engineer
ML Systems Integration Engineer

Cerebras • Toronto

On-site
CAD 90,000 - 150,000
Data Center Provisioning Engineer
Data Center Provisioning Engineer

Cerebras • Toronto

On-site
CAD 110,000 - 180,000
DevOps Engineer - New Grad 2026
DevOps Engineer - New Grad 2026

Cerebras Systems, Inc. • Toronto

On-site
CAD 70,000 - 90,000
Opportunity to work on an innovative AI platform
Diverse and inclusive work environment
Job stability with startup vitality
Senior Software Development Engineer in Test (SDET) - AI Cluster
Senior Software Development Engineer in Test (SDET) - AI Cluster

Cerebras Systems, Inc. • Toronto

On-site
CAD 140,000 - 210,000
Software Engineer - New Grad 2026
Software Engineer - New Grad 2026

Cerebras Systems, Inc. • Toronto

Hybrid
CAD 70,000 - 90,000
ML Performance Benchmarking Engineer
ML Performance Benchmarking Engineer

Cerebras • Toronto

Hybrid
CAD 120,000 - 190,000
Groundbreaking technology platform
Equal opportunity employer
Innovative and collaborative work environment