Get more replies from employers
Send a job-specific resume in minutes.
Cerebras Systems is seeking an Inference Platform SDET to join the Inference Service Quality team. You will own the quality and reliability of the infrastructure that deploys and runs the Cerebras Inference Platform — from CI/CD pipelines and Kubernetes‑based deployments to ingress, load balancing, and service discovery.
This role involves validating the platform in cloud environments and on Cerebras hardware, collaborating with the development team to ship reliable features, building testbeds
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry‑leading training and inference speeds; over 10 times faster than GPU‑based hyperscale cloud inference services.
This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real‑time iteration and increasing intelligence via additional agentic computation.
Cerebras works with the leading model labs, global enterprises, and cutting‑edge AI‑native startups. OpenAI recently announced a multi‑year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high‑speed inference.
We are looking for an Inference Platform SDET to join the Inference Service Quality team at Cerebras and work on the inference platform. This team sits at the intersection of distributed systems, cloud and cluster infrastructure, and the software stack that serves the world's fastest AI inference.
In this role, you will own the quality and reliability of the infrastructure that deploys and runs the Cerebras Inference Platform — from CI/CD pipelines and Kubernetes‑based deployments to ingress, load balancing, and service discovery. You will validate the platform both in cloud environments and on real Cerebras clusters, working side by side with the Inference Platform development team to catch issues before our customers do.
This is an excellent opportunity for engineers who enjoy infrastructure, automation, and debugging across the full deployment stack, and who want to ensure that a platform serving inference at massive scale stays fast, reliable, and production‑ready.
Toronto / Sunnyvale
Inference Service Quality
Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.
This website or its third‑party tools process personal data. For more details, click here to review our CCPA disclosure notice.