C

Cerebras

Verified Job Source

Cerebras Systems develops the fastest AI computer, the CS-3, featuring the revolutionary WSE-3 chip, accelerating deep learning tasks by unprecedented orders of magnitude.

Sunnyvale, California

Semiconductor Manufacturing
1,001–5,000 people

About

Cerebras Systems is the world's fastest AI inference. We are powering the future of generative AI. We’re a team of pioneering computer architects, deep learning researchers, and engineers building a new class of AI supercomputers from the ground up. Our flagship system, Cerebras CS-3, is powered by the Wafer Scale Engine 3—the world’s largest and fastest AI processor. CS-3s are effortlessly clustered to create the largest AI supercomputers on Earth, while abstracting away the complexity of traditional distributed computing. From sub-second inference speeds to breakthrough training performance, Cerebras makes it easier to build and deploy state-of-the-art AI—from proprietary enterprise models to open-source projects downloaded millions of times. Here’s what makes our platform different: 🔦 Sub-second reasoning – Instant intelligence and real-time responsiveness, even at massive scale ⚡ Blazing-fast inference – Up to 100x performance gains over traditional AI infrastructure 🧠 Agentic AI in action – Models that can plan, act, and adapt autonomously 🌍 Scalable infrastructure – Built to move from prototype to global deployment without friction Cerebras solutions are available in the Cerebras Cloud or on-prem, serving leading enterprises, research labs, and government agencies worldwide. 👉 Learn more: www.cerebras.ai Join us: https://cerebras.net/careers/

Open positions

Cluster Operations Software Engineer

On-site · Toronto

Manage and operate large-scale AI compute clusters using the Wafer-Scale Engine to ensure high availability and performance. Develop software solutions for monitoring, automation, and fleet management to optimize compute capacity.

Staff Software Engineer, GPU Inference

Hybrid · Toronto

You will design, build, and maintain the GPU inference stack, ensuring high performance and reliability for large-scale AI workloads. Additionally, you will establish operational practices for the GPU fleet, including deployment, monitoring, and performance tuning across distributed systems.

ML Software Engineer - Integration & Quality - New Grad

Hybrid · Canada

Integrate, test, and validate the software stack powering the Cerebras AI platform across runtime, compiler, and hardware layers. Develop automated tests and tools to improve the reliability and quality of large-scale machine learning workloads.

Software Engineer - Tools & Infrastructure / DevOps

On-site · Toronto

Develop and maintain CICD pipelines and artifact lifecycle systems to ensure efficient build and release workflows. Provision and optimize cloud infrastructure while creating internal tooling to enhance developer velocity and engineering productivity.