Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Intellectual Property Strategy, Patent Prosecution, AI Infrastructure, Trademark Law, Copyright Law
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Senior Software Development Engineer in Test (SDET) - AI Cluster Networking and Security
Bengaluru, Karnataka, India · Hybrid
Senior+$4.7B raised
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Linux System Administration, Network Configuration, x86 Server Hardware, BIOS Configuration, Bash
Senior Software Development Engineer in Test (SDET) - AI Cluster
Toronto, Ontario, Canada · Hybrid
Senior$4.7B raised
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Python, C++, Operating Systems, Computer Architecture, Linux
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: NetSuite Administration, ERP Administration, Business Process Optimization, User Provisioning, SOX Compliance
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Mechanical Engineering, Chilled Water Cooling Systems, High-Density AI Cooling, Data Center Infrastructure, Construction Management
Senior Front End Design Engineer (Microarchitecture)
Bengaluru, Karnataka, India · Hybrid
Senior+$4.7B raised
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: RTL Design, Microarchitecture, Front End Chip Integration, Synthesis, Python
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Site Reliability Engineering, Infrastructure Engineering, Platform Engineering, Capacity Management, Orchestration Systems
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
Full-time
Posted 23d ago
~40 hrs/week
Responsibilities
Lead the end-to-end release qualification process for Cerebras Cloud, Inference Platform, and Inference API to ensure stable and predictable releases. Build and manage a dedicated qualification team to handle triage, debugging, and go/no-go decisions for weekly releases.
Requirements
Requires over 7 years of experience in software integration or quality engineering with a proven track record of managing complex distributed system releases. Must possess deep expertise in Python, automation frameworks, and the ability to lead cross-functional initiatives across multiple organizations.
Full job description
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.
Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.
About The Team
The Inference Service Quality team owns the confidence behind every release shipped across Cerebras Cloud, Inference Platform, and Inference API. We work closely with platform, infrastructure, ML systems, and product engineering teams to ensure that rapid iteration never comes at the expense of customer trust. Our environment spans distributed cloud systems, multi-region deployments, APIs, orchestration layers, and hardware-backed inference services.
Release velocity is high and the systems grow more complex every quarter. We need a leader who can turn release qualification from an ad hoc, reactive scramble into a predictable, owned discipline that scales with the business.
About The Role
As the Release Qualification Team Lead, you will be the accountable owner for weekly end-to-end release qualification on behalf of our org, establishing it as a first-class, owned discipline. You will build and run a purpose-built release qualification team with planned triage capacity and the depth of context needed to triage, debug, and confidently decide what unblocks a release. You will own the full release qualification lifecycle — test planning, release planning, documentation, and tooling — turning hard-won knowledge into organized, accessible practice.
As the technical counterpart to the cross-org qualification team, you'll be the person who truly understands our tests and can speak for them. Success in this role depends on effective communication and close coordination across teams and time zones, keeping every stakeholder aligned through fast-moving release windows.
A central part of your charter is expanding end-to-end coverage across Cloud, Inference Platform, and Inference API — qualifying the steady stream of new Cerebras models, the new APIs that support them, and new cloud features, and driving qualification all the way through as the systems grow.
Responsibilities
Drive weekly end-to-end release qualification as the accountable owner from our org.
Build and run a purpose-built release qualification team so triage capacity is planned, not scrambled — ensuring strong, predictable representation from our org rather than ad hoc pulls.
Own the full release qualification lifecycle: test planning, release planning, document
Own triage, debug, and escalation during release windows — including deciding what blocks a release and routing fixes to the right owners.
Drive expanded end-to-end coverage across Cloud, Inference Platform, and Inference API,e way through despite cross-org boundaries.
Partner with the cross-org release qualification team as the technical counterpart who actually understands our tests.
Distinguish real regressions from flakiness quickly, and keep a living record of what td who to escalate to.
Mentor engineers on triage, debugging practices, and qualification methodology.
Skills & Qualifications
7+ years of relevant industry experience in software integration, development, or quality engineering, including prior release qualification experience.
Proven ownership of end-to-end release or qualification processes for complex, distribu
Strong track record of debugging complex issues across distributed, scaled-out deployments under release-blocking time pressure.
Deep expertise in automation and programming using one or more languages such as Pythongn and build reusable test frameworks from the ground up.
Demonstrated ability to lead cross-functional initiatives spanning multiple orgs, product development, product management, and field teams.
Sound judgment on coverage trade-offs, risk-based prioritization, and go/no-go release
Excellent verbal and written communication skills, with experience presenting technical findings to both engineering and leadership audiences.
Strong organizational skills, ownership mindset, and ability to drive projects to compl
Experience leading and mentoring engineers across geographically dispersed teams and time zones.
Preferred Skills & Qualifications
Hands-on experience with ML workloads including LLM and/or multimodal training or inference.
Experience designing test strategies for distributed systems, cloud infrastructure, and
Experience with microservices deployment, debugging, and orchestration at scale.
Familiarity with hardware-backed inference services and the trade-offs of hardware/soft
Prior experience building or significantly shaping a team's quality engineering culture or test infrastructure.
Why Join Cerebras
People who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:
Build a breakthrough AI platform beyond the constraints of the GPU.
Publish and open source their cutting-edge AI research.
Work on one of the fastest AI supercomputers in the world.
Enjoy job stability with startup vitality.
Our simple, non-corporate work culture that respects individual beliefs.
Find out more about what it's like to work at Cerebras here!
Apply today and become part of the forefront of groundbreaking advancements in AI!
Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.
This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.
Related keywords
AI ChipInference PlatformDistributed SystemsPythonLLMMultimodal TrainingCloud InfrastructureMicroservicesOrchestrationRelease LifecycleTest PlanningTriageRegression TestingQuality EngineeringCerebras CloudInference API
Cerebras Systems is the world's fastest AI inference. We are powering the future of generative AI. We’re a team of pioneering computer architects, deep learning researchers, and engineers building a new class of AI supercomputers from the ground up.
Our flagship system, Cerebras CS-3, is powered by the Wafer Scale Engine 3—the world’s largest and fastest AI processor. CS-3s are effortlessly clustered to create the largest AI supercomputers on Earth, while abstracting away the complexity of traditional distributed computing.
From sub-second inference speeds to breakthrough training performance, Cerebras makes it easier to build and deploy state-of-the-art AI—from proprietary enterprise models to open-source projects downloaded millions of times.
Here’s what makes our platform different:
🔦 Sub-second reasoning – Instant intelligence and real-time responsiveness, even at massive scale
⚡ Blazing-fast inference – Up to 100x performance gains over traditional AI infrastructure
🧠 Agentic AI in action – Models that can plan, act, and adapt autonomously
🌍 Scalable infrastructure – Built to move from prototype to global deployment without friction
Cerebras solutions are available in the Cerebras Cloud or on-prem, serving leading enterprises, research labs, and government agencies worldwide.
👉 Learn more: www.cerebras.ai
Join us: https://cerebras.net/careers/
Offices: 1237 E Arques Ave, Sunnyvale, California 94085, US · 150 King St W, Toronto, Ontario M5H 1J9, CA · Tokyo, JP · Bangalore, IN
artificial intelligencedeep learningnatural language processinginferencemachine learningllmAIenterprise AIand fast inferenceSemiconductor
Cerebras Systems is the world's fastest AI inference. We are powering the future of generative AI. We’re a team of pioneering computer architects, deep learning researchers, and engineers building a new class of AI supercomputers from the ground up.
Our flagship system, Cerebras CS-3, is powered by the Wafer Scale Engine 3—the world’s largest and fastest AI processor. CS-3s are effortlessly clustered to create the largest AI supercomputers on Earth, while abstracting away the complexity of traditional distributed computing.
From sub-second inference speeds to breakthrough training performance, Cerebras makes it easier to build and deploy state-of-the-art AI—from proprietary enterprise models to open-source projects downloaded millions of times.
Here’s what makes our platform different:
🔦 Sub-second reasoning – Instant intelligence and real-time responsiveness, even at massive scale
⚡ Blazing-fast inference – Up to 100x performance gains over traditional AI infrastructure
🧠 Agentic AI in action – Models that can plan, act, and adapt autonomously
🌍 Scalable infrastructure – Built to move from prototype to global deployment without friction
Cerebras solutions are available in the Cerebras Cloud or on-prem, serving leading enterprises, research labs, and government agencies worldwide.
👉 Learn more: www.cerebras.ai
Join us: https://cerebras.net/careers/
Offices: 1237 E Arques Ave, Sunnyvale, California 94085, US · 150 King St W, Toronto, Ontario M5H 1J9, CA · Tokyo, JP · Bangalore, IN
artificial intelligencedeep learningnatural language processinginferencemachine learningllmAIenterprise AIand fast inferenceSemiconductor