Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: RTL Design, ASIC Vendor Management, Front End Chip Integration, High Speed IO, TCP/IP
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Electrical Engineering, Data Center Design, Power Distribution, Medium-voltage Systems, Low-voltage Systems
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Data Center Architecture, Mechanical Engineering, Electrical Engineering, Liquid Cooling, Power Distribution
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Hardware bring-up, System integration, Board debug, Failure analysis, Oscilloscopes
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Java, QE Automation, Selenium, Page Object Model, Visual Studio Code
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Data Pipeline Architecture, ETL, Hive, Spark, Python
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
AI Inference Core - SDET Technical Lead, Release Integration Testing
Canada · Hybrid
Senior+$4.7B raised
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Python, C++, Go, Test Architecture, Distributed Systems
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
AI Inference Core - Senior SW Engineer for Platform & DevOps
Canada · Hybrid
Mid level$4.7B raised
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
AI Inference Core - Junior SDET, Release Integration Testing
Canada · Hybrid
Entry level$4.7B raised
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Python, Go, Test Automation, Debugging, Distributed Systems
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Data Center Transactions, Legal Strategy, Contract Negotiation, Digital Infrastructure, Real Estate Law
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: ML Model Compilation, Graph Optimization, High-Performance Kernels, LLVM, MLIR
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Deployment Manager – Global Data Center Build and Deploy
Sunnyvale, California, United States · Hybrid
Senior+$4.7B raised
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: AI Cluster Deployment, Rack Integration, High-Density Cabling, Data Center Infrastructure, Troubleshooting
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: Technical Program Management, Capacity Planning, Infrastructure Delivery, Demand Forecasting, Capital Planning
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
$175k–$275k/yr
Full-time
bachelor degree, postgraduate degree
Bonus, Equity
Posted 5d ago
~40 hrs/week
Responsibilities
Responsible for the bring-up and optimization of the Wafer Scale Engine, focusing on developing debug flows and refining AI systems across hardware and software constraints. Collaborates cross-functionally with design, performance, and software teams to enhance silicon performance and streamline release processes.
Requirements
Requires a degree in EE, ECE, CS, or equivalent with 7-10+ years of industry experience, including 3-5 years in Pre-silicon and Post Silicon ASIC hardware. Proficiency in Python, Verilog, and C, along with strong analytical and problem-solving skills, is essential.
Full job description
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.
Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.
The Role: In this exciting role, you will be responsible for bring up and optimizations of Cerebras’s Wafer Scale Engine (WSE). Suitable candidate will have experience delivering end to end solutions working closely with teams across chip design, system performance, software development and productization.
Responsibilities:
On Wafer Scale Engines, develop and debug flows that embed well tested and deployable optimizations in production processes to reduce time and costs
Work on refining AI Systems across H/W-S/W design constraints such as di/dt, V-F characterization space, current and temperature limits in relation to optimizations for performance.
Develop/Enhance infrastructure to enable silicon for real world workload testing
Develop self-checking metrics, as well as instrumentation for debug and coverage
Work with the silicon architects/designers, performance engineers and software engineers to enhance performance of Wafer Scale Engines.
Work across domains such as, Software, Design, Verification, Emulation & Validation to refine and optimize performance and process.
Work with CI/CD tools, git repositories, github, git actions/Jenkins, merge and release flows to streamline test and release.
Skills & Qualifications:
BS/BE/B.Tech or MS/M.Tech in EE, ECE, CS or equivalent work experience
7-10+ years of industry experience
3-5 years of experience in Pre-silicon & Post Silicon ASIC hardware
Good understanding of computer architecture and networking
Excellent Coding in languages such as Python/Verilog/System Verilog and C
Proficient in hardware/software codesign and layered architectures.
Excellent debugging, analytical, and problem-solving skills
Proficient in large scale testing and automation using pytest and python
Good presentation skills to refine diverse information and put forth optimization strategies and results.
Good interpersonal skills, ability & desire to work as a standout colleague
Proven track record of working cross-functionally learning fast and driving issues to closure
Preferred:
Previous work in AI-ML with 100+ CPU core & communication fabric-based design.
Familiarity with in-line testing and diagnostics using CPU memory and execution with self-checking.
Knowledge of chip defect profiles and mitigation strategies across the hardware and software stack
Familiarity in creating test and s/w infrastructure at large scale
Working across global time zones
Location:
Sunnyvale, California.
Bangalore, India
Toronto, Canada
The base salary range for this position is $175,000 to $275,000 annually. Actual compensation may include bonus and equity, and will be determined based on factors such as experience, skills, and qualifications.
Why Join Cerebras
People who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:
Build a breakthrough AI platform beyond the constraints of the GPU.
Publish and open source their cutting-edge AI research.
Work on one of the fastest AI supercomputers in the world.
Enjoy job stability with startup vitality.
Our simple, non-corporate work culture that respects individual beliefs.
Find out more about what it's like to work at Cerebras here!
Apply today and become part of the forefront of groundbreaking advancements in AI!
Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.
This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.
Cerebras Systems is the world's fastest AI inference. We are powering the future of generative AI. We’re a team of pioneering computer architects, deep learning researchers, and engineers building a new class of AI supercomputers from the ground up.
Our flagship system, Cerebras CS-3, is powered by the Wafer Scale Engine 3—the world’s largest and fastest AI processor. CS-3s are effortlessly clustered to create the largest AI supercomputers on Earth, while abstracting away the complexity of traditional distributed computing.
From sub-second inference speeds to breakthrough training performance, Cerebras makes it easier to build and deploy state-of-the-art AI—from proprietary enterprise models to open-source projects downloaded millions of times.
Here’s what makes our platform different:
🔦 Sub-second reasoning – Instant intelligence and real-time responsiveness, even at massive scale
⚡ Blazing-fast inference – Up to 100x performance gains over traditional AI infrastructure
🧠 Agentic AI in action – Models that can plan, act, and adapt autonomously
🌍 Scalable infrastructure – Built to move from prototype to global deployment without friction
Cerebras solutions are available in the Cerebras Cloud or on-prem, serving leading enterprises, research labs, and government agencies worldwide.
👉 Learn more: www.cerebras.ai
Join us: https://cerebras.net/careers/
Offices: 1237 E Arques Ave, Sunnyvale, California 94085, US · 150 King St W, Toronto, Ontario M5H 1J9, CA · Tokyo, JP · Bangalore, IN
artificial intelligencedeep learningnatural language processinginferencemachine learningllmAIenterprise AIand fast inferenceSemiconductor
Cerebras Systems is the world's fastest AI inference. We are powering the future of generative AI. We’re a team of pioneering computer architects, deep learning researchers, and engineers building a new class of AI supercomputers from the ground up.
Our flagship system, Cerebras CS-3, is powered by the Wafer Scale Engine 3—the world’s largest and fastest AI processor. CS-3s are effortlessly clustered to create the largest AI supercomputers on Earth, while abstracting away the complexity of traditional distributed computing.
From sub-second inference speeds to breakthrough training performance, Cerebras makes it easier to build and deploy state-of-the-art AI—from proprietary enterprise models to open-source projects downloaded millions of times.
Here’s what makes our platform different:
🔦 Sub-second reasoning – Instant intelligence and real-time responsiveness, even at massive scale
⚡ Blazing-fast inference – Up to 100x performance gains over traditional AI infrastructure
🧠 Agentic AI in action – Models that can plan, act, and adapt autonomously
🌍 Scalable infrastructure – Built to move from prototype to global deployment without friction
Cerebras solutions are available in the Cerebras Cloud or on-prem, serving leading enterprises, research labs, and government agencies worldwide.
👉 Learn more: www.cerebras.ai
Join us: https://cerebras.net/careers/
Offices: 1237 E Arques Ave, Sunnyvale, California 94085, US · 150 King St W, Toronto, Ontario M5H 1J9, CA · Tokyo, JP · Bangalore, IN
artificial intelligencedeep learningnatural language processinginferencemachine learningllmAIenterprise AIand fast inferenceSemiconductor