Deployment Manager – Global Data Center Build and Deploy
Sunnyvale, California, United States · Hybrid
Senior+$4.7B raised
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale …
Skills: AI Cluster Deployment, Rack Integration, High-Density Cabling, Data Center Infrastructure, Troubleshooting
The Region Flexibility Engineering (RFE) team builds and leverages foundational infrastructure capabilities, tools, and datasets needed to support the rapid global expansion of Amazon's SOA infrastructure. Our team focus…
Principal Conn. Systems Engr., Connectivity Systems
Sunnyvale, California, United States · On-site
$218k–$295k/yr
Senior+$35B raised
Amazon Lab126 is an inventive research and development company that designs and engineers high-profile consumer electronics. Lab126 began in 2004 as a subsidiary of Amazon.com, Inc., originally creating the best-selling …
Skills: Wireless Architecture, Communication Theory, OFDM, MIMO, RF Engineering
Meta's Silicon Engineering team is building custom silicon solutions that power the infrastructure behind Meta's AI, data center, and next-generation computing platforms. As a Silicon Engineer focused on EDA Infrastructu…
Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon Valley company is creating the digital infrastructure needed to bring intelligence to every moving…
Skills: Technical Program Management, Sprint Planning, Jira, Confluence, Stakeholder Management
Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon Valley company is creating the digital infrastructure needed to bring intelligence to every moving…
Skills: Technical Program Management, Autonomy, Robotics, Machine Learning, Performance Measurement
Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon Valley company is creating the digital infrastructure needed to bring intelligence to every moving…
Skills: Python, C++, PyTorch, Deep Learning, Computer Vision
The Role Wayve is looking for a Principal Engineer, 3D Reconstruction to lead the development of an offline 3D reconstruction capability that builds high-quality 3D world geometry from vehicle sensor data in dynamic scen…
Skills: SLAM, 3D Reconstruction, State Estimation, Sensor Fusion, Lidar-based Reconstruction
About the role We are seeking an experienced Front-end Digital Design Engineer to join our Sunnyvale based team. The ideal candidate will have a strong background in Embedded and Mixedmode SoC architectures, Verilog, and…
Skills: Digital Design, Verilog, System Verilog, Logic Synthesis, SoC Architecture
About the role We are seeking a hands-on Validation Engineer to take our custom silicon devices from first power-on to production readiness. Embedded in a cross-functional team spanning analog design, logic design and pa…
The Role Wayve is looking for a Principal Engineer, 3D Reconstruction to lead the development of an offline 3D reconstruction capability that builds high-quality 3D world geometry from vehicle sensor data in dynamic scen…
Skills: SLAM, 3D Reconstruction, State Estimation, Sensor Fusion, Lidar-based Reconstruction
AI Research Scientist: Multimodal Foundation Models - Architecture, Pre-Training & Distillation
Sunnyvale, California, United States · On-site
Senior
The Multimodal Intelligence Team is building the next generation of foundation models for Apple experiences. We are looking for a research scientist to advance the architectures, pre-training methods, and distillation te…
Skills: Multimodal Foundation Models, Pre-training, Knowledge Distillation, PyTorch, JAX
About us Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing …
Skills: Program Delivery, Technical Program Management, Systems Thinking, Risk Management, Stakeholder Management
About us Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing …
Skills: Program Delivery, Technical Program Management, Systems Thinking, Risk Management, Stakeholder Management
Description Due to the classified nature of our work, U.S. citizenship is required. Candidate must meet the eligibility to obtain and maintain a Security Clearance. Location: Sunnyvale, CA About Us Pacific Defense is a l…
Skills: Team Leadership, Software Engineering, Project Management, System Integration, DoD Acquisition Lifecycle
Onwards Together! Illumio is the leader in ransomware and breach containment, redefining how organizations contain cyberattacks and enable operational resilience. Powered by the Illumio AI Security Graph, our breach cont…
Skills: Workday Analytics, People Analytics, HRIS Management, Data Governance, Project Management
Onwards Together! Illumio is the leader in ransomware and breach containment, redefining how organizations contain cyberattacks and enable operational resilience. Powered by the Illumio AI Security Graph, our breach cont…
Position Summary... What you'll do...Role summary: Join Walmart as a Staff Software Engineer to lead the design, development, and deployment of scalable platform capabilities and AI-driven applications. This role involve…
Skills: iOS Development, Android Development, Generative AI, CI/CD, Test Automation
Onwards Together! Illumio is the leader in ransomware and breach containment, redefining how organizations contain cyberattacks and enable operational resilience. Powered by the Illumio AI Security Graph, our breach cont…
Systems Development Engineer, Edge AI Platform Infrastructure, Hardware Compute Group
Sunnyvale, California, United States · On-site
$149k–$201k/yr
Senior$35B raised
Join the Edge AI team within the Silicon and Systems Group (SSG) at Amazon. As a Systems Development Engineer, you will own the test and release infrastructure that validates custom AI accelerator silicon IP, kernel driv…
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
Full-time
bachelor degree
Posted 55d ago
~40 hrs/week
Responsibilities
Lead the end-to-end execution of AI cluster deployments, focusing on rack integration and high-density cabling from facility readiness to customer hand-off. Coordinate with facilities and networking teams to ensure safe, high-quality installation of Wafer Scale Engine clusters.
Requirements
Requires a Bachelor's degree in Engineering or IT and over 10 years of experience in hyperscale AI, ML, or HPC data center deployments. Must have deep expertise in structured cabling and the ability to manage technicians on the data center floor.
Full job description
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.
Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.
The Role
The Deployment Manager – Global Data Center Build and Deploy is responsible for the end-to-end execution of AI cluster deployments within large-scale data centers, specifically taking from facility ready for power and cooling to customer hand-off. This role leads rack integration, high-density inter-rack cabling, and close coordination with facilities, networking, and operations teams to deliver Cerebras Wafer Scale Engine based clusters safely, on schedule, and to strict quality standards.
The ideal candidate has deep experience deploying AI infrastructure, understands the operational demands of high-power, high-thermal-density environments, and excels at troubleshooting complex cabling and integration issues. This role requires providing directions to technicians on the Data Center floor, and resolve any issue blocking build and deploy in data centers. The role requires traveling and physically present at Data Center weeks at a time.
Responsibilities
AI Cluster Deployment & Rack Integration
Lead deployment of AI clusters including Cerebras Wafer Scale Engine, high-speed switches, storage, and rack-level infrastructure
Coordinate rack integration of high-density compute, specialized racks, and associated power components (PDUs, busway drops)
Ensure deployment aligns with AI cluster architecture, topology, and scaling requirements
High-Density Inter-Rack & Intra-Rack Cabling
Manage installation of high-speed interconnect cabling (fiber and copper) supporting AI fabrics (east–west traffic) in Data Centers
Coordinate inter-rack and intra-rack cabling for AI clusters, including spine-leaf and pod-level designs
Ensure proper routing, airflow clearance, labeling, and testing of all AI-related cabling
Facilities & Infrastructure Coordination
Work closely with facilities teams on power capacity, cooling readiness, containment, and grounding for dense racks
Coordinate deployment sequencing with facility commissioning milestones
Validate white-space readiness before rack and cluster deployment
Provide directions to technicians on the data center floor.
Troubleshooting & Issue Resolution
Troubleshoot cabling, connectivity, and integration issues impacting AI cluster bring-up
Lead root-cause analysis for deployment blockers related to cabling, hardware placement, or facilities dependencies
Support validation, burn-in, and handoff of AI clusters to operations teams
Cross-Functional Execution
Partner with network, server, AI platform, and operations teams to align on deployment plans and readiness
Manage multiple parallel AI cluster deployments across sites or availability zones
Communicate risks, dependencies, and milestones clearly to stakeholders
Quality, Standards & Documentation
Ensure deployments follow company design standards, structured cabling best practices, and AI deployment playbooks
Validate as-built documentation, labeling accuracy, and deployment checklists
Maintain accurate records for cluster configuration, cabling, and deployment status
Safety & Risk Management
Enforce EHS, data center safety, and access control procedures during deployment
Ensure safe handling of heavy, high-power GPU equipment
Proactively identify and mitigate deployment and operational risks
Qualifications
Bachelor’s degree in Engineering, IT, or equivalent practical experience
10+ years of experience in data center deployments, infrastructure delivery, or integration roles
Ability to provide directions to technicians on data center floor.
Hands-on experience deploying hyperscale AI, ML, or HPC infrastructure
Strong experience with structured cabling in high-density environments, and troubleshooting
Proven ability to manage complex, cross-functional deployment programs
Familiarity with high-speed fabrics (e.g., InfiniBand, high-bandwidth Ethernet)
Experience with DCIM or deployment tracking systems
Strong attention to detail and operational rigor
Ability to perform under tight timelines and production constraints
Why Join Cerebras
People who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:
Build a breakthrough AI platform beyond the constraints of the GPU.
Publish and open source their cutting-edge AI research.
Work on one of the fastest AI supercomputers in the world.
Enjoy job stability with startup vitality.
Our simple, non-corporate work culture that respects individual beliefs.
Find out more about what it's like to work at Cerebras here!
Apply today and become part of the forefront of groundbreaking advancements in AI!
Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.
This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.
Related keywords
AI InfrastructureWafer Scale EngineGPUHPCInfiniBandEthernetData CenterRack IntegrationStructured CablingDCIMPower Distribution UnitsBusway DropsSpine-Leaf ArchitecturePod-Level DesignThermal DensityEHS
Cerebras Systems is the world's fastest AI inference. We are powering the future of generative AI. We’re a team of pioneering computer architects, deep learning researchers, and engineers building a new class of AI supercomputers from the ground up.
Our flagship system, Cerebras CS-3, is powered by the Wafer Scale Engine 3—the world’s largest and fastest AI processor. CS-3s are effortlessly clustered to create the largest AI supercomputers on Earth, while abstracting away the complexity of traditional distributed computing.
From sub-second inference speeds to breakthrough training performance, Cerebras makes it easier to build and deploy state-of-the-art AI—from proprietary enterprise models to open-source projects downloaded millions of times.
Here’s what makes our platform different:
🔦 Sub-second reasoning – Instant intelligence and real-time responsiveness, even at massive scale
⚡ Blazing-fast inference – Up to 100x performance gains over traditional AI infrastructure
🧠 Agentic AI in action – Models that can plan, act, and adapt autonomously
🌍 Scalable infrastructure – Built to move from prototype to global deployment without friction
Cerebras solutions are available in the Cerebras Cloud or on-prem, serving leading enterprises, research labs, and government agencies worldwide.
👉 Learn more: www.cerebras.ai
Join us: https://cerebras.net/careers/
Offices: 1237 E Arques Ave, Sunnyvale, California 94085, US · 150 King St W, Toronto, Ontario M5H 1J9, CA · Tokyo, JP · Bangalore, IN
artificial intelligencedeep learningnatural language processinginferencemachine learningllmAIenterprise AIand fast inferenceSemiconductor
Based on 1493 listings with disclosed salaries, most software jobs in Sunnyvale, CA pay between $140k–$286k per year. Individual offers vary with seniority, company size, and specialization.
How many Software jobs are open in Sunnyvale, CA right now?
There are currently 1,773 open software positions in Sunnyvale, CA listed on Clera. New openings are added daily as companies post roles.
Which companies are hiring for Software roles in Sunnyvale, CA?
Companies currently hiring include Google, Walmart, Apple, Wayve, Applied Intuition, among others. Browse the listings above to see every active employer.
Are there remote or hybrid Software jobs in Sunnyvale, CA?
Yes — 567 of the 1773 open software positions offer remote or hybrid work (42 remote, 525 hybrid).
How do I apply for Software jobs in Sunnyvale, CA?
Each listing links directly to the employer's application page. Apply early — fresh listings get the most recruiter attention in the first two weeks.