Principle Systems Reliability Engineer Location: San Jose , CA (On-site) Ayar Labs is shattering AI data bottlenecks by moving data at the speed of light. As pioneers of co-packaged optics (CPO), we are using light inste…
Skills: System Reliability Engineering, Fleet Qualification, Test Infrastructure Design, Statistical Reliability Planning, Hardware Validation
The group you’ll be a part of The Global Workplace Solutions (GWS) team brings facilities and lab operations management, space and occupancy planning, real estate, lease and portfolio management, capital projects, busine…
Skills: IBM Tririga, Project Management, Program Management, Systems Operations, Data Analysis
The group you’ll be a part of The Global Workplace Solutions (GWS) team brings facilities and lab operations management, space and occupancy planning, real estate, lease and portfolio management, capital projects, busine…
Skills: IBM Tririga, Project Management, Program Management, Systems Administration, Data Analysis
San Francisco, California, United States · On-site
$199k–$292k/yr
Senior$5.2B raised
DoorDash Labs is an independent team within DoorDash. We explore robotics and automation to transform last mile logistics in the long term. If you have a passion for applying robotics solutions in a service used by milli…
Skills: C++, Embedded Systems, Control Theory, Motion Planning, Dynamics Modeling
Turning Space into a Transportation Layer for Earth Who We Are: Inversion builds advanced reentry systems to deliver next-generation capabilities from space. Our mission is to make Earth radically more accessible by turn…
Skills: Embedded C++, Software Architecture, Embedded Linux, Rust, C
Company Overview Docusign brings agreements to life. Over 1.5 million customers and more than a billion people in over 180 countries use Docusign solutions to accelerate the process of doing business and simplify people&…
About Rivian Rivian is on a mission to keep the world adventurous forever. This goes for the emissions-free Electric Adventure Vehicles we build, and the curious, courageous souls we seek to attract. As a company, we con…
Skills: Systems integration, Test execution, Root-cause analysis, CAN, LIN
Company Information For more than 20 years, AEG has played a pivotal role in transforming sports and live entertainment. Annually, we host more than 160 million guests, promote more than 10,000 shows and present more tha…
Skills: Office 365, Active Directory, PowerShell, Python, JAMF
Company Description It all started when engineer Fred Luddy wrote code that automated a tedious task for his coworker, Phyllis. She cried tears of joy. That moment inspired Fred to build a company that could do that for …
Los Angeles, California, United States · Remote OK
$119k–$172k/yr
Mid level$64M raised
Grow your career at Cedars-Sinai! Our EIS (Enterprise Information Services) applications department achieved Epic Honor Roll - Magna Cum Laude. EIS was also named one of the "Best Places to Work in IT 2025" by Computerwo…
Skills: Epic Optime, Anesthesia, Project management, Workflow re-engineering, System configuration
We’re in an unbelievably exciting area of tech and are fundamentally reshaping the data storage industry. Here, you lead with innovative thinking, grow along with us, and join the smartest team in the industry. This type…
At Veracyte, we offer exciting career opportunities for those interested in joining a pioneering team that is committed to transforming cancer care for patients across the globe. Working at Veracyte enables our employees…
South San Francisco, California, United States · Hybrid
$150k–$189k/yr
Senior$1.1B raised
At Veracyte, we offer exciting career opportunities for those interested in joining a pioneering team that is committed to transforming cancer care for patients across the globe. Working at Veracyte enables our employees…
MyOme’s mission is to provide clinically actionable genetic information to patients throughout their lives. We combine clinical-grade whole genome sequencing, advanced AI methods for genome interpretation, and seamless d…
Skills: Computational biology, Bioinformatics, Biostatistics, Epidemiology, Human genetics
San Francisco, California, United States · On-site
$151k–$189k/yr
Senior$16B raised
Scale's mission is to develop reliable AI systems for the world's most important decisions. We provide the high-quality data that powers the world's AI models, and we help enterprises and governments build, deploy, and o…
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblo…
Skills: Machine learning infrastructure, Engineering management, System architecture, Model training, Data pipelines
San Francisco, California, United States · On-site
$88k–$132k/yr
Senior$1.3B raised
At Klaviyo, we value the unique backgrounds, experiences and perspectives each Klaviyo (we call ourselves Klaviyos) brings to our workplace each and every day. We believe everyone deserves a fair shot at success and appr…
Skills: Customer advocacy, Community management, Customer marketing, Program management, Strategic planning
Teamwork makes the stream work. Roku is changing how the world watches TV Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pio…
Skills: Product management, Recommendations, Personalization, Search and discovery, Machine learning
Teamwork makes the stream work. Roku is changing how the world watches TV Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pio…
San Francisco, California, United States · Remote Solely
$122k–$155k/yr
Senior$2.7B raised
Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a growth mindset, natura…
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
$185k–$290k/yr
Full-time
bachelor degree, postgraduate degree
Posted 8d ago
~40 hrs/week
Responsibilities
The role involves designing and operating custom test infrastructure to validate system-level reliability for optical I/O technology. You will also serve as the primary technical contact for integration partners to manage qualification programs and fleet-level reliability reporting.
Requirements
Candidates must have at least 5 years of experience in systems reliability engineering for large-scale infrastructure and a degree in Electrical or Computer Engineering. Proficiency in statistical reliability planning, hardware interconnect standards, and direct experience with external customer qualification programs is required.
Full job description
Principle Systems Reliability Engineer
Location: San Jose , CA (On-site)
Ayar Labs is shattering AI data bottlenecks by moving data at the speed of light. As pioneers of co-packaged optics (CPO), we are using light instead of electricity to move data faster, further, and with a fraction of the energy needed to fuel the explosive growth of AI models.
Backed by industry giants like NVIDIA, AMD, Mediatek and Intel and manufactured in partnership with the world’s leading semiconductor ecosystem, Ayar Labs’ co-packaged optics solution is key to unleashing next-generation AI scale-up architectures.
About the Role
Ayar Labs builds optical I/O technology for hyperscale AI infrastructure. This role owns reliability qualification at the system and fleet level: building the test infrastructure and evidence base that proves our products are ready for large-scale deployment, and then partnering directly with customers and integration partners to carry that evidence through joint qualification programs and pilot deployments. You will be responsible for both the engineering (test infrastructure, statistical reliability planning, standards compliance) and the relationship (translating technical evidence into the specific commitments our partners need to move forward).
What You'll Own
Test infrastructure and evidence generation: Design, build, and operate custom test infrastructure that emulates real system-level electrical, thermal, and workload conditions ahead of, or independent of, any single customer's specific hardware. This includes sourcing or generating representative workload and power profiles and using them to drive continuous, long-duration test campaigns.
Reliability statistics and demonstration planning: Define statistical sample sizes and test durations needed to demonstrate specific reliability and confidence targets, and apply the right methodology to each failure population in the system rather than a single blanket target.
Standards and compliance: Ensure that test infrastructure, interfaces, and telemetry are built against relevant industry interconnect and management standards, so evidence generated internally holds up when reviewed by, or transferred to, an external partner's platform.
Staged qualification execution: Execute staged qualification gates, from lab-level interoperability and functional testing through environmental and stress testing to live pilot deployment, applying a reliability / availability / serviceability lens throughout rather than treating any one dimension in isolation.
Fault injection and lifecycle monitoring: Design and run fault-injection and failure-mode testing, including scenarios that exercise field-serviceable components under representative operating conditions, in partnership with customer and integration-partner operations teams.
Economic modeling and fleet reporting: Provide reliability and performance data into cost-of-ownership and total-cost modeling shared with customers, and own ongoing fleet reliability reporting and incident response once products reach production.
Cross-functional partnership: Serve as the primary technical point of contact with Tier-1 integration partners and hyperscale customer engineering teams across the qualification lifecycle, from early technical evidence through pilot sign-off and into steady-state operations.
Basic Qualifications
5+ years in fleet/systems reliability engineering for large-scale infrastructure, with a BS in Electrical Engineering, Computer Engineering, or a related field.
You've built or operated test infrastructure that validates a component or subsystem's behavior before it is deployed into a customer's actual system, using representative rather than production hardware.
You've defined and executed statistical reliability demonstration plans (sample sizes, test durations, confidence and reliability targets) for hardware components, and can explain the methodology behind them, not just apply a lookup table.
You've worked with die-to-die, chip-to-chip, or board-level interconnect standards and telemetry or management interfaces relevant to high-speed data center hardware.
You've partnered directly with external OEM or hyperscale customer engineering teams on a joint qualification or certification program, and are comfortable owning that relationship technically.
Comfort operating a live, continuously running test or fleet environment: on-call coverage, incident response, and building the telemetry pipeline that turns raw sensor data into fleet-level reliability statistics.
Strong written and verbal communication skills; you will regularly translate internal engineering data into evidence and documentation for external partners.
Preferred Qualifications
MS in Electrical Engineering, Computer Engineering, or a related field.
Experience with FPGA-based test or signal-generation systems.
Experience characterizing or emulating real compute or network workload behavior for use in a test or validation environment.
Background in data-center fleet operations, site-reliability engineering, or customer-facing qualification and certification processes.
Familiarity with co-packaged optics, silicon photonics, or other emerging interconnect packaging technologies.
Salary Range: $185,000 - $290,000
NOTE TO RECRUITERS: Principals only. We are not accepting resumes from recruiters for this position. Remuneration for recruiting activities is only applicable subject to a signed and executed agreement between the parties. Please don’t send candidates to Ayar Labs, and do not contact our managers.
Ayar Labs is an Equal Opportunity Employer and is strongly committed to all policies which will afford equal opportunity employment to all qualified persons without regard to age, sex, national origin, race, color, ethnicity, creed, religion, gender identity, sexual orientation, disability, veteran status, or any other characteristic protected by law. It is the policy of Ayar Labs to provide reasonable accommodation when requested by a qualified applicant or employee with a disability, unless such accommodation would cause an undue hardship. Veterans are more than welcome and encouraged to apply.
Transforming AI Architecture with Co-Packaged Optics
Industry
Computer Hardware Manufacturing
Company size
51-200 employees
Founded
2015
Headquarters
San Jose, California
LinkedIn followers
30,945
Total funding
$875M
Ayar Labs is transforming AI infrastructure by accelerating data movement. Recognizing that the complexity and size of AI models are increasing at a rate that traditional interconnect technology cannot handle, the company has developed the industry’s first co-packaged optics (CPO) solution that enables customers to maximize the compute efficiency and performance of growing AI infrastructure, while reducing costs, latency and power consumption. Based on open standards and optimized for both AI training and inference, Ayar Labs’ optical I/O solution is backed by a robust ecosystem that enables it to integrate smoothly into AI systems at scale.
Offices: 695 River Oaks Pkwy, San Jose, California 95134, US · 5909 Christie Ave, Emeryville, California 94608, US · 678 Massachusetts Ave, Suite 401, Cambridge, Massachusetts 02139, US
CPOPhotonicsand AI Scale-UpInformation TechnologySemiconductorOptical CommunicationHardwareArtificial Intelligence (AI)AI InfrastructureComputer
Based on 15411 listings with disclosed salaries, most software jobs in California pay between $125k–$275k per year. Individual offers vary with seniority, company size, and specialization.
How many Software jobs are open in California right now?
There are currently 18,711 open software positions in California listed on Clera. New openings are added daily as companies post roles.
Which companies are hiring for Software roles in California?
Companies currently hiring include Google, Apple, Amazon, NVIDIA, Qualcomm, among others. Browse the listings above to see every active employer.
Are there remote or hybrid Software jobs in California?
Yes — 8570 of the 18711 open software positions offer remote or hybrid work (2413 remote, 6157 hybrid).
How do I apply for Software jobs in California?
Each listing links directly to the employer's application page. Apply early — fresh listings get the most recruiter attention in the first two weeks.