About the Role RadixArk is seeking a Member of Technical Staff — Inference to push the limits of large-scale AI inference. You will work on the core systems that serve frontier models at scale, optimizing performance, la…
Skills: Systems Engineering, ML Infrastructure, GPU Architecture, Distributed Systems, Python
Mountain View, California, United States · On-site
$177k–$257k/yr
Senior+$26M raised
Minimum qualifications: Bachelor's degree in Computer Science, Engineering, a related field, or equivalent practical experience. 10 years of experience with data center networking architecture, operations, and power dist…
Skills: Data center networking, Robotics, Technical program management, Network architecture, Power distribution systems
Mountain View, California, United States · On-site
$174k–$252k/yr
Senior$26M raised
Minimum qualifications: Bachelor’s degree or equivalent practical experience. 5 years of experience with software development in one or more programming languages. 3 years of experience testing, maintaining, or launching…
Skills: Software Development, Distributed Systems, Infrastructure Development, System Design, Data Structures
AI Research Scientist- Multimodal Foundational Models
Sunnyvale, California, United States · On-site
$165k–$195k/yr
Mid level
Company Description The Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania, and Cambridge, Massachusetts is a part of the global Bosch Group (www.bosch.com)…
Skills: Multimodal Deep Learning, Time Series Foundation Models, Signal Processing, Python, PyTorch
Senior AI Research Scientist- Time-Series Foundational Models
Sunnyvale, California, United States · On-site
$165k–$195k/yr
Mid level
Company Description The Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania, and Cambridge, Massachusetts is a part of the global Bosch Group (www.bosch.com)…
Skills: Time Series Foundation Models, Deep Learning, Signal Processing, Python, PyTorch
Senior AI Research Scientist- Time-Series Foundational Models
Sunnyvale, California, United States · On-site
$165k–$195k/yr
Senior
Company Description The Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania, and Cambridge, Massachusetts is a part of the global Bosch Group (www.bosch.com)…
Skills: Time-Series Foundation Models, Deep Learning, Signal Processing, Python, PyTorch
About the Role RadixArk is looking for a Member of Technical Staff — Backend/API Platform Engineer to build the API layer, control plane, and platform services that power SGLang and Miles in production. You'll design and…
Hardware and Environments Test Engineer (III-Senior)
Long Beach, California, United States · On-site
$125k–$175k/yr
Senior+$1.1B raised
Space is a warfighting domain. True Anomaly seeks those with the talent and ambition to build the technology that secures it. OUR MISSION True Anomaly delivers decisive capabilities for space superiority. We build autono…
Skills: Hardware Testing, Test Campaign Leadership, Data Analysis, Systems Engineering, Failure Analysis
Description AdvancedPCB is seeking a reliable, motivated Laser Drill Operator to join our growing company. The Laser Drill Operator uses state of the art Computer Numerical Control (CNC) equipment to drill high-precision…
Description The Machine Operator will set up and operate a variety of machine tools to produce precision parts and instruments. This position may also fabricate and modify parts to make or repair machine tools or maintai…
We're Hiring: Mold Setter Location: Torrance, CA Company: Pelican Products Who We Are At Pelican, we engineer products that stand up to the world’s toughest conditions—because the people who rely on us do too…
Kindeva Drug Delivery is seeking experienced Senior Electro-Mechanical Technicians to join our growing organization. This is a critical role within our Maintenance Department with responsibilities involving troubleshooti…
Skills: Troubleshooting, Machine repair, Preventive maintenance, Corrective maintenance, Electrical systems
Thousand Oaks, California, United States · On-site
$180k–$230k/yr
Senior+
Senior Superintendent Thousand Oaks, CA At Dome Construction, our Senior Superintendents are more than field leaders. They are trusted partners, mentors, and strategic drivers of project success. We're seeking an experie…
Skills: Commercial construction, Field leadership, Project management, Site safety, Quality control
Company Description It started with a simple idea: what if surgery could be less invasive and recovery less painful? Nearly 30 years later, that question still fuels everything we do at Intuitive. As a global leader in r…
Description WHO WE ARE ZBeta is a world-class physical security design, consulting, and managed services firm that partners with some of the most dynamic, high-profile organizations and individuals in the world. We lever…
Description WHO WE ARE ZBeta is a world-class physical security design, consulting, and managed services firm that partners with some of the most dynamic, high-profile organizations and individuals in the world. We lever…
Teamwork makes the stream work. Roku is changing how the world watches TV Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pio…
Teamwork makes the stream work. Roku is changing how the world watches TV Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pio…
Skills: AdTech, MarTech, SQL, Hive, Data Visualization
Teamwork makes the stream work. Roku is changing how the world watches TV Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pio…
Description Our results-focused Technology Operations (TechOps) team delivers secure and reliable technology services across all company properties. These properties include a mix of commercial and residential locations,…
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
Full-time
Equity, Comprehensive Benefits, Flexible Work Arrangements
Posted 5d ago
~40 hrs/week
Responsibilities
Design and build large-scale inference systems for frontier AI models while optimizing latency, throughput, and GPU utilization. Collaborate with cross-functional teams to improve model serving architectures and drive the scalability of inference infrastructure.
Requirements
Requires 5+ years of experience in systems engineering or ML infrastructure with deep expertise in GPU architecture and large-scale LLM inference. Proficiency in production languages like Python, Rust, C++, or Go is essential.
Full job description
About the Role
RadixArk is seeking aMember of Technical Staff — Inferenceto push the limits of large-scale AI inference.
You will work on the core systems that serve frontier models at scale, optimizing performance, latency, throughput, and cost across thousands of GPUs. This role sits at the intersection of systems engineering, ML infrastructure, and performance optimization.
Your work will directly shape how state-of-the-art models are deployed and experienced by users worldwide.
This is a deeply technical, high-impact role for engineers who enjoy working close to the hardware–software boundary and solving performance-critical problems at scale.
Requirements
5+ years of experience in systems engineering, ML infrastructure, or performance-critical backend systems
Strong expertise in large-scale inference systems for LLMs or generative models
Deep understanding of GPU architecture and performance characteristics
Experience optimizing latency- and throughput-critical production systems
Strong knowledge of distributed systems and networking fundamentals
Proficiency in Python, Rust, C++, or Go for production systems
Experience profiling and optimizing compute-intensive workloads
Strong debugging skills across system layers (model, runtime, kernel, network)
Strong Plus
Experience with LLM serving stacks (SGLang, vLLM, TensorRT-LLM, etc.)
Open-source contributions in ML or systems infrastructure
Familiarity with CUDA, Triton, or custom kernel optimization
Experience with batching, KV-cache management, and scheduling strategies
Experience running inference at scale (1000+ GPUs)
Background in HPC or high-performance systems
Responsibilities
Design and build large-scale inference systems for frontier AI models
Optimize latency, throughput, and GPU utilization in production inference
Develop and improve model serving architectures and runtimes
Work on batching, scheduling, and memory management strategies
Collaborate with kernel, compiler, and systems teams on performance optimization
Debug performance bottlenecks across the stack
Drive reliability and scalability of inference infrastructure
Build tooling for observability, profiling, and performance analysis
Contribute to long-term inference architecture and strategy
About RadixArk
RadixArk is an infrastructure-first company built by engineers who've shipped production AI systems, created SGLang (30K+ GitHub stars, the fastest open LLM serving engine), and developed Miles (our large-scale RL framework). Founded by AI infrastructure veterans from xAI and NVIDIA, we're on a mission to democratize frontier-level AI infrastructure by building world-class open systems for inference and training. Our team has optimized kernels serving billions of tokens daily, designed distributed training systems coordinating 10,000+ GPUs, and contributed to infrastructure that powers leading AI companies and research labs.
Compensation
We offer competitive compensation with equity, comprehensive health benefits, and flexible work arrangements. Compensation is determined by location, level, and experience.
Equal Opportunity
RadixArk is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.
Based on 26292 listings with disclosed salaries, most engineering jobs in California pay between $100k–$242k per year. Individual offers vary with seniority, company size, and specialization.
How many Engineering jobs are open in California right now?
There are currently 37,286 open engineering positions in California listed on Clera. New openings are added daily as companies post roles.
Which companies are hiring for Engineering roles in California?
Companies currently hiring include Northrop Grumman, Northrop Grumman Mission Systems, Inc., Google, Apple, Applied Materials, among others. Browse the listings above to see every active employer.
Are there remote or hybrid Engineering jobs in California?
Yes — 8920 of the 37286 open engineering positions offer remote or hybrid work (1826 remote, 7094 hybrid).
How do I apply for Engineering jobs in California?
Each listing links directly to the employer's application page. Apply early — fresh listings get the most recruiter attention in the first two weeks.