About the Role RadixArk is seeking a Member of Technical Staff — Inference to push the limits of large-scale AI inference. You will work on the core systems that serve frontier models at scale, optimizing performance, la…
Skills: Systems Engineering, ML Infrastructure, GPU Architecture, Distributed Systems, Python
About the Role RadixArk is seeking a Member of Technical Staff, Developer Technology (DevTech) to make LLM inference and training dramatically faster, cheaper, and more accessible on modern GPU hardware. Our systems sit …
Mountain View, California, United States · On-site
$177k–$257k/yr
Senior+$26M raised
Minimum qualifications: Bachelor's degree in Computer Science, Engineering, a related field, or equivalent practical experience. 10 years of experience with data center networking architecture, operations, and power dist…
Skills: Data center networking, Robotics, Technical program management, Network architecture, Power distribution systems
AI Research Scientist- Multimodal Foundational Models
Sunnyvale, California, United States · On-site
$165k–$195k/yr
Mid level
Company Description The Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania, and Cambridge, Massachusetts is a part of the global Bosch Group (www.bosch.com)…
Skills: Multimodal Deep Learning, Time Series Foundation Models, Signal Processing, Python, PyTorch
Senior AI Research Scientist- Time-Series Foundational Models
Sunnyvale, California, United States · On-site
$165k–$195k/yr
Senior
Company Description The Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania, and Cambridge, Massachusetts is a part of the global Bosch Group (www.bosch.com)…
Skills: Time-Series Foundation Models, Deep Learning, Signal Processing, Python, PyTorch
Senior AI Research Scientist- Time-Series Foundational Models
Sunnyvale, California, United States · On-site
$165k–$195k/yr
Mid level
Company Description The Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania, and Cambridge, Massachusetts is a part of the global Bosch Group (www.bosch.com)…
Skills: Time Series Foundation Models, Deep Learning, Signal Processing, Python, PyTorch
About the Role RadixArk is looking for a Member of Technical Staff — Backend/API Platform Engineer to build the API layer, control plane, and platform services that power SGLang and Miles in production. You'll design and…
Hardware and Environments Test Engineer (III-Senior)
Long Beach, California, United States · On-site
$125k–$175k/yr
Senior+$1.1B raised
Space is a warfighting domain. True Anomaly seeks those with the talent and ambition to build the technology that secures it. OUR MISSION True Anomaly delivers decisive capabilities for space superiority. We build autono…
Skills: Hardware Testing, Test Campaign Leadership, Data Analysis, Systems Engineering, Failure Analysis
San Francisco, California, United States · On-site
Senior$66M raised
About the Role We're hiring a Software Engineer, Robotics Platform to help us scale our fleet of robots. This role is on-site five days a week in San Francisco — it's how our small, high-ownership team moves. The robot p…
San Bernardino, California, United States · On-site
$160k–$180k/yr
Senior+
Description JOB SUMMARY The Sr. Director of Electronic Data Interchange (EDI) provides executive leadership, strategic direction, and operational oversight for all organizational EDI systems, data workflows, trading part…
Skills: EDI Standards, X12, HIPAA Transactions, Data Mapping, Integration Architecture
Company Description It started with a simple idea: what if surgery could be less invasive and recovery less painful? Nearly 30 years later, that question still fuels everything we do at Intuitive. As a global leader in r…
Description We are seeking a skilled Computer Assembler to join our team. The successful candidate will be responsible for assembling computer components and ensuring that they function properly. Using assorted hand and …
Skills: Computer assembly, Hand tools, Power tools, Soldering, Blueprint reading
San Luis Obispo, California, United States · On-site
Mid level
Description Help Desk: Provide end user support ( on-site and remote). Record time and work log entries on ticketsActive Directory: Create user accounts for new hires. Disable accounts for departing employees.CCTV: Troub…
Skills: Active Directory, Windows Server, PowerShell, Group Policy, Network Troubleshooting
About Marqvision Protect and build a future shaped by original ideas, innovations and creativity. Threats to brand integrity are scaling fast. We're building the intelligence to scale faster. MarqVision is the brand inte…
Teamwork makes the stream work. Roku is changing how the world watches TV Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pio…
Teamwork makes the stream work. Roku is changing how the world watches TV Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pio…
Skills: AdTech, MarTech, SQL, Hive, Data Visualization
Teamwork makes the stream work. Roku is changing how the world watches TV Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pio…
Description Our results-focused Technology Operations (TechOps) team delivers secure and reliable technology services across all company properties. These properties include a mix of commercial and residential locations,…
This position is expected to be onsite in either Reno, NV or Truckee, CA. Bargaining Unit: Non Represented - Professional Rate of Pay: $111,217/annually + DOE Summary The Senior BI Analyst will play a critical role in an…
Skills: SQL, Data Analysis, Data Visualization, Tableau, Power BI
Mountain View, California, United States · On-site
$154k–$171k/yr
Senior$814M raised
Company Overview ID.me is the next-generation digital identity wallet that simplifies how individuals securely prove their identity online. Consumers can verify their identity with ID.me once and seamlessly login across …
Skills: LLM Integration, AI Agents, Python, Go, JavaScript
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
Full-time
Equity, Comprehensive Benefits, Flexible Work Arrangements
Posted 9d ago
~40 hrs/week
Responsibilities
Design and build large-scale inference systems for frontier AI models while optimizing latency, throughput, and GPU utilization. Collaborate with cross-functional teams to improve model serving architectures and drive the scalability of inference infrastructure.
Requirements
Requires 5+ years of experience in systems engineering or ML infrastructure with deep expertise in GPU architecture and large-scale LLM inference. Proficiency in production languages like Python, Rust, C++, or Go is essential.
Full job description
About the Role
RadixArk is seeking aMember of Technical Staff — Inferenceto push the limits of large-scale AI inference.
You will work on the core systems that serve frontier models at scale, optimizing performance, latency, throughput, and cost across thousands of GPUs. This role sits at the intersection of systems engineering, ML infrastructure, and performance optimization.
Your work will directly shape how state-of-the-art models are deployed and experienced by users worldwide.
This is a deeply technical, high-impact role for engineers who enjoy working close to the hardware–software boundary and solving performance-critical problems at scale.
Requirements
5+ years of experience in systems engineering, ML infrastructure, or performance-critical backend systems
Strong expertise in large-scale inference systems for LLMs or generative models
Deep understanding of GPU architecture and performance characteristics
Experience optimizing latency- and throughput-critical production systems
Strong knowledge of distributed systems and networking fundamentals
Proficiency in Python, Rust, C++, or Go for production systems
Experience profiling and optimizing compute-intensive workloads
Strong debugging skills across system layers (model, runtime, kernel, network)
Strong Plus
Experience with LLM serving stacks (SGLang, vLLM, TensorRT-LLM, etc.)
Open-source contributions in ML or systems infrastructure
Familiarity with CUDA, Triton, or custom kernel optimization
Experience with batching, KV-cache management, and scheduling strategies
Experience running inference at scale (1000+ GPUs)
Background in HPC or high-performance systems
Responsibilities
Design and build large-scale inference systems for frontier AI models
Optimize latency, throughput, and GPU utilization in production inference
Develop and improve model serving architectures and runtimes
Work on batching, scheduling, and memory management strategies
Collaborate with kernel, compiler, and systems teams on performance optimization
Debug performance bottlenecks across the stack
Drive reliability and scalability of inference infrastructure
Build tooling for observability, profiling, and performance analysis
Contribute to long-term inference architecture and strategy
About RadixArk
RadixArk is an infrastructure-first company built by engineers who've shipped production AI systems, created SGLang (30K+ GitHub stars, the fastest open LLM serving engine), and developed Miles (our large-scale RL framework). Founded by AI infrastructure veterans from xAI and NVIDIA, we're on a mission to democratize frontier-level AI infrastructure by building world-class open systems for inference and training. Our team has optimized kernels serving billions of tokens daily, designed distributed training systems coordinating 10,000+ GPUs, and contributed to infrastructure that powers leading AI companies and research labs.
Compensation
We offer competitive compensation with equity, comprehensive health benefits, and flexible work arrangements. Compensation is determined by location, level, and experience.
Equal Opportunity
RadixArk is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.
Based on 25520 listings with disclosed salaries, most technology jobs in California pay between $114k–$260k per year. Individual offers vary with seniority, company size, and specialization.
How many Technology jobs are open in California right now?
There are currently 32,707 open technology positions in California listed on Clera. New openings are added daily as companies post roles.
Which companies are hiring for Technology roles in California?
Companies currently hiring include Google, Apple, Amazon, Applied Materials, Northrop Grumman, among others. Browse the listings above to see every active employer.
Are there remote or hybrid Technology jobs in California?
Yes — 12027 of the 32707 open technology positions offer remote or hybrid work (3138 remote, 8889 hybrid).
How do I apply for Technology jobs in California?
Each listing links directly to the employer's application page. Apply early — fresh listings get the most recruiter attention in the first two weeks.