Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon Valley company is creating the digital infrastructure needed to bring intelligence to every moving…
About us Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing …
Skills: C++, Embedded Linux, Automotive sensors, LiDAR, Camera
The Video Computer Vision organization is working on exciting technologies for future Apple products. Our team delivers computer vision and machine learning algorithms that power many Apple technologies like human unders…
Skills: Computer Vision, Machine Learning, C++, Swift, Linear Algebra
The Video Computer Vision organization is working on breakthrough technologies for future Apple products. Our team delivers cutting-edge AI, machine learning, computer vision and graphics algorithms that power technologi…
Do you love understanding every detail of how new technologies work? Join the team that serves as Apple’s nerve center, our Information Systems and Technology group. There are countless ways you’ll contribute here, wheth…
Skills: Engineering Project Management, Generative AI, Data Platforms, Data Analytics, Machine Learning
Platform Software Engineer, AI & Data Platforms (AiDP)
Sunnyvale, California, United States · On-site
Senior$11B raised
The people here at Apple don’t just build products — they build the kind of wonder that’s revolutionized entire industries. It’s the diversity of those people and their ideas that encourages the innovation that runs thro…
Skills: Python, PostgreSQL, Microservices, API Development, Cloud APIs
Staff Software Engineer, Linux Kernel and Baseboard Management Controller Architecture
Sunnyvale, California, United States · On-site
$207k–$300k/yr
Senior$26M raised
Minimum qualifications: Bachelor's degree or equivalent practical experience. 8 years of experience in software development. 5 years of experience in working with embedded operating systems. 5 years of experience in test…
Skills: Linux Kernel, Baseboard Management Controller, Embedded Systems, C++, C
Software Engineering Manager, Network Service Level Objective
Sunnyvale, California, United States · On-site
$207k–$300k/yr
Senior+$26M raised
Minimum qualifications: Bachelor’s degree, or equivalent practical experience. 8 years of experience in software development. 3 years of experience with developing large-scale infrastructure, distributed systems or netwo…
Staff Software Engineer/Tech Lead, ML Network Infrastructure
Sunnyvale, California, United States · On-site
$207k–$300k/yr
Senior+$26M raised
Minimum qualifications: Bachelor's degree or equivalent practical experience. 8 years of experience programming in C++. 5 years of experience testing, and launching software products. 5 years of experience building and d…
Engineering Manager, Identity and Access Management, Capsium Serving
Sunnyvale, California, United States · On-site
$207k–$300k/yr
Senior+$26M raised
Minimum qualifications: Bachelor’s degree, or equivalent practical experience. 8 years of experience in software development. 3 years of experience in a technical leadership role. 3 years of experience with developing la…
Senior Staff Software Engineer, High Performance Networking, Platforms Infrastructure Engineering
Sunnyvale, California, United States · On-site
$262k–$364k/yr
Senior+$26M raised
Minimum qualifications: Bachelor's degree or equivalent practical experience. 8 years of experience programming in C++. 5 years of experience with design and architecture, and testing and launching software products. Exp…
Skills: C++, C, High Performance Computing, Remote Direct Memory Access, Storage Systems
Company Description It started with a simple idea: what if surgery could be less invasive and recovery less painful? Nearly 30 years later, that question still fuels everything we do at Intuitive. As a global leader in r…
Skills: Software Quality Assurance, Risk Management, ISO 13485, FDA 21 CFR Part 820, IEC 62304
We’re seeking a Senior UX Designer to lead the design of AI-native business applications and end-to-end customer experiences. This role goes beyond feature ownership, you will define experience strategy, shape product di…
Job Description The Role: General Motors is a global leader in advanced driver assistance. With Super Cruise hands-free technology in more than 500,000 Super Cruise-equipped vehicles on the road, and over 700 million han…
Senior Conversation Designer, Devices & Services Design Group
San Francisco, California, United States · On-site
$138k–$212k/yr
Senior$67M raised
Want to transform how Alexa+ makes life easier and more enjoyable? The Alexa+ Devices & Services Design Group is seeking a Senior Conversational AI Experience Designer to create and shape conversational interactions that…
As a global leader in cybersecurity, CrowdStrike protects the people, processes and technologies that drive modern organizations. Since 2011, our mission hasn’t changed — we’re here to stop breaches, and we’ve redefined …
Skills: Financial planning and analysis, Financial modeling, Accounting, Budgeting, Forecasting
Job Description The Role At General Motors, we empower Product Managers to solve challenging customer and business problems. We seek passionate and innovative team members who can collaborate effectively within product m…
The role The security of our software is paramount to the safety of our vehicles. The onboard software for our various fleets of vehicles is critical to data gathering, model training, and demonstration of our self-drivi…
Position Summary... What you'll do... About the Team: Senior Software Engineer for International Promotions and Rewards. Promo Engine is a real time transaction "rewards" service that supports acquisition and retention o…
Skills: Java, J2EE, Spring, Spring Boot, Microservices
Position Summary... What you'll do... Role summary: Walmart is seeking a Senior Software Engineer to lead the delivery of scalable, secure software solutions aligned with platform and business objectives. This role invol…
Skills: C programming, PostgreSQL, Software engineering, API development, Database optimization
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
$215k–$285k/yr
Full-time
Equity
Posted 31d ago
~40 hrs/week
Responsibilities
You will profile and optimize large-scale distributed machine learning training and high-throughput batch inference workloads to improve cluster efficiency and cost-effectiveness. This involves identifying performance bottlenecks across the stack and implementing solutions to enhance throughput and scaling efficiency.
Requirements
The role requires hands-on experience in ML performance engineering, profiling, and distributed multi-node training at scale. Candidates must be proficient in Python and C++ with a deep understanding of GPU architecture and accelerator performance concepts.
Full job description
Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon Valley company is creating the digital infrastructure needed to bring intelligence to every moving machine on the planet. Applied Intuition services the automotive, defense, trucking, construction, mining and agriculture industries in three core areas: tools and infrastructure, operating systems, and autonomy. Eighteen of the top 20 global automakers, as well as the United States military and its allies, trust the company’s solutions to deliver physical intelligence. Applied Intuition is headquartered in Sunnyvale, California, with offices in Washington, D.C.; San Diego; Ft. Walton Beach, Florida; Ann Arbor, Michigan; London; Stuttgart; Munich; Stockholm; Bangalore; Seoul; and Tokyo. Learn more at applied.co.
We are an in-office company, and our expectation is that full-time employees primarily work from their Applied Intuition office 5 days a week. However, we also recognize the importance of flexibility and trust our employees to manage their schedules responsibly. This may include occasional remote work, starting the day with morning meetings from home before heading to the office, or leaving earlier when needed to accommodate family commitments. This in-office expectation does not apply to contractor positions
About the Role
We are looking for a performance engineer who specializes in making large-scale machine learning workloads fast and cost-efficient in the datacenter. This role is focused on distributed training runs spanning many nodes, and high-throughput batch inference sweeping petabytes of real-world autonomy logs for auto-labeling, data mining, ground-truth generation, and evaluation.
The optimization target here is not tail latency on a vehicle - it is throughput, cluster goodput, and cost per unit of data processed. A training run that wastes 30% of its GPU-hours on stalled data loaders, or an offline inference sweep that takes a week instead of a day, directly slows down how fast the whole company can iterate. You will own the gap between what our fleet of accelerators is theoretically capable of and what our workloads actually achieve: profiling across the stack, finding where the compute and the wall-clock time actually go, and closing the difference.
You will work at the intersection of accelerators, ML frameworks, and large-scale data infrastructure, partnering with the teams who own each layer to land wins that show up in training time-to-result and offline processing cost. At Applied, we encourage all engineers to take ownership over technical and product decisions, closely interact with users to collect feedback, and contribute to a thoughtful, dynamic team culture.
At Applied, you will:
Profile and optimize distributed training end to end - data loading and preprocessing, augmentation, kernel execution, gradient communication, and checkpointing
Optimize large-scale offline and batch inference over petabyte-scale sensor logs: batching and scheduling strategies, quantization and low-precision execution, graph optimization, and accelerator saturation across long-running sweeps
Establish roofline and performance models for our workloads, quantify the gap between achieved and theoretical performance, and stack-rank optimization opportunities by impact and effort
Improve multi-node scaling efficiency: sharding and parallelism strategies, collective communication, interconnect utilization, and memory-bandwidth and kernel-fusion bottlenecks
Drive cluster goodput - reduce GPU idle time from input pipeline stalls, storage and network I/O, scheduling gaps, stragglers, and failure recovery on long-running jobs
Build the benchmarking, observability, and regression-detection tooling that keeps performance from silently degrading as models and code evolve
Collaborate with engineers across functions to solve complex data and compute problems at scale
Contribute to a team culture that values effective collaboration, technical excellence, and innovation
We're looking for someone who has:
Hands-on ML performance engineering experience: profiling, roofline analysis, throughput optimization, and root-cause investigation in production systems
Experience with distributed multi-node training at scale (FSDP, DeepSpeed, Megatron, NCCL, or equivalent), including diagnosing scaling inefficiency as node count grows
Deep familiarity with GPU or accelerator performance concepts - memory bandwidth, kernel launch overhead, occupancy, quantization, collective communication
Experience with high-throughput or batch inference systems (NVIDIA Triton Inference Server, TensorRT, ONNX Runtime, Ray, or similar)
Fluency in Python and proficiency in C++ or another systems language
Excellent debugging, analytical, and problem-solving skills
A deep understanding of machine learning foundations, and the ability to develop technical solutions for problems with no established playbook
Nice to have:
GPU kernel development experience: CUDA, Triton, CUTLASS, or hand-tuned attention implementations
Experience with profiling toolchains such as Nsight Systems/Compute, PyTorch Profiler, or perf
Experience with GPU scheduling and orchestration on Kubernetes, Slurm, or Ray, including multi-tenant cluster utilization
Experience with fault tolerance and elastic training for long-running jobs - checkpointing strategy, straggler mitigation, preemption recovery
Familiarity with autonomy or robotics data (ROS, OpenCV, multi-sensor log formats)
Don’t meet every single requirement? If you’re excited about this role but your past experience doesn’t align perfectly with every qualification in the job description, we encourage you to apply anyway. You may be just the right candidate for this or other roles.
Applied Intuition is an equal opportunity employer and federal contractor or subcontractor. Consequently, the parties agree that, as applicable, they will abide by the requirements of 41 CFR 60-1.4(a), 41 CFR 60-300.5(a) and 41 CFR 60-741.5(a) and that these laws are incorporated herein by reference. These regulations prohibit discrimination against qualified individuals based on their status as protected veterans or individuals with disabilities, and prohibit discrimination against all individuals based on their race, color, religion, sex, sexual orientation, gender identity or national origin. These regulations require that covered prime contractors and subcontractors take affirmative action to employ and advance in employment individuals without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status or disability. The parties also agree that, as applicable, they will abide by the requirements of Executive Order 13496 (29 CFR Part 471, Appendix A to Subpart A), relating to the notice of employee rights under federal labor laws.
FOR US-BASED ROLES: Applied Intuition is committed to providing an accessible and inclusive application and interview experience to applicants who are disabled veterans and other applicants with disabilities or medical conditions. Reasonable accommodations are available, requesting an accommodation will not affect your candidacy in any way, and you are not required to disclose the nature of your disability or medical condition in order to make a request. If you require an accommodation please contact [email protected]. We will work with you!
Applied Intuition is the physical AI company bringing intelligence to every moving machine on the planet.
Industry
Software Development
Company size
1,001-5,000 employees
Headquarters
Sunnyvale, California
LinkedIn followers
82,531
Total funding
$1.5B
Applied Intuition, Inc. is powering the future of physical AI.
Founded in 2017 and now valued at $15 billion, the Silicon Valley company is creating the digital infrastructure needed to bring intelligence to every moving machine on the planet. Applied Intuition services the automotive, defense, trucking, construction, mining and agriculture industries in three core areas: tools and infrastructure, operating systems, and autonomy.
Eighteen of the top 20 global automakers, as well as the United States military and its allies, trust the company’s solutions to deliver physical intelligence. Applied Intuition is headquartered in Sunnyvale, California, with offices in Washington, D.C.; San Diego; Ft. Walton Beach, Florida; Ann Arbor, Michigan; London; Stuttgart; Munich; Stockholm; Bangalore; Seoul; and Tokyo.
Learn more at applied.co.
Offices: 860 W California Ave, Sunnyvale, California 94086, US · 2350 Green Rd, Ann Arbor, Michigan 48105, US · 6th Floor Otemachi Bldg, 1-6-1 Otemachi, Tokyo, JP · Balanstraße 71a, München, 81541, DE · 2218 Two IFC, 10 Gukjegeumyung-ro, Seoul, KR
AIPhysical AIVehicle IntelligenceSelf-Driving CarsAutonomous VehiclesSelf-Driving SystemsArtificial IntelligenceDefense TechnologyNational SecurityVehicle OS
Based on 1493 listings with disclosed salaries, most software jobs in Sunnyvale, CA pay between $140k–$286k per year. Individual offers vary with seniority, company size, and specialization.
How many Software jobs are open in Sunnyvale, CA right now?
There are currently 1,773 open software positions in Sunnyvale, CA listed on Clera. New openings are added daily as companies post roles.
Which companies are hiring for Software roles in Sunnyvale, CA?
Companies currently hiring include Google, Walmart, Apple, Wayve, Applied Intuition, among others. Browse the listings above to see every active employer.
Are there remote or hybrid Software jobs in Sunnyvale, CA?
Yes — 569 of the 1773 open software positions offer remote or hybrid work (42 remote, 527 hybrid).
How do I apply for Software jobs in Sunnyvale, CA?
Each listing links directly to the employer's application page. Apply early — fresh listings get the most recruiter attention in the first two weeks.