The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
Skills: Post-silicon validation, System bring-up, High-speed IO, Memory subsystems, PCIe
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
Skills: Network architecture, Data center engineering, L2/L3 Ethernet, Networking protocols, Linux kernel
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
Skills: DevOps, Site Reliability Engineering, Cloud Infrastructure, Python, Go
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
Skills: Machine learning, Large language models, Deep learning, Hardware-software co-design, Inference optimization
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
Skills: New product introduction, Manufacturing engineering, Hardware development, Design for manufacturability, PCB design
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
Skills: Rack system integration, Datacenter deployment, Network equipment, Power systems, Compute equipment
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
Skills: Network Architecture, AI Compute Platforms, Hyperscale Infrastructure, InfiniBand, RoCEv2
Teamwork makes the stream work. Roku is changing how the world watches TV Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pio…
Skills: Machine learning, Ranking systems, Recommendation systems, Deep learning, LLMs
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
Senior Product Manager – Data Collaboration & Measurement
San Jose, California, United States · Hybrid
$125k–$262k/yr
Senior$221M raised
Teamwork makes the stream work. Roku is changing how the world watches TV Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pio…
Skills: Product Management, Advertising Technology, Data Collaboration, Cleanroom Technologies, SQL
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform …
Credo is looking for a Principal AI System Architect to join our team in San Jose, CA, reporting to AVP, XPU system AI interface. This role needs to solve the system-level interconnection challenges that connect XPUs (NP…
Skills: AI system architecture, NPU hardware, ASIC tapeouts, Interconnect design, Distributed training
Veeam is the Data and AI Trust Company, specializing in helping organizations ensure their data and AI are fully understood, secured, and resilient to enable the acceleration of safe AI at scale. As the market leader in …
Skills: Product management, Data modeling, Graph databases, SQL, Data analytics
WHAT YOU DO AT AMD CHANGES EVERYTHING At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in …
The Corporate Strategy group at Credo is looking for a Software Engineer to join our team in San Jose, CA and drive automation of internal workflows and toolsets, reporting to the SVP of Product. In this role, you become…
Skills: Python, SQL, Data modeling, Systems thinking, LLM APIs
Number of Position(s): 1 Duration: 6 - 12 months Dates: January 2027 - June 2027 - December 2027 Location: Onsite in San Jose, CA Educational Requirements Currently a candidate for a master's degree or PhD in computer sc…
Skills: Python, AI Agent Design, Prompt Engineering, Retrieval-Augmented Generation, Git
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
$245k–$325k/yr
Full-time
Medical insurance, Dental insurance, Vision insurance, Health Savings Account, Short-term disability insurance, Long-term disability insurance
Posted 5d ago
~40 hrs/week
Responsibilities
Define and drive the technical strategy for inference-systems performance, including workload capture, benchmarking, and performance modeling. Serve as the senior technical voice across engineering teams to optimize distributed inference pipelines and inform system planning.
Requirements
Requires 12+ years of experience in performance engineering with a proven record of technical leadership on large-scale, complex systems. Candidates must possess deep expertise in end-to-end performance analysis, simulation, and the ability to localize bottlenecks in distributed environments.
Full job description
The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform their businesses and operations at scale.
SambaNova Suite™ is the first full-stack, generative AI platform, from chip to model, optimized for enterprise and government organizations. Powered by the intelligent SN40L chip, the SambaNova Suite is a fully integrated platform, delivered on-premises or in the cloud, combined with state-of-the-art open-source models that can be easily and securely fine-tuned using customer data for greater accuracy. Once adapted with customer data, customers retain model ownership in perpetuity, so they can turn generative AI into one of their most valuable assets.
About the Team
The Inference Systems Performance team answers how fast SambaNova's systems can serve large language models and what it takes to get there. We capture real production traffic, benchmark it faithfully, model and simulate configurations that do not exist yet, and profile the distributed serving pipeline across host, accelerator, and fabric. Our work feeds serving optimization today and hardware and capacity planning for the next generation. We sit close to model optimization, systems, hardware, and product, and the whole company depends on our numbers.
About the Role
You will be the architect for end-to-end inference performance: how a request moves through tokenization, prefill, decode, and the fabric between them, and how a deployment is sized against customer SLOs. The work has two coupled pillars. One is reproducible workload capture and benchmarking, building replayable representations of real and increasingly agentic traffic so measurements reflect production rather than a naive load script. The other is performance modeling and simulation, turning measurement into a what-if capability for configurations and hardware that do not exist yet.
The frontier you will help define is heterogeneous, disaggregated inference, with GPU on prefill and the RDU on decode. That opens hard problems in networking, storage, prompt caching, and tail-latency-bound data movement. Inference systems performance is a young field and most answers are still being discovered, so you will spend your time on ambiguous problems with no established solution, and you will set the technical direction that others build on.
Responsibilities
Define and drive the technical strategy for inference-systems performance including workload capture, benchmarking, modeling, and simulation, while developing and architecture that enables many potential futures
Build the workload-capture and agentic-benchmarking capability - capture representative production traffic and enforce the discipline of interrogating results, spotting artificial contention or misleadingly high cache-hit rates that never occur in real use
Own the performance-modeling and simulation practice - models that predict how a configuration change moves the output, informing capacity planning against customer SLOs and next-generation system and hardware planning
Attack the end-to-end profiling gap - drive tooling that produces accurate, actionable profiles of a distributed inference pipeline so bottlenecks can be localized across host, accelerator, and fabric
Serve as the senior technical voice across model-optimization, systems, hardware, and product, tying together multiple engineering activities and teams, and weighing trade-offs of reliability, scalability, operational cost, and ease of adoption
Act as a resource for the entire organization including representing SambaNova's performance story to customers and partners
Mentor and multiply by raising the capability of principal and senior engineers, building the systems, tools, and patterns that make everyone more productive
Drive the resolution of the most ambiguous, novel challenges that span organizational boundaries or have no established answer in the field yet
Required qualifications
B.S. in Computer Science, Computer Engineering, or Related Field
12+ years of experience in performance engineering, with a demonstrated record of technical leadership on large-scale, complex systems
Deep expertise in end-to-end performance analysis of distributed systems with many moving parts and the ability to localize bottlenecks that others cannot
Proven command of realistic workload generation and simulation and of performance modeling, including calibrating models against real, variable workloads
Demonstrated ability to enter an unfamiliar domain and apply core performance methods with transferable discipline expertise
Ability to lead cross-functional efforts, mentor senior engineers, and influence organizational direction
Experience representing an organizations credibly to customers and partners
Track record of independently scoping and delivering high-complexity, high-ambiguity work with significant impact on products or roadmap
Preferred qualifications
M.S. or PHD in Computer Science, Computer Engineering, or Related Field
Direct experience with LLM inference serving - continuous batching, prompt/KV caching, prefill/decode disaggregation, tail-latency SLOs
Familiarity with inference simulation frameworks or agentic benchmarking efforts
A public technical voice - talks, writing, or community presence on systems performance
Base Salary Range:
Base Pay Range
$245,000—$325,000 USD
Submission Guidelines Please note that in order to be considered an applicant for any position at SambaNova Systems, you must submit an application form for each position for which you believe you are qualified.
EEO Policy SambaNova Systems is an Equal Opportunity/Affirmative Action Employer. All qualified applicants will receive consideration for employment without regard basis of age (40 and over), color, disability, gender identity, genetic information, marital status, military or veteran status, national origin/ancestry, race, religion, creed, sex (including pregnancy, childbirth, breastfeeding), sexual orientation, and any other applicable status protected by federal, state, or local laws.
Benefits Summary for US-Based, Full-Time Employment Positions SambaNova offers a competitive total rewards package, including the base salary, plus equity and benefits. We cover 95% premium coverage for employee medical insurance, and 77% premium coverage for dependents and offer a Health Savings Account (HSA) with employer contribution. We also offer Dental, Vision, Short/Long term Disability, Basic Life, Voluntary Life, and AD&D insurance plans in addition to Flexible Spending Account (FSA) options like Health Care, Limited Purpose, and Dependent Care. Our library of well-being benefits available to you and your dependents includes a full subscription to Headspace, Gympass+ membership with access to physical gyms, One Medical membership, counseling services with an Employee Assistance Program, and much more.
Transforming AI with efficiency, security, and sovereignty - driven by our relentless pursuit of intelligence.
Industry
Computer Hardware Manufacturing
Company size
201-500 employees
Founded
2017
Headquarters
San Jose, California
LinkedIn followers
98,810
Total funding
$2.5B
Welcome to SambaNova: Revolutionizing AI Capacity
At SambaNova, we're empowering developers, enterprises, governments, and data centers to unlock their full AI potential. Our full-stack infrastructure, from chips to models, enables lightning-fast performance, low power consumption, and high-efficiency computing.
Our Mission
To give every developer, enterprise, government and data center absolute sovereignty over their own data, models and AI infrastructure – to future-proof the AI workloads that will power and scale tomorrow.
Our Technology
We give our customers the optionality to experience SambaNova through the cloud or on-premise.
Samba Cloud delivers the fastest inferences on the largest open source models like Llama 4 and DeepSeek. Developers can get started building in minutes with our OpenAI compatible APIs. All customers start on the developer tier and when they need more capacity can scale into our enterprise tier.
SambaStack is our on-premise offering which includes the system, the platform, and foundation models. These components combine into a powerful technology stack that delivers unparalleled performance, ease of use, accuracy, data privacy, and the ability to power every use case across the world's largest organizations.
SambaManaged is a modular and ready-to-deploy AI cloud designed to deliver unmatched efficiency for data centers and cloud service providers. This solution allows organizations to quickly deploy advanced AI inference services—without the need for costly infrastructure upgrades or specialized expertise—in as little as 90 days.
At the heart of SambaNova innovation is the Reconfigurable Dataflow Unit (RDU). Purpose built for AI workloads, the RDU takes advantage of a dataflow architecture and a three-tiered memory design. The three tiers of memory enable the platform to run hundreds of models on a single node and to switch between them in microseconds. In 2023, SambaNova released its 4th generation RDU chip, the SN40L.
Offices: 2460 N First Street, 100, San Jose, California 95131, US
How much do Data & Analytics jobs in San Jose, CA pay?
Based on 939 listings with disclosed salaries, most data & analytics jobs in San Jose, CA pay between $110k–$293k per year. Individual offers vary with seniority, company size, and specialization.
How many Data & Analytics jobs are open in San Jose, CA right now?
There are currently 1,040 open data & analytics positions in San Jose, CA listed on Clera. New openings are added daily as companies post roles.
Which companies are hiring for Data & Analytics roles in San Jose, CA?
Companies currently hiring include Adobe, Capital One, AMD, Google, WD, among others. Browse the listings above to see every active employer.
Are there remote or hybrid Data & Analytics jobs in San Jose, CA?
Yes — 366 of the 1040 open data & analytics positions offer remote or hybrid work (68 remote, 298 hybrid).
How do I apply for Data & Analytics jobs in San Jose, CA?
Each listing links directly to the employer's application page. Apply early — fresh listings get the most recruiter attention in the first two weeks.