About the Role We are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly technical role, you will operate at the intersection of cutting-edge inferen…
GenBio AI develops multiscale foundation models to decode and simulate human biology. Our team is accelerating towards an ambitious future where scientists can unlock humanity's biggest challenges in drug discovery, heal…
Skills: AI for Structural Biology, Machine Learning, Computational Biology, Protein Engineering, Antibody Discovery
About Subsense Subsense is a deep-tech company developing the world’s first non-surgical, bidirectional brain-computer interface powered by plasmonic and magnetoelectric nanoparticles. Our mission is to unlock direct com…
Skills: Python, Instrument Control, PyVISA/SCPI, PyQt, Data Acquisition
If you're ready to be part of our legacy of hope and innovation, we encourage you to take the first step and explore our current job openings. Your best is waiting to be discovered. Day - 08 Hour (United States of Americ…
If you're ready to be part of our legacy of hope and innovation, we encourage you to take the first step and explore our current job openings. Your best is waiting to be discovered. Day - 08 Hour (United States of Americ…
Skills: Product Management, Digital Health, Roadmap Development, Wireframing, Lean Six Sigma
WindBorne Systems is supercharging weather forecasts with a unique proprietary data source: a global constellation of next-generation smart weather balloons targeting the most critical atmospheric data. We design, manufa…
Director, Applied Science, Alexa for Shopping (Rufus)
Palo Alto, California, United States · On-site
$263k–$350k/yr
Senior+$35B raised
Alexa for Shopping (Rufus) is Amazon's new AI-powered shopping assistant that combines the capabilities of Rufus and Alexa+ to provide a more personalized and intelligent shopping experience. We are building the future o…
Skills: Large Language Models, Multi-agent Systems, Reinforcement Learning, RLHF, DPO
If you're ready to be part of our legacy of hope and innovation, we encourage you to take the first step and explore our current job openings. Your best is waiting to be discovered. Day - 08 Hour (United States of Americ…
About Rivian Rivian is on a mission to keep the world adventurous forever. This goes for the emissions-free Electric Adventure Vehicles we build, and the curious, courageous souls we seek to attract. As a company, we con…
Skills: Quantized deep learning, Hardware acceleration, Autonomous systems, Perception model design, Embedded compute platforms
Model AI
Founding Agent Harness Engineer
Palo Alto, California, United States · On-site
Mid level
Founding Agent Systems Engineer Location: Onsite in Palo Alto Compensation: Competitive Salary + Equity About Model AI Model AI is building the infrastructure and application stack for the next generation of agentic AI s…
Skills: Python, Systems Engineering, LLM Agents, Evaluation Harnesses, Developer Tools
Model AI
Founding Machine Learning Infrastructure Engineer
Palo Alto, California, United States · On-site
Senior
Founding Machine Learning Infrastructure Engineer Location: Onsite in Palo Alto Compensation: Competitive Salary + Equity About Model AI Model AI is building the infrastructure and application stack for the next generati…
Skills: ML Systems, Distributed Systems, High-Performance Computing, LLM Inference, CUDA
WindBorne Systems is supercharging weather forecasts with a unique proprietary data source: a global constellation of next-generation smart weather balloons targeting the most critical atmospheric data. We design, manufa…
Skills: Ruby on Rails, Postgres, Full-stack Development, Infrastructure Architecture, Data Pipeline Management
If you are a current Jazz employee please apply via the Internal Career site Jazz Pharmaceuticals is a global biopharma company whose purpose is to innovate to transform the lives of patients and their families. We are d…
About Mind: Mind Robotics is building Physical AI for real-world industrial deployment, starting with the factory floor. We believe the hardest problems in AI are solved when researchers and engineers are hands-on with t…
Skills: Control Theory, C++, Rust, MATLAB/Simulink, Python
About Mind: Mind Robotics is building Physical AI for real-world industrial deployment, starting with the factory floor. We believe the hardest problems in AI are solved when researchers and engineers are hands-on with t…
APP Nurse Practitioner/Physician Assistant- Plastic Surgery- FT Days
Palo Alto, California, United States · Hybrid
$89/hr–$117/hr
Mid level
If you're ready to be part of our legacy of hope and innovation, we encourage you to take the first step and explore our current job openings. Your best is waiting to be discovered. Day - 10 Hour (United States of Americ…
Associate Medical Director, Clinical Development - Job ID: 1912, 1913, 1914
Palo Alto, California, United States · Hybrid
$255k–$265k/yr
Mid level$1.9B raised
Ascendis Pharma is a dynamic, fast-growing global biopharmaceutical company with locations in Denmark, Europe, and the United States. Today, we're advancing programs in Endocrinology Rare Disease and Oncology. Here at As…
Skills: Clinical Trial Design, Medical Monitoring, Data Analysis, Regulatory Submissions, Clinical Development Planning
About Us DELFI Diagnostics, Inc. (DELFI Diagnostics) is developing next-generation, blood-based tests that are reliable, accessible and deliver a new way to help detect cancer. Employing advanced machine-learning methods…
Senior Drive Unit Lubrication and Thermals Engineer
Palo Alto, California, United States · Hybrid
$139k–$233k/yr
SeniorVisa sponsorship$16B raised
We made history and now we work to transform the future – for our customers, our communities and our families. You'll see your work on the road every day, helping people move freely and pursue their dreams. At Ford, you …
Urologic Oncologist - Assistant Professor of Urology
Palo Alto, California, United States · On-site
$372k–$399k/yr
Mid level
The Department of Urology at Stanford University and Veterans Affairs Palo Alto Health Care System (VAPAHCS) seeks a urologic oncologist to join the Department as Assistant Professor in the University Medical Line and a …
Skills: Open Surgery, Robotic Surgery, Urologic Oncology, Clinical Research, Translational Research
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
$185k–$300k/yr
Full-time
Competitive salary, Equity, Comprehensive health benefits, Monthly stipends, Company retreats
Posted 81d ago
~40 hrs/week
Responsibilities
Lead the implementation of advanced inference acceleration and GPU parallelism strategies to optimize Pika's AI-driven video and language models. Collaborate with research and engineering teams to deploy high-performance computing kernels and scalable production pipelines.
Requirements
Requires 5+ years of engineering experience with deep expertise in CUDA, NCCL, and distributed inference techniques like TP, SP, and PP. Candidates should have a proven track record in model quantization and familiarity with video generation or LLMs.
Full job description
About the Role
We are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly technical role, you will operate at the intersection of cutting-edge inference acceleration, GPU parallelism, advanced model deployment, and video generation technologies. Your expertise will drive significant improvements to model speed and efficiency, ensuring our creative AI systems deliver industry-leading user experiences at scale.
You will design and optimize inference pipelines, implement state-of-the-art acceleration techniques, and work closely with researchers and engineers across the team to push the boundaries of what’s possible in real-time AI deployment. Your efforts will play a foundational role in powering the next generation of Pika’s video and language models.
What You’ll Do
Accelerate Inference: Lead and implement advanced inference acceleration techniques, including attention optimization and quantization for efficient model serving.
Maximize GPU Parallelism: Engineer and optimize GPU strategies across tensor, sequence, and pipeline parallelism (TP, SP, PP) for maximal efficiency and scalability.
Programming for Performance: Develop and optimize high-performance computing kernels and distributed workloads using CUDA and NCCL.
Advance AI Deployment: Collaborate with research and engineering teams to bring state-of-the-art videogen and large language models into production.
Improve Training Efficiency: (Bonus) Contribute to improvements in model training speed, stability, and resource utilization as part of our deployment lifecycle.
Technical Excellence: Drive rigorous code reviews, participate in technical discussions, and mentor fellow engineers on best practices in inference and GPU programming.
What We’re Looking For
Experience: 5+ years engineering experience, with a strong track record in inference acceleration and model deployment at scale.
Inference Mastery: Proven expertise in inference optimization, including quantization, attention acceleration, and deep learning compiler stacks.
GPU & Parallelism: Deep knowledge of GPU programming (CUDA, NCCL) and experience with SP, TP, PP, and other forms of parallelism for distributed inference.
AI Domain Knowledge: Familiarity with video generation (videogen) models and large language models (LLMs).
Collaboration: Strong cross-discipline communication skills; able to drive shared goals across research and engineering functions.
Ownership Mindset: Self-driven, solutions-oriented, and capable of managing ambiguity in a fast-paced startup environment.
Bonus: Experience in enhancing training efficiency, stability, or resource optimization for large models.
Nice to Have
Experience with high-throughput video or real-time streaming model deployment
Familiarity with distributed training and optimization toolkits
Contributions to open source projects in AI infrastructure or deep learning compilers
Startup or rapid prototyping experience
What We Offer
Competitive salary in the AI industry
Equity in a fast-growing startup shaping the future of AI
Comprehensive health benefits, monthly stipends, company retreats
A supportive and collaborative office culture—we’re all building and launching together
About Pika
At Pika, we're crafting a future where video creation is seamless, intuitive, and universally accessible. Our mission is to empower creativity by breaking down technical barriers using the transformative power of AI. We’re a tight-knit, energetic team based in Palo Alto, CA, valuing efficiency, curiosity, and the ambition to make a meaningful impact on the world.
We work from our Palo Alto office 3–5 days a week and welcome applicants who are eager to contribute onsite.
How many Science & Research jobs are open in Palo Alto, CA right now?
There are currently 496 open science & research positions in Palo Alto, CA listed on Clera. New openings are added daily as companies post roles.
Which companies are hiring for Science & Research roles in Palo Alto, CA?
Companies currently hiring include Stanford Health Care, CONCEPT Continuing & Professional Studies Division, Palo Alto University, Stanford Medicine Children's Health, Amazon, Summit Therapeutics, Inc., among others. Browse the listings above to see every active employer.
Are there remote or hybrid Science & Research jobs in Palo Alto, CA?
Yes — 138 of the 496 open science & research positions offer remote or hybrid work (22 remote, 116 hybrid).
How do I apply for Science & Research jobs in Palo Alto, CA?
Each listing links directly to the employer's application page. Apply early — fresh listings get the most recruiter attention in the first two weeks.