Member of Technical Staff

Los Angeles +2 · On-site$200k – $200k + EquityVisa Sponsorship Available

About this role

Member of Technical Staff

Our mission at Wafer is to maximize intelligence per watt by using AI to optimize AI infrastructure, achieving orders of magnitude better energy and cost efficiency per token.

We believe cheap intelligence is the most essential piece of technology for a future of abundance. We care about building a future where intelligence is "too cheap to meter."

Wafer commercializes these efforts by serving serverless and dedicated inference for open source LLMs at the best performance per dollar. Our core bet is doing this through autonomous optimization of heterogeneous hardware.

What you'll do

Ship day-zero support for new open-source models, tuned for latency and throughput

Optimize the serving stack: batching, KV cache, speculative decoding, quantization

Write and tune kernels in CUDA, HIP, and Triton for NVIDIA, AMD, TPU, Trainium, D-Matrix, and more.

Design, deploy, and operate heterogeneous clusters across vendors

Run production inference across a mixed fleet: reliability, observability, and cost per token at scale

How we evaluate

We score every candidate on seven values

Infinitely Resourceful

Exceptionalism

Unreasonable Standards

Company Over Self

High EQ

Learns Quickly

First Principles Thinker

How we work

On-site in San Francisco, five days a week. Small team with massive surface area and ownership. You operate with complete autonomy of how to solve problems, and work with the team to set the direction of your work. We don't see engineers as code writers, but as problem solvers. You will do everything from talking to customers to writing custom GPU kernels in esoteric hardware.

Company at a glance

Wafer is an AI infrastructure company that optimizes GPU resources and accelerates inference across diverse technology stacks, helping organizations reduce computational costs and improve AI performance.

Team Size1-10
WorkspaceOn-site
StageSeed
IndustryAI/ML
Locations
Los Angeles, CA, United States ·New York City, NY, United States ·San Francisco, CA, United States
Websitewafer.ai
LinkedInLinkedIn

What happens next

Skip the application pile. I get you in front of the people who decide.

Confirm the fit

A few questions to make sure this role is the right shape for you. Two minutes.

I pitch you to the company

I write the intro, send it to the founder, and handle the back-and-forth.

A meeting lands on your calendar

When the company wants to meet, I get the call on your calendar. You just show up.

Culture & values

Team works directly together to define technical direction

Focus on building core systems collaboratively

Culture centered around building the future of AI infrastructure

Know someone who'd be great for this?