AI Researcher (Multimodal Audio/Video Generation)

Los Angeles +2

About this role

AI Researcher (Multimodal Audio/Video Generation)

Tavus is a research lab pioneering human computing through AI Humans—a new interface enabling real-time, face-to-face conversations between people and machines. We're building neural avatars that see, hear, respond, and look real, combining emotional intelligence with 24/7 availability across languages. As a Series A company backed by Sequoia Capital and Y Combinator, we're shaping the future of human–AI interaction at scale.

What you'll do

  • Lead research efforts on audio-visual generation for conversational avatars, including neural avatars and talking-head models.
  • Design models that couple conversation flow with synchronized verbal and non-verbal signal generation.
  • Drive innovation in diffusion models, long-video generation, and multimodal audio-visual modeling.
  • Translate research into production by collaborating with Applied ML and engineering teams.
  • Mentor researchers, set research directions, and publish impactful work in top-tier venues.

What Tavus is looking for

  • PhD or equivalent research experience, plus 2–3+ years applying generative models at scale.
  • Deep expertise in diffusion models and knowledge of efficiency techniques.
  • Proven experience in multimodal generation spanning video, audio, and language.
  • Track record of innovation in long-video or audio generation.
  • Strong programming skills with fluency in PyTorch and GPU-optimized workflows.
  • Published research in top-tier venues (CVPR, NeurIPS, BMVC, ICASSP, etc.).
  • Experience leading research activities or mentoring teams.
  • Nice-to-haves: 3D graphics, Gaussian splatting, large-scale training setups, or broad exposure to generative AI models.

Location: San Francisco (hybrid) or London preferred; remote within U.S. or Europe considered for exceptional candidates.

Company at a glance

Tavus is a research lab pioneering AI humans as a new human-computer interface through real-time human simulation models, founded by Hassaan Raza and Quinn Favret. The company is rapidly advancing technology that creates innovative AI-powered human representations for next-generation interactions.

Founded2020
Team Size11-50
StageSeries B
IndustryAI/ML
Locations
Los Angeles, CA, United States ·New York City, NY, United States ·San Francisco, CA, United States
Investors
Accel ·CRV ·Flex Capital ·HubSpot Ventures ·Index Ventures ·Lightspeed Venture Partners ·Scale Venture Partners ·Sequoia Capital ·Y Combinator ·YC Continuity
Websitetavus.io
LinkedInLinkedIn

Top Benefits

  • Unlimited PTO
  • Extremely competitive healthcare stipends
  • Gear stipends

What happens next

Skip the application pile. I get you in front of the people who decide.

Confirm the fit

A few questions to make sure this role is the right shape for you. Two minutes.

I pitch you to the company

I write the intro, send it to the founder, and handle the back-and-forth.

A meeting lands on your calendar

When the company wants to meet, I get the call on your calendar. You just show up.

Culture & values

Diverse and supportive team

Work is driven by people and success is shared by all

Place to learn, directly drive impact, and be with a team you love

Looking for culture creators, not cultural fits

Diversity drives success and is at the core of how they hire, communicate, and work

Combine diverse backgrounds, skill sets, and thinking to build best experiences

Flexible work schedule

Know someone who'd be great for this?