AI Researcher (Multimodal Audio/Video Generation)
Tavus · San Francisco, California
- Location
- San Francisco, California, United States
- Employment
- Full time
- Level
- Mid level
- Posted
- 11 months ago
About this role
About Us
Tavus is a research lab pioneering human computing. We’re building AI Humans: a new interface that closes the gap between people and machines, free from the friction of today’s systems. Our real-time human simulation models let machines see, hear, respond, and even look real - enabling meaningful, face-to-face conversations. AI Humans combine the emotional intelligence of humans with the reach and reliability of machines, making them capable, trusted agents available 24/7, in every language, on our terms.
Imagine a therapist anyone can afford. A personal trainer that adapts to your schedule. A fleet of medical assistants that can give every patient the attention they need. With Tavus, individuals, enterprises, and developers can all build AI Humans to connect, understand, and act with empathy at scale.
We’re a Series A company backed by world-class investors including Sequoia Capital, Y Combinator, and Scale Venture Partners.
Be part of shaping a future where humans and machines truly understand each other.
The Role
We’re hiring a Senior AI Researcher to lead research in audio-visual avatar generation. This role is for someone who thrives in ambiguity, has a track record of pushing generative models to new frontiers, and wants to define what human - AI interaction looks like in practice.
Your Mission 🚀
Lead research efforts on audio-visual generation for avatars (Neural Avatars, Talking-Heads), with a focus on conversational settings.
Design models that are coupled with conversation flow - capturing and generating verbal + non-verbal signals in sync.
Drive innovation in diffusion models, long-video generation, and audio-visual modeling.
Translate research into production by partnering with Applied ML and engineering.
Mentor researchers, set research directions, and publish impactful work.
You’ll Bring:
A PhD or equivalent research experience, plus 2 - 3+ years of hands-on experience applying generative models at scale.
Expertise in diffusion models and awareness of the latest efficiency techniques.
Experience in multimodal generation - spanning video, audio, and language.
Proven innovation in long-video generation and/or audio generation.
Excellent programming skills - fluent in PyTorch and GPU-optimized workflows.
Track record of publications in top-tier venues (CVPR, NeurIPS, BMVC, ICASSP, etc.).
Experience leading research activities or mentoring teams.
Nice-to-Haves:
Skills in 3D graphics, Gaussian splatting, or large-scale training setups.
Broad exposure to generative AI models beyond your specialty.
Familiarity with software development best practices.
Location:
Preferred: San Francisco (hybrid) or London.
Remote within U.S. or Europe considered for exceptional candidates.
As published by Tavus. Applications are handled on their site.
Skills this posting mentions
About Tavus
Tavus is a research lab pioneering human computing. We’re building AI humans: a new interface that closes the gap between us and machines, free from the friction of today’s systems. Our real-time human simulation models let machines see, hear, respond, and even look real, enabling meaningful face-to-face conversations with people. AI Humans connect and act with precision and empathy, making them capable, trusted agents. It’s the best of both worlds: the emotional intelligence of humans, with the reach and reliability of machines. They’re available 24/7, in every language, on our terms. Imagine a therapist that anyone can afford. A personal trainer that adapts to your schedule. A fleet of medical assistants that can give every patient the attention they need. Tavus: teaching machines how to be human.
All 16 openings at TavusOne click, then it is written
Apply to Tavus with a resume written for this role.
Queue AI Researcher (Multimodal Audio/Video Generation) and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in Tavus’s form for you.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- 25 sent a week, free
- No card
- Nothing sent until you say so
More roles at Tavus
See all- 2 months ago
- 2 months ago
- 2 months ago
- 2 months ago
- 2 months ago
- 2 months ago
Similar roles elsewhere
See more- Yesterday
Member of Technical Staff - Applied AI
DoorDashSan Francisco, CA +1
$204k - $299k/yrStaffScience and research - Yesterday
Member of Technical Staff, Lead Researcher
DoorDashSan Francisco, CA +1
$204k - $299k/yrStaffScience and research - Yesterday
Associate Manager, AI Research Lab Strategy & Operations
DoorDashSan Francisco, CA +4
$124k - $155k/yrManagerScience and research - Yesterday
AI Research Fellowship, (Summer and Fall 2026)
DoorDashSan Francisco, CA
$107k - $158k/yrPrincipalScience and research - Last week
Applied AI Scientist, Small Language Model and AI Training
PostmanSan Francisco, California
$219k - $288k/yrMid levelScience and research
Put this to work
Paste your career in once. Every application after that is written for you.
Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- New matches ranked and written before you are up.
- Every bullet stays inside what your history supports. Nothing invented.
- Queued, submitted, interviewing, offer: one screen, not a spreadsheet.
500 free credits on sign-up. No card. Nothing is sent until you say so.
Listed from the job board Tavus publishes. Refolk is not the employer and does not handle their hiring. Applications go to Tavus directly.