Software Engineer, Infrastructure
Tavus · United States
- Location
- United States
- Workplace
- Remote
- Employment
- Full time
- Level
- Mid level
- Posted
- 5 months ago
About this role
About Us
Tavus is a research lab pioneering human computing. We’re building AI Humans: a new interface that closes the gap between people and machines, free from the friction of today’s systems. Our real-time human simulation models let machines see, hear, respond, and even look real - enabling meaningful, face-to-face conversations. AI Humans combine the emotional intelligence of humans with the reach and reliability of machines, making them capable, trusted agents available 24/7, in every language, on our terms.
Imagine a therapist anyone can afford. A personal trainer that adapts to your schedule. A fleet of medical assistants that can give every patient the attention they need. With Tavus, individuals, enterprises, and developers can all build AI Humans to connect, understand, and act with empathy at scale.
We’re a Series B company backed by world-class investors including Sequoia Capital, Y Combinator, and Scale Venture Partners.
Be part of shaping a future where humans and machines truly understand each other.
The Role
We're hiring a Senior Software Engineer (Infrastructure) to own the systems behind CVI, our real-time conversational product. Every live conversation between a person and a PAL runs on infrastructure your team owns. You'll take goals like uptime, latency, and cost and chase them wherever they lead, including into backend services and product code.
What you'll own
CVI's inference deployments. The GPU infrastructure serving live conversations across multiple providers and regions. You'll join as an early senior member of a growing infra team, working on projects like tuning the newest GPU generations and cutting cold-start and model load times so users wait less.
Expanding our GPU footprint. You'll bring on new providers and regions, stand up clusters on EKS, and build the routing, scheduling, and throughput needed for fast weight loading.
Uptime. You'll be one of the people pushing our uptime bar higher, along with the security and SOC2 work that keeps our infrastructure trustworthy.
Fix what you find. When you see a problem, you have the trust and the mandate to fix it or flag it. Reworking our deploy pipeline so shipping is fast and boring is exactly the kind of thing you'd take on.
What this role has shipped
Multi-provider, multi-region inference infrastructure: routes live conversations across GPU providers and regions, so one provider's outage never becomes a user's problem
CUDA optimizations for Phoenix, our video rendering model: doubled the frame rate by tracing and optimizing hot paths with our researchers
Parallel conversations on a single GPU: several live conversations sharing one card, multiplying what the fleet can serve
Who you are
You own outcomes. You don't stop where "infrastructure" ends. If the fix lives in backend code or the CVI stack, you dive in, and you don't wait for a ticket to do it.
You're energized by unfamiliar problems. If the next thing that matters is standing up a training deployment you've never touched, you jump in and learn on the fly.
You adapt as priorities evolve. In a space moving this fast, the most important thing to build can change as we learn. When it does, you adjust course without losing momentum.
You care about this problem. Keeping large-scale, real-time systems fast and reliable is something you think about unprompted.
Requirements
Hands-on GPU inference experience. You've deployed and optimized inference workloads on GPUs and know what it takes to build reliable systems on top of GPU cloud providers.
Kubernetes and EKS depth, including routing and scheduling. You're comfortable designing how work gets placed across a fleet, and writing the services that make it happen.
Deep AWS experience. You're at home spinning up new services and turning them into simple, repeatable processes others can build on.
A senior track record of ownership. You've set technical direction, made decisions others built on, and carried ambiguous work over the finish line. You explain complex ideas clearly, to engineers and non-engineers alike.
Nice to have
Experience with GCP
Experience with video streaming infrastructure
Experience with training infrastructure or LLM serving
Experience with SOC2 or security compliance
If you don't check every box but this sounds like the work you want to be doing, apply anyway.
As published by Tavus. Applications are handled on their site.
Skills this posting mentions
About Tavus
Tavus is a research lab pioneering human computing. We’re building AI humans: a new interface that closes the gap between us and machines, free from the friction of today’s systems. Our real-time human simulation models let machines see, hear, respond, and even look real, enabling meaningful face-to-face conversations with people. AI Humans connect and act with precision and empathy, making them capable, trusted agents. It’s the best of both worlds: the emotional intelligence of humans, with the reach and reliability of machines. They’re available 24/7, in every language, on our terms. Imagine a therapist that anyone can afford. A personal trainer that adapts to your schedule. A fleet of medical assistants that can give every patient the attention they need. Tavus: teaching machines how to be human.
All 16 openings at TavusOne click, then it is written
Apply to Tavus with a resume written for this role.
Queue Software Engineer, Infrastructure and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in Tavus’s form for you.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- 25 sent a week, free
- No card
- Nothing sent until you say so
More roles at Tavus
See all- 2 months ago
- 2 months ago
- 2 months ago
- 2 months ago
- 2 months ago
- 2 months ago
Similar roles elsewhere
See more- Today
Senior Frontend Engineer, Ads Creative
RedditRemote - United StatesRemote
$191k - $267k/yrSeniorEngineering - Today
- Today
Staff Technical Product Manager, Ads ML Platform
RedditRemote - United StatesRemote
$217k - $304k/yrStaffEngineering - Today
Senior Software Engineer, Agentic Ads Experience
RedditRemote - United StatesRemote
$217k - $304k/yrSeniorEngineering
Put this to work
Paste your career in once. Every application after that is written for you.
Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- New matches ranked and written before you are up.
- Every bullet stays inside what your history supports. Nothing invented.
- Queued, submitted, interviewing, offer: one screen, not a spreadsheet.
500 free credits on sign-up. No card. Nothing is sent until you say so.
Listed from the job board Tavus publishes. Refolk is not the employer and does not handle their hiring. Applications go to Tavus directly.