Skip to main content

Senior Software Engineer at Tavus

You will be responsible for building the compute foundation for real-time multimodal AI agents that converse like humans. This high-impact role involves scaling massive Kubernetes clusters and managing GPU workloads across providers like AWS and CoreWeave to ensure low-latency performance. Join a world-class team of 50 engineers in San Francisco, London, or New York to solve the hardest infrastructure challenges in the generative AI space today.

Want to apply for this role?

Tavus

Jack finds you jobs at companies like Tavus. Talk to Jack to get considered for roles that fit what you're great at.

Location

San Francisco, United States / Remote

Compensation

Not Disclosed

Company

Tavus

Talk to Jack

Role overview

You will own the end-to-end GPU and cloud infrastructure powering real-time AI agents. By provisioning clusters, managing Kubernetes environments, and optimizing inference workloads, you ensure the reliability and scalability of Tavus’s core compute foundation. This is a high-impact, hands-on role working at the intersection of machine learning, real-time video, and distributed systems.

Tavus is a $64M Series B AI research lab backed by Sequoia, CRV, and YC, pioneering real-time multimodal AI agents. Based in San Francisco with a team of 50, Tavus builds “AI Humans” that see, hear, and converse, requiring cutting-edge GPU infrastructure to power low-latency video and audio inference at scale.

What you will do

  • Architect and scale GPU infrastructure across cloud providers like AWS, CoreWeave, and Lambda Labs to support training and real-time inference.
  • Design and maintain production-grade Kubernetes (EKS) clusters, ensuring high availability, uptime, and efficient resource orchestration for model serving.
  • Collaborate cross-functionally with research and product teams to optimize developer experience and internal platforms for deploying cutting-edge ML models.

Who this is a fit for

  • Extensive hands-on experience operating and provisioning GPU clusters in production environments, moving beyond simple resource consumption to infrastructure ownership.
  • Deep proficiency in Kubernetes and AWS ecosystems, with a track record of scaling distributed systems and managing multi-cloud or hybrid GPU strategies.
  • A startup-ready mindset focused on ownership and reliability, ideally with experience in AI-native environments or high-performance machine learning infrastructure.

Why this role is remarkable

  • Work at a Sequoia and CRV-backed Series B startup with $64M in funding, tackling the most complex compute challenges in generative video AI.
  • Own the mission-critical GPU infrastructure that serves as the foundation for real-time multimodal AI agents used by enterprise customers.
  • Join a high-caliber engineering team of 50 where you have high autonomy to build and scale infrastructure from scratch.

How Jack & Jill work together

Jack
I get to know what you’re great at, then find roles you’d never find yourself.
Jill
I recruit from Jack’s network and make the intro when I spot a great match.
Thumbnail for Meet Jack

Jack gets to know what you're great at and what you want next, then searches 15 million jobs daily and helps you discover roles at companies like this.

Meet Jack

What happens next?

Jack’s an AI agent for job searching and career coaching. He works for you.

Jill is the AI recruiter working for the company. She recruits from Jack’s network.

If your profile’s a match and Tavus wants to meet, Jill will make the intro. In the meantime, Jack will send you excellent alternatives.

Learn about Jack