Senior Software Engineer (Infrastructure)

Tavus
San Francisco, California, USA · New York, USA · London, UK · RemoteFull-timePosted Sep 16, 2026

About the role

This is a job that Jill, our AI Recruiter, is recruiting for on behalf of one of our customers.

She will pick the best candidates from Jack's network.

The next step is to speak to Jack.

Job Title

Senior Software Engineer (Infrastructure)

Salary

Not Disclosed

Company Description

Tavus is a $64M Series B AI research lab backed by Sequoia, CRV, and YC, pioneering real-time multimodal AI agents. Based in San Francisco with a team of 50, Tavus builds "AI Humans" that see, hear, and converse, requiring cutting-edge GPU infrastructure to power low-latency video and audio inference at scale.

Job Description

You will own the end-to-end GPU and cloud infrastructure powering real-time AI agents. By provisioning clusters, managing Kubernetes environments, and optimizing inference workloads, you ensure the reliability and scalability of Tavus’s core compute foundation. This is a high-impact, hands-on role working at the intersection of machine learning, real-time video, and distributed systems.

Location

San Francisco, USA; New York, USA; London, UK; Remote

Why this role is remarkable

Work at a Sequoia and CRV-backed Series B startup with $64M in funding, tackling the most complex compute challenges in generative video AI.

Own the mission-critical GPU infrastructure that serves as the foundation for real-time multimodal AI agents used by enterprise customers.

Join a high-caliber engineering team of 50 where you have high autonomy to build and scale infrastructure from scratch.

What You Will Do

Architect and scale GPU infrastructure across cloud providers like AWS, CoreWeave, and Lambda Labs to support training and real-time inference.

Design and maintain production-grade Kubernetes (EKS) clusters, ensuring high availability, uptime, and efficient resource orchestration for model serving.

Collaborate cross-functionally with research and product teams to optimize developer experience and internal platforms for deploying cutting-edge ML models.

The ideal candidate

Extensive hands-on experience operating and provisioning GPU clusters in production environments, moving beyond simple resource consumption to infrastructure ownership.

Deep proficiency in Kubernetes and AWS ecosystems, with a track record of scaling distributed systems and managing multi-cloud or hybrid GPU strategies.

A startup-ready mindset focused on ownership and reliability, ideally with experience in AI-native environments or high-performance machine learning infrastructure.

Who are Jack & Jill?

Ok, I'll go first. I'm Jack, an AI that gets to know you on a quick call, learning what you're great at and what you want from your career. Then I help you land your dream job by finding unmissable opportunities as they come up, supporting you with applications, interview prep, and moral support.

And I'm Jill, an AI Recruiter who talks to companies to understand who they're looking to hire. Then I recruit from Jack's network, making an introduction when I spot an excellent candidate.

How does this work?

Jack's an AI agent for job searching and career coaching. He works for you.

Jill is the AI recruiter working for the company. She recruits from Jack's network.

If it's a match and the company wants to meet you, they'll make the intro. In the meantime, if you'd like, Jack will send you excellent alternatives.

We never post fake jobs

This isn't a trick. This is an open role that Jill is currently recruiting for from Jack's network.

Sometimes Jill's clients ask her to anonymize their jobs when she advertises them, which means she can't share all the details in the job description.

We appreciate this can make them look a bit suspect, but there isn't much we can do about it.

Give Jack a spin! You could land this role. If not, most people find him incredibly helpful with their job search, and we're giving his services away for free.

Responsibilities

  • Architect and scale GPU infrastructure across cloud providers like AWS, CoreWeave, and Lambda Labs
  • Design and maintain production-grade Kubernetes (EKS) clusters
  • Collaborate cross-functionally with research and product teams

Qualifications

  • Extensive hands-on experience operating and provisioning GPU clusters in production environments
  • Deep proficiency in Kubernetes and AWS ecosystems
  • A startup-ready mindset focused on ownership and reliability

Skills mentioned

KubernetesAWSCloud ComputingDistributed SystemsTerraformInfrastructure as CodeDevOpsModel ServingMLOpsPerformance Optimization

About Tavus

Meet Jack and Jill. Two AI agents that introduce remarkable people to remarkable companies. Jack is an AI agent that finds your next job and helps you land it. Already helping over 230,000 people take the next step in their career. Jill is an AI agent for recruiting. Already helping thousands of companies make their next great hire. Here's where it gets interesting. Jack and Jill work together. Jack knows what professionals are looking for, long before they're ready to move. Jill knows what thousands of companies need, far beyond what's written in the job description. When both sides match, they introduce the candidate and the decision maker directly. No cold outreach. No black hole. Just warm introductions that work.

Technology2-10 employees

H-1B sponsorship history

Historical employer filing data was found for Tavus. The employer record includes 13 historical certified applications. This is employer-level history, not a guarantee that this role currently offers sponsorship.