Senior Software Engineer (Infrastructure)
About the role
This is a job that Jill, our AI Recruiter, is recruiting for on behalf of one of our customers.
She will pick the best candidates from Jack's network.
The next step is to speak to Jack.
Job Title
Senior Software Engineer (Infrastructure)
Salary
Not Disclosed
Company Description
Tavus is a $64M Series B AI research lab backed by Sequoia, CRV, and YC, pioneering real-time multimodal AI agents. Based in San Francisco with a team of 50, Tavus builds "AI Humans" that see, hear, and converse, requiring cutting-edge GPU infrastructure to power low-latency video and audio inference at scale.
Job Description
You will own the end-to-end GPU and cloud infrastructure powering real-time AI agents. By provisioning clusters, managing Kubernetes environments, and optimizing inference workloads, you ensure the reliability and scalability of Tavus’s core compute foundation. This is a high-impact, hands-on role working at the intersection of machine learning, real-time video, and distributed systems.
Location
San Francisco, USA; New York, USA; London, UK; Remote
Why this role is remarkable
Work at a Sequoia and CRV-backed Series B startup with $64M in funding, tackling the most complex compute challenges in generative video AI.
Own the mission-critical GPU infrastructure that serves as the foundation for real-time multimodal AI agents used by enterprise customers.
Join a high-caliber engineering team of 50 where you have high autonomy to build and scale infrastructure from scratch.
What You Will Do
Architect and scale GPU infrastructure across cloud providers like AWS, CoreWeave, and Lambda Labs to support training and real-time inference.
Design and maintain production-grade Kubernetes (EKS) clusters, ensuring high availability, uptime, and efficient resource orchestration for model serving.
Collaborate cross-functionally with research and product teams to optimize developer experience and internal platforms for deploying cutting-edge ML models.
The ideal candidate
Extensive hands-on experience operating and provisioning GPU clusters in production environments, moving beyond simple resource consumption to infrastructure ownership.
Deep proficiency in Kubernetes and AWS ecosystems, with a track record of scaling distributed systems and managing multi-cloud or hybrid GPU strategies.
A startup-ready mindset focused on ownership and reliability, ideally with experience in AI-native environments or high-performance machine learning infrastructure.
Who are Jack & Jill?
Ok, I'll go first. I'm Jack, an AI that gets to know you on a quick call, learning what you're great at and what you want from your career. Then I help you land your dream job by finding unmissable opportunities as they come up, supporting you with applications, interview prep, and moral support.
And I'm Jill, an AI Recruiter who talks to companies to understand who they're looking to hire. Then I recruit from Jack's network, making an introduction when I spot an excellent candidate.
How does this work?
Jack's an AI agent for job searching and career coaching. He works for you.
Jill is the AI recruiter working for the company. She recruits from Jack's network.
If it's a match and the company wants to meet you, they'll make the intro. In the meantime, if you'd like, Jack will send you excellent alternatives.
We never post fake jobs
This isn't a trick. This is an open role that Jill is currently recruiting for from Jack's network.
Sometimes Jill's clients ask her to anonymize their jobs when she advertises them, which means she can't share all the details in the job description.
We appreciate this can make them look a bit suspect, but there isn't much we can do about it.
Give Jack a spin! You could land this role. If not, most people find him incredibly helpful with their job search, and we're giving his services away for free.
Responsibilities
- Architect and scale GPU infrastructure across cloud providers like AWS, CoreWeave, and Lambda Labs
- Design and maintain production-grade Kubernetes (EKS) clusters
- Collaborate cross-functionally with research and product teams
Qualifications
- Extensive hands-on experience operating and provisioning GPU clusters in production environments
- Deep proficiency in Kubernetes and AWS ecosystems
- A startup-ready mindset focused on ownership and reliability
Skills mentioned
About Tavus
Meet Jack and Jill. Two AI agents that introduce remarkable people to remarkable companies. Jack is an AI agent that finds your next job and helps you land it. Already helping over 230,000 people take the next step in their career. Jill is an AI agent for recruiting. Already helping thousands of companies make their next great hire. Here's where it gets interesting. Jack and Jill work together. Jack knows what professionals are looking for, long before they're ready to move. Jill knows what thousands of companies need, far beyond what's written in the job description. When both sides match, they introduce the candidate and the decision maker directly. No cold outreach. No black hole. Just warm introductions that work.
H-1B sponsorship history
Historical employer filing data was found for Tavus. The employer record includes 13 historical certified applications. This is employer-level history, not a guarantee that this role currently offers sponsorship.