Voice AI Engineer

Stealth Startup
RemoteFull-timePosted Sep 16, 2026

About the role

Voice AI Engineer — Senior / CTO / Co-Founder

Location: Remote

Employment Type: Full-Time

Level: Senior Engineer / CTO / Co-Founder, depending on experience

About the Role

We are looking for an exceptional Voice AI Engineer to build the next generation of AI-powered voice experiences.

This is a fully remote, high-ownership role for someone who can work across voice agents, LLMs, speech technologies, real-time systems, and production infrastructure.

We are open to hiring at the Senior Engineer level for strong candidates. For truly exceptional candidates with deep technical expertise, entrepreneurial experience, and a track record of building products from the ground up, there is an opportunity to join at the CTO / Co-Founder level.

What You’ll Do

Design, build, and deploy production-grade AI voice agents and conversational systems.

Develop real-time voice pipelines involving speech-to-text (STT), LLM reasoning, and text-to-speech (TTS).

Integrate and optimize leading LLM, voice AI, speech, and telephony technologies.

Build low-latency, highly reliable conversational experiences.

Develop systems for conversation management, memory, context, interruptions, tool calling, and agent orchestration.

Experiment with emerging AI models and technologies and rapidly turn promising ideas into production features.

Build APIs, services, and infrastructure required to operate voice AI systems at scale.

Optimize latency, reliability, accuracy, cost, and overall conversational quality.

Establish evaluation frameworks and continuously improve voice-agent performance.

Work closely with product and business teams to translate ideas into working AI products.

Operate independently in a fast-moving, remote environment with a high degree of ownership.

Required Qualifications

Strong software engineering background with experience building production systems.

Strong proficiency in Python and/or another modern programming language.

Hands-on experience building applications using LLMs and generative AI.

Experience with voice AI, conversational AI, speech recognition, speech synthesis, or AI agents.

Strong understanding of APIs, distributed systems, databases, cloud infrastructure, and production deployment.

Experience building real-time or low-latency applications.

Ability to independently take a technical problem from concept → architecture → implementation → production.

Strong debugging and systems-thinking abilities.

Excellent written and verbal communication skills.

Ability to work effectively and independently as part of a fully remote team.

Strongly Preferred

Experience with technologies such as OpenAI, Anthropic, Gemini, ElevenLabs, Deepgram, Cartesia, Twilio, Retell, Vapi, or similar platforms.

Experience building customer-facing or enterprise voice agents.

Experience with WebSockets, WebRTC, SIP, telephony infrastructure, or streaming audio.

Experience with RAG, vector databases, agent frameworks, function/tool calling, and conversational memory.

Experience with AWS, GCP, Azure, Docker, Kubernetes, CI/CD, and observability.

Experience with audio processing, speech models, NLP, or machine learning.

Experience taking an AI product from 0 → 1.

Startup or early-stage company experience.

CTO / Co-Founder Opportunity

For an exceptional candidate, this role can evolve into a CTO / Co-Founder position with significant ownership over the company's technology and product direction.

We are particularly interested in candidates who:

Have founded or led an early-stage AI or technology company.

Have built and shipped an AI product from scratch.

Can define technical strategy and architecture while remaining highly hands-on.

Can build, mentor, and lead an engineering/AI team.

Have strong product instincts and understand how to turn technology into business value.

Are comfortable operating with significant autonomy and ambiguity.

Think like a founder and are excited about building something from the ground up.

What Success Looks Like

Within the first 3–6 months, you will be expected to:

Build and ship production-quality voice AI capabilities.

Establish scalable architecture for voice agents and supporting services.

Improve conversational quality, latency, reliability, and cost.

Rapidly prototype and validate new AI capabilities.

Help define the long-term technical roadmap for the company's voice AI platform.

Ideal Candidate

You are a builder first. You don't simply research AI technologies—you know how to turn them into reliable, scalable products.

You are comfortable experimenting with new models and APIs while also understanding production engineering, scalability, latency, reliability, and user experience.

You thrive in a remote, high-autonomy environment, move quickly, and take ownership from idea to execution.

For the right person, this is much more than an engineering position. Exceptional candidates will have the opportunity to shape the product, technology, team, and potentially the company itself.

Responsibilities

  • Design, build, and deploy production-grade AI voice agents and conversational systems
  • Develop real-time voice pipelines involving speech-to-text (STT), LLM reasoning, and text-to-speech (TTS)
  • Integrate and optimize leading LLM, voice AI, speech, and telephony technologies
  • Build low-latency, highly reliable conversational experiences
  • Develop systems for conversation management, memory, context, interruptions, tool calling, and agent orchestration
  • Experiment with emerging AI models and technologies and rapidly turn promising ideas into production features
  • Build APIs, services, and infrastructure required to operate voice AI systems at scale
  • Optimize latency, reliability, accuracy, cost, and overall conversational quality

Qualifications

  • Strong software engineering background with experience building production systems
  • Strong proficiency in Python and/or another modern programming language
  • Hands-on experience building applications using LLMs and generative AI
  • Experience with voice AI, conversational AI, speech recognition, speech synthesis, or AI agents
  • Strong understanding of APIs, distributed systems, databases, cloud infrastructure, and production deployment
  • Experience building real-time or low-latency applications
  • Ability to independently take a technical problem from concept → architecture → implementation → production
  • Strong debugging and systems-thinking abilities

Skills mentioned

PythonGenerative AILarge Language ModelsAI AgentsWebSocketsDistributed SystemsAPI IntegrationCloud ComputingDockerKubernetes

About Stealth Startup

A network for entrepreneurs building in stealth. Submit your information here so investors can find you: harmonic.ai/get-discovered

Technology11-50 employeesSan Francisco, California

H-1B sponsorship history

Historical employer filing data was found for Stealth Startup. The employer record includes 11 historical certified applications. This is employer-level history, not a guarantee that this role currently offers sponsorship.