Software Engineer Voice AI Inference Runtime

Baseten is an AI inference platform for deploying, optimizing, and scaling custom, open-source, and fine-tuned models in production.

San Francisco, United States
About Baseten

Baseten provides model runtimes, inference infrastructure, developer workflows, and deployment options including managed cloud, self-hosted, and hybrid environments.

View jobs by Baseten

Skills

About the Role

You will own Voice AI product areas from architecture through production operations. You will design, build, and operate real-time model serving systems for speech-to-text, text-to-speech, and voice agents; improve performance; coordinate delivery across engineering teams; and mentor teammates through technical reviews.

Requirements

  • Bachelor's degree or higher in Computer Science or a related field
  • Track record owning production-grade real-time, large-scale systems with tail-latency requirements
  • Proficient coding ability in a popular programming or scripting language
  • Product judgment for developer-oriented tools
  • Interest in ML/AI infrastructure
  • Collaboration and communication skills
  • Experience using AI coding assistants

Responsibilities

  • Own and lead Voice AI product areas from architecture through production operations
  • Design, build, and operate real-time model serving systems
  • Build systems for STT, TTS, and voice-agent workloads
  • Drive cross-team collaboration and end-to-end delivery
  • Mentor teammates through code reviews, design documents, and technical leadership

Benefits

  • Equity
  • Medical, dental, and vision insurance for U.S. employees and dependents
  • Flexible PTO and company-wide Winter Break
  • Paid parental leave
  • Fertility and family-building stipend through Carrot
  • Company-facilitated 401(k) for U.S. employees