Staff Machine Learning Engineer, Voice AI
togetherai · San Francisco
- Posted
- 43 days ago
- Last confirmed live
- 2 days ago
What this role involves
Together AI is building inference infrastructure for voice applications. This Staff ML Engineer role focuses on optimizing model serving for voice workloads, including speech-to-text and text-to-speech, using inference engines like TRT-LLM and SGLang. The role involves owning the voice inference roadmap, driving performance, building evaluation frameworks, and collaborating with model partners.
Skills this posting asks for
- trt-llm
- sglang
- whisper
- parakeet
- orpheus
- kokoro
- snac
- encodec
- streaming inference
- gpu optimization
- batching strategies
- memory management
- real-time audio
- latency optimization
- serverless endpoints
- dedicated endpoints
- evaluation frameworks
- wer
- tts
- stt
- speech-to-speech
- audio-native llms
- codec-based architectures
- model deployment
Requirements
- Level: staff
From the employer’s posting
About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications — serving speech-to-text and text-to-speech models with best-in-class latency and reli…
Read the full description on togetherai’s careers pageApply without filling the form
Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.
Other roles at togetherai
- Senior Product Manager, Model APIs & Developer ExperienceSan Francisco
- Senior Software Engineer - Together Cloud InfrastructureSan Francisco
- Software Engineer, Customer InsightsSan Francisco
- Senior Product Engineer, FullstackSan Francisco
- Staff Software Engineer, Inference / Compute Infrastructure EngineeringSan Francisco
- Senior Network EngineerSan Francisco