Staff Machine Learning Engineer, Voice AI
$220k - $280kTogether
About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications — serving speech-to-text and text-to-speech models with best-in-class latency and reliability. We're looking for a Staff ML Engineer to drive the model serving layer for voice workloads. You'll work hands-on with inference engines like TRT-LLM and SGLang to optimize how we serve models like Whisper, Parakeet, Orpheus, and Kokoro — pushing latency and throughput to the frontier. You'll profile GPU utilization, design batching strategies for streaming audio, and ensure new model architectures can go from research to production quickly. This is a foundational hire on a small, high-impact team. Voice inference has unique challenges — streaming audio, tokenization, real-time latency budgets — that require dedicated ML engineering focus. You'll shape how Together serves voice models as the industry moves from pipeline architectures (ASR → LLM → TTS) toward end-to-end speech-to-speech. Own the model serving stack that powers Together's voice platform across STT, TTS, and speech-to-speech. Work directly with state-of-the-art accelerators (H100s, H200s, B200s) to optimize voice model inference. Collaborate with model partners (Cartesia, Deepgram, Rime, and others) to bring their models to production on Together's infrastructure. Build quality evaluation frameworks that guide model selection for customers and inform the roadmap. Join a small, early-stage team with outsized impact on a fast-growing product area. Responsibilities Own the voice inference roadmap end-to-end — define and execute the technical strategy for optimizing STT, TTS, and speech-to-speech models across Together's infrastructure, with a clear-eyed view of where the field is heading and how to position the platform ahead of it. Drive best-in-class inference performance — architect and implement systems targeting leading TTFB, throughput, and GPU utilization for voice workloads; set the performance bar others in the industry measure against, not just catch up to. Lead productionization of voice models at scale — design the serving architecture for serverless and dedicated endpoints, including batching strategies, streaming inference pipelines, and memory management tailored to real-time audio; own reliability and latency SLAs. Build the voice evaluation platform — design a rigorous, extensible evaluation framework covering WER across accents, languages, and noise conditions for STT; naturalness, latency, and pronunciation fidelity for TTS; establish the internal benchmark methodology that informs model selection and roadmap decisions. Shape the architecture for next-generation model support — anticipate and enable emerging model paradigms — audio-native LLMs, codec-based architectures (SNAC, Encodec), and end-to-end speech-to-speech systems — before they're mainstream, not after. Serve as the technical DRI for model partner integrations — lead deep collaboration with partners such as Cartesia, Deepgram, and Rime; own the full lifecycle from integration to optimization to ongoing performance accountability. Diagnose and resolve the hardest performance problems in the stack — conduct systematic profiling and root-cause analysis from GPU kernel behavior to framework-level bottlenecks; drive shipped improvements with documented, measurable impact. Influence platform architecture across the organization — partner with platform engineering leadership to ensure the serving layer is built for the latency and reliability demands of real-time voice APIs; your technical decisions should raise the ceiling for the whole team. Define and scale voice fine-tuning capabilities — lead the technical direction for enabling customers to fine-tune STT and TTS models on Together's infrastructure, establishing the primitives for differentiated voice experiences. Lay technical foundations for a category-defining product surface — architect systems with enough foresight that they support multiple new voice products with minimal rework; think in terms of platforms, not point solutions. Requirements 8+ years of ML engineering experience, with a demonstrated focus on model serving, inference optimization, or ML infrastructure at production scale — including systems you've owned from design through live traffic. Deep, practical expertise in LLM serving engines (vLLM, SGLang, TensorRT-LLM, or equivalent) — you've modified engine internals, debugged edge cases under load, and contributed improvements back; you don't stop at the API surface. Expert-level Python and PyTorch proficiency, with a strong command of GPU optimization — CUDA kernels, memory hierarchies, profiling toolchains — and a track record of turning that knowledge into shipped latency or throughput wins. Proven system design judgment — you've made architectural decisions that held up at scale and influenced how a team or platform evolved; you can articulate the tradeoffs you made and why. Strong technical leadership — you operate with high autonomy, define the right problems before solving them, and raise the bar for engineering quality around you without requiring process overhead. Sharp product intuition for developer tooling — you understand what voice application developers actually need to ship great products, and you let that shape your technical priorities, not just the other way around. Proven ability to move fast in ambiguous environments — you've thrived on early-stage or platform teams where scope is wide, ownership is deep, and the roadmap you build is the one you execute. Strong foundation in speech and audio ML (ASR/TTS architectures, audio signal processing) — directly relevant experience is strongly preferred; exceptional ML engineering fundamentals with genuine curiosity about the domain is also considered. Familiarity with audio codec and tokenization schemes (SNAC, Encodec, DAC) is a meaningful plus at this level. Experience training or fine-tuning speech models at scale is a significant advantage. Bachelor's or Master's in Computer Science, Electrical Engineering, or related field — or equivalent depth demonstrated through your work. About Together AI Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure. Compensation We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $220,000 - $280,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Accepted file types: pdf, doc, docx, txt, rtf Enter manually Accepted file types: pdf, doc, docx, txt, rtf #J-18808-Ljbffr Together
- Weave is looking for a Staff GenAI Engineer to join the Machine Learning Team specialised in voice and audio modalities , where you will be at the forefront of enabling product innovation and building AI-powered applications. In this role, you will help design how teams...SuggestedRemote work
$200k - $280k
A leading voice AI startup is seeking a Founding Senior Machine Learning Engineer to fine-tune and deploy human-like voice agents, handling millions of real-time calls. This role offers a salary range of $200,000 to $280,000 per year, with significant equity and benefits...Suggested- AppFolio in San Francisco is hiring a Senior Machine Learning Engineer to design and ship the next generation of voice and conversational AI agents within Realm-X. You’ll help define production voice and chat agent pipelines, working at the intersection of LLM agent frameworks...Suggested
- Together AI is building the best inference infrastructure for voice applications. We seek a Staff ML Engineer to own the model serving stack and optimize latency and throughput for real-time voice workloads. You'll work with state-of-the-art accelerators and collaborate...Suggested
- ...Inference—the gateway for developers to access the best models for voice AI through a single integration. The PM team is small and high-... ...owner. You will set the vision and roadmap, partner with engineering, manage model providers and deployment, and aim to make LiveKit...SuggestedRemote job
$209k - $313k
...themselves, live in the moment, learn about the world, and... ...digital services.Snap Engineering teams build fun and... ....We’re looking for a Machine Learning Engineer to join... ...driven featuresUtilize AI tools to design and... ...diverse backgrounds and voices working together will enable...Full timeLive inWork at officeLocal area$173k - $259k
...themselves, live in the moment, learn about the world, and... ...digital services.Snap Engineering teams build fun and... ....We’re looking for a Machine Learning Engineer to join... ...driven featuresUtilize AI tools and high velocity... ...backgrounds and voices working together will enable...Full timeLive inWork at officeLocal area- Confidential client, a seed-stage startup building the simulation and evaluation layer for Voice AI agents, seeks a Senior Member of Technical Staff. You will architect and develop systems for conversational AI, scaling AWS infrastructure to handle millions of real-time...
$229k - $343k
..., live in the moment, learn about the world, and have... ...digital services.Snap Engineering teams build fun and... ....We’re looking for a Staff Machine Learning Engineer to join... ...ranking, generative AI, LLM-based ranking,... ...diverse backgrounds and voices working together will...Full timeLive inWork at officeLocal area- ...large-scale infrastructure. About the Role We’re hiring Machine Learning Engineers to build and improve the AI systems that help strategic partners adapt the... ...encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity...
- ...Seeking Founding Data Scientists and Machine Learning EngineersImagine multiplying your impact.... ...past; teams need their next move. Palladio AI is the intelligence layer between raw... ...set the standard than inherit one.Product voice. You can walk a PM through trade-offs as...
$195k - $365k
...Plaud is building the real-world AI interface for professionals to... ...and privacy protection. To learn more about Plaud, please visit... ...intersection of research and engineering, eager to design novel sequence... ...architectures specifically for speech and voice generation. Alignment &...Full timeWork at officeWorldwide- ...2025, we started Handshake AI and built the fastest-growing... ...institutions Work alongside engineers, scientists, operators, and... ...Role Handshake is hiring a Staff Machine Learning Engineer for the Network and... ...metrics, and you'll be a key voice in how Handshake approaches...Full timeWork at officeRemote workFlexible hours
- ...than just a software company — we're building the AI-native platform where the real estate industry... .... Who We Are Looking For We're hiring a Senior Machine Learning Engineer to design and ship the next generation of voice and conversational AI agents within Realm-X. This...Full timeFlexible hours
- Wispr is building the first capable voice interface to reach billions of users. We are hiring an ML Engineer to prototype and ship features for our voice interface and to build scalable infrastructure for sub-500ms LLM inference across a global user base. You will help...
$342k - $399k
...Devices group. We study how AI systems perceive people and their... ...products. Our work spans machine learning, sensing, and hardware, with... ...looking for a machine learning engineer to help shape how future AI systems... ...many different perspectives, voices, and experiences that form...Work at officeRelocation package3 days per week- ...build and orchestrate AI workforces. Our AI workers... ...autonomously across voice, email, and enterprise... ...runs the real economy. Learn more about our vision in... ...Collaborate with product and engineering teams to integrate and... ...Strong experience in machine learning, deep learning...WorldwideShift work
$200k - $280k
Founding Senior Machine Learning Engineer Join to apply for the Founding Senior Machine Learning Engineer role at Retell AI. Base pay range $200,000 - $280,000 per year About Retell AI... ...reimagine the call center with cutting‑edge voice AI. We believe voice is still the most...H1bWork at office$266k
...operational stability. About The Role As a Machine Learning Engineer in OpenAI's Integrity team, you will... ...with some of the brightest minds in AI. You’ll work on state-of-the-art models... ...value the many different perspectives, voices, and experiences that form the full spectrum...- ...programming in Python. The company seeks language agnostic engineers who are able and interested to learn new technologies as needed. Time spent will be roughly 50% building a simulation and evaluation platform for AI agents, and 50% working on the framework used to build...
$295k
Machine Learning Engineer, Distributed Data Systems OpenAI | OpenAI | Posted Mar 2 Full-time Unknown... ...integrating multimodal functionalities into our AI products, ensuring they are reliable,... ...value the many different perspectives, voices, and experiences that form the full...Full timeWork at officeRelocation package$180k - $250k
...A cutting-edge AI company in San Francisco seeks a talented engineer to build infrastructure for voice AI conversations. The role involves creating distributed data pipelines and optimizing systems for large datasets. Ideal candidates should have experience with AI/ML...$163.42k - $285.98k
...work. Creating a career you love? It’s Possible.At Pinterest, AI isn't just a feature, it's a powerful partner that augments... ...around the world and 300 billion ideas saved, Pinterest Machine Learning engineers build personalized experiences to help Pinners create a life...Local areaRelocation package$200k - $235k
...our community. The Trust Frontier AI team is where new AI technology for Trust... ...product managers, data scientists, software engineers, fraud intelligence, and operations... ...Difference You Will Make: As a Senior Machine Learning Engineer on the Trust Frontier AI team,...Work experience placementCasual workLive inWork at officeRemote work$190.2k - $345.65k
...operate as a startup founding engineer while having the resources of... ...the next generation agentic AI platform working with other engineers... ...We are looking for a Staff Backend Engineer with 10+ years... ...software development using Machine Learning and Agentic AI systems to help...Full timeTemporary workLocal areaWorldwide- ...artificial intelligence and advanced ML, deep learning techniques to power decision-making in... .... About the RoleWe’re looking for a Machine Learning Engineer to help design, build, optimize and... ...a related field.Proficiency in using AI coding tools (e.g., Claude Code, Codex...Hourly payWork at officeLocal areaRemote workFlexible hours
$161.26k - $332.01k
...? It’s Possible.At Pinterest, AI isn't just a feature, it's a powerful... ..., multimodal representation learning, heterogeneous graph neural... ...pod is a small group (~6 engineers along with a product prototyping... ...vision experience.M.S. or PhD in Machine Learning, Computer Science, or...Currently hiringWork at officeLocal areaRemote workRelocationRelocation package- ...SuperhumanGrammarly is now part of Superhuman, the AI productivity platform on a mission to... ...busywork and focus on what matters. Learn more at superhuman.com and about our... ...Superhuman ubiquitous UI. As a Machine Learning Engineer on this team, you will be at the heart...WorldwideHome officeFlexible hours
$160k - $240k
...As a Fortune 500 company and a leading AI platform for managing people, money, and... ...challenging problems at the intersection of machine learning, agentic reasoning, and enterprise-scale... ....About the RoleAs a Machine Learning Engineer on the AI Core team, you will develop tailored...Full timeWork at officeRemote workHome officeFlexible hours$172.5k - $306.63k
...creative ecosystem. Our mission is to employ machine learning to enhance our comprehension of the... ...What You'll DoAs a Senior Machine Learning Engineer on the Content Intelligence team, you... ...with the latest advancements in ML and AI to ensure our solutions remain at the forefront...Full timeTemporary workLocal areaWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Machine Learning Engineer, Voice AI. Be the first to apply!
- staff engineer San Francisco, CA
- assistant engineer San Francisco, CA
- research assistant engineering San Francisco, CA
- staff design engineer San Francisco, CA
- staff security engineer San Francisco, CA
- engineering aide San Francisco, CA
- senior staff engineer San Francisco, CA
- senior staff systems engineer San Francisco, CA
- assistant chief engineer San Francisco, CA
- assistant electrical engineer San Francisco, CA

