Audio AI Researcher: Build Steerable Speech Systems
$350kJobleads-US
A leading AI research company in San Francisco is hiring for a position on their Audio team. The role involves developing and training advanced audio models, optimizing performance, and working collaboratively across teams. Ideal candidates will have strong expertise in JAX or PyTorch, a background in audio machine learning, and a passion for creating safe, reliable AI systems. The compensation ranges from $350,000 to $500,000, and the position offers flexible working hours and a supportive team environment. #J-18808-Ljbffr Jobleads-US
- Kotoba's speech models are licensed to Fortune 50 companies and... ...users a day. We're hiring an AI Researcher to build the next generation of real-... ...to-text, and text-to-speech systems — models that don't just... ...lingual modeling Knowledge of audio tokenization, neural audio codecs...Audio
- ...Applied Scientist / Research Engineer â Speech AI Â HIGHLIGHTS Location: Â San... ...speech and conversational AI systems. This person will work... ...modern machine learning for audio, speech, and language, and... ...Speech Generation Systems Build and improve machine...AudioRemote work
- Goaly is seeking an AI Researcher to lead research on agentic AI, focusing on training specialized models and building orchestration stacks. You'll design experiments, publish results, and advance the field. Ideal candidates have a Ph.D. or Master's in relevant fields,...Suggested
$150k - $250k
About Distyl AI Distyl is an applied AI technology company partnering... ...social organizations.We research and deploy technologies that power... ...into self-constructing systems, the development of the most reliable... ...built.Who You AreExperience Building with Models, Not Just Building...SuggestedWork at office3 days per week- Thinking Machines Lab in San Francisco is actively researching audio capabilities, blending rigorous theory with practical engineering to build multimodal AI systems. You will work across pre-training, post-training, and product to develop models that understand and generate...Audio
$180k - $270k
...Plaud Inc. Plaud is building the world's most trusted AI work companion for... ...models or foundational speech models.... ...Token (or Time-To-First-Audio) in real-time streaming... ...genuinely enjoy the systems-engineering challenge... ...collaborative, fast-paced research. Gear & Perks:...AudioFull timeWork at officeWorldwide- ...Description tl;dr: We're building a 3D first-person... ...game where an AI companion is the... ...LLM storytelling system that blends AI,... ..., and top-tier AI researchers. As an early member... ...text-to-motion or audio-to-motion models.... ...have worked on speech-driven 3D facial animation...AudioWork at officeVisa sponsorship
- ...this. We're creating an AI-powered experience that replicates... ...develop cutting-edge speech recognition models that... ...Expanding our ASR systems to new languages and markets Building and maintaining data infrastructure... ...with speech or audio Office ~ San Francisco...AudioFull timeLive inWork at officeWorldwide
- Distyl in San Francisco is seeking an Applied AI Researcher for the System Self-Construction team. You will design architectures enabling autonomous... ...research credentials, experience with meta-learning, and building prototypes to validate ideas in enterprise settings. #J-1...Work at office3 days per week
- Distyl AI is seeking researchers to push the boundaries of enterprise AI. You will explore novel AI system architectures, build prototypes, and demonstrate measurable improvements in real-world workflows. The role emphasizes AI-native thinking, cross-domain connections...Work at office
- ...a Product Manager focused on AI speech (text-to-speech, speech-to-text... ...Speech Evaluation Frameworks: Build and maintain comprehensive frameworks... ...in a product role at a voice/audio software company (e.g.... ...language models (e.g. developing systems leveraging LLMs as a component...Audio
$34 per hour
...join our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role focused on building the datasets that power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are...AudioWork experience placementRemote work$26 - $28 per hour
...join our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role focused on building the datasets that power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are...AudioFull timeWork experience placementRemote workVisa sponsorship$26 - $28 per hour
...Finnish Data Labeling Analyst(Speech & Voice) Overview... ..., supporting speech and voice AI systems. This is a high-impact production role focused on building the datasets that power real-world... ...systems. You'll be working with audio, speech, and language data -...AudioFull timeWork experience placementRemote workVisa sponsorship$204k - $300k
...Advanced Technology Group (ATG) is the research division of the company. ATG’s... ...electrical engineering, such as AI/ML, algorithms, digital signal processing, audio engineering, image processing,... ...science & analytics, distributed systems, cloud, edge & mobile computing,...AudioFull timeLocal areaWorldwideFlexible hours- ...technology firm in San Francisco is seeking a Senior Applied Researcher in Audio Understanding to tackle complex audio perception tasks. You will lead projects that push the boundaries of traditional speech recognition, emphasizing large-scale model development and innovative...AudioRelocation package
- Cartesia is hiring for an Audio Post-Training role to build and improve capabilities that shape how users interact with our generative audio models. This team spans research to production, designing evaluations, data pipelines, finetuning, RL, and model evaluation. You...Audio
- Postdoctoral Researcher, Computer Vision (PhD), New Grad Join... ...New Grad role at Jobright.ai Postdoctoral Researcher,... ...interdisciplinary teams to build contextually aware AI systems. Responsibilities: •... ...modalities (images, video, text, audio, speech and other modalities)...AudioFull timePart timeWork experience placementInternship
$320k
..., interpretable, and steerable AI systems. We want AI to be safe... ...group of committed researchers, engineers, policy experts... ...working together to build beneficial AI systems... ...bring Anthropic's audio models from research... ...team, who train the speech understanding and generation...AudioFull timeWork at officeVisa sponsorshipFlexible hours- ...Francisco, California. The Role: As a Research Engineer - Audio & Speech Models , you will be a core contributor on Zyphra’s Audio Team, building the next generation of open-source... ...all enjoy what we do and love discussing AI Benefits and Perks:...AudioWork at officeRelocation package
$117.2k - $313.7k
...AI Research Scientist And Research Engineer Salesforce... ...workflows, self-evolving agent systems, computer use agents,... ...AI agents. Speech Intelligence: Voice intelligence... ..., and low-latency audio processing. Core Modeling... ...of the systems you build. Unleash Your...Audio$289.1k - $408.5k
...Technology Group (ATG) is the research division of the company.... ...engineering, such as AI/ML, algorithms, digital signal processing, audio engineering, image... ...analytics, distributed systems, cloud, edge & mobile computing... ...with near-term impact.Build and mentor world-class...AudioFull timeLocal areaWorldwideFlexible hoursShift work$117.2k - $223.9k
...SalesforceSalesforce is the #1 AI CRM, where humans with agents... ...capabilities—processing images, audio, and video—to deliver... ...You will get an opportunity to build a platform on GCP to enable agentic... ...distributed and scalable distributed systems. This role requires hands-on...AudioFull time- ...Bot Bracket Bot is building low-cost, general-purpose... ...vertically-integrated system, and we’re assembling a... ...Logan Kilpatrick (Google AI), Mohith Mothukuri (... ...to own the full audio pipeline — from microphones... ...models (e.g. real-time speech systems) Low-latency...AudioFull time
$114.2k - $306.6k
...duplicating efforts.*Salesforce Research advances state-of-the-art AI techniques, developing... ...for GUI agents.* **Speech Intelligence:** Voice... ..., and low-latency audio processing.* **Efficient Systems:** Scalable deep... ...enterprise ecosystem.* Design, build, and maintain...AudioFull time- ...partner closely with Research to bring the next... ...boundaries of what AI can do. We’re expanding... ...inference, building the infrastructure needed... ...that handle image, audio, and other non-text... ...build and optimize the systems that let users generate speech, understand images,...AudioFull time
$216.3k - $280.8k
...received.Meet the TeamAt Foundation AI, we are leading frontier AI research across Cisco. Our mission is to advance... ...AI, multimodal learning, reasoning systems, scalable training algorithms,... ...leading academic institutions, and build production systems that solve challenging...Full timeTemporary workLocal areaFlexible hours- ...site in the specified location(s).As an AI Researcher within Schwab’s AI Strategy &... ...applied research and production‑grade AI systems that support millions of investors, advisors... ...practical experience.7+ years of experience building, deploying, and operating machine...Full timeWork at office
$197.3k - $313.7k
...SalesforceSalesforce is the #1 AI CRM, where humans with agents... ..., diverse team of researchers at Agentforce Operations. The... ...and shipping well-engineered systems.Partner with engineering, product... ...Better If...You have experience building production-grade ML pipelines...Full timeImmediate startRemote work- An innovative technology startup seeks an ML Engineer to design and deploy production-grade machine learning systems. Responsibilities include developing AI-powered solutions and handling the entire AI lifecycle. Candidates should have a Bachelor's or Master's degree in...Audio
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Audio AI Researcher: Build Steerable Speech Systems. Be the first to apply!
- senior researcher San Francisco, CA
- machine learning researcher San Francisco, CA
- researcher San Francisco, CA
- senior design researcher San Francisco, CA
- design researcher San Francisco, CA
- qualitative researcher San Francisco, CA
- data collection researcher San Francisco, CA
- product researcher San Francisco, CA
- survey researcher San Francisco, CA
- legal researcher San Francisco, CA



