Foundational Audio Research Scientist: Speech Multimodal ML
Jobleads-US
Google Research in Mountain View seeks a Research Scientist to set up large-scale tests, deploy ideas rapidly, and manage deadlines while applying the latest theories to develop new products and technologies. You will design experiments, prototype architectures, and publish findings through collaborations with universities and institutes worldwide.
Individual pay includes a base salary, 15% bonus target, equity, and comprehensive benefits.
#J-18808-Ljbffr Jobleads-US$174k - $252k
...maintains a portfolio of research projects driven by fundamental... ...of work. As a Research Scientist, you'll setup large-scale... ...evaluation benchmarks in sequence, speech, audio, and multimodal modeling. Co-design and... ..., and post-training foundational audio, speech, or...Audio- ...The work sits between research and product. You'll formulate... .... What You'll Work On Multimodal intelligence. Build... ...content across video, audio, image, and text. Ranking... ...feedback. Foundation models and agents. Apply... ...which advances in AI and ML are useful for Agentio...AudioWork at officeLocal areaFlexible hours
$174k - $252k
...Google maintains a portfolio of research projects driven by... ...types of work. As a Research Scientist, you’ll setup large-scale tests... ...health insights and builds foundation models (wearables and Electronic... ...Design architectures for multimodal AI models and set up tests to...Suggested- ...Amazon is hiring a Senior Data Scientist for the Vertical Ads group to lead innovation in Ads Foundational Model (Ads-FM). You will apply advanced ML techniques across image, video, and audio content to improve advertiser performance and customer experience. This is...Audio
- ...capabilities, building the foundation for intelligent, multimodal patient interactions... ..., including speech‑to‑text, natural language... ...that power AI or ML workflows Hands‑... ...that combine text, audio, and visual inputs... ...clinicians, and AI researchers to build something with...Audio
- ...real-time APIs for speech-to-text (STT),... ...Deepgram’s voice-native foundation models are... ...over 50,000 years of audio and transcribed more... ...partner with QA, Research, Product, Data, and... ...pipelines, agents, or multimodal models, including... ...benchmarking, or ML infrastructure...Audio
- ...Audible, a leading audio storytelling company, seeks a Senior Applied Scientist to solve real-world problems at scale in ML, NLP, GenAI and distributed systems. You will design end-to-end, production-ready models spanning Content Understanding and recommendations, collaborating...Audio
- ...the role Flam builds multimodal AI systems that ship to... ...remove them. Take research to production: read the... ...Required 2+ years building ML systems that ran in... ...in one or more of: speech ASR/TTS, diffusion and... ...systems WebRTC, low-latency audio/video pipelines)....Audio
- ...providing real-time APIs for speech-to-text (STT), text-to-speech... ...Box. Deepgram’s voice-native foundation models are accessed through cloud... ...processed over 50,000 years of audio and transcribed more than 1... ...Experience with GPU infrastructure, ML platforms, or multi‑tenant...Audio
$150.7k - $251.2k
...same, and with constant research and development, you'... ...team of scientists and engineers who love... ...-value production AI/ML solutions on a major... ...# Proficient in foundation models, generative AI... ...processing techniques and/or speech and audio processing and analysis...AudioFull timeWork experience placementImmediate start- ...the CTO, you will help build the next foundational model for gut activity: aself-supervised, multimodal architecture trained on continuous... ...learning systems in production, not research alone. ~ Deep expertise in time series, audio or biosignal modelling. ~ Strong...Audio
- # Research EngineerElevenlabsRemoteUnited StatesData Science & AnalyticsPosted Sep 24,... ...creators and marketers to generate and edit speech, music, image, and video across 70+... ...gives developers access to our leading AI audio foundational models. Everything we do is the result...AudioImmediate startRemote work
$49 - $69 per hour
...with confidence and enjoy a better quality of life. Home Health Speech Language Pathologist Evaluate, direct and provide speech/language... ...tests and applications of therapeutic treatments including audio logic screening. Observe, record and report changes in the patient...AudioDaily paidFull time$54 - $76 per hour
...confidence and enjoy a better quality of life. As a Home Health Speech Therapist, you will: Evaluate, direct and provide speech/... ...tests and applications of therapeutic treatments including audio logic screening. Observe, record and report changes in the patient...AudioDaily paid$54 - $76 per hour
...confidence and enjoy a better quality of life. As a Home Health Speech Therapist , you will: Evaluate, direct and provide speech/... ...tests and applications of therapeutic treatments including audio logic screening. Observe, record and report changes in the patient...AudioDaily paidFull timeApprenticeship- ...About David AI David AI is the first audio data research company. We bring an R&D approach to data... ..., and we believe audio is the gateway. Speech is versatile, accessible, and human—it fits... ...software engineering, data engineering, ML, and signal processing. 6+ years of...AudioWork at office
$54 - $76 per hour
...confidence and enjoy a better quality of life. As a Home Health Speech Language Pathologist , you will: Evaluate, direct and... ...diagnostic tests and applications of therapeutic treatments including audio logic screening. Observe, record and report changes in the...AudioDaily paidFull timeApprenticeshipRelief$400k
...Principal Research & Engineering, Realtime Voice AI... ...powered by Inflection AI’s foundation model, proving that AI... ...Voice AI stack across speech models, streaming... ...train strategies for core audio, speech, and realtime... ..., audio tokenization, multimodal models, barge‑in, low‑...AudioWork at officeFlexible hours- ...group and seeking exceptional researchers to join our dynamic team. As... ...Researcher, you will apply advanced ML techniques to a wide range of... ...We’re looking for research scientists with a proven track record of... ...time series data Strong foundation in mathematics, statistics,...Full time
$130k - $190k
...We are seeking a Staff ML Engineer to lead the design... ..., partnering with AI Scientists and Product to define... ...clearly to engineering, research, product, and... ...frameworks.* Experience with speech-to-text, speaker diarization... ...conversational audio data.* Experience deploying...AudioFor contractorsLocal area- ...innovative Machine Learning Scientists to join our AI team, focusing... ...Robotics. As a key member of our research and development efforts, you... ...Language Models (LLMs), Multimodal Large Language Models (MLLMs)... ...PhD and with +5 years for ML Scientist, +8 years for Sr. ML...Work at officeRemote work
- ...bringing together world-class researchers and engineers to turn... ...We are seeking a Research Scientist, Robotics & Embodied AI to... ...learning, imitation learning, multimodal foundation models, motion intelligence... ...and experience with modern ML and robotics frameworks....
$175.1k - $236.9k
...creators to produce and share audio storytelling with our... ...We are seeking a data scientist builder to join the... ...and scale econometric/ML models and quantitative... ...measurement, and adoption Research and evaluate emerging... ...in mathematical foundations of statistics, machine...AudioWork experience placementFlexible hours$54 - $76 per hour
...confidence and enjoy a better quality of life. Role as a Home Health Speech Language Pathologist As a Home Health Speech Language... ...diagnostic tests and applications of therapeutic treatments including audio logic screening. Observe, record and report changes in the...AudioDaily paid$170.6k
...Details: Job Description: The Graphics Research Organization at Intel is seeking a highly motivated Research Scientist Intern to join its world-class team. This role... ...vision conferences, or journals. Strong foundation in linear algebra, numerical optimization, probability...Remote jobInternshipLocal areaImmediate startWork from homeWorldwideShift work$170k - $400k
...One that is proactive, multimodal, and capable of... ...with the world through speech, text, vision, and persistent... ...and build the native foundations of Hark's desktop app... ...end, work closely with researchers, model engineers,... ...including voice input, audio, screen understanding,...AudioFull time$54 - $76 per hour
...confidence and enjoy a better quality of life. As a Home Health Speech Language Pathologist, you will: Evaluate, direct and provide... ...tests and applications of therapeutic treatments including audio logic screening. Observe, record and report changes in the patient...AudioDaily paidFull time- ...data infrastructure Collaborate with ML and backend teams to deliver high-quality... ...into a coherent semantic graph. This foundation delivers verified facts and rich context... ...Charrier- Founding Engineer at Niland, an audio search company (acquired by Spotify), then...AudioRemote workFlexible hours
$130.33k - $195.5k
...will partner closely with AI Scientists, Data Engineers, Product... ...data, as well as multimodal data such as documents, images, and audio. Design for environments where... ..., and alerting.* Apply AI/ML to production systems (applied, not research). Build and serve inferencing...AudioFull timeWork experience placementRemote work$70 - $97 per hour
...confidence and enjoy a better quality of life. As a Home Health Speech Language Pathologist , you will: Evaluate, direct and provide... ...diagnostic tests and applications of therapeutic treatments including audio logic screening. Observe, record and report changes in the...AudioFull timeTemporary workApprenticeship
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Foundational Audio Research Scientist: Speech Multimodal ML. Be the first to apply!


