Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Foundational Audio Research Scientist: Speech Multimodal ML

Jobleads-US

Google Research in Mountain View seeks a Research Scientist to set up large-scale tests, deploy ideas rapidly, and manage deadlines while applying the latest theories to develop new products and technologies. You will design experiments, prototype architectures, and publish findings through collaborations with universities and institutes worldwide.

Individual pay includes a base salary, 15% bonus target, equity, and comprehensive benefits.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Foundational Audio Research Scientist: Speech Multimodal ML in Mountain View, CA vacancy
  • $165k - $195k

     ...The Bosch Research and Technology Center North America with offices in Sunnyvale, California...  ...research in Silicon Valley focuses on Foundation Models, Natural Language Processing,...  ...AI & big data analytic solutions (e.g., audio, images, sensor logs) for a range of domains... 
    Audio
    Full time
    Work experience placement
    Local area
    Worldwide

    Bosch Group

    Sunnyvale, CA
    14 days ago
  • $174k - $252k

    Define and lead research agendas, experimental designs, and evaluation...  ...benchmarks in sequence, speech, audio, and multimodal modeling.Co-design and...  ..., and post-training foundational audio, speech, or multimodal...  ...types of work. As a Research Scientist, you'll setup large-scale... 
    Audio

    Google

    Mountain View, CA
    2 days ago
  • $117.2k - $313.7k

     ...ExperienceSalesforce AI Research is a global...  ...BLIP, a foundational breakthrough in multimodal AI; trained state...  ...entrepreneurial Research Scientists who want to...  ...multimodal AI agents.Speech Intelligence:...  ...and low-latency audio processing.Core...  ...and scaling ML models for real-... 
    Audio
    Full time
    Worldwide

    Salesforce

    Palo Alto, CA
    3 days ago
  • $117.2k - $313.7k

     ...Role Salesforce AI Research is seeking outstanding AI Research Scientists / Research Engineers to...  ...autonomous workflows Multimodal & Computer Vision – Vision...  ...for GUI agents Speech Intelligence – Voice intelligence...  ...‑taking, low‑latency audio processing Efficient... 
    Audio
    Full time

    100 Salesforce, Inc.

    Palo Alto, CA
    4 days ago
  •  ...Join the pioneering AI Foundations Research team at IBM Research...  .... Our group of scientists, engineers, and designers...  ...multi-modal (vision, speech, language, code) and...  ...Experience with Modern ML: Hands-on experience...  ...generative AI (LLMs) and multimodal models, from training... 
    Suggested
    Full time
    Contract work
    Part time
    Fixed term contract
    Internship
    Shift work

    IBM

    San Jose, CA
    2 days ago
  • $174k - $252k

    Conduct AI research and development in the field of health technology...  ....Design architectures for multimodal AI models and set up tests...  ...machine learning research, foundation models (e.g. LLMs, VLMs, time...  ...types of work. As a Research Scientist, you'll setup large-scale... 

    Google

    Mountain View, CA
    2 days ago
  • $184k - $287.5k

     ...Deep Learning Scientists to advance our...  ...streaming and agentic multimodal AI. You will demonstrate foundational expertise in...  ...and applied research to develop, train...  ...encompassing audio-visual reasoning...  ...Strong knowledge of ML/DL techniques...  ...in audio/speech AI, especially... 
    Audio
    Full time
    Work experience placement

    NVIDIA

    Santa Clara, CA
    4 days ago
  •  ...Amazon is hiring a Senior Data Scientist for the Vertical Ads group to lead innovation in Ads Foundational Model (Ads-FM). You will apply advanced ML techniques across image, video, and audio content to improve advertiser performance and customer experience. This is... 
    Audio

    Jobleads-US

    Palo Alto, CA
    2 days ago
  • $165k - $185k

     ...The Bosch Research and Technology Center North America with offices in Sunnyvale, California...  ...research in Silicon Valley focuses on Foundation Models, Big Data Visual Analytics,...  ...AI & big data analytic solutions (e.g., audio, images, sensor logs) for a range of domains... 
    Audio
    Full time
    Work experience placement
    Worldwide

    Bosch Group

    Sunnyvale, CA
    25 days ago
  • $165k - $185k

     ...Company DescriptionThe Bosch Research and Technology Center North America with offices in...  ...AI research in Silicon Valley focuses on Foundation Models, Big Data Visual Analytics,...  ...AI & big data analytic solutions (e.g., audio, images, sensor logs) for a range of domains... 
    Audio
    Work experience placement
    Worldwide

    Robert Bosch

    Sunnyvale, CA
    2 days ago
  •  ...Description Salesforce AI Research is looking for outstanding AI Research Scientists / Research Engineers. Do...  ...Generative AI, Agentic Systems, Speech Intelligence, and Multimodal Deep Learning? Our team...  ...-taking, and low-latency audio processing. Efficient Systems... 
    Audio
    Full time

    Salesforce

    Palo Alto, CA
    more than 2 months ago
  •  .... We are seeking a Senior Speech Software Engineer to drive both...  ...the intersection of speech research, model optimization, and production...  ...work closely with Speech Scientists, ML Researchers, and...  ...thousands of concurrent real-time audio streams Design and operate... 
    Audio

    ASAPP

    Mountain View, CA
    5 days ago
  • $180k

     ...THE ROLE: You will join the multimodal team to push toward superhuman...  ...across modalities—image, video, audio, and text—spanning the full...  ...the-art performance. Build research tooling, user-friendly interfaces...  ...large-scale distributed ML systems (training/inference optimization... 
    Audio
    Temporary work

    SpaceXAI

    Palo Alto, CA
    a month ago
  •  ...Senior / Staff AI Research Scientist, Foundation Models RoboForce is an AI robotics company developing...  ...data sources (vision, language, speech, etc.) to enable natural human-...  .... Decent understanding of multimodal models, modern ML architectures (transformers,... 
    Work at office
    Visa sponsorship

    Embedding VC

    Milpitas, CA
    4 days ago
  •  ...Archetype AI, based in Silicon Valley, is seeking an AI Researcher to design, train, and interpret large-scale models that...  ...streams. You will work on representation learning, multimodal learning, and foundation models across diverse sensing modalities. You will collaborate... 

    Jobleads-US

    Palo Alto, CA
    6 days ago
  •  ...intelligence into the real world. Our foundation model, Newton, understands...  .... We are looking for an AI Researcher to design, train, and...  ...of representation learning, multimodal learning, foundation models,...  ...experience developing advanced ML/AI systems, with a focus on real... 

    Jobleads-US

    Palo Alto, CA
    6 days ago
  •  ...passion for conducting impactful research in support of creating better...  ...manual annotation of data (audio, video, sensor, and/or text) according...  ...-generated transcripts (e.g., speech-to-text, activity/event...  ...validation, or dataset QA in a research, ML, or operations context
    Audio
    For contractors

    Mindlance

    Sunnyvale, CA
    2 days ago
  • $207k - $300k

    Identify and maintain ML training and serving benchmarks that...  ...Google Product teams and researchers to solve their performance problems...  ...or more of the following: speech/audio (e.g., technology...  ...Core team builds the technical foundation behind Google’s flagship products... 
    Audio

    Google

    Sunnyvale, CA
    2 days ago
  • $206.3k - $388k

     ...ABOUT THE ROLE We’re looking for a Principal ML Engineer to architect and scale the multimodal data processing pipelines and infrastructure behind Adobe Firefly’s multimodal foundation models (image, video, audio). In this role, you’ll sit at the intersection of data... 
    Audio
    Temporary work
    Local area
    Worldwide

    Jobleads-US

    San Jose, CA
    2 days ago
  •  ...and respond with natural speech and expressive motion....  ...diffusion, video, and multimodal models, to low‑latency...  ...talking‑head synthesis, audio‑driven facial animation...  ...backend, product, and research teams to ship avatar features...  ...~3+ years of hands‑on ML engineering experience... 
    Audio

    LiveX AI Inc.

    Palo Alto, CA
    1 day ago
  • $185k - $215k

    Company DescriptionThe Bosch Research and Technology Center North...  ...in Silicon Valley focuses on Foundation Models, Big Data Visual...  ...DescriptionAs a Senior Research Scientist- Robotics AI, you contribute...  ...experience building and applying multimodal transformer-based sequence-... 
    Work experience placement
    Local area
    Worldwide

    Robert Bosch

    Sunnyvale, CA
    1 day ago
  •  ...of belonging for one another that is the foundation of our culture. We want each team member...  ...journey. The mission of the Waymo Research team is to develop machine learning...  ...this role, you'll: Work on open-ended ML research problems for realistic simulation... 
    Internship
    Summer internship
    Local area

    Waymo

    Mountain View, CA
    5 days ago
  • $26 - $28 per hour

     ...team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-...  ...world AI systems. You’ll be working with audio, speech, and language data — helping...  ...scale systems and the opportunity to build foundational experience in data, language, and AI workflows... 
    Audio
    Full time
    Remote work
    Visa sponsorship

    Welo Global

    Sunnyvale, CA
    4 days ago
  • $120k - $225k

     ...the-shelf datasets spanning image, video, multimodal, reasoning, 3D, and more, as well as...  ...data engineering, Abaka AI provides the foundation for building high-performance AI systems...  ...data modalities, such as text, images, audio, video, or 3D point clouds. Experience... 
    Audio
    Flexible hours

    Jobleads-US

    Mountain View, CA
    2 days ago
  •  ...build, and deploy production‑grade ML systems with end‑to‑end ownership of...  ...‑powered solutions enabling natural speech interaction and real‑time audio understanding. Develop and optimize...  ...cross‑functional teams to shape the foundations of the AI stack, improve tooling, and... 
    Audio
    Full time

    Catalyst Labs

    Palo Alto, CA
    6 days ago
  •  ...Google DeepMind seeks a Research Scientist/Engineer to advance autonomous security through large-scale experiments, new architectures, and...  ...AI pipelines in Mountain View. You will train and evaluate foundation models, explore automated reasoning for cybersecurity, and publish... 

    Jobleads-US

    Mountain View, CA
    3 days ago
  •  ...Us: We are looking for an exceptional Research Scientist to develop next-generation AI technologies...  ...advances representation learning, multimodal understanding, and transformer-based modeling...  ...learning, semantic embeddings, and foundation-model applications. * Design,... 
    Full time
    Internship
    Local area

    Predactiv

    Palo Alto, CA
    more than 2 months ago
  •  ...Google LLC is seeking a Tech Lead for YouTube Shorts Discovery, ML Recommendations in Mountain View, CA. You will drive technical...  ...expertise in ML design, deployment, and optimization, including reinforcement learning and speech/audio domains. #J-18808-Ljbffr Jobleads-US
    Audio

    Jobleads-US

    Mountain View, CA
    2 days ago
  •  ...scalable, production-grade advanced ML solutions across natural language processing, speech recognition, recommendation...  ...independent time in learning, researching, and experimenting with new innovations...  ...harnesses) Strong foundation in machine learning, deep learning... 

    Chase

    Palo Alto, CA
    2 days ago
  • Planet Pharma is seeking a hands-on Contract Scientist, Machine Learning AI to accelerate project work across protein structure, phenotypic, and multimodal data. You will partner with computational biology, data science, translational, and discovery teams to develop, evaluate... 
    Contract work

    Planet Pharma

    Redwood City, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Foundational Audio Research Scientist: Speech Multimodal ML. Be the first to apply!