Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Scientist, Gemini Audio i18n, DeepMind

$174k - $252k

Google

Create comprehensive evaluation sets and benchmarks to measure audio-to-audio (A2A) model performance across international languages, accents, and regional dialects.Propose, prototype, and evaluate novel modeling techniques to improve A2A audio understanding, dialog, and audio generation capabilities with a focus on scalability.Identify performance gaps in current multilingual audio models and collaborate with cross-functional research and engineering teams to deploy solutions.Minimum qualifications:Bachelor's degree in Computer Science, Speech Recognition, Computational Linguistics, a related technical field, or equivalent practical experience.Experience conducting research or development in Speech Recognition, Text-to-Speech (TTS), or Large Language Models (LLMs).Experience coding in Python or C++ and using deep learning frameworks such as PyTorch, JAX, or TensorFlow.Experience working with audio data, speech processing, or multilingual datasets.Preferred qualifications:Master's degree or Ph.D. in Automatic Speech Recognition (ASR), Text-to-Speech (TTS), Computational Linguistics, or a related field.3 years of experience in Large Language Models (LLMs) or multimodal foundation models.Experience developing audio-to-audio (A2A) architectures, end-to-end speech models, or spoken dialog systems.Experience scaling speech models across international languages, accents, or low-resource locales.Publication record in speech or machine learning conferences (e.g., ICASSP, INTERSPEECH, NeurIPS, or ACL).As an organization, Google maintains a portfolio of research projects driven by fundamental research, new product innovation, product contribution and infrastructure goals, while providing individuals and teams the freedom to emphasize specific types of work. As a Research Scientist, you'll setup large-scale tests and deploy promising ideas quickly and broadly, managing deadlines and deliverables while applying the latest theories to develop new and improved products, processes, or technologies. From creating experiments and prototyping implementations to designing new architectures, our research scientists work on real-world problems that span the breadth of computer science, such as machine (and deep) learning, data mining, natural language processing, hardware and software performance analysis, improving compilers for mobile platforms, as well as core search and much more.As a Research Scientist, you'll also actively contribute to the wider research community by sharing and publishing your findings, with ideas inspired by internal projects as well as from collaborations with research programs at partner universities and technical institutes all over the world.Artificial intelligence will be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.We are pushing the boundaries across multiple domains. Our global teams offer diverse learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort.Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $174000 - $252000 (USD) + 15% bonus target + equity + benefitsLearn more about benefits at Google.Note: By applying to this position you will have an opportunity to share your preferred working location from the following: Mountain View, CA, USA; New York, NY, USA.Bachelor's degree in Computer Science, Speech Recognition, Computational Linguistics, a related technical field, or equivalent practical experience.Experience conducting research or development in Speech Recognition, Text-to-Speech (TTS), or Large Language Models (LLMs).Experience coding in Python or C++ and using deep learning frameworks such as PyTorch, JAX, or TensorFlow.Experience working with audio data, speech processing, or multilingual datasets.

Vacancy posted 22 hours ago
Similar jobs that could be interesting for youBased on the Research Scientist, Gemini Audio i18n, DeepMind in Mountain View, CA vacancy
  •  ...useful inventions. At Google DeepMind, we’re a team of scientists, engineers, machine...  ...the highest priority.The Gemini Safety team is accountable...  ...Gemini models. The role of the Research Scientist / Research Engineer...  ...text-to-text, image/video/audio-to-text modalities and... 
    Audio

    DeepMind

    Mountain View, CA
    3 days ago
  • $207k - $301k

     ...and ML efficiency. Drive new research ideas from conception, experimentation...  ...types of work. As a Research Scientist, you'll setup large-scale...  .... benchmarking and improving Gemini models across modalities to...  ...inventions. At Google DeepMind, we are a pioneering AI lab with... 
    Suggested
    Shift work

    Google

    Mountain View, CA
    1 day ago
  • $174k - $253k

     ...model calibration, ensuring Gemini accurately identifies its own...  ...years of experience leading a research agenda.Preferred qualifications...  ...researchers.As a Research Scientist, you will be responsible for...  ...transformative inventions. At Google DeepMind, we are a pioneering AI lab... 
    Suggested

    Google

    Mountain View, CA
    3 days ago
  • $174k - $253k

     ...maintain training infrastructure to support Gemini Audio encoder and pretraining models,...  ...audio related components.Collaborate with research teams across Gemini Audio to improve the...  ...transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional... 
    Audio

    Google

    Mountain View, CA
    3 days ago
  • $147k - $211k

    SnapshotWe are seeking strong Research Scientists with expertise in AI...  ...research effort within Google DeepMind's Frontier AI unit. This role...  ...breakthrough improvements to models (Gemini) and products. Our work is...  ...data types (e.g., vision, audio, text).The US base salary... 
    Audio
    Full time

    DeepMind

    Mountain View, CA
    4 days ago
  • $207k - $300k

    Partner with the Gemini or DeepMind teams to design, develop, and deploy novel multimodal conversational agents.Develop audio-first models capable of orchestrating and planning complex...  ....g., NeurIPS, ICML, ICLR, AAAI, CVPR).Research background in NLP/Generative AI.At Google... 
    Audio

    Google

    Mountain View, CA
    3 days ago
  • $174k - $253k

    Be a part of team of scientists, engineers, machine learning experts,...  ...experience in deep learning research and development, including generative AI, audio and video synthesis, diffusion...  ...that is fundamental to Google DeepMind’s work on Gemini’s audio capabilities, on... 
    Audio

    Google

    Mountain View, CA
    2 days ago
  • $323k

     ...SnapshotAt Google DeepMind, we’re a team of scientists, engineers, machine learning experts...  ...ambitious goals.The Area: Gemini AppThe Gemini App is at...  ...management, engineering, and research teams to deliver an...  ...deterministic AI outputs (streaming audio/visuals) with zero jitter... 
    Audio
    Full time

    DeepMind

    Mountain View, CA
    1 day ago
  • $185k - $400k

     ...intelligent agentic platforms. We are looking for a staff or lead-level Research Engineer, Data to architect and scale data engineering systems...  ...support model training and research workflows for text, image, audio, and video datasetsPartner with research and engineering teams... 
    Audio
    Remote work

    Pika

    Palo Alto, CA
    1 day ago
  • $117.2k - $313.7k

     ...the future of Salesforce.The ExperienceSalesforce AI Research is looking for outstanding AI Research Scientists and Research Engineers. Our team discovers new...  ...(TTS/ASR), human-like turn-taking, and low-latency audio processing.Core Modeling and Post-Training: Machine... 
    Audio
    Full time

    Salesforce

    Palo Alto, CA
    4 days ago
  • $262k - $364k

     ...software engineering skills to complement the research background.Build recommender/search...  ...specific types of work. As a Research Scientist, you'll setup large-scale tests and deploy...  ...transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional... 

    Google

    Mountain View, CA
    3 days ago
  • $185k - $400k

     ...and intelligent agentic platforms. We are seeking accomplished Research Scientists in Foundation Models with expertise in pre-training and mid-...  ...-scale multimodal pre-training/mid-training (text, image, audio, and video), and drive innovative approaches for foundational... 
    Audio
    Remote work

    Pika

    Palo Alto, CA
    3 days ago
  • $165k - $195k

    Company DescriptionThe Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania...  ...and transparent AI & big data analytic solutions (e.g., audio, images, sensor logs) for a range of domains, including Industry... 
    Audio
    Full time
    Work experience placement
    Local area
    Worldwide

    Robert Bosch

    Sunnyvale, CA
    2 days ago
  • $158.8k - $218.1k

     ...medical devices, AI, and health services. We research, develop, and commercialize innovative...  ...Ph.D.-level Senior, Applied AI Health Scientist with solid ML/AI technology background...  ...signals from sensors, including light, audio, physiological, and inertial data collected... 
    Audio

    PVH (Tommy Hilfiger/Calvin Klein)

    Mountain View, CA
    5 days ago
  •  ...translation fluency under real-world disfluency. We’re looking for a Research Scientist who can define what "better" actually means across all of...  ...years of research or applied research experience in speech, audio, or NLP, with a demonstrated focus on evaluation methodology... 
    Audio

    Sanas

    Palo Alto, CA
    1 day ago
  • $165k - $185k

    Company DescriptionThe Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania...  ...and transparent AI & big data analytic solutions (e.g., audio, images, sensor logs) for a range of domains, including Industry... 
    Audio
    Work experience placement
    Worldwide

    Robert Bosch

    Sunnyvale, CA
    3 days ago
  • $117.2k - $313.7k

    About the Role Salesforce AI Research is seeking outstanding AI Research Scientists / Research Engineers to build and deploy high‑impact AI solutions at scale....  ...intelligence, TTS/ASR, human‑like turn‑taking, low‑latency audio processing Efficient Systems - Scalable deep... 
    Audio
    Full time

    100 Salesforce, Inc.

    Palo Alto, CA
    4 days ago
  • $192k - $304.75k

    We're now looking for a Senior Research Scientist, Multi-Modal Language Models!NVIDIA is seeking a Senior Research Scientist passionate about...  ...mix multiple modalities together, such as text, image, video, audio, etc …Design solutions that improve pareto... 
    Audio
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $192k - $304.75k

    We are now looking for a Senior Research Scientist for Generative AI!NVIDIA is searching for a world-class researcher in generative AI to join...  ...image generation, video generation, 3D generation, and audio generation. You will be working with a team of world-class researchers... 
    Audio
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $137.7k - $275.4k

    What you can expectZoom's AI Incubation team seeks a Research Scientist to advance AI-driven innovation in healthcare. You will join a world-...  ...understanding across structured and unstructured data — including text, audio, and imaging modalities.Advancing LLM post-training,... 
    Audio
    Full time
    Work at office
    Remote work
    Worldwide

    Zoom

    San Jose, CA
    2 days ago
  • $244.8k

     ...ByteDance Seed team is pursuing groundbreaking research in speech, audio, music, and multimodal learning. The team welcomes talents to tackle complex challenges with global collaboration across the US, China, and Singapore. Open roles focus on cutting-edge ML research,... 
    Audio

    ByteDance

    San Jose, CA
    4 days ago
  •  ...foundation models that understand everything from video and audio to text and artwork at a semantic level. By treating these powerful...  ...to enhance personalizationAbout the RoleWe are looking for a Research Scientist specializing in embeddings and representation learning to... 
    Audio
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    2 days ago
  • $192.2k - $260k

    We are looking for a Senior Applied Scientist to help drive the research and development of real-time multimodal conversational AI. You will contribute...  ...focus areas: advancing foundation models for speech and audio, and building the post-training systems (reward modeling... 
    Audio
    Local area
    Flexible hours

    Amazon

    Sunnyvale, CA
    1 day ago
  • $192.2k - $260k

     ...Description Amazon Music is an immersive audio entertainment service that deepens...  ...team is seeking an experienced Applied Scientist who will join a team of experts in the field...  ...environment where you can pursue applied research, with many peta-bytes of data, work on problems... 
    Audio
    Local area
    Flexible hours

    Amazon

    Sunnyvale, CA
    4 days ago
  • $174k - $252k

    Snapshot At Google DeepMind, our research team is dedicated to tackling the most...  ...flagship models, including Gemini.About us Artificial...  ...DeepMind, we’re a team of scientists, engineers, machine learning...  ...trustworthiness of media (images, audio, and videos) on the... 
    Audio
    Full time

    DeepMind

    Mountain View, CA
    1 day ago
  • $221.35k - $253k

    Build and enhance serving solutions for Gemini models, tailoring configurations to meet diverse...  ...as large-scale streaming and specialized audio logic within the orchestration framework;...  ...transformative inventions. At Google DeepMind, we are a pioneering AI lab with... 
    Audio
    Full time
    Temporary work
    Work at office

    Google

    Mountain View, CA
    2 days ago
  • $147k - $211k

     ..., debugging).Experience with core GenAI concepts (LLM, Multi-Modal, Large Vision Models) and experience with text, image, video, or audio generation.Preferred qualifications:Master's degree or PhD in Computer Science, or a related technical field.Familiar with modern Large... 
    Audio

    Google

    Sunnyvale, CA
    2 days ago
  • $307k - $427k

     ...Experience analyzing the challenges of dual-use research of concern (DURC), biological weapons...  ...transformative inventions. At Google DeepMind, we are a pioneering AI lab with...  ...innovation. Partner closely with Research Scientists, Engineers, and ethics experts to embed... 

    Socket

    Mountain View, CA
    5 days ago
  • $207k - $301k

     ..., including Android ML, ML Compiler, and DeepMind, to co-design performance and evaluation...  ...with one or more of the following: Speech/audio (e.g., technology duplicating and responding...  ...device deployment of key models, such as Gemini Nano and Gemma, across various... 
    Audio
    Shift work

    Google

    Sunnyvale, CA
    2 days ago
  • $252k - $274k

     ...generative AI. Move beyond tactical testing to lead foundational research that informs the multi-year product roadmap.Define the signals...  ...be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary... 

    Google

    Mountain View, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Scientist, Gemini Audio i18n, DeepMind. Be the first to apply!