Research Scientist, Gemini Audio i18n, DeepMind
$174k - $252kCreate comprehensive evaluation sets and benchmarks to measure audio-to-audio (A2A) model performance across international languages, accents, and regional dialects.Propose, prototype, and evaluate novel modeling techniques to improve A2A audio understanding, dialog, and audio generation capabilities with a focus on scalability.Identify performance gaps in current multilingual audio models and collaborate with cross-functional research and engineering teams to deploy solutions.Minimum qualifications:Bachelor's degree in Computer Science, Speech Recognition, Computational Linguistics, a related technical field, or equivalent practical experience.Experience conducting research or development in Speech Recognition, Text-to-Speech (TTS), or Large Language Models (LLMs).Experience coding in Python or C++ and using deep learning frameworks such as PyTorch, JAX, or TensorFlow.Experience working with audio data, speech processing, or multilingual datasets.Preferred qualifications:Master's degree or Ph.D. in Automatic Speech Recognition (ASR), Text-to-Speech (TTS), Computational Linguistics, or a related field.3 years of experience in Large Language Models (LLMs) or multimodal foundation models.Experience developing audio-to-audio (A2A) architectures, end-to-end speech models, or spoken dialog systems.Experience scaling speech models across international languages, accents, or low-resource locales.Publication record in speech or machine learning conferences (e.g., ICASSP, INTERSPEECH, NeurIPS, or ACL).As an organization, Google maintains a portfolio of research projects driven by fundamental research, new product innovation, product contribution and infrastructure goals, while providing individuals and teams the freedom to emphasize specific types of work. As a Research Scientist, you'll setup large-scale tests and deploy promising ideas quickly and broadly, managing deadlines and deliverables while applying the latest theories to develop new and improved products, processes, or technologies. From creating experiments and prototyping implementations to designing new architectures, our research scientists work on real-world problems that span the breadth of computer science, such as machine (and deep) learning, data mining, natural language processing, hardware and software performance analysis, improving compilers for mobile platforms, as well as core search and much more.As a Research Scientist, you'll also actively contribute to the wider research community by sharing and publishing your findings, with ideas inspired by internal projects as well as from collaborations with research programs at partner universities and technical institutes all over the world.Artificial intelligence will be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.We are pushing the boundaries across multiple domains. Our global teams offer diverse learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort.Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $174000 - $252000 (USD) + 15% bonus target + equity + benefitsLearn more about benefits at Google.Note: By applying to this position you will have an opportunity to share your preferred working location from the following: Mountain View, CA, USA; New York, NY, USA.Bachelor's degree in Computer Science, Speech Recognition, Computational Linguistics, a related technical field, or equivalent practical experience.Experience conducting research or development in Speech Recognition, Text-to-Speech (TTS), or Large Language Models (LLMs).Experience coding in Python or C++ and using deep learning frameworks such as PyTorch, JAX, or TensorFlow.Experience working with audio data, speech processing, or multilingual datasets.
$174k - $252k
Research Scientist, Gemini Audio i18n, DeepMind Mountain View, CA, USA; New York, NY, USA Note: By applying to this position you will have an opportunity to share your preferred working location from the following: Mountain View, CA, USA; New York, NY, USA . Minimum qualifications...Audio- ...useful inventions. At Google DeepMind, we’re a team of scientists, engineers, machine... ...the highest priority.The Gemini Safety team is accountable... ...Gemini models. The role of the Research Scientist / Research Engineer... ...text-to-text, image/video/audio-to-text modalities and...Audio
$207k - $300k
Develop audio-first models capable of orchestrating and planning... ...AI.Experience in an applied research setting.Experience working with... ...types of work. As a Research Scientist, you'll setup large-scale... ...transformative inventions. At Google DeepMind, we are a pioneering AI lab...Audio- Google DeepMind in Mountain View, CA seeks a Research Scientist for Gemini Audio i18n to advance multilingual audio models. You will design large-scale experiments, prototype architectures, and evaluate A2A speech systems across languages. You will publish findings and...Audio
$174k - $252k
Research Scientist, Frontier Health, DeepMind DeepMind Mountain View, CA, USA Minimum qualifications: PhD degree in Computer Science, a related field, or... ...tool use, and multimodal understanding (text, image, audio/video). Information collected and processed as part of...AudioFull time$147k - $211k
SnapshotWe are seeking strong Research Scientists with expertise in AI... ...research effort within Google DeepMind's Frontier AI unit. This role... ...breakthrough improvements to models (Gemini) and products. Our work is... ...data types (e.g., vision, audio, text).The US base salary...AudioFull time$141k - $244k
...Research Scientist - Gemini Personal Intelligence Mountain View, California, US At Google DeepMind, we value diversity of experience, knowledge, backgrounds and perspectives and harness... ...across diverse modalities (Text, Image, Audio, Video). Agency: Empowering AI models to...AudioFull timeRelocationFlexible hoursShift work$207k - $300k
...misbehavior and misuse end-to-end.Research and develop cross-context... ...teams and data scientists to scale your work and regularly... ...training and evaluation for Gemini and GenMedia models. The Generative... ...transformative inventions. At DeepMind, we are a pioneering AI lab with...$174k - $253k
...visual, language, and multimodal research. 1 year of experience... ...types of work. As a Research Scientist, you'll setup large-scale tests... ...transformative inventions. At Google DeepMind, we are a pioneering AI lab... ...research ideas into Gemini production codebase by solving...$174k - $252k
Be a part of team of scientists, engineers, machine learning experts,... ...experience in deep learning research and development, including generative AI, audio and video synthesis, diffusion... ...that is fundamental to Google DeepMind’s work on Gemini’s audio capabilities, on...Audio$207k - $300k
Partner with the Gemini or DeepMind teams to design, develop, and deploy novel multimodal conversational agents.Develop audio-first models capable of orchestrating and planning complex... ....g., NeurIPS, ICML, ICLR, AAAI, CVPR).Research background in NLP/Generative AI.At Google...Audio- Google DeepMind is seeking a Research Scientist to design and deploy large-scale experiments in speech, audio, and multilingual models. You will prototype architectures, run benchmarks, and publish findings with internal and external collaborators. The role emphasizes...Audio
$323k
...SnapshotAt Google DeepMind, we’re a team of scientists, engineers, machine learning experts... ...ambitious goals.The Area: Gemini AppThe Gemini App is at... ...management, engineering, and research teams to deliver an... ...deterministic AI outputs (streaming audio/visuals) with zero jitter...AudioFull time$174k - $252k
Conduct fundamental and applied ML research to develop physiological and behavioral world... ...emphasize specific types of work. As a Research Scientist, you'll setup large-scale tests and... ...’s most transformative inventions. At DeepMind, we are a pioneering AI lab with...InternshipShift work$185k - $400k
...intelligent agentic platforms. We are looking for a staff or lead-level Research Engineer, Data to architect and scale data engineering systems... ...support model training and research workflows for text, image, audio, and video datasetsPartner with research and engineering teams...AudioRemote work- Google Research is seeking a Senior Research Scientist to drive bold research initiatives in Gemini and related areas in Mountain View. You will design experiments, prototype architectures, and evaluate large language model capabilities, collaborating with teams across...
$174k - $252k
..., Flax, or Gemma).Experience conducting research and development, including experimental... ...from different data types (e.g., vision, audio, text).Proven expertise in working with,... ...emphasize specific types of work. As a Research Scientist, you'll setup large-scale tests and...Audio$200k - $240k
...Archetype AI is hiring a Senior AI Research Scientist as a backfill for Jamie. You will join the core research team and work on machine learning... ...data or other non-image data. ~ Experience in robotics, audio, or another physical-world ML domain. ~ A track record of taking...AudioWork at office2 days per week3 days per week$117.2k - $313.7k
...Salesforce.The ExperienceSalesforce AI Research is a global leader in Enterprise AI, driving... ...looking for entrepreneurial Research Scientists who want to build, ship, and scale the next... ...speech recognition, and low-latency audio processing.Core Modeling and Post-Training...AudioFull timeWorldwide$185k - $400k
...and intelligent agentic platforms. We are seeking accomplished Research Scientists in Foundation Models with expertise in pre-training and mid-... ...-scale multimodal pre-training/mid-training (text, image, audio, and video), and drive innovative approaches for foundational...AudioRemote work$262k - $364k
...software engineering skills to complement the research background.Build recommender/search... ...specific types of work. As a Research Scientist, you'll setup large-scale tests and deploy... ...transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional...$300k - $333k
...data to test hypotheses.Manage various stakeholders, including DeepMind Research Engineers for technical delivery, clients for project... ...(fine-tuning, agent harnesses), generative image, video, and audio models, and technical software development to build products...Audio$165k - $195k
Company DescriptionThe Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania... ...and transparent AI & big data analytic solutions (e.g., audio, images, sensor logs) for a range of domains, including Industry...AudioFull timeWork experience placementLocal areaWorldwide$165k - $185k
Company DescriptionThe Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania... ...and transparent AI & big data analytic solutions (e.g., audio, images, sensor logs) for a range of domains, including Industry...AudioWork experience placementWorldwide$158.8k - $218.1k
...medical devices, AI, and health services. We research, develop, and commercialize innovative... ...Ph.D.-level Senior, Applied AI Health Scientist with solid ML/AI technology background... ...signals from sensors, including light, audio, physiological, and inertial data collected...AudioContract workFor contractorsFor subcontractorWork at officeLocal area- Bosch USA in Sunnyvale, CA, is seeking a senior AI researcher to lead projects on multi-sensor data fusion, time-series foundation models... ...You will design, implement, and validate advanced models across audio, radar, lidar, and video data, collaborating with teams in AI...Audio
$117.2k - $313.7k
About the Role Salesforce AI Research is seeking outstanding AI Research Scientists / Research Engineers to build and deploy high‑impact AI solutions at scale.... ...intelligence, TTS/ASR, human‑like turn‑taking, low‑latency audio processing Efficient Systems - Scalable deep...AudioFull time$117.2k - $313.7k
Salesforce AI Research is looking for outstanding AI Research Scientists and Research Engineers to discover new research problems, develop novel models, and bridge... ...(TTS/ASR), human‑like turn‑taking, and low‑latency audio processing. Core Modeling and Post‑Training:...Audio- ...communication. Founded by a team of Stanford researchers and entrepreneurs with deep industry... .... We're looking for a Research Scientist who can define what "better" actually means... ...applied research experience in speech, audio, or NLP, with a demonstrated focus on evaluation...Audio
$192k - $304.75k
We are now looking for a Senior Research Scientist for Generative AI!NVIDIA is searching for a world-class researcher in generative AI to join... ...image generation, video generation, 3D generation, and audio generation. You will be working with a team of world-class researchers...AudioFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist, Gemini Audio i18n, DeepMind. Be the first to apply!
- molecular biology scientist Mountain View, CA
- water quality scientist Mountain View, CA
- machine learning scientist Mountain View, CA
- image scientist Mountain View, CA
- machine learning research scientist Mountain View, CA
- research associate scientist Mountain View, CA
- health scientist Mountain View, CA
- scientist Mountain View, CA
- quality control scientist Mountain View, CA
- scientist biology Mountain View, CA


