Audio AI Researcher: Multimodal & Low-Latency
$350kThinkingmachines
Thinking Machines in San Francisco is seeking researchers to advance audio capabilities in AI. This role combines fundamental research with practical engineering, where you'll collaborate with top-tier researchers and engineers. The ideal candidate possesses a strong background in machine learning and proficiency in Python. You will significantly shape AI's foundational abilities, ensuring effective communication and collaboration through innovative audio models. The position offers a salary range of $350,000 - $475,000 annually, along with generous benefits. #J-18808-Ljbffr Thinkingmachines
- ...Machines Lab in San Francisco is actively researching audio capabilities, blending rigorous... ...with practical engineering to build multimodal AI systems. You will work across pre-training... ...audio with high fidelity and low latency. The role requires deep experimentation...Audio
$204k - $300k
...Advanced Technology Group (ATG) is the research division of the company. ATG’s... ...electrical engineering, such as AI/ML, algorithms, digital signal processing, audio engineering, image processing,... ...a senior research leader in the Multimodal Experiences Lab, you will shape the...AudioFull timeLocal areaWorldwideFlexible hours- OpenAI is at the center of high-impact multimodal AI. The Chat and Multimodal Safety team builds safe, scalable models and evaluations for text, vision, and audio tasks. As a Researcher on the Chat and Multimodal Safety team in San Francisco, you will shape model perception...AudioWork at officeRelocation package
- ...day. We’re hiring an AI Researcher to build the next generation... ...timing, prosody, and latency. Your work runs the... .... At our core is a low-latency, high-accuracy... ...language processing, multimodal AI, human‑computer interaction... ...modeling Knowledge of audio tokenization, neural...Audio
- ...About the Team API Multimodal builds the developer-facing... ...bring OpenAI’s image, audio, and real-time model capabilities... ...speech generation, and low-latency voice interactions. We partner closely with Research and Inference to bring... ...help make multimodal AI useful at scale. Model...AudioFull timeInternship
- We are seeking an Edge AI Research Scientist to develop next-generation speech and audio AI systems that run efficiently on smartphones... ...of operating under strict latency, memory, and power constraints.... ...state-of-the-art approaches for: Low-rank adaptation and compression...Audio
- ...center of some of the highest-impact multimodal work in AI. ChatGPT serves a massive global audience... ...these experiences. We develop the research, training methods, and evaluations needed... ...advanced safety for image, video, or audio systems. In this role, you will: Define...AudioWork at officeRelocation package
- ...building the human layer of AI. Our mission is to... ...through pioneering research in multimodal AI for modeling human-... ...communication (language, audio, and video), as well... ...benchmark trade-offs across latency, cost, and quality... ...architectures such as low-rank adapters Strong...AudioRemote workRelocation packageFlexible hours
- ...As a Data Engineer - Multimodal Systems , you will be... ...variety of modalities (text, audio, image) Designing and... ...others in a high-paced research setting Can rapidly... ...the bar to impact as low as possible We all... ...do and love discussing AI Benefits and Perks...AudioWork at officeRelocation package
- AI Researcher (Computer Vision/Multimodal/Generative AI) About the Role We are hiring ML Researchers to develop novel approaches that advance the frontier of multimodal vision AI and create product-defining capabilities for SpreeAI. This role exists because current generative...
$84.13 - $91.34 per hour
AI Researcher - Efficient AI (Contractor) Step into the innovative world... ...that make modern LLMs, VLMs, multimodal models, and AI agents faster,... ...methods (PTQ, QAT, pruning, low-rank approximation, etc) for... ...decoding, constrained decoding, low-latency generation, and kernel-level...Full timeContract workTemporary workFor contractorsLocal areaImmediate start- An innovative AI company based in California is seeking an experienced AI Researcher focusing on computer vision and multimodal AI. The candidate will develop novel architectures that improve various aspects of generative models, with direct implications for real-world...
- As a Research Scientist , you'll lead cutting-edge research that advances the state of generative AI for long-form storytelling. You'll work at the intersection of LLMs, multimodal AI, agentic systems, and scalable machine... ...unify text, speech, audio, and other modalities...AudioWorldwide
$117.2k - $313.7k
Salesforce AI Research is looking for outstanding AI Research Scientists and Research Engineers... ...research agents, and coding agents. Multimodal and Computer Vision: Vision‑language... ...TTS/ASR), human‑like turn‑taking, and low‑latency audio processing. Core Modeling and Post‑...Audio$114.2k - $306.6k
...duplicating efforts.*Salesforce Research advances state-of-the-art AI techniques, developing models and... ...learning, and autonomous workflows.* **Multimodal & Computer Vision:** Vision-... ...ASR, human-like turn-taking, and low-latency audio processing.* **Efficient Systems:...AudioFull time- ...we partner closely with Research to bring the next generation... ...the boundaries of what AI can do. We’re expanding into multimodal inference, building the... ...that handle image, audio, and other non-text modalities... ...for high-throughput, low-latency delivery of image and audio...AudioFull time
- Fearn is seeking an experienced AI researcher to join our team in San Francisco. You will train models, design inference pipelines, and push... ...teams to deploy production-ready solutions, optimize for latency and cost, and contribute to scalable AI infrastructure with a...
- ...Applied Scientist / Research Engineer â Speech AI Â HIGHLIGHTS Location: Â... ...modern machine learning for audio, speech, and language, and... ...output quality while keeping latency low. Improve performance... ..., generative modeling, multimodal systems, or large-scale model...AudioRemote work
- ...including GPT-5 and a growing set of multimodal capabilities across text, image, audio, and video. Our team also... ...These systems are a combination of low-latency, high scale, high reliability,... ...About OpenAI OpenAI is an AI research and deployment company dedicated...AudioFull time
- ...and implementing infrastructure for large-scale multimodal models, focusing on high-performance delivery of audio and image inputs. You'll collaborate closely with researchers and product teams to push the boundaries of AI technology, ensuring reliable production...Audio
- ...a Google Deepmind veteran behind Project Astra, and top-tier AI researchers. As an early member of this team, you will have significant... ...You have hands-on experience in training text-to-motion or audio-to-motion models. You might have previously published papers...AudioWork at officeVisa sponsorship
$206.3k - $388k
...architect and scale the multimodal data processing... ...models (image, video, audio). In this role, you’ll... ...quantization, serving, throughput/latency tradeoffs) Experience... ...across the full stack, from low-level systems and GPU... ...impact, powered by AI and driven by human ingenuity...AudioFull timeTemporary workLocal areaWorldwide$350k
A leading AI research company in San Francisco is hiring for a position on their Audio team. The role involves developing and training advanced audio models, optimizing performance, and working collaboratively across teams. Ideal candidates will have strong expertise in...AudioFlexible hours- Huxley in San Francisco is seeking an Edge AI Research Scientist to advance speech and audio AI for devices with limited resources. You will design compact models, compress and optimize inference, and enable real-time speech experiences directly on phones and wearables....Audio
- A technology firm in San Francisco is seeking a Senior Applied Researcher in Audio Understanding to tackle complex audio perception tasks. You will lead projects that push the boundaries of traditional speech recognition, emphasizing large-scale model development and innovative...AudioRelocation package
- ...Engineer in San Francisco to build and scale AI systems used daily by many users. You will own ML projects end-to-end, from research through deployment and iteration, and develop multimodal models across video, text, images, and audio. You’ll work with LLMs and external AI...Audio
$342k - $399k
...Team The Future of Computing Research team is an applied research team... ...Devices group. We study how AI systems perceive people and their... ...in it. The role focuses on multimodal perception and authentication,... ...authentication methods across visual, audio, and other sensing signals....AudioFull timeWork at officeRelocation package3 days per week- ...improve dataset quality. You'll work directly with frontier AI labs to tackle challenging multimodal data problems, fine-tune models, build evaluation... ...measurable improvements in dataset quality across video, audio, images, and text. This is an onsite role requiring...AudioFull time
$320k
...interpretable, and steerable AI systems. We want AI to... ...group of committed researchers, engineers, policy... ...that bring Anthropic's audio models from research... ...intersection of real-time media, low-latency inference, and... ...Interpretability, Multimodal Neurons, Scaling Laws,...AudioFull timeWork at officeVisa sponsorshipFlexible hours$171.2k - $214k
Scale is the leading AI data foundry, helping fuel the most exciting... ...an AI Product Manager to support Multimodal & Coding AI data verticals, including audio, image, video, and world models.... ...across operations, engineering, research, and go-to-market teams to ensure...AudioFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Audio AI Researcher: Multimodal & Low-Latency. Be the first to apply!
- postdoctoral researcher cosmetic science San Francisco, CA
- senior researcher San Francisco, CA
- researcher San Francisco, CA
- online researcher San Francisco, CA
- music researcher San Francisco, CA
- senior design researcher San Francisco, CA
- visiting researcher San Francisco, CA
- remote researcher San Francisco, CA
- product researcher San Francisco, CA
- security researcher San Francisco, CA



