Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Audio AI Researcher: Multimodal, Low-Latency Modeling

Mosaic.tech

Thinking Machines Lab in San Francisco is actively researching audio capabilities, blending rigorous theory with practical engineering to build multimodal AI systems. You will work across pre-training, post-training, and product to develop models that understand and generate audio with high fidelity and low latency. The role requires deep experimentation, code writing, and collaboration with researchers, engineers, and designers to push the foundations of how AI learns and communicates. #J-18808-Ljbffr Mosaic.tech

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Audio AI Researcher: Multimodal, Low-Latency Modeling in San Francisco, CA vacancy
  • OpenAI is at the center of high-impact multimodal AI. The Chat and Multimodal Safety team builds safe, scalable models and evaluations for text, vision, and audio tasks. As a Researcher on the Chat and Multimodal Safety team in San Francisco, you will shape model perception... 
    Audio
    Work at office
    Relocation package

    United States Digital Space LLC

    San Francisco, CA
    5 days ago
  • This is a job that Jill, our AI Recruiter, is recruiting for on behalf of one of...  ...is to speak to Jack. Job Title AI Researcher (Multimodal Audio/Video Generation) Salary Not Disclosed...  ...humans. You will design diffusion-based models for high-fidelity talking heads and neural... 
    Audio

    Jack & Jill

    San Francisco, CA
    2 days ago
  • Kotoba's speech models are licensed to Fortune...  .... We're hiring an AI Researcher to build the next...  ..., prosody, and latency. Your work runs the...  ...At our core is a low-latency, high-accuracy...  ...processing, multimodal AI, human-computer...  ...modeling Knowledge of audio tokenization,... 
    Audio

    Kotoba

    San Francisco, CA
    2 days ago
  • $204k - $300k

     ...Advanced Technology Group (ATG) is the research division of the company. ATG’s...  ...electrical engineering, such as AI/ML, algorithms, digital signal processing, audio engineering, image processing,...  ...a senior research leader in the Multimodal Experiences Lab, you will shape the... 
    Audio
    Full time
    Local area
    Worldwide
    Flexible hours

    Dolby

    San Francisco, CA
    5 days ago
  •  ...and implementing infrastructure for large-scale multimodal models, focusing on high-performance delivery of audio and image inputs. You'll collaborate closely with researchers and product teams to push the boundaries of AI technology, ensuring reliable production... 
    Audio

    Jobleads-US

    San Francisco, CA
    5 days ago
  •  ...About the Team API Multimodal builds the developer...  ...bring OpenAI’s image, audio, and real-time model capabilities into...  ...speech generation, and low-latency voice interactions....  ...closely with Research and Inference to bring...  ...help make multimodal AI useful at scale. Model... 
    Audio
    Full time
    Internship

    OpenAI

    San Francisco, CA
    1 day ago
  •  ...some of the highest-impact multimodal work in AI. ChatGPT serves a massive global...  ...interactive surfaces grow, models also need to adapt to...  ...experiences. We develop the research, training methods, and evaluations...  ...safety for image, video, or audio systems. In this role, you... 
    Audio
    Work at office
    Relocation package

    United States Digital Space LLC

    San Francisco, CA
    5 days ago
  • A cutting-edge AI startup is searching for an experienced AI Researcher eager to advance generative AI. This role requires a PhD and 5+ years of research experience, focusing on developing innovative models that harness earth observation. The ideal candidate will demonstrate... 
    Remote job

    hum.ai

    San Francisco, CA
    4 days ago
  •  ...based in San Francisco is seeking a Machine Learning Researcher to enhance K-12 education through AI. This role combines advanced technical skills with a...  ...Candidates should possess expertise in generative AI models, Python programming, and have a graduate degree in a... 
    Flexible hours

    Kiddom

    San Francisco, CA
    3 days ago
  •  ...person adventure game where an AI companion is the core of the...  ...Project Astra, and top-tier AI researchers. As an early member of this...  ...best image and video generation model teams in the world on data...  ...in training text-to-motion or audio-to-motion models. You might have... 
    Audio
    Work at office
    Visa sponsorship

    Spellbrush

    San Francisco, CA
    19 days ago
  • $117.2k - $313.7k

     ...AI Research Scientist And Research Engineer Salesforce is the...  ...research problems, develops novel models, and bridges the gap...  ...agents, and coding agents. Multimodal and Computer Vision: Vision...  ...-like turn-taking, and low-latency audio processing. Core Modeling... 
    Audio

    Salesforce

    San Francisco, CA
    3 days ago
  • An innovative AI company based in California is seeking an experienced AI Researcher focusing on computer vision and multimodal AI. The candidate will develop novel architectures that improve various aspects of generative models, with direct implications for real-world... 

    SpreeAI

    San Francisco, CA
    5 days ago
  • $84.13 - $91.34 per hour

    AI Researcher - Efficient AI (Contractor) Step into the innovative world...  ...make modern LLMs, VLMs, multimodal models, and AI agents faster, smaller...  ...methods (PTQ, QAT, pruning, low-rank approximation, etc) for...  ..., constrained decoding, low-latency generation, and kernel-level... 
    Full time
    Contract work
    Temporary work
    For contractors
    Local area
    Immediate start

    LG Electronics

    San Francisco, CA
    5 days ago
  • AI Researcher (Computer Vision/Multimodal/Generative AI) About the Role We are hiring ML Researchers to develop novel approaches that advance the frontier...  ...This role exists because current generative and vision models are not designed for photorealistic human... 

    SpreeAI

    San Francisco, CA
    5 days ago
  • Fearn is seeking an experienced AI researcher to join our team in San Francisco. You will train models, design inference pipelines, and push the frontier of visual-language...  ...production-ready solutions, optimize for latency and cost, and contribute to scalable AI... 

    Kindredventures

    San Francisco, CA
    1 day ago
  • $114.2k - $306.6k

     ...duplicating efforts.*Salesforce Research advances state-of-the-art AI techniques, developing models and prototypes that pave the...  ...autonomous workflows.* **Multimodal & Computer Vision:** Vision-...  ...human-like turn-taking, and low-latency audio processing.* **Efficient... 
    Audio
    Full time

    Niebles

    San Francisco, CA
    4 days ago
  • $206.3k - $388k

     ...architect and scale the multimodal data processing...  ...multimodal foundation models (image, video, audio). In this role, you’ll...  ...serving, throughput/latency tradeoffs) Experience...  ...the full stack, from low-level systems and GPU...  ...into impact, powered by AI and driven by human... 
    Audio
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    5 days ago
  •  ...our most advanced models - including our GPT...  ...partner closely with Research to bring the next...  ...boundaries of what AI can do. We’re expanding into multimodal inference, building...  ...that handle image, audio, and other non-text...  ...for high-throughput, low-latency delivery of image... 
    Audio
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  • About Luma AI Luma’s mission is to build multimodal AGI. Through our research on video, 3D, and now multimodal models at Luma, we believe that AI needs to be jointly trained over all signal modalities - text, video, audio, images - analogous to the human brain. To advance... 
    Audio
    Worldwide

    Socket.dev

    San Francisco, CA
    1 day ago
  •  ...the OpenAI API, powering models including GPT-5 and a growing set of multimodal capabilities across text, image, audio, and video. Our team also...  ...systems are a combination of low-latency, high scale, high...  ...About OpenAI OpenAI is an AI research and deployment company dedicated... 
    Audio
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  • $350k

    A leading AI research company in San Francisco is hiring for a position on their Audio team. The role involves developing and training advanced audio models, optimizing performance, and working collaboratively across teams. Ideal candidates will have strong expertise in... 
    Audio
    Flexible hours

    Jobleads-US

    San Francisco, CA
    2 days ago
  •  ...technology firm in San Francisco is seeking a Senior Applied Researcher in Audio Understanding to tackle complex audio perception tasks. You will...  ...of traditional speech recognition, emphasizing large-scale model development and innovative research. The ideal candidate will... 
    Audio
    Relocation package

    Cartesia

    San Francisco, CA
    1 day ago
  •  ...dataset quality. You'll work directly with frontier AI labs to tackle challenging multimodal data problems, fine-tune models, build evaluation systems, and deliver...  ...improvements in dataset quality across video, audio, images, and text. This is an onsite role requiring... 
    Audio
    Full time

    Sieve, Inc.

    San Francisco, CA
    1 day ago
  •  ...Solutions Pvt. Ltd. is seeking an Applied Research Engineer to design scalable...  ...understanding. You will work on multimodal AI applications, including CV, audio, and NLP tasks, building production...  ...performance, leverage foundation models and external APIs, and translate customer... 
    Audio

    Guidant Solutions Pvt. Ltd.

    San Francisco, CA
    1 day ago
  •  ...As a Data Engineer - Multimodal Systems , you will be...  ...variety of modalities (text, audio, image) Designing and...  ...others in a high-paced research setting Can rapidly...  ...the bar to impact as low as possible We all...  ...do and love discussing AI Benefits and Perks... 
    Audio
    Work at office
    Relocation package

    Zyphra

    San Francisco, CA
    more than 2 months ago
  • $320k

     ...interpretable, and steerable AI systems. We want AI...  ...group of committed researchers, engineers, policy...  ...bring Anthropic's audio models from research into production...  ...of real-time media, low-latency inference, and...  ...Based Interpretability, Multimodal Neurons, Scaling Laws... 
    Audio
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    1 day ago
  • LG Electronics is seeking a Contract AI Researcher focusing on Efficient AI in Santa Clara, CA, hybrid work arrangement. You will explore model compression, quantization, efficient inference, and architectures to make LLMs/VLMs faster and more deployable on devices. You... 
    Contract work

    LG Electronics

    San Francisco, CA
    5 days ago
  •  ...financial services firm in San Francisco is seeking an Applied Researcher II to develop innovative AI systems. In this role, you will collaborate with a cross-functional team to build AI foundation models, engage in impactful research, and translate complex work into business... 

    Capital One

    San Francisco, CA
    1 day ago
  • $171.2k - $214k

    Scale is the leading AI data foundry, helping fuel...  ...AI, including frontier model training, enterprise adoption...  ...Manager to support Multimodal & Coding AI data verticals, including audio, image, video, and world...  ...operations, engineering, research, and go-to-market teams... 
    Audio
    Full time

    Scale AI

    San Francisco, CA
    5 days ago
  • Capital One is seeking an Applied Researcher I in San Francisco to join the AI Foundations team. You will collaborate with data scientists, software engineers...  ...platforms. The role emphasizes applied research, model development, and translating research into business outcomes... 

    Capital One

    San Francisco, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Audio AI Researcher: Multimodal, Low-Latency Modeling. Be the first to apply!