Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Audio AI Researcher: Multimodal & Low-Latency

$350k

Thinkingmachines

Thinking Machines in San Francisco is seeking researchers to advance audio capabilities in AI. This role combines fundamental research with practical engineering, where you'll collaborate with top-tier researchers and engineers. The ideal candidate possesses a strong background in machine learning and proficiency in Python. You will significantly shape AI's foundational abilities, ensuring effective communication and collaboration through innovative audio models. The position offers a salary range of $350,000 - $475,000 annually, along with generous benefits. #J-18808-Ljbffr Thinkingmachines

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Audio AI Researcher: Multimodal & Low-Latency in San Francisco, CA vacancy
  •  ...Machines Lab in San Francisco is actively researching audio capabilities, blending rigorous...  ...with practical engineering to build multimodal AI systems. You will work across pre-training...  ...audio with high fidelity and low latency. The role requires deep experimentation... 
    Audio

    Mosaic.tech

    San Francisco, CA
    5 days ago
  • $204k - $300k

     ...Advanced Technology Group (ATG) is the research division of the company. ATG’s...  ...electrical engineering, such as AI/ML, algorithms, digital signal processing, audio engineering, image processing,...  ...a senior research leader in the Multimodal Experiences Lab, you will shape the... 
    Audio
    Full time
    Local area
    Worldwide
    Flexible hours

    Dolby Laboratories, Inc.

    Brisbane, CA
    2 days ago
  • OpenAI is at the center of high-impact multimodal AI. The Chat and Multimodal Safety team builds safe, scalable models and evaluations for text, vision, and audio tasks. As a Researcher on the Chat and Multimodal Safety team in San Francisco, you will shape model perception... 
    Audio
    Work at office
    Relocation package

    United States Digital Space LLC

    San Francisco, CA
    5 days ago
  •  ...day. We’re hiring an AI Researcher to build the next generation...  ...timing, prosody, and latency. Your work runs the...  .... At our core is a low-latency, high-accuracy...  ...language processing, multimodal AI, human‑computer interaction...  ...modeling Knowledge of audio tokenization, neural... 
    Audio

    Kotoba

    San Francisco, CA
    5 days ago
  •  ...About the Team API Multimodal builds the developer-facing...  ...bring OpenAI’s image, audio, and real-time model capabilities...  ...speech generation, and low-latency voice interactions. We partner closely with Research and Inference to bring...  ...help make multimodal AI useful at scale. Model... 
    Audio
    Full time
    Internship

    OpenAI

    San Francisco, CA
    1 day ago
  • We are seeking an Edge AI Research Scientist to develop next-generation speech and audio AI systems that run efficiently on smartphones...  ...of operating under strict latency, memory, and power constraints....  ...state-of-the-art approaches for: Low-rank adaptation and compression... 
    Audio

    Huxley

    San Francisco, CA
    3 days ago
  •  ...center of some of the highest-impact multimodal work in AI. ChatGPT serves a massive global audience...  ...these experiences. We develop the research, training methods, and evaluations needed...  ...advanced safety for image, video, or audio systems. In this role, you will: Define... 
    Audio
    Work at office
    Relocation package

    United States Digital Space LLC

    San Francisco, CA
    5 days ago
  •  ...building the human layer of AI. Our mission is to...  ...through pioneering research in multimodal AI for modeling human-...  ...communication (language, audio, and video), as well...  ...benchmark trade-offs across latency, cost, and quality...  ...architectures such as low-rank adapters Strong... 
    Audio
    Remote work
    Relocation package
    Flexible hours

    Neura Market

    San Francisco, CA
    3 days ago
  •  ...As a Data Engineer - Multimodal Systems , you will be...  ...variety of modalities (text, audio, image) Designing and...  ...others in a high-paced research setting Can rapidly...  ...the bar to impact as low as possible We all...  ...do and love discussing AI Benefits and Perks... 
    Audio
    Work at office
    Relocation package

    Zyphra

    San Francisco, CA
    more than 2 months ago
  • AI Researcher (Computer Vision/Multimodal/Generative AI) About the Role We are hiring ML Researchers to develop novel approaches that advance the frontier of multimodal vision AI and create product-defining capabilities for SpreeAI. This role exists because current generative... 

    SpreeAI

    San Francisco, CA
    5 days ago
  • $84.13 - $91.34 per hour

    AI Researcher - Efficient AI (Contractor) Step into the innovative world...  ...that make modern LLMs, VLMs, multimodal models, and AI agents faster,...  ...methods (PTQ, QAT, pruning, low-rank approximation, etc) for...  ...decoding, constrained decoding, low-latency generation, and kernel-level... 
    Full time
    Contract work
    Temporary work
    For contractors
    Local area
    Immediate start

    LG Electronics

    San Francisco, CA
    5 days ago
  • An innovative AI company based in California is seeking an experienced AI Researcher focusing on computer vision and multimodal AI. The candidate will develop novel architectures that improve various aspects of generative models, with direct implications for real-world... 

    SpreeAI

    San Francisco, CA
    5 days ago
  • As a Research Scientist , you'll lead cutting-edge research that advances the state of generative AI for long-form storytelling. You'll work at the intersection of LLMs, multimodal AI, agentic systems, and scalable machine...  ...unify text, speech, audio, and other modalities... 
    Audio
    Worldwide

    Pocket FM

    San Francisco, CA
    3 days ago
  • $117.2k - $313.7k

    Salesforce AI Research is looking for outstanding AI Research Scientists and Research Engineers...  ...research agents, and coding agents. Multimodal and Computer Vision: Vision‑language...  ...TTS/ASR), human‑like turn‑taking, and low‑latency audio processing. Core Modeling and Post‑... 
    Audio

    salesforce.com, inc.

    San Francisco, CA
    2 days ago
  • $114.2k - $306.6k

     ...duplicating efforts.*Salesforce Research advances state-of-the-art AI techniques, developing models and...  ...learning, and autonomous workflows.* **Multimodal & Computer Vision:** Vision-...  ...ASR, human-like turn-taking, and low-latency audio processing.* **Efficient Systems:... 
    Audio
    Full time

    Niebles

    San Francisco, CA
    4 days ago
  •  ...we partner closely with Research to bring the next generation...  ...the boundaries of what AI can do. We’re expanding into multimodal inference, building the...  ...that handle image, audio, and other non-text modalities...  ...for high-throughput, low-latency delivery of image and audio... 
    Audio
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  • Fearn is seeking an experienced AI researcher to join our team in San Francisco. You will train models, design inference pipelines, and push...  ...teams to deploy production-ready solutions, optimize for latency and cost, and contribute to scalable AI infrastructure with a... 

    Kindredventures

    San Francisco, CA
    1 day ago
  •  ...Applied Scientist / Research Engineer — Speech AI   HIGHLIGHTS Location: Â...  ...modern machine learning for audio, speech, and language, and...  ...output quality while keeping latency low. Improve performance...  ..., generative modeling, multimodal systems, or large-scale model... 
    Audio
    Remote work

    GTN Technical Staffing

    San Francisco, CA
    more than 2 months ago
  •  ...including GPT-5 and a growing set of multimodal capabilities across text, image, audio, and video. Our team also...  ...These systems are a combination of low-latency, high scale, high reliability,...  ...About OpenAI OpenAI is an AI research and deployment company dedicated... 
    Audio
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  •  ...and implementing infrastructure for large-scale multimodal models, focusing on high-performance delivery of audio and image inputs. You'll collaborate closely with researchers and product teams to push the boundaries of AI technology, ensuring reliable production... 
    Audio

    Jobleads-US

    San Francisco, CA
    5 days ago
  •  ...a Google Deepmind veteran behind Project Astra, and top-tier AI researchers. As an early member of this team, you will have significant...  ...You have hands-on experience in training text-to-motion or audio-to-motion models. You might have previously published papers... 
    Audio
    Work at office
    Visa sponsorship

    Spellbrush

    San Francisco, CA
    2 days ago
  • $206.3k - $388k

     ...architect and scale the multimodal data processing...  ...models (image, video, audio). In this role, you’ll...  ...quantization, serving, throughput/latency tradeoffs) Experience...  ...across the full stack, from low-level systems and GPU...  ...impact, powered by AI and driven by human ingenuity... 
    Audio
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    5 days ago
  • $350k

    A leading AI research company in San Francisco is hiring for a position on their Audio team. The role involves developing and training advanced audio models, optimizing performance, and working collaboratively across teams. Ideal candidates will have strong expertise in... 
    Audio
    Flexible hours

    Jobleads-US

    San Francisco, CA
    3 days ago
  • Huxley in San Francisco is seeking an Edge AI Research Scientist to advance speech and audio AI for devices with limited resources. You will design compact models, compress and optimize inference, and enable real-time speech experiences directly on phones and wearables.... 
    Audio

    Huxley

    San Francisco, CA
    3 days ago
  • A technology firm in San Francisco is seeking a Senior Applied Researcher in Audio Understanding to tackle complex audio perception tasks. You will lead projects that push the boundaries of traditional speech recognition, emphasizing large-scale model development and innovative... 
    Audio
    Relocation package

    Cartesia

    San Francisco, CA
    1 day ago
  •  ...Engineer in San Francisco to build and scale AI systems used daily by many users. You will own ML projects end-to-end, from research through deployment and iteration, and develop multimodal models across video, text, images, and audio. You’ll work with LLMs and external AI... 
    Audio

    Twenty80 llc

    San Francisco, CA
    2 days ago
  • $342k - $399k

     ...Team The Future of Computing Research team is an applied research team...  ...Devices group. We study how AI systems perceive people and their...  ...in it. The role focuses on multimodal perception and authentication,...  ...authentication methods across visual, audio, and other sensing signals.... 
    Audio
    Full time
    Work at office
    Relocation package
    3 days per week

    OpenAI

    San Francisco, CA
    1 day ago
  •  ...improve dataset quality. You'll work directly with frontier AI labs to tackle challenging multimodal data problems, fine-tune models, build evaluation...  ...measurable improvements in dataset quality across video, audio, images, and text. This is an onsite role requiring... 
    Audio
    Full time

    Sieve, Inc.

    San Francisco, CA
    1 day ago
  • $320k

     ...interpretable, and steerable AI systems. We want AI to...  ...group of committed researchers, engineers, policy...  ...that bring Anthropic's audio models from research...  ...intersection of real-time media, low-latency inference, and...  ...Interpretability, Multimodal Neurons, Scaling Laws,... 
    Audio
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    1 day ago
  • $171.2k - $214k

    Scale is the leading AI data foundry, helping fuel the most exciting...  ...an AI Product Manager to support Multimodal & Coding AI data verticals, including audio, image, video, and world models....  ...across operations, engineering, research, and go-to-market teams to ensure... 
    Audio
    Full time

    Scale AI

    San Francisco, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Audio AI Researcher: Multimodal & Low-Latency. Be the first to apply!