Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Audio AI Researcher: Build Steerable Speech Systems

$350k

Jobleads-US

A leading AI research company in San Francisco is hiring for a position on their Audio team. The role involves developing and training advanced audio models, optimizing performance, and working collaboratively across teams. Ideal candidates will have strong expertise in JAX or PyTorch, a background in audio machine learning, and a passion for creating safe, reliable AI systems. The compensation ranges from $350,000 to $500,000, and the position offers flexible working hours and a supportive team environment. #J-18808-Ljbffr Jobleads-US

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Audio AI Researcher: Build Steerable Speech Systems in San Francisco, CA vacancy
  •  ...revenue-generating voice AI company empowering enterprises to build AI phone agents at...  .... No speculative research track here. If you...  ...goal: a fully speech-to-speech conversational...  ...-to-speech, neural audio codecs, and getting...  ...human Build STT systems that stay accurate... 
    Audio
    Permanent employment
    Full time
    Immediate start

    DeepRec.ai

    San Francisco, CA
    3 days ago
  • Kotoba's speech models are licensed to Fortune 50 companies and...  ...users a day. We're hiring an AI Researcher to build the next generation of real-...  ...to-text, and text-to-speech systems — models that don't just...  ...lingual modeling Knowledge of audio tokenization, neural audio codecs... 
    Audio

    Kotoba

    San Francisco, CA
    2 days ago
  • Goaly is seeking an AI Researcher to lead research on agentic AI, focusing on training specialized models and building orchestration stacks. You'll design experiments, publish results, and advance the field. Ideal candidates have a Ph.D. or Master's in relevant fields,... 
    Suggested

    Goaly

    San Francisco, CA
    4 days ago
  • $150k - $250k

    About Distyl AI Distyl is an applied AI technology company partnering...  ...social organizations.We research and deploy technologies that power...  ...into self-constructing systems, the development of the most reliable...  ...System Self-Improvement team builds architectures that... 
    Suggested
    Work at office
    3 days per week

    Distyl AI

    San Francisco, CA
    3 days ago
  • Thinking Machines Lab in San Francisco is actively researching audio capabilities, blending rigorous theory with practical engineering to build multimodal AI systems. You will work across pre-training, post-training, and product to develop models that understand and generate... 
    Audio

    Mosaic.tech

    San Francisco, CA
    5 days ago
  •  ...impact multimodal work in AI. ChatGPT serves a...  ...interactions via text, speech, and visuals. As interactive...  .... We develop the research, training methods, and...  ...for image, video, or audio systems. In this role, you will...  ...a track record of building or advancing multimodal... 
    Audio
    Work at office
    Relocation package

    United States Digital Space LLC

    San Francisco, CA
    5 days ago
  •  ...Description tl;dr: We're building a 3D first-person...  ...game where an AI companion is the...  ...LLM storytelling system that blends AI,...  ..., and top-tier AI researchers. As an early member...  ...text-to-motion or audio-to-motion models....  ...have worked on speech-driven 3D facial animation... 
    Audio
    Work at office
    Visa sponsorship

    Spellbrush

    San Francisco, CA
    19 days ago
  • $180k - $270k

     ...Plaud Inc. Plaud is building the world's most trusted AI work companion for...  ...models or foundational speech models....  ...Token (or Time-To-First-Audio) in real-time streaming...  ...genuinely enjoy the systems-engineering challenge...  ...collaborative, fast-paced research. Gear & Perks:... 
    Audio
    Full time
    Work at office
    Worldwide

    Plaud

    San Francisco, CA
    1 day ago
  •  ...this. We're creating an AI-powered experience that replicates...  ...develop cutting-edge speech recognition models that...  ...Expanding our ASR systems to new languages and markets Building and maintaining data infrastructure...  ...with speech or audio Office ~ San Francisco... 
    Audio
    Full time
    Live in
    Work at office
    Worldwide

    Speak

    San Francisco, CA
    1 day ago
  • $215k - $290k

    A leading Voice AI startup is seeking a Senior Software Engineer - Backend to shape the core systems for millions of AI-driven conversations. This role involves working with real-time audio and global telephony, leading complex projects, and collaborating across teams.... 
    Audio

    Retell AI

    San Francisco, CA
    3 days ago
  • AI agents are changing how enterprises operate. Companies want to move...  ...engineers, shipping fast. We are building the infrastructure that will make...  ...how AI agents fail in real systems and turn that work into clear security research. Research agent security risks across... 

    CodeIntegrity

    San Francisco, CA
    2 days ago
  • Distyl in San Francisco is seeking an Applied AI Researcher for the System Self-Construction team. You will design architectures enabling autonomous...  ...research credentials, experience with meta-learning, and building prototypes to validate ideas in enterprise settings. #J-1... 
    Work at office
    3 days per week

    SupportFinity™

    San Francisco, CA
    2 days ago
  • Distyl AI is seeking researchers to push the boundaries of enterprise AI. You will explore novel AI system architectures, build prototypes, and demonstrate measurable improvements in real-world workflows. The role emphasizes AI-native thinking, cross-domain connections... 
    Work at office

    Distyl

    San Francisco, CA
    5 days ago
  • Axiom is building the first accurate AI systems to replace animal testing with human-relevant predictive models. You will work on end-to-end ML and agent systems across wet-lab data, multimodal inputs, and mechanistic reasoning to predict human toxicity. Join a founding... 

    Axiom

    San Francisco, CA
    3 days ago
  •  ...a Product Manager focused on AI speech (text-to-speech, speech-to-text...  ...Speech Evaluation Frameworks: Build and maintain comprehensive frameworks...  ...in a product role at a voice/audio software company (e.g....  ...language models (e.g. developing systems leveraging LLMs as a component... 
    Audio

    Artificial Analysis, Inc.

    San Francisco, CA
    5 days ago
  • $34 per hour

     ...join our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role focused on building the datasets that power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are... 
    Audio
    Work experience placement
    Remote work

    Welocalize

    San Francisco, CA
    2 days ago
  • $26 - $28 per hour

     ...join our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role focused on building the datasets that power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are... 
    Audio
    Full time
    Work experience placement
    Remote work
    Visa sponsorship

    Welocalize

    San Francisco, CA
    3 days ago
  • $26 - $28 per hour

     ...Finnish Data Labeling Analyst(Speech & Voice) Overview...  ..., supporting speech and voice AI systems. This is a high-impact production role focused on building the datasets that power real-world...  ...systems. You'll be working with audio, speech, and language data -... 
    Audio
    Full time
    Work experience placement
    Remote work
    Visa sponsorship

    Welocalize

    San Francisco, CA
    3 days ago
  • $204k - $300k

     ...Advanced Technology Group (ATG) is the research division of the company. ATG’s...  ...electrical engineering, such as AI/ML, algorithms, digital signal processing, audio engineering, image processing,...  ...science & analytics, distributed systems, cloud, edge & mobile computing,... 
    Audio
    Full time
    Local area
    Worldwide
    Flexible hours

    Dolby

    San Francisco, CA
    9 hours ago
  •  ...Applied Scientist / Research Engineer — Speech AI   HIGHLIGHTS Location:  San...  ...speech and conversational AI systems. This person will work...  ...modern machine learning for audio, speech, and language, and...  ...Speech Generation Systems Build and improve machine... 
    Audio
    Remote work

    GTN Technical Staffing

    San Francisco, CA
    a month ago
  • $150k - $250k

    Distyl AI is seeking creative researchers to join its Multi-Agent Systems team in San Francisco. This role requires expertise in designing architectures for multi-...  ...including investigating interaction patterns and building systems for structured communication. Offering a... 

    Distyl AI

    San Francisco, CA
    1 day ago
  •  ...technology firm in San Francisco is seeking a Senior Applied Researcher in Audio Understanding to tackle complex audio perception tasks. You will lead projects that push the boundaries of traditional speech recognition, emphasizing large-scale model development and innovative... 
    Audio
    Relocation package

    Cartesia

    San Francisco, CA
    1 day ago
  • OpenAI is at the center of high-impact multimodal AI. The Chat and Multimodal Safety team builds safe, scalable models and evaluations for text, vision, and audio tasks. As a Researcher on the Chat and Multimodal Safety team in San Francisco, you will shape model perception... 
    Audio
    Work at office
    Relocation package

    United States Digital Space LLC

    San Francisco, CA
    5 days ago
  • Cartesia is hiring for an Audio Post-Training role to build and improve capabilities that shape how users interact with our generative audio models. This team spans research to production, designing evaluations, data pipelines, finetuning, RL, and model evaluation. You... 
    Audio

    Cartesia

    San Francisco, CA
    4 days ago
  • This is a job that Jill, our AI Recruiter, is recruiting for on behalf of one of...  ...is to speak to Jack. Job Title AI Researcher (Multimodal Audio/Video Generation) Salary Not Disclosed...  ...avatar generation at a cutting-edge lab building real-time conversational humans. You will... 
    Audio

    Jack & Jill

    San Francisco, CA
    2 days ago
  •  ...Cartesia Our mission is to build the next generation of AI: ubiquitous,...  ...year-long stream of audio, video and text—1B text...  ...innovation and systems engineering paired with...  ...As a Senior Applied Researcher in Audio Understanding...  ...beyond traditional speech recognition to... 
    Audio
    Work at office
    Relocation package

    Cartesia

    San Francisco, CA
    3 days ago
  • Postdoctoral Researcher, Computer Vision (PhD), New Grad Join...  ...New Grad role at Jobright.ai Postdoctoral Researcher,...  ...interdisciplinary teams to build contextually aware AI systems. Responsibilities: •...  ...modalities (images, video, text, audio, speech and other modalities)... 
    Audio
    Full time
    Part time
    Work experience placement
    Internship

    Jobright.ai

    San Francisco, CA
    5 days ago
  •  .... locations. The role focuses on high-volume labeling and annotation of speech and language data to power AI systems. You will follow detailed guidelines, maintain data quality, and work with audio and language data, including transcription and tagging. Requirements include... 
    Audio

    Welocalize

    San Francisco, CA
    5 days ago
  •  ...mission is to architect AI that learns from...  ...innovation and systems engineering paired...  ...engineering team to build and ship cutting edge...  ...About the Role On the Audio Post-Training team,...  ...needs meet research, and covers the full...  ...generative models (speech, text, or multimodal... 
    Audio
    Work at office
    Visa sponsorship
    Flexible hours

    Cartesia

    San Francisco, CA
    4 days ago
  •  ...dataset quality. You'll work directly with frontier AI labs to tackle challenging multimodal data problems, fine-tune models, build evaluation systems, and deliver measurable improvements in dataset quality across video, audio, images, and text. This is an onsite role... 
    Audio
    Full time

    Sieve, Inc.

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Audio AI Researcher: Build Steerable Speech Systems. Be the first to apply!