Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Engineer - Audio & Speech Models

Zyphra

Job Description

Job Description

Zyphra is an artificial intelligence company based in San Francisco, California.

The Role:

As a Research Engineer - Audio & Speech Models , you will be a core contributor on Zyphra’s Audio Team, building the next generation of open-source autoencoders, ASR, TTS, SSL, and speech-to-speech models. You will be deeply involved in the entire model training process, from data gathering and processing to designing novel architectures and training methodologies.

You’ll Work Across:
  • Large-scale audio training runs

  • Performance optimization of our training stack

  • Audio dataset collection, processing, and evaluation

  • Architecture and training methodology ablations and improvements

What We're Looking For / Requirements:
  • Strong research taste and intuition. The ability to work through a research project from conception to execution to write-up.

  • Strong implementation and prototyping ability (can take an idea from conception to experimentation quickly)

  • The ability to work well with others in a high-paced research setting

  • Excellent communication and collaboration skills, and can work effectively on both research and engineering implementation at scale

Qualifications / Additional Skills:
  • Expertise and intuition for training models in the audio domain, including text-to-speech, ASR, speech-to-speech, speech-emotion-recognition, or other models

  • Experience in training audio autoencoders

  • Understanding of signal processing, especially of audio signals

  • Experience with diffusion models, consistency models, or GANs

  • Experience with training on large-scale (multi-node) GPU clusters

  • Strong grasp of proper experimental methodology for running rigorous ablations and other hypothesis testing

  • Understanding of and interest in large-scale, highly parallel data processing pipelines

  • Proficiency with PyTorch and Python

  • Experience contributing to large pre-existing codebases and rapidly getting up to speed

  • Previously published machine learning research in well-respected venues

  • Postgraduate degree in a scientific subject (Computer Science, EE/EECS, Mathematics, Physics, Machine Learning)

Why Work at Zyphra:
  • Our research methodology is grounded in methodical, step-by-step approaches to ambitious goals. Both deep research and engineering excellence are equally valued

  • We strongly value new and crazy ideas and are very willing to bet big on new ideas

  • We move as quickly as we can; we aim to minimize the bar to impact as low as possible

  • We all enjoy what we do and love discussing AI

Benefits and Perks:
  • Comprehensive medical, dental, vision, and FSA plans

  • Competitive compensation and 401(k) plan

  • Relocation and immigration support on a case-by-case basis

  • In-office snacks and meals provided

  • Unlimited PTO and company holidays

  • In-person team in San Francisco with a collaborative, high-energy environment

Vacancy posted more than 2 months ago
Similar jobs that could be interesting for youBased on the Research Engineer - Audio & Speech Models in San Francisco, CA vacancy
  •  ...Applied Scientist / Research Engineer — Speech AI   HIGHLIGHTS Location:  San Francisco, CA OR...  ...person will work across applied research, model development, training infrastructure,...  ...with modern machine learning for audio, speech, and language, and who enjoys... 
    Audio
    Remote work

    GTN Technical Staffing

    San Francisco, CA
    more than 2 months ago
  • $110.7k - $379.2k

    Position Summary Research Engineer — Post-Training & Small Language Models (SLMs), Healthcare AI Three hundred fifty million Americans rely on a healthcare...  ..., or Ollama. • Experience with multimodal models, speech models, or domain-specific foundation models;... 
    Suggested
    Local area
    Visa sponsorship

    Deloitte

    San Francisco, CA
    4 days ago
  •  ...Phonic is a product and research lab focused on powering the...  ...pursuit of this goal, from models to product, to create voice...  ...About The Role As a Research Engineer at Phonic, you'll sit at the...  ...models across the voice stack (speech, audio, language, and beyond), while... 
    Audio
    Work at office

    Phonic

    San Francisco, CA
    3 days ago
  •  ...looking for an experienced Machine Learning Engineer to join our team and help develop cutting-edge speech recognition models that help teach language fluency. In this role...  ...experience Bonus Experience with speech or audio Office ~ San Francisco, CA... 
    Audio
    Full time
    Live in
    Work at office
    Worldwide

    Speak

    San Francisco, CA
    1 day ago
  • $164.6k - $313.3k

     ...SODA) is looking for a driven Data/ML engineer to push the boundaries of audio GenAI. Join the team behind Firefly...  ...Sound Effects and multiple AI models that have shipped in Adobe products...  ...small, collaborative and efficient research team looking for highly motivated candidates... 
    Audio
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    2 days ago
  •  ...Archive Human Archive is a research lab backed by Y Combinator focused on modeling human embodied intelligence. Humans...  ...Opportunity As a Research Engineer, you'll work on multimodal...  ...capture, IMUs, tactile sensing, audio, and wearable systems, and study... 
    Audio
    Shift work

    Human Archive

    San Francisco, CA
    5 days ago
  • $155k - $269k

     ...driving stack is powered by Waabi World, which delivers realistic, scalable, controllable, and efficient simulation. As a Research Engineer in the World Models team, you will develop algorithms and productionize the next generation of World Models that can reason about... 
    Full time
    Work at office
    Work from home
    Flexible hours

    Waabi

    San Francisco, CA
    more than 2 months ago
  •  ...Description Zyphra is an artificial intelligence company based in San Francisco, California. The Role: As a Research Engineer - Brain Computer Interface Models , you will be a core contributor to Zyphra’s BCI work, building the next generation of open-source EEG and... 
    Work at office
    Relocation package

    Zyphra

    San Francisco, CA
    more than 2 months ago
  •  ...Job Description Zyphra is an artificial intelligence company based in San Francisco, California. The Role: As a Research Engineer - Model Architectures , you will be a core contributor to Zyphra’s AI Architecture Research Team. This will involve designing and... 
    Work at office
    Relocation package

    Zyphra

    San Francisco, CA
    more than 2 months ago
  •  ...organizations and operates in a research-driven, high-growth environment where engineers address complex technical challenges...  ...challenges in computer vision, audio processing, and natural language...  ...processing. You will work with modern AI models and APIs, optimizing performance... 
    Audio
    Full time
    H1b
    Visa sponsorship

    Guidant Solutions Pvt. Ltd.

    San Francisco, CA
    4 days ago
  • $180k - $270k

     ...throughput, ultra-low-latency inference engines for large language models or foundational speech models. Understand the...  ...To-First-Token (or Time-To-First-Audio) in real-time streaming environments...  ...highly collaborative, fast-paced research. Gear & Perks: Choice of top-... 
    Audio
    Full time
    Work at office
    Worldwide

    Plaud

    San Francisco, CA
    1 day ago
  •  ...AI Research Engineer Opportunity Poly is building a better file storage platform for everyone...  ...re maximally multimodal. We support any audio, video, document, office file, slide...  ...baseline knowledge of core statistical modeling principles, including the ability to write... 
    Audio
    Work at office

    Poly

    San Francisco, CA
    4 days ago
  •  ...Astra, and top-tier AI researchers. As an early member...  ...AI researchers and engineers in the world in a small...  ...and video generation model teams in the world on...  ...training text-to-motion or audio-to-motion models....  ...if you have worked on speech-driven 3D facial animation... 
    Audio
    Work at office
    Visa sponsorship

    Spellbrush

    San Francisco, CA
    4 days ago
  •  ...based in San Francisco, is building autonomous factories with state-of-the-art perception and foundation-model driven robotics. As a Member of Technical Staff, Research, you will advance perception models, data pipelines, and scalable training workflows for real-world... 

    Industrial Next (YC W22)

    San Francisco, CA
    4 days ago
  •  ...Product Manager focused on AI speech (text-to-speech, speech-to-...  ..., collaborating closely with engineering and analysis colleagues, as well...  ...to benchmark their latest models. If you're excited about cutting...  ...in a product role at a voice/audio software company (e.g.... 
    Audio

    Artificial Analysis, Inc.

    San Francisco, CA
    2 days ago
  • $26 - $28 per hour

     ...join our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact...  ...power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are trained on accurate, well-structured, and representative... 
    Audio
    Full time
    Work experience placement
    Remote work
    Visa sponsorship

    Welocalize

    San Francisco, CA
    5 days ago
  • $26 - $28 per hour

     ...About the job Czech Data Labeling Analyst(Speech & Voice ) Overview Welo Data is looking for detail...  ...real-world AI systems. You'll be working with audio, speech, and language data - helping ensure models are trained on accurate, well-structured, and representative... 
    Audio
    Full time
    Work experience placement
    Remote work
    Visa sponsorship

    Welocalize

    San Francisco, CA
    5 days ago
  •  ...offline and live evals that keep our speech and multimodal models honest in production. Harness the...  ...models, not slide decks — partner with research and infra to prototype, train, and...  ...Expert-level PyTorch. Proven software engineer who loves ML; comfortable writing... 
    Full time
    Contract work
    Flexible hours
    Shift work

    SESAME

    San Francisco, CA
    3 days ago
  •  ...Research Engineer Lotus Health is a groundbreaking primary care app that integrates your medical...  ...product engineering: dataset curation, model training and evaluation, retrieval and...  ...tool use Experience building speech or multimodal pipelines for medical settings... 

    Lotus Health

    San Francisco, CA
    4 days ago
  • $34 per hour

     ...join our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact...  ...power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are trained on accurate, well-structured, and representative... 
    Audio
    Work experience placement
    Remote work

    Welocalize

    San Francisco, CA
    5 days ago
  • $150k - $350k

    David Joseph & Company is looking for an applied research engineer in San Francisco, CA, to build high-performance pipelines for video understanding...  ...challenging research problems across computer vision and audio processing. The ideal candidate has over 3 years of... 
    Audio

    David Joseph & Company

    San Francisco, CA
    2 days ago
  •  ...company is seeking a talented software engineer to join their dynamic Inference team...  ...for large-scale multimodal models, focusing on high-performance delivery of audio and image inputs. You'll collaborate closely with researchers and product teams to push the boundaries... 
    Audio

    Jobleads-US

    San Francisco, CA
    2 days ago
  • $350k

    A leading AI research company in San Francisco is hiring for a position on their Audio team. The role involves developing and training advanced audio models, optimizing performance, and working collaboratively across teams. Ideal candidates will have strong expertise in... 
    Audio
    Flexible hours

    Jobleads-US

    San Francisco, CA
    4 days ago
  • Thinking Machines Lab in San Francisco is actively researching audio capabilities, blending rigorous theory with practical engineering to build multimodal AI systems. You will...  ..., post-training, and product to develop models that understand and generate audio with high... 
    Audio

    Mosaic.tech

    San Francisco, CA
    2 days ago
  •  ...Applied Research Engineer - San Francisco As an applied research engineer at Confidential, you...  ...be working in the computer vision, audio processing, and text processing domains...  ...fit if you’re comfortable working with models + APIs and squeezing every drop of performance... 
    Audio
    Full time

    My HR Solutions

    San Francisco, CA
    more than 2 months ago
  • Luma AI in the San Francisco Bay Area and beyond is seeking a Research Scientist / Engineer to shape the foundation of multimodal AI. You will bridge frontier research with shipped products like Dream Machine and Ray3, tackling problems without playbooks. The role emphasizes... 
    Remote job

    Luma AI

    San Francisco, CA
    2 days ago
  •  ...Job Description Zyphra is an artificial intelligence company based in San Francisco, California. The Role: As a Research Engineer - Language Model Pre-Training , you'll shape our language model roadmap through end-to-end pretraining development. You will work... 
    Work at office
    Relocation package

    Zyphra

    San Francisco, CA
    18 days ago
  • Stealth Startup in San Francisco seeks a Founding Machine Learning Research Engineer to advance real-time AI by exploring state-of-the-art LLMs, speech models, and multimodal AI for human-like voice agents operating in complex environments. You’ll design evaluation frameworks... 

    Stealth Startup

    San Francisco, CA
    3 days ago
  • $197.3k - $313.7k

     ...looking for talented software and platform engineers to embed in our AI team to bridge the...  ...engineering skills directly enable world-class research and products used by millions?At...  ...alongside Research Scientists to bring models to life. This position is dedicated to a... 
    Full time

    Salesforce

    San Francisco, CA
    4 days ago
  •  ...Senior Product Engineer @ Kato About the Job Position: Senior Product Engineer (...  ...Have: Experience with real-time audio processing or WebRTC Background in Golang...  ..., Docker AI/ML: Large Language Models, Speech To Text, Text To Speech DevOps: GitHub... 
    Audio
    Full time
    Start working today

    pear.ai

    San Francisco, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Engineer - Audio & Speech Models. Be the first to apply!