Research Engineer - Audio & Speech Models
Zyphra
Job Description
Job Description
Zyphra is an artificial intelligence company based in San Francisco, California.
The Role:As a Research Engineer - Audio & Speech Models , you will be a core contributor on Zyphra’s Audio Team, building the next generation of open-source autoencoders, ASR, TTS, SSL, and speech-to-speech models. You will be deeply involved in the entire model training process, from data gathering and processing to designing novel architectures and training methodologies.
You’ll Work Across:Large-scale audio training runs
Performance optimization of our training stack
Audio dataset collection, processing, and evaluation
Architecture and training methodology ablations and improvements
Strong research taste and intuition. The ability to work through a research project from conception to execution to write-up.
Strong implementation and prototyping ability (can take an idea from conception to experimentation quickly)
The ability to work well with others in a high-paced research setting
Excellent communication and collaboration skills, and can work effectively on both research and engineering implementation at scale
Expertise and intuition for training models in the audio domain, including text-to-speech, ASR, speech-to-speech, speech-emotion-recognition, or other models
Experience in training audio autoencoders
Understanding of signal processing, especially of audio signals
Experience with diffusion models, consistency models, or GANs
Experience with training on large-scale (multi-node) GPU clusters
Strong grasp of proper experimental methodology for running rigorous ablations and other hypothesis testing
Understanding of and interest in large-scale, highly parallel data processing pipelines
Proficiency with PyTorch and Python
Experience contributing to large pre-existing codebases and rapidly getting up to speed
Previously published machine learning research in well-respected venues
Postgraduate degree in a scientific subject (Computer Science, EE/EECS, Mathematics, Physics, Machine Learning)
Our research methodology is grounded in methodical, step-by-step approaches to ambitious goals. Both deep research and engineering excellence are equally valued
We strongly value new and crazy ideas and are very willing to bet big on new ideas
We move as quickly as we can; we aim to minimize the bar to impact as low as possible
We all enjoy what we do and love discussing AI
Comprehensive medical, dental, vision, and FSA plans
Competitive compensation and 401(k) plan
Relocation and immigration support on a case-by-case basis
In-office snacks and meals provided
Unlimited PTO and company holidays
In-person team in San Francisco with a collaborative, high-energy environment
- ...Applied Scientist / Research Engineer â Speech AI Â HIGHLIGHTS Location: Â San Francisco, CA OR... ...person will work across applied research, model development, training infrastructure,... ...with modern machine learning for audio, speech, and language, and who enjoys...AudioRemote work
$110.7k - $379.2k
Position Summary Research Engineer — Post-Training & Small Language Models (SLMs), Healthcare AI Three hundred fifty million Americans rely on a healthcare... ..., or Ollama. • Experience with multimodal models, speech models, or domain-specific foundation models;...SuggestedLocal areaVisa sponsorship- ...Phonic is a product and research lab focused on powering the... ...pursuit of this goal, from models to product, to create voice... ...About The Role As a Research Engineer at Phonic, you'll sit at the... ...models across the voice stack (speech, audio, language, and beyond), while...AudioWork at office
- ...looking for an experienced Machine Learning Engineer to join our team and help develop cutting-edge speech recognition models that help teach language fluency. In this role... ...experience Bonus Experience with speech or audio Office ~ San Francisco, CA...AudioFull timeLive inWork at officeWorldwide
$164.6k - $313.3k
...SODA) is looking for a driven Data/ML engineer to push the boundaries of audio GenAI. Join the team behind Firefly... ...Sound Effects and multiple AI models that have shipped in Adobe products... ...small, collaborative and efficient research team looking for highly motivated candidates...AudioFull timeTemporary workLocal areaWorldwide- ...Archive Human Archive is a research lab backed by Y Combinator focused on modeling human embodied intelligence. Humans... ...Opportunity As a Research Engineer, you'll work on multimodal... ...capture, IMUs, tactile sensing, audio, and wearable systems, and study...AudioShift work
$155k - $269k
...driving stack is powered by Waabi World, which delivers realistic, scalable, controllable, and efficient simulation. As a Research Engineer in the World Models team, you will develop algorithms and productionize the next generation of World Models that can reason about...Full timeWork at officeWork from homeFlexible hours- ...Description Zyphra is an artificial intelligence company based in San Francisco, California. The Role: As a Research Engineer - Brain Computer Interface Models , you will be a core contributor to Zyphra’s BCI work, building the next generation of open-source EEG and...Work at officeRelocation package
- ...Job Description Zyphra is an artificial intelligence company based in San Francisco, California. The Role: As a Research Engineer - Model Architectures , you will be a core contributor to Zyphra’s AI Architecture Research Team. This will involve designing and...Work at officeRelocation package
- ...organizations and operates in a research-driven, high-growth environment where engineers address complex technical challenges... ...challenges in computer vision, audio processing, and natural language... ...processing. You will work with modern AI models and APIs, optimizing performance...AudioFull timeH1bVisa sponsorship
$180k - $270k
...throughput, ultra-low-latency inference engines for large language models or foundational speech models. Understand the... ...To-First-Token (or Time-To-First-Audio) in real-time streaming environments... ...highly collaborative, fast-paced research. Gear & Perks: Choice of top-...AudioFull timeWork at officeWorldwide- ...AI Research Engineer Opportunity Poly is building a better file storage platform for everyone... ...re maximally multimodal. We support any audio, video, document, office file, slide... ...baseline knowledge of core statistical modeling principles, including the ability to write...AudioWork at office
- ...Astra, and top-tier AI researchers. As an early member... ...AI researchers and engineers in the world in a small... ...and video generation model teams in the world on... ...training text-to-motion or audio-to-motion models.... ...if you have worked on speech-driven 3D facial animation...AudioWork at officeVisa sponsorship
- ...based in San Francisco, is building autonomous factories with state-of-the-art perception and foundation-model driven robotics. As a Member of Technical Staff, Research, you will advance perception models, data pipelines, and scalable training workflows for real-world...
- ...Product Manager focused on AI speech (text-to-speech, speech-to-... ..., collaborating closely with engineering and analysis colleagues, as well... ...to benchmark their latest models. If you're excited about cutting... ...in a product role at a voice/audio software company (e.g....Audio
$26 - $28 per hour
...join our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact... ...power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are trained on accurate, well-structured, and representative...AudioFull timeWork experience placementRemote workVisa sponsorship$26 - $28 per hour
...About the job Czech Data Labeling Analyst(Speech & Voice ) Overview Welo Data is looking for detail... ...real-world AI systems. You'll be working with audio, speech, and language data - helping ensure models are trained on accurate, well-structured, and representative...AudioFull timeWork experience placementRemote workVisa sponsorship- ...offline and live evals that keep our speech and multimodal models honest in production. Harness the... ...models, not slide decks — partner with research and infra to prototype, train, and... ...Expert-level PyTorch. Proven software engineer who loves ML; comfortable writing...Full timeContract workFlexible hoursShift work
- ...Research Engineer Lotus Health is a groundbreaking primary care app that integrates your medical... ...product engineering: dataset curation, model training and evaluation, retrieval and... ...tool use Experience building speech or multimodal pipelines for medical settings...
$34 per hour
...join our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact... ...power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are trained on accurate, well-structured, and representative...AudioWork experience placementRemote work$150k - $350k
David Joseph & Company is looking for an applied research engineer in San Francisco, CA, to build high-performance pipelines for video understanding... ...challenging research problems across computer vision and audio processing. The ideal candidate has over 3 years of...Audio- ...company is seeking a talented software engineer to join their dynamic Inference team... ...for large-scale multimodal models, focusing on high-performance delivery of audio and image inputs. You'll collaborate closely with researchers and product teams to push the boundaries...Audio
$350k
A leading AI research company in San Francisco is hiring for a position on their Audio team. The role involves developing and training advanced audio models, optimizing performance, and working collaboratively across teams. Ideal candidates will have strong expertise in...AudioFlexible hours- Thinking Machines Lab in San Francisco is actively researching audio capabilities, blending rigorous theory with practical engineering to build multimodal AI systems. You will... ..., post-training, and product to develop models that understand and generate audio with high...Audio
- ...Applied Research Engineer - San Francisco As an applied research engineer at Confidential, you... ...be working in the computer vision, audio processing, and text processing domains... ...fit if you’re comfortable working with models + APIs and squeezing every drop of performance...AudioFull time
- Luma AI in the San Francisco Bay Area and beyond is seeking a Research Scientist / Engineer to shape the foundation of multimodal AI. You will bridge frontier research with shipped products like Dream Machine and Ray3, tackling problems without playbooks. The role emphasizes...Remote job
- ...Job Description Zyphra is an artificial intelligence company based in San Francisco, California. The Role: As a Research Engineer - Language Model Pre-Training , you'll shape our language model roadmap through end-to-end pretraining development. You will work...Work at officeRelocation package
- Stealth Startup in San Francisco seeks a Founding Machine Learning Research Engineer to advance real-time AI by exploring state-of-the-art LLMs, speech models, and multimodal AI for human-like voice agents operating in complex environments. You’ll design evaluation frameworks...
$197.3k - $313.7k
...looking for talented software and platform engineers to embed in our AI team to bridge the... ...engineering skills directly enable world-class research and products used by millions?At... ...alongside Research Scientists to bring models to life. This position is dedicated to a...Full time- ...Senior Product Engineer @ Kato About the Job Position: Senior Product Engineer (... ...Have: Experience with real-time audio processing or WebRTC Background in Golang... ..., Docker AI/ML: Large Language Models, Speech To Text, Text To Speech DevOps: GitHub...AudioFull timeStart working today
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Engineer - Audio & Speech Models. Be the first to apply!
- research programmer San Francisco, CA
- research engineer San Francisco, CA
- senior research engineer San Francisco, CA
- junior machine learning research engineer San Francisco, CA
- deep learning research engineer San Francisco, CA
- research software engineer San Francisco, CA
- research assistant engineering San Francisco, CA
- ai research engineer San Francisco, CA
- audio mixing San Francisco, CA
- audio visual San Francisco, CA



