Edge AI Researcher (Speech & Audio Models)
Huxley
We are seeking an Edge AI Research Scientist to develop next-generation speech and audio AI systems that run efficiently on smartphones, wearables, and other resource-constrained devices. This role sits at the intersection of machine learning research, systems optimization, and production engineering. You will design compact model architectures, develop advanced compression techniques, and optimize inference pipelines that enable real-time speech AI experiences directly on end-user devices. You will work across a broad range of voice technologies, including automatic speech recognition (ASR), text-to-speech (TTS), speech translation, speech-to-speech systems, and neural audio codecs. The ideal candidate combines strong research credentials with hands-on implementation skills and has a deep understanding of efficient deep learning, model optimization, and hardware-aware machine learning. What You'll Do Research and Model Development Drive research in efficient machine learning and edge AI for speech and audio applications. Design compact model architectures capable of operating under strict latency, memory, and power constraints. Develop and improve state-of-the-art approaches for: Low-rank adaptation and compression Hardware-aware architectures Efficient training and inference techniques Contribute to speech-to-speech, speech recognition, speech translation, text-to-speech, and audio generation systems. Inference Optimization Build highly optimized inference pipelines for mobile and embedded hardware. Improve performance across CPUs, GPUs, NPUs, and other acceleration hardware. Optimize: Operator execution Scheduling strategies Caching mechanisms End-to-end system latency Integrate models with production runtimes and deployment frameworks. Performance Evaluation Develop rigorous benchmarking methodologies for edge AI systems. Measure and improve: Real-time factor Time-to-first-audio Thermal behavior Speech quality and accuracy Validate performance directly on target devices rather than relying solely on simulator environments. Cross-Functional Collaboration Partner with machine learning researchers, mobile engineers, and systems engineers to bring research into production. Translate research prototypes into scalable products and customer-facing technologies. Communicate findings through internal documentation, technical publications, conference papers, and open-source contributions where appropriate. Required Qualifications Master's degree, Ph.D., or equivalent industry experience in: Computer Science Efficient Deep Learning Demonstrated expertise in model compression, efficient inference, or edge AI through research publications, production systems, or both. Strong understanding of one or more of the following: Hardware-aware optimization Strong software engineering skills with: PyTorch or JAX C/C++ or equivalent systems-level programming experience Experience optimizing neural networks for resource-constrained hardware. Practical knowledge of: GPUs NPUs Memory systems Numerical precision tradeoffs Ability to make informed tradeoffs between model quality, latency, memory footprint, power consumption, and deployment portability. Professional proficiency in English. Ability to thrive in a fast-moving, research-driven environment. Preferred Qualifications Experience working with: Automatic Speech Recognition (ASR) Speech-to-Speech Models Familiarity with deployment frameworks such as: Core ML ExecuTorch ONNX Runtime LiteRT / TensorFlow Lite TensorRT Experience with acceleration technologies including: Metal Vulkan CUDA QNN
XNNPACK
Custom kernels and operator fusion Knowledge of: Experience building streaming and low-latency audio systems. Experience deploying machine learning models to: Embedded systems Experience training, distilling, or evaluating large-scale foundation models using distributed GPU infrastructure. Previous experience in industrial research labs, AI startups, or leading technology companies. Publication record at relevant conferences or meaningful open-source contributions #J-18808-Ljbffr Huxley- Miso Labs in San Francisco is hiring founding researchers to advance speech-to-speech modeling beyond the current STT->LLM->TTS paradigm, training full‑duplex systems. You will join a small, founding team and relocate to SF as we cover moving costs, contributing from day...SuggestedRelocationRelocation package
$350k
A leading AI research company in San Francisco is hiring for a position on their Audio team. The role involves developing and training advanced audio models, optimizing performance, and working collaboratively across teams. Ideal candidates will have strong expertise in...AudioFlexible hours- ...Machines Lab in San Francisco is actively researching audio capabilities, blending rigorous theory... ...practical engineering to build multimodal AI systems. You will work across pre-... ..., post-training, and product to develop models that understand and generate audio with...Audio
- ...adventure game where an AI companion is the core... ...Astra, and top-tier AI researchers. As an early member... ...image and video generation model teams in the world on... ...training text-to-motion or audio-to-motion models. You... ...if you have worked on speech-driven 3D facial...AudioWork at officeVisa sponsorship
- LG Electronics is seeking a Contract AI Researcher focusing on Efficient AI in Santa Clara, CA, hybrid work arrangement. You will explore model compression, quantization, efficient inference, and architectures to make LLMs/VLMs faster and more deployable on devices. You...SuggestedContract work
- ...revenue-generating voice AI company empowering... ...they are building the models and infrastructure that... ...immediately. No speculative research track here. If you want... ...goal: a fully speech-to-speech conversational... ...text-to-speech, neural audio codecs, and getting LLMs...AudioPermanent employmentFull timeImmediate start
$180k - $270k
...world's most trusted AI work companion for professionals... ...exposure to cutting-edge AI for Pro tools and... ...for large language models or foundational speech models. Understand... ...(or Time-To-First-Audio) in real-time... ...collaborative, fast-paced research. Gear & Perks:...AudioFull timeWork at officeWorldwide- ...journey to fix this. We're creating an AI-powered experience that replicates the... ...our team and help develop cutting-edge speech recognition models that help teach language fluency. In this... ...Bonus Experience with speech or audio Office ~ San Francisco, CA...AudioFull timeLive inWork at officeWorldwide
- A cutting-edge AI startup is searching for an experienced AI Researcher eager to advance generative AI. This role requires a PhD and 5+ years of research experience, focusing on developing innovative models that harness earth observation. The ideal candidate will demonstrate...Remote job
- ...the leading independent AI benchmarking and insights... ...measure the cutting edge of AI, they are actively... ...Product Manager focused on AI speech (text-to-speech, speech-... ...coverage of speech models stays up to date by benchmarking... ...product role at a voice/audio software company or in a...Audio
- ...for a Product Manager focused on AI speech (text-to-speech, speech-to-text, voice... ...field to benchmark their latest models. If you're excited about cutting‑edge speech models and care deeply... ...working in a product role at a voice/audio software company (e.g. ElevenLabs,...Audio
- ...based in San Francisco, California. The Role: As a Research Engineer - Audio & Speech Models , you will be a core contributor on Zyphra’s Audio Team... ...possible We all enjoy what we do and love discussing AI Benefits and Perks: Comprehensive medical, dental...AudioWork at officeRelocation package
$204k - $300k
...Advanced Technology Group (ATG) is the research division of the company. ATG’s... ...electrical engineering, such as AI/ML, algorithms, digital signal processing, audio engineering, image processing, computer... ..., distributed systems, cloud, edge & mobile computing, computer networking...AudioFull timeLocal areaWorldwideFlexible hours- Miso Labs is building next-generation speech models and inviting you to join its founding research team in San Francisco. You’ll work across the model training stack... .... We value researchers who publish, present at top AI/ML conferences, and own long-term outcomes end-to-...Relocation
- OpenAI is at the center of high-impact multimodal AI. The Chat and Multimodal Safety team builds safe, scalable models and evaluations for text, vision, and audio tasks. As a Researcher on the Chat and Multimodal Safety team in San Francisco, you will shape model perception...AudioWork at officeRelocation package
- ...impact multimodal work in AI. ChatGPT serves a... ...interactions via text, speech, and visuals. As interactive surfaces grow, models also need to adapt to emerging... ...experiences. We develop the research, training methods, and... ...for image, video, or audio systems. In this role, you...AudioWork at officeRelocation package
- ...mission is to architect AI that learns from... ...'re pioneering the model architectures that... ...and ship cutting edge models and experiences... ...the Role On the Audio Post-Training team,... ...customer needs meet research, and covers the full... ...generative models (speech, text, or...AudioWork at officeVisa sponsorshipFlexible hours
$117.2k - $313.7k
...SalesforceSalesforce is the #1 AI CRM, where humans... ...AI Research is looking for outstanding... ..., develops novel models, and bridges the gap... ...multimodal AI agents.Speech Intelligence: Voice... ..., and low-latency audio processing.Core... ...translate cutting-edge ideas into practical...AudioFull time- ...institution based in San Francisco is looking for an Applied Researcher to work on AI-powered products. This role involves delivering innovative... ...Responsibilities include partnering with teams to build AI models and conducting impactful research. #J-18808-Ljbffr Capital...
- ...financial services firm in San Francisco is seeking an Applied Researcher II to develop innovative AI systems. In this role, you will collaborate with a cross-functional team to build AI foundation models, engage in impactful research, and translate complex work into business...
- causal is building a Large Physics foundation Model focused on weather with a mission to enable verifiable cause and effect in AI systems. We seek researchers to tackle unsolved problems across language, vision, robotics, biology, physics, and weather, training ground-...
- Acceler8 Talent is seeking an ML Researcher focused on World Models to join an early-stage lab-backed venture in San Francisco. You will explore world model approaches and contribute to research that informs real-world voice products with production-ready implications....
- Capital One is seeking an Applied Researcher I in San Francisco to join the AI Foundations team. You will collaborate with data scientists, software engineers... ...platforms. The role emphasizes applied research, model development, and translating research into business outcomes...
- ...building the human layer of AI. Our mission is to make human... ...achieve this through pioneering research in multimodal AI for modeling human-to-human communication (language, audio, and video), as well as... ...path. Your Mission Take cutting-edge research models and make them...AudioRemote workRelocation packageFlexible hours
$144k - $187k
...skilled professional to join the Factors Research Platform and Governance team. This role involves... ...research infrastructure with a focus on AI integration. The ideal candidate holds an... ...include developing tools for model delivery and collaborating with cross-functional...Flexible hours- ...based in San Francisco is seeking a Machine Learning Researcher to enhance K-12 education through AI. This role combines advanced technical skills with a... ...Candidates should possess expertise in generative AI models, Python programming, and have a graduate degree in a...Flexible hours
$34 per hour
...our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role... ...power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are trained on accurate, well-structured, and...AudioWork experience placementRemote work$26 - $28 per hour
...job Finnish Data Labeling Analyst(Speech & Voice) Overview... ...Analysts, supporting speech and voice AI systems. This is a high-impact... ...systems. You'll be working with audio, speech, and language data - helping ensure models are trained on accurate, well-structured...AudioFull timeWork experience placementRemote workVisa sponsorship$26 - $28 per hour
...our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role... ...power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are trained on accurate, well-structured, and...AudioFull timeWork experience placementRemote workVisa sponsorship$197.3k - $313.7k
...DetailsAbout SalesforceSalesforce is the #1 AI CRM, where humans with agents... ...a collaborative, diverse team of researchers at Agentforce Operations. The Foundational Models Team develops the next generation... ...bridges the gap between cutting-edge research and customer value....Full timeImmediate startRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Edge AI Researcher (Speech & Audio Models). Be the first to apply!
- senior researcher San Francisco, CA
- machine learning researcher San Francisco, CA
- researcher San Francisco, CA
- senior design researcher San Francisco, CA
- design researcher San Francisco, CA
- qualitative researcher San Francisco, CA
- data collection researcher San Francisco, CA
- product researcher San Francisco, CA
- survey researcher San Francisco, CA
- legal researcher San Francisco, CA


