Edge AI Researcher (Speech & Audio Models)
Huxley
We are seeking an Edge AI Research Scientist to develop next-generation speech and audio AI systems that run efficiently on smartphones, wearables, and other resource-constrained devices. This role sits at the intersection of machine learning research, systems optimization, and production engineering. You will design compact model architectures, develop advanced compression techniques, and optimize inference pipelines that enable real-time speech AI experiences directly on end-user devices. You will work across a broad range of voice technologies, including automatic speech recognition (ASR), text-to-speech (TTS), speech translation, speech-to-speech systems, and neural audio codecs. The ideal candidate combines strong research credentials with hands-on implementation skills and has a deep understanding of efficient deep learning, model optimization, and hardware-aware machine learning. What You'll Do Research and Model Development Drive research in efficient machine learning and edge AI for speech and audio applications. Design compact model architectures capable of operating under strict latency, memory, and power constraints. Develop and improve state-of-the-art approaches for: Low-rank adaptation and compression Hardware-aware architectures Efficient training and inference techniques Contribute to speech-to-speech, speech recognition, speech translation, text-to-speech, and audio generation systems. Inference Optimization Build highly optimized inference pipelines for mobile and embedded hardware. Improve performance across CPUs, GPUs, NPUs, and other acceleration hardware. Optimize: Operator execution Scheduling strategies Caching mechanisms End-to-end system latency Integrate models with production runtimes and deployment frameworks. Performance Evaluation Develop rigorous benchmarking methodologies for edge AI systems. Measure and improve: Real-time factor Time-to-first-audio Thermal behavior Speech quality and accuracy Validate performance directly on target devices rather than relying solely on simulator environments. Cross-Functional Collaboration Partner with machine learning researchers, mobile engineers, and systems engineers to bring research into production. Translate research prototypes into scalable products and customer-facing technologies. Communicate findings through internal documentation, technical publications, conference papers, and open-source contributions where appropriate. Required Qualifications Master's degree, Ph.D., or equivalent industry experience in: Computer Science Efficient Deep Learning Demonstrated expertise in model compression, efficient inference, or edge AI through research publications, production systems, or both. Strong understanding of one or more of the following: Hardware-aware optimization Strong software engineering skills with: PyTorch or JAX C/C++ or equivalent systems-level programming experience Experience optimizing neural networks for resource-constrained hardware. Practical knowledge of: GPUs NPUs Memory systems Numerical precision tradeoffs Ability to make informed tradeoffs between model quality, latency, memory footprint, power consumption, and deployment portability. Professional proficiency in English. Ability to thrive in a fast-moving, research-driven environment. Preferred Qualifications Experience working with: Automatic Speech Recognition (ASR) Speech-to-Speech Models Familiarity with deployment frameworks such as: Core ML ExecuTorch ONNX Runtime LiteRT / TensorFlow Lite TensorRT Experience with acceleration technologies including: Metal Vulkan CUDA QNN
XNNPACK
Custom kernels and operator fusion Knowledge of: Experience building streaming and low-latency audio systems. Experience deploying machine learning models to: Embedded systems Experience training, distilling, or evaluating large-scale foundation models using distributed GPU infrastructure. Previous experience in industrial research labs, AI startups, or leading technology companies. Publication record at relevant conferences or meaningful open-source contributions #J-18808-Ljbffr Huxley- Huxley in San Francisco is seeking an Edge AI Research Scientist to advance speech and audio AI for devices with limited resources. You will design compact models, compress and optimize inference, and enable real-time speech experiences directly on phones and wearables...Audio
- ...Project Astra, and top-tier AI researchers. As an early member of this... ...best image and video generation model teams in the world on data... ...in training text-to-motion or audio-to-motion models. You... ...good fit if you have worked on speech-driven 3D facial animation....AudioWork at officeVisa sponsorship
- ...Machines Lab in San Francisco is actively researching audio capabilities, blending rigorous theory... ...practical engineering to build multimodal AI systems. You will work across pre-... ..., post-training, and product to develop models that understand and generate audio with...Audio
- DeepRec.ai is a fast-growing voice AI company in SF, delivering production-ready AI phone... ...-time role focuses on building end-to-end speech technologies across TTS, STT, and neural... ...push theory to production, train on massive audio datasets, and collaborate with engineering...AudioFull time
- Kotoba’s speech models are licensed to Fortune 50 companies and US big... ...users a day. We’re hiring an AI Researcher to build the next generation... ...from the data center to edge devices, and we license this... ...lingual modeling Knowledge of audio tokenization, neural audio codecs...Audio
- LG Electronics is seeking a Contract AI Researcher focusing on Efficient AI in Santa Clara, CA, hybrid work arrangement. You will explore model compression, quantization, efficient inference, and architectures to make LLMs/VLMs faster and more deployable on devices. You...Contract work
- ...based in San Francisco, California. The Role: As a Research Engineer - Audio & Speech Models , you will be a core contributor on Zyphra’s Audio Team... ...possible We all enjoy what we do and love discussing AI Benefits and Perks: Comprehensive medical, dental...AudioWork at officeRelocation package
- ...revenue-generating voice AI company empowering... ...they are building the models and infrastructure that... ...immediately. No speculative research track here. If you want... ...goal: a fully speech-to-speech conversational... ...text-to-speech, neural audio codecs, and getting LLMs...AudioPermanent employmentFull timeImmediate start
$180k - $270k
...world's most trusted AI work companion for professionals... ...exposure to cutting-edge AI for Pro tools and... ...for large language models or foundational speech models. Understand... ...(or Time-To-First-Audio) in real-time... ...collaborative, fast-paced research. Gear & Perks:...AudioFull timeWork at officeWorldwide- A cutting-edge AI startup is searching for an experienced AI Researcher eager to advance generative AI. This role requires a PhD and 5+ years of research experience, focusing on developing innovative models that harness earth observation. The ideal candidate will demonstrate...Remote job
- ...for a Product Manager focused on AI speech (text-to-speech, speech-to-text, voice... ...field to benchmark their latest models. If you're excited about cutting‑edge speech models and care deeply... ...working in a product role at a voice/audio software company (e.g. ElevenLabs,...Audio
$204k - $300k
...Advanced Technology Group (ATG) is the research division of the company. ATG’s... ...electrical engineering, such as AI/ML, algorithms, digital signal processing, audio engineering, image processing, computer... ..., distributed systems, cloud, edge & mobile computing, computer networking...AudioFull timeLocal areaWorldwideFlexible hours$195k - $365k
...is building the real-world AI interface for professionals... ...interaction. Gain exposure to cutting-edge AI for Pro tools and play a... ...and training large-scale audio or speech models from the ground up, whether... ...at the intersection of research and engineering, eager to design...AudioFull timeWork at officeWorldwide- Plaud Inc. is seeking senior AI researchers to join our SpeechLLM lab in San Francisco. You will help build and train large-scale audio/speech models and push the boundaries of human-AI interaction. We value hands-on experience with PyTorch or JAX, distributed training...Audio
- ...mission is to architect AI that learns from... ...'re pioneering the model architectures that... ...and ship cutting edge models and experiences... ...the Role On the Audio Post-Training team,... ...customer needs meet research, and covers the full... ...generative models (speech, text, or...AudioWork at officeVisa sponsorshipFlexible hours
- ...building the future of voice AI operating systems for... ...inflection point where advances in speech, language models, and clinical AI can... ...are hiring two ML Engineers / Researchers to help build the next generation... ...one of two areas: Speech & Audio: Build state-of-the-art...AudioFull time
$117.2k - $313.7k
...SalesforceSalesforce is the #1 AI CRM, where humans... ...AI Research is a global leader... ...art large language models including CodeGen;... ...translate cutting-edge research into production... ...multimodal AI agents.Speech Intelligence: Voice... ...recognition, and low-latency audio processing.Core...AudioFull timeWorldwide- Capital One is seeking an Applied Researcher I in San Francisco to join the AI Foundations team. You will collaborate with data scientists, software engineers... ...platforms. The role emphasizes applied research, model development, and translating research into business outcomes...
- ...institution based in San Francisco is looking for an Applied Researcher to work on AI-powered products. This role involves delivering innovative... ...Responsibilities include partnering with teams to build AI models and conducting impactful research. #J-18808-Ljbffr Capital...
- Acceler8 Talent is seeking an ML Researcher focused on World Models to join an early-stage lab-backed venture in San Francisco. You will explore world model approaches and contribute to research that informs real-world voice products with production-ready implications....
- As a Research Scientist , you'll lead cutting-edge research that advances the state of generative AI for long-form storytelling. You'll work at the... ...on large language models and multimodal foundation... ...methods that unify text, speech, audio, and other modalities into...AudioWorldwide
$117.2k - $313.7k
Salesforce AI Research is looking for outstanding AI Research... ..., develop novel models, and bridge the gap between... ...multimodal AI agents. Speech Intelligence: Voice intelligence... ..., and low‑latency audio processing. Core... ...and translate cutting‑edge ideas into practical, high...Audio- Verita AI is seeking an Applied AI Researcher to work with clients on model evaluation and data strategy. You will assess model performance, identify failure modes, and design data-driven solutions, collaborating with operations and engineering to implement scalable data...
$34 per hour
...our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role... ...power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are trained on accurate, well-structured, and...AudioWork experience placementRemote work$26 - $28 per hour
...our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role... ...power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are trained on accurate, well-structured, and...AudioWork experience placementRemote work$26 - $28 per hour
...job Czech Data Labeling Analyst(Speech & Voice ) Overview... ...Analysts, supporting speech and voice AI systems. This is a high-impact... ...systems. You'll be working with audio, speech, and language data - helping ensure models are trained on accurate, well-structured...AudioFull timeWork experience placementRemote workVisa sponsorship$150k - $200k
..., our team is tackling cutting-edge engineering challenges to bring... ...Senior Firmware Engineer, Edge AI / NPU Runtime to help architect... ...product. You’ll help define how models run on-device, how sensor data... ...as biosignals, sensor fusion, audio, gesture recognition, keyword spotting...AudioVisa sponsorship- ...implementing infrastructure for large-scale multimodal models, focusing on high-performance delivery of audio and image inputs. You'll collaborate closely with researchers and product teams to push the boundaries of AI technology, ensuring reliable production services. If...Audio
- A technology company specializing in AI is seeking an experienced DSP Engineer to design and develop innovative audio processing algorithms. You will work on deep learning models for tasks such as speech enhancement and voice identification, while contributing to real-...Audio
$216.3k - $280.8k
Meet the TeamAt Foundation AI, we are leading frontier AI research across Cisco. Our mission is to advance the state... ...AI.Our research spans foundation models, agentic AI, multimodal learning,... ...work at the intersection of cutting-edge research and real-world impact, developing...Full timeTemporary workLocal areaFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Edge AI Researcher (Speech & Audio Models). Be the first to apply!
- field researcher San Francisco, CA
- product researcher San Francisco, CA
- security researcher San Francisco, CA
- lead researcher San Francisco, CA
- data collection researcher San Francisco, CA
- machine learning researcher San Francisco, CA
- court researcher San Francisco, CA
- researcher San Francisco, CA
- postdoctoral researcher cosmetic science San Francisco, CA
- senior researcher San Francisco, CA



