Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Edge AI Researcher (Speech & Audio Models)

Huxley

We are seeking an Edge AI Research Scientist to develop next-generation speech and audio AI systems that run efficiently on smartphones, wearables, and other resource-constrained devices. This role sits at the intersection of machine learning research, systems optimization, and production engineering. You will design compact model architectures, develop advanced compression techniques, and optimize inference pipelines that enable real-time speech AI experiences directly on end-user devices. You will work across a broad range of voice technologies, including automatic speech recognition (ASR), text-to-speech (TTS), speech translation, speech-to-speech systems, and neural audio codecs. The ideal candidate combines strong research credentials with hands-on implementation skills and has a deep understanding of efficient deep learning, model optimization, and hardware-aware machine learning. What You'll Do Research and Model Development Drive research in efficient machine learning and edge AI for speech and audio applications. Design compact model architectures capable of operating under strict latency, memory, and power constraints. Develop and improve state-of-the-art approaches for: Low-rank adaptation and compression Hardware-aware architectures Efficient training and inference techniques Contribute to speech-to-speech, speech recognition, speech translation, text-to-speech, and audio generation systems. Inference Optimization Build highly optimized inference pipelines for mobile and embedded hardware. Improve performance across CPUs, GPUs, NPUs, and other acceleration hardware. Optimize: Operator execution Scheduling strategies Caching mechanisms End-to-end system latency Integrate models with production runtimes and deployment frameworks. Performance Evaluation Develop rigorous benchmarking methodologies for edge AI systems. Measure and improve: Real-time factor Time-to-first-audio Thermal behavior Speech quality and accuracy Validate performance directly on target devices rather than relying solely on simulator environments. Cross-Functional Collaboration Partner with machine learning researchers, mobile engineers, and systems engineers to bring research into production. Translate research prototypes into scalable products and customer-facing technologies. Communicate findings through internal documentation, technical publications, conference papers, and open-source contributions where appropriate. Required Qualifications Master's degree, Ph.D., or equivalent industry experience in: Computer Science Efficient Deep Learning Demonstrated expertise in model compression, efficient inference, or edge AI through research publications, production systems, or both. Strong understanding of one or more of the following: Hardware-aware optimization Strong software engineering skills with: PyTorch or JAX C/C++ or equivalent systems-level programming experience Experience optimizing neural networks for resource-constrained hardware. Practical knowledge of: GPUs NPUs Memory systems Numerical precision tradeoffs Ability to make informed tradeoffs between model quality, latency, memory footprint, power consumption, and deployment portability. Professional proficiency in English. Ability to thrive in a fast-moving, research-driven environment. Preferred Qualifications Experience working with: Automatic Speech Recognition (ASR) Speech-to-Speech Models Familiarity with deployment frameworks such as: Core ML ExecuTorch ONNX Runtime LiteRT / TensorFlow Lite TensorRT Experience with acceleration technologies including: Metal Vulkan CUDA QNN

XNNPACK

Custom kernels and operator fusion Knowledge of: Experience building streaming and low-latency audio systems. Experience deploying machine learning models to: Embedded systems Experience training, distilling, or evaluating large-scale foundation models using distributed GPU infrastructure. Previous experience in industrial research labs, AI startups, or leading technology companies. Publication record at relevant conferences or meaningful open-source contributions #J-18808-Ljbffr Huxley

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Edge AI Researcher (Speech & Audio Models) in San Francisco, CA vacancy
  • Huxley in San Francisco is seeking an Edge AI Research Scientist to advance speech and audio AI for devices with limited resources. You will design compact models, compress and optimize inference, and enable real-time speech experiences directly on phones and wearables... 
    Audio

    Huxley

    San Francisco, CA
    2 days ago
  •  ...Project Astra, and top-tier AI researchers. As an early member of this...  ...best image and video generation model teams in the world on data...  ...in training text-to-motion or audio-to-motion models. You...  ...good fit if you have worked on speech-driven 3D facial animation.... 
    Audio
    Work at office
    Visa sponsorship

    Spellbrush

    San Francisco, CA
    1 day ago
  •  ...Machines Lab in San Francisco is actively researching audio capabilities, blending rigorous theory...  ...practical engineering to build multimodal AI systems. You will work across pre-...  ..., post-training, and product to develop models that understand and generate audio with... 
    Audio

    Mosaic.tech

    San Francisco, CA
    4 days ago
  • DeepRec.ai is a fast-growing voice AI company in SF, delivering production-ready AI phone...  ...-time role focuses on building end-to-end speech technologies across TTS, STT, and neural...  ...push theory to production, train on massive audio datasets, and collaborate with engineering... 
    Audio
    Full time

    DeepRec.ai

    San Francisco, CA
    4 days ago
  • Kotoba’s speech models are licensed to Fortune 50 companies and US big...  ...users a day. We’re hiring an AI Researcher to build the next generation...  ...from the data center to edge devices, and we license this...  ...lingual modeling Knowledge of audio tokenization, neural audio codecs... 
    Audio

    Kotoba

    San Francisco, CA
    4 days ago
  • LG Electronics is seeking a Contract AI Researcher focusing on Efficient AI in Santa Clara, CA, hybrid work arrangement. You will explore model compression, quantization, efficient inference, and architectures to make LLMs/VLMs faster and more deployable on devices. You... 
    Contract work

    LG Electronics

    San Francisco, CA
    4 days ago
  •  ...based in San Francisco, California. The Role: As a Research Engineer - Audio & Speech Models , you will be a core contributor on Zyphra’s Audio Team...  ...possible We all enjoy what we do and love discussing AI Benefits and Perks: Comprehensive medical, dental... 
    Audio
    Work at office
    Relocation package

    Zyphra

    San Francisco, CA
    10 days ago
  •  ...revenue-generating voice AI company empowering...  ...they are building the models and infrastructure that...  ...immediately. No speculative research track here. If you want...  ...goal: a fully speech-to-speech conversational...  ...text-to-speech, neural audio codecs, and getting LLMs... 
    Audio
    Permanent employment
    Full time
    Immediate start

    DeepRec.ai

    San Francisco, CA
    2 days ago
  • $180k - $270k

     ...world's most trusted AI work companion for professionals...  ...exposure to cutting-edge AI for Pro tools and...  ...for large language models or foundational speech models. Understand...  ...(or Time-To-First-Audio) in real-time...  ...collaborative, fast-paced research. Gear & Perks:... 
    Audio
    Full time
    Work at office
    Worldwide

    Plaud

    San Francisco, CA
    1 day ago
  • A cutting-edge AI startup is searching for an experienced AI Researcher eager to advance generative AI. This role requires a PhD and 5+ years of research experience, focusing on developing innovative models that harness earth observation. The ideal candidate will demonstrate... 
    Remote job

    hum.ai

    San Francisco, CA
    3 days ago
  •  ...for a Product Manager focused on AI speech (text-to-speech, speech-to-text, voice...  ...field to benchmark their latest models. If you're excited about cutting‑edge speech models and care deeply...  ...working in a product role at a voice/audio software company (e.g. ElevenLabs,... 
    Audio

    Artificial Analysis, Inc.

    San Francisco, CA
    4 days ago
  • $204k - $300k

     ...Advanced Technology Group (ATG) is the research division of the company. ATG’s...  ...electrical engineering, such as AI/ML, algorithms, digital signal processing, audio engineering, image processing, computer...  ..., distributed systems, cloud, edge & mobile computing, computer networking... 
    Audio
    Full time
    Local area
    Worldwide
    Flexible hours

    Dolby

    San Francisco, CA
    4 days ago
  • $195k - $365k

     ...is building the real-world AI interface for professionals...  ...interaction. Gain exposure to cutting-edge AI for Pro tools and play a...  ...and training large-scale audio or speech models from the ground up, whether...  ...at the intersection of research and engineering, eager to design... 
    Audio
    Full time
    Work at office
    Worldwide

    Plaud

    San Francisco, CA
    4 days ago
  • Plaud Inc. is seeking senior AI researchers to join our SpeechLLM lab in San Francisco. You will help build and train large-scale audio/speech models and push the boundaries of human-AI interaction. We value hands-on experience with PyTorch or JAX, distributed training... 
    Audio

    Plaud

    San Francisco, CA
    4 days ago
  •  ...mission is to architect AI that learns from...  ...'re pioneering the model architectures that...  ...and ship cutting edge models and experiences...  ...the Role On the Audio Post-Training team,...  ...customer needs meet research, and covers the full...  ...generative models (speech, text, or... 
    Audio
    Work at office
    Visa sponsorship
    Flexible hours

    Cartesia

    San Francisco, CA
    3 days ago
  •  ...building the future of voice AI operating systems for...  ...inflection point where advances in speech, language models, and clinical AI can...  ...are hiring two ML Engineers / Researchers to help build the next generation...  ...one of two areas: Speech & Audio: Build state-of-the-art... 
    Audio
    Full time

    Knowtex

    San Francisco, CA
    1 day ago
  • $117.2k - $313.7k

     ...SalesforceSalesforce is the #1 AI CRM, where humans...  ...AI Research is a global leader...  ...art large language models including CodeGen;...  ...translate cutting-edge research into production...  ...multimodal AI agents.Speech Intelligence: Voice...  ...recognition, and low-latency audio processing.Core... 
    Audio
    Full time
    Worldwide

    Salesforce

    San Francisco, CA
    2 days ago
  • Capital One is seeking an Applied Researcher I in San Francisco to join the AI Foundations team. You will collaborate with data scientists, software engineers...  ...platforms. The role emphasizes applied research, model development, and translating research into business outcomes... 

    Capital One

    San Francisco, CA
    4 days ago
  •  ...institution based in San Francisco is looking for an Applied Researcher to work on AI-powered products. This role involves delivering innovative...  ...Responsibilities include partnering with teams to build AI models and conducting impactful research. #J-18808-Ljbffr Capital... 

    Capital One

    San Francisco, CA
    2 days ago
  • Acceler8 Talent is seeking an ML Researcher focused on World Models to join an early-stage lab-backed venture in San Francisco. You will explore world model approaches and contribute to research that informs real-world voice products with production-ready implications.... 

    Acceler8 Talent

    San Francisco, CA
    4 days ago
  • As a Research Scientist , you'll lead cutting-edge research that advances the state of generative AI for long-form storytelling. You'll work at the...  ...on large language models and multimodal foundation...  ...methods that unify text, speech, audio, and other modalities into... 
    Audio
    Worldwide

    Pocket FM

    San Francisco, CA
    2 days ago
  • $117.2k - $313.7k

    Salesforce AI Research is looking for outstanding AI Research...  ..., develop novel models, and bridge the gap between...  ...multimodal AI agents. Speech Intelligence: Voice intelligence...  ..., and low‑latency audio processing. Core...  ...and translate cutting‑edge ideas into practical, high... 
    Audio

    salesforce.com, inc.

    San Francisco, CA
    1 day ago
  • Verita AI is seeking an Applied AI Researcher to work with clients on model evaluation and data strategy. You will assess model performance, identify failure modes, and design data-driven solutions, collaborating with operations and engineering to implement scalable data... 

    Verita AI

    San Francisco, CA
    3 days ago
  • $34 per hour

     ...our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role...  ...power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are trained on accurate, well-structured, and... 
    Audio
    Work experience placement
    Remote work

    Welocalize

    San Francisco, CA
    1 day ago
  • $26 - $28 per hour

     ...our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role...  ...power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are trained on accurate, well-structured, and... 
    Audio
    Work experience placement
    Remote work

    Welocalize

    San Francisco, CA
    3 days ago
  • $26 - $28 per hour

     ...job Czech Data Labeling Analyst(Speech & Voice ) Overview...  ...Analysts, supporting speech and voice AI systems. This is a high-impact...  ...systems. You'll be working with audio, speech, and language data - helping ensure models are trained on accurate, well-structured... 
    Audio
    Full time
    Work experience placement
    Remote work
    Visa sponsorship

    Welocalize

    San Francisco, CA
    2 days ago
  • $150k - $200k

     ..., our team is tackling cutting-edge engineering challenges to bring...  ...Senior Firmware Engineer, Edge AI / NPU Runtime to help architect...  ...product. You’ll help define how models run on-device, how sensor data...  ...as biosignals, sensor fusion, audio, gesture recognition, keyword spotting... 
    Audio
    Visa sponsorship

    Tacit

    San Francisco, CA
    2 days ago
  •  ...implementing infrastructure for large-scale multimodal models, focusing on high-performance delivery of audio and image inputs. You'll collaborate closely with researchers and product teams to push the boundaries of AI technology, ensuring reliable production services. If... 
    Audio

    Jobleads-US

    San Francisco, CA
    4 days ago
  • A technology company specializing in AI is seeking an experienced DSP Engineer to design and develop innovative audio processing algorithms. You will work on deep learning models for tasks such as speech enhancement and voice identification, while contributing to real-... 
    Audio

    femtoAI

    San Bruno, CA
    2 days ago
  • $216.3k - $280.8k

    Meet the TeamAt Foundation AI, we are leading frontier AI research across Cisco. Our mission is to advance the state...  ...AI.Our research spans foundation models, agentic AI, multimodal learning,...  ...work at the intersection of cutting-edge research and real-world impact, developing... 
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Francisco, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Edge AI Researcher (Speech & Audio Models). Be the first to apply!