Research Engineer, Audio and Speech
$200k - $400kDecagon
About Decagon
Decagon is the leading conversational AI platform empowering every brand to deliver concierge customer experiences.
Our technology enables industry-defining enterprises like Avis Budget Group, Block’s Cash App and Square, Chime, Oura Health, and Hunter Douglas to deploy AI agents that power personalized, deeply satisfying interactions across voice, chat, email, SMS, and every other channel.
We’re building a future where customer experiences are being redefined from support tickets and hold music to faster resolutions, richer conversations, and deeper relationships. We’re proud to be backed by world-class investors who share that vision, including a16z, Accel, Bain Capital Ventures, Coatue, and Index Ventures, along with many others.
We’re an in-office company, driven by a shared commitment to excellence and velocity. Our values — Just Get It Done, Invent What Customers Want, Winner’s Mindset, and The Polymath Principle — shape how we work and grow as a team.
About the Team
Read more about the Speech Research Team's work:
The Research team develops the model and decision-making stack that powers Decagon’s conversational agents for enterprise support. We research, adapt, and implement state-of-the-art techniques in model training, prompting, orchestration, and evaluation in order to make our agents more accurate, robust, and efficient in real-world deployments.
Our goal is to push the frontier of applied conversational AI: agents that reliably understand nuanced intent, track long context, and take the right actions under uncertainty. We measure success the way customers feel it: higher resolution rates, better user satisfaction, and consistent behavior at scale.
About the Role
As a Research Engineer focused on Audio and Speech, you’ll be responsible for building the models and agent harnesses that power Decagon’s real-time voice agents and taking them all the way from idea to production. Your work will advance multimodal and full-duplex systems that can listen, reason, speak, and respond naturally in real time.
We’re looking for strong engineers who want to build the next generation of AI voice agents. People here own their work end-to-end, ship real improvements, and are trusted to make high-impact technical decisions.
In this role, you will
Design and build next-generation agent harnesses optimized for streaming speech, turn-taking, interruptions, overlapping speech, and continuous interaction
Research and train multimodal and full-duplex models that jointly understand audio, reason, and generate speech
Improve speech recognition, voice activity detection, endpointing, and speech generation across diverse speakers, environments, domains, and languages
Build evaluations and use production calls to ship measurable improvements in accuracy, latency, naturalness, and task outcomes
Optimize end-to-end inference for responsiveness, throughput, stability, and cost, partnering with Voice Platform and Infrastructure teams to deploy at scale
Your background looks something like this
2+ years of experience in speech, audio ML, multimodal ML, or production machine learning
Experience developing or adapting autoregressive, diffusion, flow-matching, or codec-based speech models
Hands-on experience with streaming agent systems, low-latency inference, production model serving, and evaluation on real-world audio
Fluency in Python and a modern deep-learning framework such as PyTorch, with strong foundations in machine learning and signal processing
A track record of taking research ideas from prototype to reliable, measurable production impact
Even better if you have
Familiarity with speech-to-speech or full-duplex models
Experience with telephony, multilingual speech, noisy-channel robustness, speaker adaptation, or expressive speech generation
Compensation
$200K – $400K + Offers Equity
Benefits
We proudly offer the following benefits for our full-time employees:
Medical, Dental, and Vision benefits for you and your family
Life Insurance and Disability Benefits
Retirement Plan (e.g., 401K, pension)
Parental Leave
Fertility and family building benefits through Carrot
Monthly stipend to support your wellness, lifestyle, and work-life balance
Daily lunches and snacks in the office to keep you at your best
Take what you need vacation policy (subject to local requirements; UK employees receive 25 days of statutory leave)
These benefits are described in more detail in Decagon’s policies, may vary by location, and can change at any time according to applicable compensation and benefits plans.
$35 - $45 per hour
A tech company focused on AI is seeking an AI Tutor specialized in multilingual audio capabilities. In this role, you'll train and enhance the AI's ability in voice interactions and speech recognition, ensuring natural spoken interactions across diverse languages. Candidates...AudioHourly payRemote work$35 - $45 per hour
...technology firm is seeking an AI Tutor specialized in multilingual audio capabilities. This position focuses on training Grok to excel in... ...based on experience. Remote work is possible, aiming to bridge language barriers and improve AI's speech processing. #J-18808-LjbffrAudioHourly payFull timePart timeRemote work$270k - $310k
...inference calls monthly, process 1M+ hours of audio daily, and power 2 billion+ end-user... ...the Role We're looking for a Senior Research Engineer to join our Research team, developing... .... Optimize production inference for speech language models, both from a serving...Audio$250k - $300k
Hudson River Trading (HRT) is seeking an AI Research Engineer (Inference) to join the HAIL team. HAIL (HRT AI Labs) is the team at HRT responsible... ..., for any domain: robotics, biology, chemistry, physics, audio, video, recommendations, etc.Experience translating methods...AudioWork experience placementWork at officeLocal areaImmediate start- ...is seeking an AI Tutor to train Grok for multilingual audio, focusing on voice interactions, speech recognition, and cross-language performance. You will... ..., careful transcription, and collaboration with engineers to improve tools and manage diverse accents and noises...AudioRemote jobWorldwide
$113.7k - $211.9k
...responsibilities Conduct cutting‑edge research and development in Generative AI Develop... ...Collaborate with world‑class researchers and engineers to bring research ideas to production... ...models Strong publication record in audio/image/video generation and audio/image/video...AudioTemporary work- An innovative technology firm is seeking an AI Tutor to specialize in multilingual audio capabilities. You will be responsible for training speech recognition systems, enhancing interactions, and curating audio data for better accessibility. Ideal candidates possess native...AudioRemote jobFlexible hours
- ...seeking an AI Tutor specialized in multilingual audio capabilities to train Grok for voice interactions and speech processing across languages and accents. The role... ...of prosodic features, and collaborate with engineers to enhance annotation tools and #J-18808-Ljbffr...AudioRemote jobWorldwide
- Position: Voice Actor / Narrator AI Text-to-Speech Voice Capture Type: Short-Term Contract Location: Remote Commitment: Full-time (30 4... ...consistency in tone, pronunciation, and pacing across recordings Submit audio files in the required format within project timelines...AudioRemote jobFull timeTemporary workImmediate start
$45 per hour
...Remote Commitment: 2-10 hours per week for 1-2 weeks Responsibilities Record high-quality Hebrew audio recordings to support AI model training Deliver clear, natural speech with a neutral-to-upbeat “customer service” tone Follow transcripts precisely or adjust them slightly...AudioRemote jobHourly payContract work10 hours per week$26 - $28 per hour
...individuals to join our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role... ...datasets that power real-world AI systems. You’ll be working with audio, speech, and language data — helping ensure models are trained on...AudioFull timeRemote workVisa sponsorship$110.7k - $379.2k
Position Summary Research Engineer — Post-Training & Small Language Models (SLMs), Healthcare AI Three hundred fifty million Americans... ...-LLM, TGI, or Ollama. • Experience with multimodal models, speech models, or domain-specific foundation models; experience using...Local areaVisa sponsorship$34 per hour
...individuals to join our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role... ...that power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are trained on...AudioWork experience placementRemote work$26 - $28 per hour
...individuals to join our team as Data Labeling Analysts, supporting speech and voice AI systems. This is a high-impact production role... ...that power real-world AI systems. You'll be working with audio, speech, and language data — helping ensure models are trained on...AudioWork experience placementRemote work- ...automate offline and live evals that keep our speech and multimodal models honest in... ...models, not slide decks — partner with research and infra to prototype, train, and deploy... ...Qualifications:Expert-level PyTorch.Proven software engineer who loves ML; comfortable writing...Full timeContract workShift work
$35 - $45 per hour
...A technology firm in the United States is seeking candidates to manage multilingual audio data for AI projects. The ideal candidate must be a native Danish speaker with strong English skills, and the ability to transcribe and annotate audio accurately. The position offers...AudioHourly payRemote workFlexible hours$250k - $300k
...company, is hiring a Head of Research to join the team in New York.... ...as a recognised authority in audio data research and evaluation... ...recognized leader in audio AI, speech intelligence and AI... ...Collaborate with product and engineering teams to define AI data strategies...AudioFull time- ...Executive Director / Research Director In Ai ResearchJPMorgan Chase AI Research is a global team of research scientists, engineers and product managers that develops novel AI capabilities... ...modalities e.g. websites, news, audio, speech, documents, images, etc. Formal...AudioWork at officeShift work
- ...investors. I'll share more once we meet.About the RoleAs an ML Research Engineer at Maple, you'll be a part of our core product team... ...intentionally, and fix them just as fast.What You'll DoOptimize speech recognition (ASR), large language models (LLMs), and text-to-...Work at officeLocal area
- JPMorganChase AI Research is a global team of research scientists, engineers and product managers that develops novel AI capabilities and partners across the... ...spanning multiple modalities e.g. websites, news, audio, speech, documents, images, etc. Formal methods (e.g., temporal...AudioWork at officeShift work
$197.2k - $266.8k
...around the world.About the role...We are looking for a Senior AI Research Engineer to join our Duolingo Video Call team. The ideal candidate... ...machine learning techniques including large language models, speech models, benchmarking, and/or personalization. They will have...Work experience placement$180k - $240k
...inference calls monthly, process 1M+ hours of audio daily, and power 2 billion+ end-user... ...role: We're looking for a Senior Design Engineer to own the craft and feel of AssemblyAI's... ...before applying. Keep Exploring AssemblyAI: Speech-to-text Streaming speech-to-text Speech...AudioLive in$197.3k - $313.7k
...TeamSalesforce AI is looking for talented software and platform engineers to embed in our AI team to bridge the gap between frontier AI... ...where your engineering skills directly enable world-class research and products used by millions?At Salesforce, we are driving the...Full time- Flow Traders is looking for a Senior Research Engineer, successful candidate will be relocated to our Hong Kong Office. This is a unique opportunity to join a leading proprietary trading firm with an entrepreneurial and innovative culture at the heart of its business. We...Work at officeRelocation
$174k - $252k
...training recipes.Integrate novel ML and agentic techniques into research prototypes.Build evaluation pipelines, benchmarks, and... ...applied research projects.At Google, research-focused Software Engineers are embedded throughout the company, allowing them to setup large...- ...comPhone: (***) ***-****Job TitleResearch EngineerLocationNew York, NYAbout the OpportunityWe are seeking a disciplined and creative Research Engineer to join an elite systematic trading group. In this role, you will collaborate with seasoned technologists and quantitative...Immediate start
- ...automate the entire primary research workflow, from study design and... ...for a Senior Applied AI Engineer to help build the LLM and retrieval... ...unstructured, messy data - audio, transcripts, financial... ...systems. Familiarity with speech-to-text transcription models...AudioFull timeFlexible hours
$175k - $225k
...team that intimately collaborates with traders and quantitative researchers to implement, refine and deploy alpha signals, evaluate and... ...products that stand the test of time.About the RoleAs a Research Engineer, you will be an integral member of a systematic trading team...Temporary workFlexible hours$127k - $235k
...solutions for customers? Then come and apply your skills and passion for technology at Thomson Reuters Labs.We are seeking a Senior Research Engineer who will bring expertise in AI and ML and is interested in building data-driven capabilities that transform the way legal,...Full timeWork at officeLocal areaRemote workFlexible hours2 days per week3 days per week$175k - $250k
...FINRA, and other organizations.What you’ll doAs a Machine Learning Engineer - Applied Scientist you will play a critical role in developing... ..., and time series analysis. You will manage all aspects of the research process including methodology selection, data collection and...Work experience placement
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Engineer, Audio and Speech. Be the first to apply!
- engineering business analyst New York, NY
- research programmer New York, NY
- senior research engineer New York, NY
- junior machine learning research engineer New York, NY
- deep learning research engineer New York, NY
- research engineer New York, NY
- ai research engineer New York, NY
- research assistant engineering New York, NY
- research software engineer New York, NY
- audio video New York, NY



