Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Machine Learning Engineer (Real-Time Speech Translation)

$120k - $161.43k
Full-time

LILT

About LILT

AI is changing how the world communicates — and LILT is leading that transformation.

We're on a mission to make the world's information accessible to everyone , regardless of the language they speak. We use cutting-edge AI, machine translation, and human-in-the-loop expertise to translate content faster, more accurately, and more cost-effectively without compromising on brand, voice, or quality.

At LILT, we empower our teammates with leading tools, global collaboration, and growth opportunities to do their best work. Our company virtues— Work together, win together; Find a way or make one; Dance in the customer's shoes; Quicker than they expect; Quality is Job 1—guide everything we do. We are trusted by Intel Corporation, Canva, the United States Department of Defense, the United States Air Force, ASICS, and hundreds of global Enterprises. Backed by Sequoia, Intel Capital, and Redpoint, we’re building a category-defining company in a $50B+ global translation market being redefined by AI.

Role Summary

We are building a new live translation product. We are looking for an ML Engineer to build the real-time speech translation backend that powers it.

You will own the real-time speech translation backend end-to-end, from live audio input to translated output. You will build on LILT's production model serving platform (Ray Serve on GPU Kubernetes clusters) and our in-house adaptive machine translation models, working closely with the senior architects of that platform and with our language processing researchers. The ASR and MT models exist. Your job is to make them work together as a low-latency streaming system that holds up in production.

This is a hands-on backend engineering role for those looking to own a real-time ML system from end to end, supported by expert guidance and a clear product vision. Our team adopts an AI-first approach, leveraging agentic coding and AI-driven PR reviews to accelerate development. We combine this with deep technical expertise, requiring not only expert Python proficiency but also a comprehensive understanding of the entire ML stack, from optimizing neural network architectures to managing production infrastructure on Kubernetes, and making informed, cost-aware decisions on hardware selection.

Location & eligibility: This position requires US citizenship and residence in the United States. Preferred locations are Washington, D.C.; Boston, MA; and Indianapolis, IN (East Coast / ET timezone preferred).

Key Responsibilities

  • Real-time pipeline architecture: Build and manage services for high-throughput, real-time audio and text streaming. Handle signal processing, session lifecycles, and concurrency management to ensure robust operation under load.

  • ML model integration: Integrate and serve streaming speech recognition and machine translation models, collaborating with research teams to ensure models operate within required latency budgets.

  • Quality and confidence workflows: Develop logic for model-based confidence scoring, routing segments for human intervention as needed, and broadcasting real-time updates and corrections to end-users.

  • Infrastructure and scale: Architect and scale production ML infrastructure on GPU-accelerated Kubernetes clusters. Implement batching, load balancing, and autoscaling strategies to maintain performance and cost-efficiency.

  • Latency engineering: Establish comprehensive instrumentation for real-time performance. Identify bottlenecks, optimize system throughput, and drive down end-to-end latency metrics to meet production standards.

  • Interface and API definition: Define technical contracts and interfaces for audio ingestion and downstream service integrations. Partner with frontend and platform engineering teams to maintain clean, robust integration points.

  • Collaboration and technical leadership: Drive cross-team alignment by defining clear API interfaces and technical contracts, facilitating effective communication between engineering and product teams to ensure seamless system integration.

Required Qualifications

  • BS or MS in Computer Science or a related field, or equivalent practical experience.

  • 3+ years building production backend or ML serving systems in Python, including strong async programming (asyncio) skills.

  • Hands-on experience with real-time streaming transport: WebSocket or gRPC bidirectional streaming, session state, backpressure, and connection lifecycle handling.

  • Experience serving ML models in production on GPUs (Ray Serve, Triton, vLLM, or similar), with Docker and Kubernetes.

  • Experience integrating speech or NLP models into production systems, ideally streaming ASR (partial hypotheses, endpointing, VAD).

  • A latency-engineering mindset: you have profiled, instrumented, and optimized a real-time or low-latency system and can reason in per-stage budgets.

  • Effective use of AI coding agents (Claude Code, Codex, or similar) on top of fundamentals learned the hard way: you let agents do the typing, but you can debug, review, and reason about every line without them, and you know when not to trust them.

  • US citizenship and residence in the United States (contract requirement).

Preferred Qualifications

  • Ray Serve specifically, including streaming responses and model multiplexing.

  • Familiarity with simultaneous or incremental MT concepts (retranslation, prefix stability, wait-k policies).

  • Machine translation quality estimation (COMET/CometKiwi class models) or other confidence estimation in production.

  • Message brokers for real-time fan-out and state distribution (RabbitMQ or similar).

  • Streaming text-to-speech integration and time-to-first-audio optimization.

  • WebRTC and SFU concepts, or voice pipeline frameworks (LiveKit Agents, Pipecat).

  • Handling of CJK and other non-Latin text in NLP pipelines (our first languages are Japanese, Korean, and English).

  • Observability tooling (Datadog, Prometheus) for production ML systems.

Our Story

Our founders, Spence and John met at Google working on Google Translate. As researchers at Stanford and Berkeley, they both worked on language technology to make information accessible to everyone. While together at Google, they were amazed to learn that Google Translate wasn’t used for enterprise products and services inside the company.The quality just wasn’t there. So they set out to build something better. LILT was born.

LILT has been a machine learning company since its founding in 2015. At the time, machine translation didn’t meet the quality standard for enterprise translations, so LILT assembled a cutting-edge research team tasked with closing that gap. While meeting customer demand for translation services, LILT has prioritized investments in Large Language Models, human-in-the-loop systems, and now agentic AI.

With AI innovation accelerating and enterprise demand growing, the next phase of LILT’s journey is just beginning.

Our Tech

What sets our platform apart:

  • Brand-aware AI that learns your voice, tone, and terminology to ensure every translation is accurate and consistent

  • Agentic AI workflows that automate the entire translation process from content ingestion to quality review to publishing

  • 100+ native integrations with systems like Adobe Experience Manager, Webflow, Salesforce, GitHub, and Google Drive to simplify content translation

  • Human-in-the-loop reviews via our global network of professional linguists, for high-impact content that requires expert review


LILT in the News

  • Featured in The Software Report’s Top 100 Software Companies!

  • LILT makes it onto the Inc. 5000 List .

  • LILT’s continues to be an intellectual powerhouse, holding numerous patents that help power the most efficient and sophisticated AI and language models in the industry.

  • Check out all our news on our website .

Information collected and processed as part of your application process, including any job applications you choose to submit, is subject to LILT's Privacy Policy at .

At LILT, we are committed to a fair, inclusive, and transparent hiring process. As part of our recruitment efforts, we may use artificial intelligence (AI) and automated tools to assist in the evaluation of applications, including résumé screening, assessment scoring, and interview analysis. These tools are designed to support human decision-making and help us identify qualified candidates efficiently and objectively. All final hiring decisions are made by people. If you have any concerns, require accommodations, or would like to opt-out of the use of AI in our hiring process, please let us know at View email address on aiapply.co.

LILT is an equal opportunity employer. We extend equal opportunity to all individuals without regard to an individual’s race, religion, color, national origin, ancestry, sex, sexual orientation, gender identity, age, physical or mental disability, medical condition, genetic characteristics, veteran or marital status, pregnancy, or any other classification protected by applicable local, state or federal laws. We are committed to the principles of fair employment and the elimination of all discriminatory practices.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Machine Learning Engineer (Real-Time Speech Translation) in Washington DC vacancy
  •  ...WA, we are a team of engineers and technologists from...  ...computer vision, automatic speech recognition (ASR), and...  ...perform reliably in real-world environments....  ...will work closely with machine learning and data engineering teams...  ...using relational and time-series databases... 
    Suggested

    VTI Aerospace

    Washington DC
    a month ago
  •  ...BigBear.ai is seeking Data Engineers to build source adapters and normalization logic translating heterogeneous data into a common risk-signal schema. You will ingest...  ..., ensuring reliable ETL/ELT pipelines and near-real-time signal delivery. This role is remote with... 
    Suggested
    Remote job

    Jobleads-US

    McLean, VA
    2 days ago
  • $150k - $172k

     ...Description Job Description Title: Sr. Machine Learning Engineer, Autonomy Company: Heven AeroTech...  ..., verification, and delivery. Translate mission needs into defined autonomous...  ...-quality C++ and Python software for real-time and near-real-time execution.... 
    Suggested
    Permanent employment
    Temporary work
    Flexible hours

    Heven AeroTech

    Washington DC
    12 days ago
  • $120k - $140k

     ...The role We’re adding an Machine Learning Engineer to our team to help us build...  ...based on feedback from team and real-world performance...  ...engineers, product managers to help translate your work into production-...  ...Validation, Deep Learning and Time Series Analysis, Logistic... 
    Suggested
    Work at office

    Grid

    Washington DC
    a month ago
  •  ...Enterprise Readiness. Our mission is to establish readiness as a real-time condition that is continuously achieved. Today, a...  ...Job Description We are seeking an experienced Senior Machine Learning Engineer to join our AI/ML team and build the infrastructure that powers... 
    Suggested
    Full time
    Work at office
    Remote work

    Air

    Arlington, VA
    27 days ago
  • $170k

     ...a global networking leader, learn why there’s no better time to join the Extreme team. Senior Software Engineer (GenAI, ML): Experience:...  ...cutting edge of Generative AI, Machine Learning, Big Data, and...  ...perceive, learn, and act in real time. At Extreme, innovation... 
    Worldwide
    Shift work

    Extreme Networks

    Washington DC
    14 days ago
  •  ...Description Description: Type: Full-Time (W2) On-Site/Hybrid,...  ...solutions using cutting edge machine learning techniques to transform baseband...  ...-time Machine Learning DSP Engineer who will help combine...  ...systems in simulation and in real deployed products at scale. This... 
    Full time
    Flexible hours

    DeepSig Inc Careers Page

    Arlington, VA
    a month ago
  • Infinitive Inc in the United States is seeking an experienced Data Engineer to design, build, and scale event-driven data platforms. You will bridge real-time streaming with durable workflows using Kafka and Temporal, focusing on strict data contracts and robust schema... 

    Infinitive Inc

    Mc Lean, VA
    4 days ago
  • $77.6k - $176k

     ...Job Number: R0245042 Machine Learning Engineer The Opportunity As an experienced AI and ML engineer...  ...frameworks. Work with us to solve real-world challenges and define AI strategy...  ...demonstration of our values. Full-time and part-time employees working at least... 
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Washington DC
    4 days ago
  •  ...video, AI, radar, RF sensing, and AIS to deliver real-time maritime domain awareness at global scale. With 6...  ...are seeking a versatile and pragmatic Applied ML Engineer to contribute across a broad range of machine learning and perception tasks that power our edge-intelligent... 
    Remote work
    Flexible hours
    Shift work

    Quartermaster AI Inc

    Arlington, VA
    1 day ago
  • $115k - $140k

     ...develops advanced analytics and machine learning-based solutions to solve...  ...of passionate and motivated engineers with advanced degrees in engineering...  ...and unstructured text, time series, geospatial, and imagery...  ...technologies we develop have real world impact and US Government... 
    Full time
    Local area
    Night shift

    STR

    Arlington, VA
    25 days ago
  • $62k - $100k

     ...meaningful and rewarding career. With real flexibility to balance work...  ...About the Role As an AI Engineering team member, you will be...  ...ethically and with integrity at all times. Crowe accepts applications...  ...total rewards package. Learn more about what working at Crowe... 
    Full time
    Local area
    Worldwide

    Crowe

    Washington DC
    6 days ago
  • $150k - $200k

     ...Extreme! As a global networking leader, learn why there is no better time to join the Extreme team....  ...details Title of position: Principal Machine Learning Engineer Position type: Full time...  ...that perceive, learn, and act in real time. At Extreme, innovation is not... 
    Full time
    H1b
    Local area
    Work from home
    Work visa
    Flexible hours
    Shift work

    Extreme Networks

    Washington DC
    a month ago
  •  ...confidential search for a Principal Machine Learning Engineer to serve as the senior technical authority...  ...deep learning and LLM systems, and translate emerging techniques into practical,...  ...Background in feature stores and real-time inference systems Track record of... 
    Full time

    CONFIDENTIAL Scovai

    Washington DC
    15 days ago
  •  ...intelligence approaches to dynamic, real-world problems. We consider...  ...on projects related to machine learning, artificial intelligence,...  ...Science, Electrical and Computer Engineering, or related field...  ...office, with the remainder of time in Kitware's office Preferred... 
    Contract work
    Temporary work
    Work at office
    Worldwide
    Flexible hours
    2 days per week

    Kitware

    Arlington, VA
    a month ago
  •  ...the food they love and more time to enjoy it together. Where...  ...regular in-person events. Learn more about our flexible approach...  ...About the Role: As a Machine Learning Engineer, you will have the opportunity...  ...using machine learning to solve real-world problems with large... 
    Permanent employment
    Full time
    Work at office
    Remote work
    Work from home
    Flexible hours

    Instacart

    Adelphi, MD
    4 days ago
  • $197.3k - $225.1k

    Machine Learning Engineer 4 Join the Dealer Tech division within Capital One's Financial Services Technology...  ...doers and disruptors who love to solve real problems and meet real customer needs....  ..., or proof of concept At this time, Capital One will not sponsor a new applicant... 
    Full time
    Part time
    H1b
    Local area

    Capital One

    McLean, VA
    19 hours ago
  •  ...Description Job Description Machine Learning Engineer Location: McLean, VA (...  ...what AI/ML can do in the real world.  About the Role...  ...engineering team in McLean over time. You'll lead end-to-end...  ..., mentoring engineers, and translating complex technical work into... 
    Temporary work
    Work at office
    Local area
    Flexible hours
    Shift work

    CoVar

    McLean, VA
    a month ago
  • $142k - $194k

     ...years of experience in data science and/or machine learning, software development, computer science...  ...Science, Machine Learning, Computer Engineering, Mathematics, Physics, or a related...  ...frameworks to develop and optimize models for real-world business use cases. We value... 
    Full time

    Oracle

    Washington DC
    6 days ago
  • $95k - $141k

     ...Protagonist, you'll work on compelling projects that make a real difference. We seek talented individuals eager to...  ...Job Description Protagonist is looking for a Senior Machine Learning Engineer who builds production-ready systems for mission-critical work... 

    Protagonist

    Washington DC
    a month ago
  •  ...Mission Starts Here TheIncLab engineers and delivers intelligent...  ...We are looking for a Senior Machine Learning Engineer to that will focus on...  ...learning models to solve complex, real-world problems. We encourage...  ...production (batch or real-time inference) Background in research... 
    Flexible hours

    TheIncLab

    McLean, VA
    a month ago
  • $175k - $215k

     ...Job Description Job Description Who we are The real world is the next frontier, and at Metropolis, we are creating the...  ...help us create it. Who you are We are seeking a Senior Machine Learning Engineer to play a key role to join our growing team. As a key member... 
    Temporary work
    Work experience placement
    Work at office
    Local area

    Metropolis

    Washington DC
    20 days ago
  • $65 - $67 per hour

     ...Description Position Title: ML Data Engineer Location: Bethesda, MD -...  ...work end to end, and wants real upside growth potential in a...  ...analytics, reporting, and machine learning use cases. Implement ELT/...  ...with streaming or near–real-time data pipelines. Knowledge... 
    Contract work
    Local area
    Relocation
    Monday to Friday

    Addison Group

    Bethesda, MD
    a month ago
  • $314.8k - $359.3k

    Senior Staff Machine Learning Engineer Do you love building and pioneering in the AI and technology space...  ...doers and disruptors who love to solve real problems and meet real customer needs....  .... The minimum and maximum full-time annual salaries for this role are listed... 
    Full time
    Part time
    Local area

    Capital One

    McLean, VA
    13 hours ago
  • $229.9k - $262.4k

    Senior Manager, Machine Learning Engineer Do you love building and pioneering in the AI and technology...  ...doers and disruptors who love to solve real problems and meet real customer needs....  ...position. The minimum and maximum full-time annual salaries for this role are... 
    Full time
    Part time
    Local area

    Capital One

    McLean, VA
    13 hours ago
  •  ...industry experience with deep learning frameworks in Python, such as...  ...large, complex data sets for machine learning, including capture...  ...infrastructure focused software engineers. PyTorch or similar AI/ML...  .... Working with complex, real-world multimodal data. Audio... 
    Work experience placement

    SGS Consulting

    Washington DC
    more than 2 months ago
  •  ...Analytica is seekinga highly skilled Senior MLOps Engineer to lead the design, development, and deployment of machine learning operations infrastructure for defense...  ...implement robust MLOps pipelines that support real‑time decision‑making capabilities for defense applications... 

    Analytica

    Washington DC
    2 days ago
  •  ...transforming your career. Principal Machine Learning Engineer What you will do Let’s do this....  ...co-owning model-performance KPIs. Translate domain needs (R&D, Manufacturing, Commercial...  ...clustering, dimensionality reduction, time-series models, deep-learning... 
    Flexible hours

    MedDevicejobs.com

    Washington DC
    1 day ago
  • $160.6k - $200.8k

     ...manufacturing, data processing, and software engineering, our office is a truly inspiring mix of...  ...focuses on novel geospatial and time series analytics for customers requiring...  ...generative AI capabilities. As a Senior Machine Learning Engineer, you will drive hands‑on... 
    Full time
    Temporary work
    For contractors
    Work experience placement
    Work at office
    Local area
    Remote work
    Home office
    3 days per week

    Jobleads-US

    Arlington, VA
    2 days ago
  • $206k - $248k

     ...Machine Learning Engineer (TS/SCI Clearance With Full Scope Polygraph Required) Laurel, United States | Posted on 10/01/2026 Position: Software...  ...Level 3 Location: Laurel, MD Employment Type: Full-Time Clearance Requirement: Active TS/SCI with Polygraph... 
    Full time
    Local area

    Jobleads-US

    Laurel, MD
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Machine Learning Engineer (Real-Time Speech Translation). Be the first to apply!