Machine Learning Engineer (Real-Time Speech Translation)
$120k - $161.43kLILT
About LILT
AI is changing how the world communicates — and LILT is leading that transformation.
We're on a mission to make the world's information accessible to everyone , regardless of the language they speak. We use cutting-edge AI, machine translation, and human-in-the-loop expertise to translate content faster, more accurately, and more cost-effectively without compromising on brand, voice, or quality.
At LILT, we empower our teammates with leading tools, global collaboration, and growth opportunities to do their best work. Our company virtues— Work together, win together; Find a way or make one; Dance in the customer's shoes; Quicker than they expect; Quality is Job 1—guide everything we do. We are trusted by Intel Corporation, Canva, the United States Department of Defense, the United States Air Force, ASICS, and hundreds of global Enterprises. Backed by Sequoia, Intel Capital, and Redpoint, we’re building a category-defining company in a $50B+ global translation market being redefined by AI.
Role Summary
We are building a new live translation product. We are looking for an ML Engineer to build the real-time speech translation backend that powers it.
You will own the real-time speech translation backend end-to-end, from live audio input to translated output. You will build on LILT's production model serving platform (Ray Serve on GPU Kubernetes clusters) and our in-house adaptive machine translation models, working closely with the senior architects of that platform and with our language processing researchers. The ASR and MT models exist. Your job is to make them work together as a low-latency streaming system that holds up in production.
This is a hands-on backend engineering role for those looking to own a real-time ML system from end to end, supported by expert guidance and a clear product vision. Our team adopts an AI-first approach, leveraging agentic coding and AI-driven PR reviews to accelerate development. We combine this with deep technical expertise, requiring not only expert Python proficiency but also a comprehensive understanding of the entire ML stack, from optimizing neural network architectures to managing production infrastructure on Kubernetes, and making informed, cost-aware decisions on hardware selection.
Location & eligibility: This position requires US citizenship and residence in the United States. Preferred locations are Washington, D.C.; Boston, MA; and Indianapolis, IN (East Coast / ET timezone preferred).
Key Responsibilities
Real-time pipeline architecture: Build and manage services for high-throughput, real-time audio and text streaming. Handle signal processing, session lifecycles, and concurrency management to ensure robust operation under load.
ML model integration: Integrate and serve streaming speech recognition and machine translation models, collaborating with research teams to ensure models operate within required latency budgets.
Quality and confidence workflows: Develop logic for model-based confidence scoring, routing segments for human intervention as needed, and broadcasting real-time updates and corrections to end-users.
Infrastructure and scale: Architect and scale production ML infrastructure on GPU-accelerated Kubernetes clusters. Implement batching, load balancing, and autoscaling strategies to maintain performance and cost-efficiency.
Latency engineering: Establish comprehensive instrumentation for real-time performance. Identify bottlenecks, optimize system throughput, and drive down end-to-end latency metrics to meet production standards.
Interface and API definition: Define technical contracts and interfaces for audio ingestion and downstream service integrations. Partner with frontend and platform engineering teams to maintain clean, robust integration points.
Collaboration and technical leadership: Drive cross-team alignment by defining clear API interfaces and technical contracts, facilitating effective communication between engineering and product teams to ensure seamless system integration.
Required Qualifications
BS or MS in Computer Science or a related field, or equivalent practical experience.
3+ years building production backend or ML serving systems in Python, including strong async programming (asyncio) skills.
Hands-on experience with real-time streaming transport: WebSocket or gRPC bidirectional streaming, session state, backpressure, and connection lifecycle handling.
Experience serving ML models in production on GPUs (Ray Serve, Triton, vLLM, or similar), with Docker and Kubernetes.
Experience integrating speech or NLP models into production systems, ideally streaming ASR (partial hypotheses, endpointing, VAD).
A latency-engineering mindset: you have profiled, instrumented, and optimized a real-time or low-latency system and can reason in per-stage budgets.
Effective use of AI coding agents (Claude Code, Codex, or similar) on top of fundamentals learned the hard way: you let agents do the typing, but you can debug, review, and reason about every line without them, and you know when not to trust them.
US citizenship and residence in the United States (contract requirement).
Preferred Qualifications
Ray Serve specifically, including streaming responses and model multiplexing.
Familiarity with simultaneous or incremental MT concepts (retranslation, prefix stability, wait-k policies).
Machine translation quality estimation (COMET/CometKiwi class models) or other confidence estimation in production.
Message brokers for real-time fan-out and state distribution (RabbitMQ or similar).
Streaming text-to-speech integration and time-to-first-audio optimization.
WebRTC and SFU concepts, or voice pipeline frameworks (LiveKit Agents, Pipecat).
Handling of CJK and other non-Latin text in NLP pipelines (our first languages are Japanese, Korean, and English).
Observability tooling (Datadog, Prometheus) for production ML systems.
Our Story
Our founders, Spence and John met at Google working on Google Translate. As researchers at Stanford and Berkeley, they both worked on language technology to make information accessible to everyone. While together at Google, they were amazed to learn that Google Translate wasn’t used for enterprise products and services inside the company.The quality just wasn’t there. So they set out to build something better. LILT was born.
LILT has been a machine learning company since its founding in 2015. At the time, machine translation didn’t meet the quality standard for enterprise translations, so LILT assembled a cutting-edge research team tasked with closing that gap. While meeting customer demand for translation services, LILT has prioritized investments in Large Language Models, human-in-the-loop systems, and now agentic AI.
With AI innovation accelerating and enterprise demand growing, the next phase of LILT’s journey is just beginning.
Our Tech
What sets our platform apart:
Brand-aware AI that learns your voice, tone, and terminology to ensure every translation is accurate and consistent
Agentic AI workflows that automate the entire translation process from content ingestion to quality review to publishing
100+ native integrations with systems like Adobe Experience Manager, Webflow, Salesforce, GitHub, and Google Drive to simplify content translation
Human-in-the-loop reviews via our global network of professional linguists, for high-impact content that requires expert review
LILT in the News
Featured in The Software Report’s Top 100 Software Companies!
LILT makes it onto the Inc. 5000 List .
LILT’s continues to be an intellectual powerhouse, holding numerous patents that help power the most efficient and sophisticated AI and language models in the industry.
Check out all our news on our website .
Information collected and processed as part of your application process, including any job applications you choose to submit, is subject to LILT's Privacy Policy at .
At LILT, we are committed to a fair, inclusive, and transparent hiring process. As part of our recruitment efforts, we may use artificial intelligence (AI) and automated tools to assist in the evaluation of applications, including résumé screening, assessment scoring, and interview analysis. These tools are designed to support human decision-making and help us identify qualified candidates efficiently and objectively. All final hiring decisions are made by people. If you have any concerns, require accommodations, or would like to opt-out of the use of AI in our hiring process, please let us know at View email address on aiapply.co.
LILT is an equal opportunity employer. We extend equal opportunity to all individuals without regard to an individual’s race, religion, color, national origin, ancestry, sex, sexual orientation, gender identity, age, physical or mental disability, medical condition, genetic characteristics, veteran or marital status, pregnancy, or any other classification protected by applicable local, state or federal laws. We are committed to the principles of fair employment and the elimination of all discriminatory practices.
- ...WA, we are a team of engineers and technologists from... ...computer vision, automatic speech recognition (ASR), and... ...perform reliably in real-world environments.... ...will work closely with machine learning and data engineering teams... ...using relational and time-series databases...Suggested
- ...BigBear.ai is seeking Data Engineers to build source adapters and normalization logic translating heterogeneous data into a common risk-signal schema. You will ingest... ..., ensuring reliable ETL/ELT pipelines and near-real-time signal delivery. This role is remote with...SuggestedRemote job
$150k - $172k
...Description Job Description Title: Sr. Machine Learning Engineer, Autonomy Company: Heven AeroTech... ..., verification, and delivery. Translate mission needs into defined autonomous... ...-quality C++ and Python software for real-time and near-real-time execution....SuggestedPermanent employmentTemporary workFlexible hours$120k - $140k
...The role We’re adding an Machine Learning Engineer to our team to help us build... ...based on feedback from team and real-world performance... ...engineers, product managers to help translate your work into production-... ...Validation, Deep Learning and Time Series Analysis, Logistic...SuggestedWork at office- ...Enterprise Readiness. Our mission is to establish readiness as a real-time condition that is continuously achieved. Today, a... ...Job Description We are seeking an experienced Senior Machine Learning Engineer to join our AI/ML team and build the infrastructure that powers...SuggestedFull timeWork at officeRemote work
$170k
...a global networking leader, learn why there’s no better time to join the Extreme team. Senior Software Engineer (GenAI, ML): Experience:... ...cutting edge of Generative AI, Machine Learning, Big Data, and... ...perceive, learn, and act in real time. At Extreme, innovation...WorldwideShift work- ...Description Description: Type: Full-Time (W2) On-Site/Hybrid,... ...solutions using cutting edge machine learning techniques to transform baseband... ...-time Machine Learning DSP Engineer who will help combine... ...systems in simulation and in real deployed products at scale. This...Full timeFlexible hours
- Infinitive Inc in the United States is seeking an experienced Data Engineer to design, build, and scale event-driven data platforms. You will bridge real-time streaming with durable workflows using Kafka and Temporal, focusing on strict data contracts and robust schema...
$77.6k - $176k
...Job Number: R0245042 Machine Learning Engineer The Opportunity As an experienced AI and ML engineer... ...frameworks. Work with us to solve real-world challenges and define AI strategy... ...demonstration of our values. Full-time and part-time employees working at least...Full timeContract workPart timeWork at officeLocal areaRemote work- ...video, AI, radar, RF sensing, and AIS to deliver real-time maritime domain awareness at global scale. With 6... ...are seeking a versatile and pragmatic Applied ML Engineer to contribute across a broad range of machine learning and perception tasks that power our edge-intelligent...Remote workFlexible hoursShift work
$115k - $140k
...develops advanced analytics and machine learning-based solutions to solve... ...of passionate and motivated engineers with advanced degrees in engineering... ...and unstructured text, time series, geospatial, and imagery... ...technologies we develop have real world impact and US Government...Full timeLocal areaNight shift$62k - $100k
...meaningful and rewarding career. With real flexibility to balance work... ...About the Role As an AI Engineering team member, you will be... ...ethically and with integrity at all times. Crowe accepts applications... ...total rewards package. Learn more about what working at Crowe...Full timeLocal areaWorldwide$150k - $200k
...Extreme! As a global networking leader, learn why there is no better time to join the Extreme team.... ...details Title of position: Principal Machine Learning Engineer Position type: Full time... ...that perceive, learn, and act in real time. At Extreme, innovation is not...Full timeH1bLocal areaWork from homeWork visaFlexible hoursShift work- ...confidential search for a Principal Machine Learning Engineer to serve as the senior technical authority... ...deep learning and LLM systems, and translate emerging techniques into practical,... ...Background in feature stores and real-time inference systems Track record of...Full time
- ...intelligence approaches to dynamic, real-world problems. We consider... ...on projects related to machine learning, artificial intelligence,... ...Science, Electrical and Computer Engineering, or related field... ...office, with the remainder of time in Kitware's office Preferred...Contract workTemporary workWork at officeWorldwideFlexible hours2 days per week
- ...the food they love and more time to enjoy it together. Where... ...regular in-person events. Learn more about our flexible approach... ...About the Role: As a Machine Learning Engineer, you will have the opportunity... ...using machine learning to solve real-world problems with large...Permanent employmentFull timeWork at officeRemote workWork from homeFlexible hours
$197.3k - $225.1k
Machine Learning Engineer 4 Join the Dealer Tech division within Capital One's Financial Services Technology... ...doers and disruptors who love to solve real problems and meet real customer needs.... ..., or proof of concept At this time, Capital One will not sponsor a new applicant...Full timePart timeH1bLocal area- ...Description Job Description Machine Learning Engineer Location: McLean, VA (... ...what AI/ML can do in the real world. About the Role... ...engineering team in McLean over time. You'll lead end-to-end... ..., mentoring engineers, and translating complex technical work into...Temporary workWork at officeLocal areaFlexible hoursShift work
$142k - $194k
...years of experience in data science and/or machine learning, software development, computer science... ...Science, Machine Learning, Computer Engineering, Mathematics, Physics, or a related... ...frameworks to develop and optimize models for real-world business use cases. We value...Full time$95k - $141k
...Protagonist, you'll work on compelling projects that make a real difference. We seek talented individuals eager to... ...Job Description Protagonist is looking for a Senior Machine Learning Engineer who builds production-ready systems for mission-critical work...- ...Mission Starts Here TheIncLab engineers and delivers intelligent... ...We are looking for a Senior Machine Learning Engineer to that will focus on... ...learning models to solve complex, real-world problems. We encourage... ...production (batch or real-time inference) Background in research...Flexible hours
$175k - $215k
...Job Description Job Description Who we are The real world is the next frontier, and at Metropolis, we are creating the... ...help us create it. Who you are We are seeking a Senior Machine Learning Engineer to play a key role to join our growing team. As a key member...Temporary workWork experience placementWork at officeLocal area$65 - $67 per hour
...Description Position Title: ML Data Engineer Location: Bethesda, MD -... ...work end to end, and wants real upside growth potential in a... ...analytics, reporting, and machine learning use cases. Implement ELT/... ...with streaming or near–real-time data pipelines. Knowledge...Contract workLocal areaRelocationMonday to Friday$314.8k - $359.3k
Senior Staff Machine Learning Engineer Do you love building and pioneering in the AI and technology space... ...doers and disruptors who love to solve real problems and meet real customer needs.... .... The minimum and maximum full-time annual salaries for this role are listed...Full timePart timeLocal area$229.9k - $262.4k
Senior Manager, Machine Learning Engineer Do you love building and pioneering in the AI and technology... ...doers and disruptors who love to solve real problems and meet real customer needs.... ...position. The minimum and maximum full-time annual salaries for this role are...Full timePart timeLocal area- ...industry experience with deep learning frameworks in Python, such as... ...large, complex data sets for machine learning, including capture... ...infrastructure focused software engineers. PyTorch or similar AI/ML... .... Working with complex, real-world multimodal data. Audio...Work experience placement
- ...Analytica is seekinga highly skilled Senior MLOps Engineer to lead the design, development, and deployment of machine learning operations infrastructure for defense... ...implement robust MLOps pipelines that support real‑time decision‑making capabilities for defense applications...
- ...transforming your career. Principal Machine Learning Engineer What you will do Let’s do this.... ...co-owning model-performance KPIs. Translate domain needs (R&D, Manufacturing, Commercial... ...clustering, dimensionality reduction, time-series models, deep-learning...Flexible hours
$160.6k - $200.8k
...manufacturing, data processing, and software engineering, our office is a truly inspiring mix of... ...focuses on novel geospatial and time series analytics for customers requiring... ...generative AI capabilities. As a Senior Machine Learning Engineer, you will drive hands‑on...Full timeTemporary workFor contractorsWork experience placementWork at officeLocal areaRemote workHome office3 days per week$206k - $248k
...Machine Learning Engineer (TS/SCI Clearance With Full Scope Polygraph Required) Laurel, United States | Posted on 10/01/2026 Position: Software... ...Level 3 Location: Laurel, MD Employment Type: Full-Time Clearance Requirement: Active TS/SCI with Polygraph...Full timeLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Machine Learning Engineer (Real-Time Speech Translation). Be the first to apply!
- senior ml engineer Washington DC
- machine learning software engineer Washington DC
- machine learning engineer Washington DC
- machine learning ai engineer Washington DC
- ai ml engineer Washington DC
- computer vision machine learning engineer Washington DC
- internship machine learning Washington DC
- machine learning Washington DC
- artificial intelligence - machine learning intern Washington DC
- machine learning scientist Washington DC





