Staff / Principal Machine Learning Engineer, Serving - USA [Remote]
$270kInworld AI
- Remote job
Inworld is a research lab of top researchers and engineers, building the world’s top-ranked realtime voice models.
Today our models are the #1 ranked realtime voice models in the world. They are used to power the largest consumer-facing AI applications available, across categories like health, fitness, learning, therapy, companions, customer experience and media; representing 100s of millions of end users. Our work spans areas like research and development of state-of-the-art models, optimizing realtime inference, and creating best-in-class APIs and products that allow developers to engage their users.
We’ve raised more than $125M from Lightspeed, Section 32, Kleiner Perkins, Microsoft’s M12 venture fund, Founders Fund, Meta and Stanford, among others. Our technology has powered experiences from companies such as NVIDIA, Microsoft Xbox, Niantic, Logitech Streamlabs, Wishroll, Little Umbrella and Bible Chat. We’ve also been recognized by CB Insights as one of the 100 most promising AI companies globally and have been named one of LinkedIn’s Top 10 Startups in the USA.
Who We're Looking ForA year ago, reliably working agentic systems and sub-second multimodal inference at scale barely existed. Nobody has a decade of experience here. So we're not screening for a resume template — we're looking for strong people from varied backgrounds who learn fast, thrive in ambiguity, and can show us what they've built, broken, and understood.
Experience We Find Useful You don't need all of this. But you need enough to make a case.
- Inference Optimization. Deep understanding of modern serving frameworks and techniques like vLLM or TRT-LLM.
- Model Acceleration . Hands-on experience with quantization, distillation, caching strategies , continuous batching, paged attention, and speculative decoding.
- High-Performance Systems. Proficiency in C++, CUDA, Rust, or highly optimized Python. You know how to profile code and squeeze every ounce of performance out of NVIDIA GPUs.
- Distributed Systems & Scaling. Experience with Kubernetes, Ray, custom load balancing, multi-GPU/multi-node inference, and reliably handling thousands of concurrent connections.
- Public work. Non-trivial systems programming projects, open-source contributions to major inference engines, or deep-dive technical write-ups.
- Full-cycle ownership. You can take a model from the research team, containerize it, optimize its serving, and ensure it runs reliably in production.
- Background. PhD in CS, Physics, Math, or equivalent practical experience building backend or ML systems.
- You don’t need a roadmap to start walking; you’re comfortable picking a direction and building the map as you go.
- You believe engineering isn't finished until it’s shipped and stable. You have a bias for impact over purely theoretical optimizations.
- You don't just ship code; you obsess over the why. You’re the first to question an architecture if you think there’s a better way to solve the core latency or throughput problem.
- You aren't satisfied with "the PM said so." You thrive on deep context and want to understand the fundamental logic behind every decision we make.
We hand you unclear problems and expect you to make them clear. We value engineers who say "I don't know yet" and then design the benchmark or prototype that finds out. We treat performance, latency, and reliability as first-class product features, not a box to check before launch. Impact comes before everything else, though we support sharing work and open-source contributions that move the field forward. Your work should be visible. Flat structure, fast iterations, minimal process theater.
We believe in the power of in-person collaboration to solve the hardest problems and foster a strong team culture. We offer relocation assistance and look forward to you joining us in our Mountain View office.
The base salary range for this full-time position is $270,000 - $500,000+ bonus + equity + benefits.
[Inworld Jobs Privacy](
$166k - $244k
...experts in energy, AI, software engineering, and products to build tools... ...matters at global scale. Learn more about our team and our mission... ...looking for an early career Machine Learning Engineer to join our... ...ML model training at serving at enterprise scale Stay abreast...SuggestedFull timeFlexible hours- ...Siri and Apple products more intelligent for our users? The Machine Learning Platform Technology team is building groundbreaking technology... ...Spotlight, Safari, Siri and upcoming ever exciting Apple products serving millions of queries every day with incredible low latencies,...Suggested
- ...interaction) Team: Inference & Reinforcement Learning Platform About the Role We’re looking for a Machine Learning Engineer (MLE) to work directly with customers and... ...inference stacks (e.g. vLLM / SGLang / Ray Serve–style architectures) Optimize latency, throughput...Suggested
- ...organizations that keep the world running. Our Team's Vision: Our Engineering team is shaping the future of cybersecurity. We thrive on... ...offer made will be contingent upon the applicant’s capacity to serve in compliance with U.S. export controls #LI-TD1 #LI-ONSITE...SuggestedImmediate start
$196k - $294k
...you’ll still consider applying. Want to learn more about life at Klaviyo? Visit... ...creators to own their own destiny. Sr. Machine Learning Engineer Palo Alto, CA (Onsite 5x a week)... ...the infrastructure and application that serve as the interface between businesses and...Suggested$173k - $259k
...themselves, live in the moment, learn about the world, and have fun together... ...other digital services. Snap Engineering teams build fun and technically... ...you’ll do Build and deploy machine learning models that power core products, serving millions of Snapchatters...Work experience placementLive inLocal area- ...Full time Location Type On-site Department Engineering Role Overview As a Machine Learning Engineer, you will play a central role in... ...infrastructure for LLMs/VLMs, including expertise in model serving frameworks like vLLM , TGI. Proficient in Python...Full time
$218.4k - $327.6k
...Location Mountain View, CA On-site Seniority Staff Compared with 41 other Engineering roles: Common across similar roles Unique to this role... ...to production: training, fine-tuning, distillation, serving Apply compression, quantization, pruning, and...$115k - $230k
...Rewards, and Great Careers. Senior Machine Learning Engineer, AI Research GEICO | Hybrid | Palo... ...workflows and applications. ~ Optimize ML serving systems for speed, reliability, and... ...with growth opportunities toward Staff and Senior Staff roles ~ Contribute...Hourly payFull timeWork experience placementLocal area$210k - $275k
...ML and work alongside industry-veteran scientists and engineers. As a Staff Machine Learning Engineer, you’ll bring your strong software engineering... ...systems in production across training, inference, and serving infrastructure, including model versioning, rollback strategies...Permanent employmentImmediate start$165.6k - $250.5k
...ecosystem, a forward-looking evolution of our machine learning infrastructure and models. We are seeking a Senior Machine Learning Engineer to join our Vector Ads Modeling team. In... ...product offering, advertisers goals, ads serving funnel, and rich real-time and dynamic...Full timeWork at officeWorldwide$105k - $215k
...trusted signals that enable downstream automation and decision-making across multiple lines of business. As a Staff Machine Learning Engineer, you will serve as a technical lead through the design, development, and deployment of advanced machine learning solutions...Hourly payFull timeWork experience placementLocal area$190k - $300k
...information, visit About the Role We are looking for a Machine Learning Engineer to build intelligent systems that connect consumers with relevant... ..., from data preparation and model training to online serving and monitoring. Partner with product, engineering, and data...Full timeInternshipLocal areaWork from home$270k
...is a research lab of top researchers and engineers, building the world’s top-ranked realtime... ...across categories like health, fitness, learning, therapy, companions, customer experience... ...one of LinkedIn’s Top 10 Startups in the USA. Who We're Looking For A year ago, reliably...Full timeWork at officeRelocation package$235k - $414k
...themselves, live in the moment, learn about the world, and have... ...digital services. Snap Engineering teams build fun and technically... .... We’re looking for a Principal Machine Learning Engineer to join our... ..., reinforce our values, and serve our community, customers and...Full timeLive inWork at officeLocal area$170k - $190k
...and a relentless focus on outcomes. ASAPP’s AI Engineering team is seeking an enterprising, talented and curious machine learning engineer. The AI Engineering team is... ...deploy them in a production setting designed to serve our customers at scale. We are looking for a...Work at office- ...intersection of natural language processing, machine learning, ML ops, and cloud computing. Each... ...stack data scientist/machine learning engineer, and you will be given ownership across... ...availed for stakeholders and users to serve above use-cases. Natural Language Processing...Remote work
$186k - $220k
...possible in AI. As a software engineer, you will work coactively... ...techniques in data-centric AI (active learning, multimodal intelligent... ...algorithms built on top of machine learning models to enable... ...infrastructure to enable training and serving large scale machine learning...Night shift- ...you, we'd love to talk. Role Overview: As a Senior MLOps Engineer, you will own the infrastructure that takes Nace.AI's models... ...Language Models (SLMs) in real time - which means our training, serving, and evaluation infrastructure isn't an afterthought; it is the...Full time
$206.4k - $379.1k
...surfaces and products, including Firefly, Photoshop, Illustrator, Express, Stock, and Premiere. We are hiring a Principal Machine Learning Engineer to serve as the technical lead for our GenAI Services area. This is not a model-training or research role — it is the senior...Temporary workLocal areaWorldwide- ...production stack during the Auriga migration window. Monitor system health, triage and resolve incidents, and address data/model-serving issues. Apply security updates, dependency patches, and ensure SLA continuity for downstream consumers. Lead migration of...
- ...MLOps Engineer Duration: 12+ Months Location: Sunnyvale CA - hybrid Rate: DOE Key Responsibilities: Design and implement scalable model serving platforms for both batch and real-time inference Build model deployment pipelines with automated testing and validation...
- ...you’ll still consider applying. Want to learn more about life at Klaviyo? Visit... ...creators to own their own destiny. Sr. Machine Learning Engineer Palo Alto, CA (Onsite 5x a week)... ...the infrastructure and application that serve as the interface between businesses and...Full time
$175k - $275k
...curated datasets, or full-cycle data engineering, Abaka AI provides the foundation for... ...the Role We’re hiring our first Machine Learning Engineer in the United States, a foundational... ...pipelines, data loaders, model serving, and evaluation frameworks. ~ Experience...Full timeImmediate startFlexible hours$150k - $230k
...information, visit About the Role We are looking for a hands-on Machine Learning Engineer to drive the post-training of our large language models,... ...model quality, prevent regressions, and ensure training–serving consistency. Stay current with post-training research...Full timeLocal areaWork from home$188.5k - $282.7k
...SAGE , Rubrik's Semantic AI Governance Engine, which is the first system designed to... ...curating data, training small models, serving them at production latency, and closing... ...degree (or higher) in Computer Science, Machine Learning, Computer Engineering, Statistics, or a...Permanent employmentFull timeLocal area- Overview Atlassian is looking for a Senior Machine Learning Engineer to join our Search & Intelligence organization. We build AI-native experiences... ...large language models.Experience with ML platforms, model serving, data pipelines, observability, or responsible AI....Work at officeLocal area
$195k - $230k
...information, visit About the Role We are looking for a Senior Machine Learning Engineer to help evolve our large-scale recommendation systems... ...will work on core feed, retrieval, and ranking systems serving tens of millions of users, while also participating in AI-...Full timeLocal areaWork from home- ...Ads Experimentation Platform team is looking for a senior machine learning engineer to lead the evolution of how we validate and optimize our global... ...models at scale. Cross-Functional Technical Leadership: Serve as the lead subject matter expert on experimentation for ML...Full timeTemporary workWork at officeWorldwideRelocation package
- ...The opportunity We are looking for a Senior Machine Learning Engineer to join our Ads Conversion Modeling team. In this role, you will design... ...state-of-the-art conversion rate prediction model. This model serves as the core intelligence behind delivering the right ad to...Full timeWork at officeWorldwideRelocation package
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff / Principal Machine Learning Engineer, Serving - USA [Remote]. Be the first to apply!
- software engineer staff Mountain View, CA
- technology administrator Mountain View, CA
- assistant engineer Mountain View, CA
- staff engineer Mountain View, CA
- senior staff systems engineer Mountain View, CA
- senior staff engineer Mountain View, CA
- engineering aide Mountain View, CA
- principal developer Mountain View, CA
- data center chief engineer Mountain View, CA
- engineering director Mountain View, CA


