Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff / Principal Machine Learning Engineer, Serving - USA [Remote]

$270k
Full-time

Inworld AI

About Inworld

Inworld is a research lab of top researchers and engineers, building the world’s top-ranked realtime voice models.

Today our models are the #1 ranked realtime voice models in the world. They are used to power the largest consumer-facing AI applications available, across categories like health, fitness, learning, therapy, companions, customer experience and media; representing 100s of millions of end users. Our work spans areas like research and development of state-of-the-art models, optimizing realtime inference, and creating best-in-class APIs and products that allow developers to engage their users.

We’ve raised more than $125M from Lightspeed, Section 32, Kleiner Perkins, Microsoft’s M12 venture fund, Founders Fund, Meta and Stanford, among others. Our technology has powered experiences from companies such as NVIDIA, Microsoft Xbox, Niantic, Logitech Streamlabs, Wishroll, Little Umbrella and Bible Chat. We’ve also been recognized by CB Insights as one of the 100 most promising AI companies globally and have been named one of LinkedIn’s Top 10 Startups in the USA.

Who We're Looking For

A year ago, reliably working agentic systems and sub-second multimodal inference at scale barely existed. Nobody has a decade of experience here. So we're not screening for a resume template — we're looking for strong people from varied backgrounds who learn fast, thrive in ambiguity, and can show us what they've built, broken, and understood.

Experience We Find Useful

You don't need all of this. But you need enough to make a case.

  • Inference Optimization. Deep understanding of modern serving frameworks and techniques like vLLM or TRT-LLM.

  • Model Acceleration . Hands-on experience with quantization, distillation, caching strategies , continuous batching, paged attention, and speculative decoding.

  • High-Performance Systems. Proficiency in C++, CUDA, Rust, or highly optimized Python. You know how to profile code and squeeze every ounce of performance out of NVIDIA GPUs.

  • Distributed Systems & Scaling. Experience with Kubernetes, Ray, custom load balancing, multi-GPU/multi-node inference, and reliably handling thousands of concurrent connections.

  • Public work. Non-trivial systems programming projects, open-source contributions to major inference engines, or deep-dive technical write-ups.

  • Full-cycle ownership. You can take a model from the research team, containerize it, optimize its serving, and ensure it runs reliably in production.

  • Background. PhD in CS, Physics, Math, or equivalent practical experience building backend or ML systems.

Who Thrives Here
  • You don’t need a roadmap to start walking; you’re comfortable picking a direction and building the map as you go.
  • You believe engineering isn't finished until it’s shipped and stable. You have a bias for impact over purely theoretical optimizations.
  • You don't just ship code; you obsess over the why. You’re the first to question an architecture if you think there’s a better way to solve the core latency or throughput problem.
  • You aren't satisfied with "the PM said so." You thrive on deep context and want to understand the fundamental logic behind every decision we make.
What Working Here Is Like

We hand you unclear problems and expect you to make them clear. We value engineers who say "I don't know yet" and then design the benchmark or prototype that finds out. We treat performance, latency, and reliability as first-class product features, not a box to check before launch. Impact comes before everything else, though we support sharing work and open-source contributions that move the field forward. Your work should be visible. Flat structure, fast iterations, minimal process theater.

We believe in the power of in-person collaboration to solve the hardest problems and foster a strong team culture. We offer relocation assistance and look forward to you joining us in our Mountain View office.

The base salary range for this full-time position is $270,000 - $500,000+ bonus + equity + benefits.

[Inworld Jobs Privacy](

Vacancy posted a month ago
Similar jobs that could be interesting for youBased on the Staff / Principal Machine Learning Engineer, Serving - USA [Remote] in Mountain View, CA vacancy
  • $166k - $244k

     ...experts in energy, AI, software engineering, and products to build tools...  ...matters at global scale. Learn more about our team and our mission...  ...looking for an early career Machine Learning Engineer to join our...  ...ML model training at serving at enterprise scale Stay abreast... 
    Suggested
    Full time
    Flexible hours

    Tapestry

    Mountain View, CA
    1 hour ago
  •  ...Siri and Apple products more intelligent for our users? The Machine Learning Platform Technology team is building groundbreaking technology...  ...Spotlight, Safari, Siri and upcoming ever exciting Apple products serving millions of queries every day with incredible low latencies,... 
    Suggested

    Applecart

    Sunnyvale, CA
    2 days ago
  •  ...interaction) Team: Inference & Reinforcement Learning Platform About the Role We’re looking for a Machine Learning Engineer (MLE) to work directly with customers and...  ...inference stacks (e.g. vLLM / SGLang / Ray Serve–style architectures) Optimize latency, throughput... 
    Suggested

    GMI Cloud

    Mountain View, CA
    20 hours ago
  •  ...organizations that keep the world running. Our Team's Vision: Our Engineering team is shaping the future of cybersecurity. We thrive on...  ...offer made will be contingent upon the applicant’s capacity to serve in compliance with U.S. export controls #LI-TD1 #LI-ONSITE... 
    Suggested
    Immediate start

    Illumio

    Sunnyvale, CA
    4 days ago
  • $196k - $294k

     ...you’ll still consider applying. Want to learn more about life at Klaviyo? Visit...  ...creators to own their own destiny. Sr. Machine Learning Engineer Palo Alto, CA (Onsite 5x a week)...  ...the infrastructure and application that serve as the interface between businesses and... 
    Suggested

    Klaviyo

    Palo Alto, CA
    3 days ago
  • $173k - $259k

     ...themselves, live in the moment, learn about the world, and have fun together...  ...other digital services. Snap Engineering teams build fun and technically...  ...you’ll do Build and deploy machine learning models that power core products, serving millions of Snapchatters... 
    Work experience placement
    Live in
    Local area

    Snap

    Palo Alto, CA
    2 days ago
  •  ...Full time Location Type On-site Department Engineering Role Overview As a Machine Learning Engineer, you will play a central role in...  ...infrastructure for LLMs/VLMs, including expertise in model serving frameworks like vLLM , TGI. Proficient in Python... 
    Full time

    NACE

    Palo Alto, CA
    2 days ago
  • $218.4k - $327.6k

     ...Location Mountain View, CA On-site Seniority Staff Compared with 41 other Engineering roles: Common across similar roles Unique to this role...  ...to production: training, fine-tuning, distillation, serving Apply compression, quantization, pruning, and... 

    Workman Labs

    Mountain View, CA
    2 days ago
  • $115k - $230k

     ...Rewards, and Great Careers. Senior Machine Learning Engineer, AI Research GEICO | Hybrid | Palo...  ...workflows and applications. ~ Optimize ML serving systems for speed, reliability, and...  ...with growth opportunities toward Staff and Senior Staff roles ~ Contribute... 
    Hourly pay
    Full time
    Work experience placement
    Local area

    GEICO

    Palo Alto, CA
    1 hour ago
  • $210k - $275k

     ...ML and work alongside industry-veteran scientists and engineers. As a Staff Machine Learning Engineer, you’ll bring your strong software engineering...  ...systems in production across training, inference, and serving infrastructure, including model versioning, rollback strategies... 
    Permanent employment
    Immediate start

    Cacheflow

    Mountain View, CA
    2 days ago
  • $165.6k - $250.5k

     ...ecosystem, a forward-looking evolution of our machine learning infrastructure and models. We are seeking a Senior Machine Learning Engineer to join our Vector Ads Modeling team. In...  ...product offering, advertisers goals, ads serving funnel, and rich real-time and dynamic... 
    Full time
    Work at office
    Worldwide

    Unity Technologies

    Mountain View, CA
    2 days ago
  • $105k - $215k

     ...trusted signals that enable downstream automation and decision-making across multiple lines of business. As a Staff Machine Learning Engineer, you will serve as a technical lead through the design, development, and deployment of advanced machine learning solutions... 
    Hourly pay
    Full time
    Work experience placement
    Local area

    GEICO

    Palo Alto, CA
    2 days ago
  • $190k - $300k

     ...information, visit  About the Role We are looking for a Machine Learning Engineer to build intelligent systems that connect consumers with relevant...  ..., from data preparation and model training to online serving and monitoring. Partner with product, engineering, and data... 
    Full time
    Internship
    Local area
    Work from home

    NewsBreak

    Mountain View, CA
    10 days ago
  • $270k

     ...is a research lab of top researchers and engineers, building the world’s top-ranked realtime...  ...across categories like health, fitness, learning, therapy, companions, customer experience...  ...one of LinkedIn’s Top 10 Startups in the USA. Who We're Looking For A year ago, reliably... 
    Full time
    Work at office
    Relocation package

    Inworld AI

    Mountain View, CA
    a month ago
  • $235k - $414k

     ...themselves, live in the moment, learn about the world, and have...  ...digital services. Snap Engineering teams build fun and technically...  .... We’re looking for a Principal Machine Learning Engineer to join our...  ..., reinforce our values, and serve our community, customers and... 
    Full time
    Live in
    Work at office
    Local area

    Snap Inc.

    Palo Alto, CA
    1 day ago
  • $170k - $190k

     ...and a relentless focus on outcomes. ASAPP’s AI Engineering team is seeking an enterprising, talented and curious machine learning engineer. The AI Engineering team is...  ...deploy them in a production setting designed to serve our customers at scale. We are looking for a... 
    Work at office

    ASAPP

    Mountain View, CA
    13 hours ago
  •  ...intersection of natural language processing, machine learning, ML ops, and cloud computing. Each...  ...stack data scientist/machine learning engineer, and you will be given ownership across...  ...availed for stakeholders and users to serve above use-cases. Natural Language Processing... 
    Remote work

    McKinsey & Company

    San Jose, CA
    1 day ago
  • $186k - $220k

     ...possible in AI. As a software engineer, you will work coactively...  ...techniques in data-centric AI (active learning, multimodal intelligent...  ...algorithms built on top of machine learning models to enable...  ...infrastructure to enable training and serving large scale machine learning... 
    Night shift

    Coactive Systems Corporation

    San Jose, CA
    2 days ago
  •  ...you, we'd love to talk. Role Overview: As a Senior MLOps Engineer, you will own the infrastructure that takes Nace.AI's models...  ...Language Models (SLMs) in real time - which means our training, serving, and evaluation infrastructure isn't an afterthought; it is the... 
    Full time

    Nace AI

    Palo Alto, CA
    4 days ago
  • $206.4k - $379.1k

     ...surfaces and products, including Firefly, Photoshop, Illustrator, Express, Stock, and Premiere. We are hiring a Principal Machine Learning Engineer to serve as the technical lead for our GenAI Services area. This is not a model-training or research role — it is the senior... 
    Temporary work
    Local area
    Worldwide

    Adobe

    San Jose, CA
    20 hours ago
  •  ...production stack during the Auriga migration window. Monitor system health, triage and resolve incidents, and address data/model-serving issues. Apply security updates, dependency patches, and ensure SLA continuity for downstream consumers. Lead migration of... 

    VBeyond

    Sunnyvale, CA
    2 days ago
  •  ...MLOps Engineer Duration: 12+ Months Location: Sunnyvale CA - hybrid Rate: DOE Key Responsibilities: Design and implement scalable model serving platforms for both batch and real-time inference Build model deployment pipelines with automated testing and validation... 

    Redolent

    Sunnyvale, CA
    4 days ago
  •  ...you’ll still consider applying. Want to learn more about life at Klaviyo? Visit...  ...creators to own their own destiny. Sr. Machine Learning Engineer Palo Alto, CA (Onsite 5x a week)...  ...the infrastructure and application that serve as the interface between businesses and... 
    Full time

    Klaviyo

    Palo Alto, CA
    26 days ago
  • $175k - $275k

     ...curated datasets, or full-cycle data engineering, Abaka AI provides the foundation for...  ...the Role   We’re hiring our first Machine Learning Engineer in the United States, a foundational...  ...pipelines, data loaders, model serving, and evaluation frameworks. ~ Experience... 
    Full time
    Immediate start
    Flexible hours

    Abaka Ai

    Palo Alto, CA
    more than 2 months ago
  • $150k - $230k

     ...information, visit  About the Role We are looking for a hands-on Machine Learning Engineer to drive the post-training of our large language models,...  ...model quality, prevent regressions, and ensure training–serving consistency. Stay current with post-training research... 
    Full time
    Local area
    Work from home

    NewsBreak

    Mountain View, CA
    22 days ago
  • $188.5k - $282.7k

     ...SAGE , Rubrik's Semantic AI Governance Engine, which is the first system designed to...  ...curating data, training small models, serving them at production latency, and closing...  ...degree (or higher) in Computer Science, Machine Learning, Computer Engineering, Statistics, or a... 
    Permanent employment
    Full time
    Local area

    Rubrik

    Palo Alto, CA
    more than 2 months ago
  • Overview Atlassian is looking for a Senior Machine Learning Engineer to join our Search & Intelligence organization. We build AI-native experiences...  ...large language models.Experience with ML platforms, model serving, data pipelines, observability, or responsible AI.... 
    Work at office
    Local area

    Atlassian

    Mountain View, CA
    19 hours ago
  • $195k - $230k

     ...information, visit  About the Role We are looking for a Senior Machine Learning Engineer to help evolve our large-scale recommendation systems...  ...will work on core feed, retrieval, and ranking systems serving tens of millions of users, while also participating in AI-... 
    Full time
    Local area
    Work from home

    NewsBreak

    Mountain View, CA
    23 days ago
  •  ...Ads Experimentation Platform team is looking for a senior machine learning engineer to lead the evolution of how we validate and optimize our global...  ...models at scale. Cross-Functional Technical Leadership: Serve as the lead subject matter expert on experimentation for ML... 
    Full time
    Temporary work
    Work at office
    Worldwide
    Relocation package

    Unity Technologies

    Mountain View, CA
    more than 2 months ago
  •  ...The opportunity We are looking for a Senior Machine Learning Engineer to join our Ads Conversion Modeling team. In this role, you will design...  ...state-of-the-art conversion rate prediction model. This model serves as the core intelligence behind delivering the right ad to... 
    Full time
    Work at office
    Worldwide
    Relocation package

    Unity Technologies

    Mountain View, CA
    more than 2 months ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff / Principal Machine Learning Engineer, Serving - USA [Remote]. Be the first to apply!