Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Machine Learning Engineer (Inference Platform)

$200k - $250k
Full-time

Wizard

About Wizard AI


At Wizard AI, we’re building a high-performing AI Shopping Agent that helps people discover the best products across the web with speed, accuracy, and trust. Our ML systems sit at the core of that experience, and we’re looking for a Senior MLOps Engineer to help us run them reliably and efficiently in production.

The Role


As a Senior MLOps Engineer at Wizard, you’ll own the end-to-end lifecycle of our ML systems — from packaging and deployment to monitoring, performance, and scaling — across a custom-built inference platform powering a live conversational product.

This isn’t a typical “pipeline” role. Our platform runs multiple specialized inference engines (LLMs, embeddings, and extraction models), each with different performance and scaling characteristics. A big part of the role is thinking through tradeoffs — latency vs. cost, throughput vs. reliability — and helping us evolve the system as we grow.

You’ll work closely with ML, Data, and DevOps, and have real input into how the platform is designed — not just how it’s maintained.

What You’ll Do



  • Build and improve production ML pipelines, making it easy to move models from experimentation to reliable production use

  • Help own and evolve our multi-engine inference platform (LLMs, embeddings, and extraction), improving how different workloads are served and scaled

  • Put strong foundations in place for model versioning, rollouts, and rollbacks so systems stay reproducible and safe to iterate on

  • Define and monitor key system metrics like latency, availability, and GPU utilization, and set clear expectations around performance

  • Improve overall system performance — whether that’s reducing latency, increasing throughput, or making better use of GPU resources

  • Design systems that are resilient and cost-aware, with thoughtful approaches to autoscaling, failure isolation, and graceful degradation

  • Bring solid engineering practices (testing, CI/CD, observability) into ML workflows to help the team move faster without sacrificing reliability

  • Partner closely with ML, Data, Product, and DevOps to turn ideas into production-ready systems and help guide technical decisions

What We’re Looking For



  • 5–8+ years of experience in software, ML, platform, or infrastructure engineering, with hands-on ownership of production ML systems

  • Experience deploying and running LLMs or other deep learning models in real-world environments

  • Strong Python skills and a solid foundation in software engineering

  • Familiarity with cloud platforms (AWS, GCP, Azure) and common ML tooling (model registries, experiment tracking, etc.)

  • A good understanding of inference performance — batching, memory usage, quantization, and how systems behave across CPU and GPU

  • Experience working with (or curiosity about) systems that serve different types of models with different constraints

  • Ability to think through tradeoffs between speed, cost, and reliability in a practical way

  • Comfort working in a fast-moving environment where things evolve quickly

What Success Looks Like


Reliable, Scalable Systems
Our ML systems run smoothly with clear visibility into performance, and can scale as demand grows without constant firefighting.

End-to-End Ownership
You’re able to take a model from idea to production and keep it running well, while making it easier for others to do the same.

Real Impact
You help shape how our ML platform evolves — improving performance, reducing costs, and making the overall system stronger over time.

Compensation & Benefits


The expected base salary range for this role is $200,000 – $250,000 USD, and will vary based on skills, experience, role level, and geographic location. Final compensation will be determined by considering these factors alongside overall role scope and responsibilities.

In addition to base salary, Wizard offers:


  • Equity in the form of stock options

  • Medical, dental, and vision coverage

  • 401(k) plan

  • Flexible PTO and company holidays

  • Fully remote work within the United States

  • Periodic company offsites and team gatherings

Wizard is committed to fair, transparent, and competitive compensation practices.

Vacancy posted 18 hours ago
Similar jobs that could be interesting for youBased on the Senior Machine Learning Engineer (Inference Platform) in United States vacancy
  • $200k - $250k

     ...experience, and we’re looking for a Senior MLOps Engineer to help us run them reliably and efficiently...  ...and scaling — across a custom-built inference platform powering a live conversational...  ...and running LLMs or other deep learning models in real-world environments... 
    Senior
    Full time
    Remote work
    Flexible hours

    Wizard

    United States
    2 days ago
  •  ...growth and superior returns, as we deliver rare value and impact across our businesses.  The Role As a Senior ML Engineer for AWS and Real-Time Inference, you'll own the fast path: ingesting live trading data and scoring it in near real time. It's a systems-heavy... 
    Senior
    Full time

    Twg Global Ai

    Remote
    2 days ago
  • NVIDIA is building the software stack for fast LLM inference on edge AI hardware in Westford, MA. This role focuses on evaluating open-source inference frameworks and mapping architectures to NVIDIA GPUs to maximize throughput and minimize latency. You will own validation... 
    Senior

    NVIDIA

    Westford, MA
    3 days ago
  • $228.7k - $306.7k

     ...is a global organization of engineers, product developers,...  ...building the products and platforms that will power our media,...  ...As our team's Sr Principal Machine Learning Engineer (IC leadership role...  ...librariesModel optimization and inference (TensorRT, ONNX, DeepSpeed)... 
    Senior

    Disney Interactive

    Seattle, WA
    2 days ago
  • $150k - $210k

     ...actionable recommendations. Our AI platform is central to this mission...  ...day. WHOOP is hiring a Senior AI/ML Engineer to help scale the...  ...of experience in applied machine learning, AI engineering, or ML-focused...  ...deployments with inference optimization, observability... 
    Senior
    Full time
    Work at office
    Relocation

    WHOOP

    Boston, MA
    3 days ago
  • $228.7k - $306.7k

    Job Posting Title:Senior Principal Machine Learning Engineer, Ad PlatformsReq ID:10150390Job Description:Technology...  ...and building the products and platforms that will power our media, advertising...  ...librariesModel optimization and inference (TensorRT, ONNX, DeepSpeed)Ad Tech... 
    Senior
    Full time

    Hulu

    Seattle, WA
    2 days ago
  • $295k - $405.5k

     ...a technology wholesale platform built on the belief that...  ...of tech, data, and machine learning to connect this thriving...  ....About this roleAs the Senior Staff Machine Learning Platform Engineer, you will own the technical...  ...including training, inference, feature management, governanceEstablish... 
    Senior
    Work experience placement
    Work at office
    Local area
    Remote work
    Monday to Friday
    Flexible hours
    3 days per week

    Faire

    San Francisco, CA
    3 days ago
  • Haus Analytics in Seattle is seeking a senior ML engineer to drive high-impact projects on the cMMM space, blending ML, causal inference, and scalable production code. You will collaborate with applied scientists, data engineers, and cross-functional teams to deliver trustworthy... 
    Senior
    Flexible hours

    Haus Analytics

    Seattle, WA
    3 days ago
  • NVIDIA AI is seeking an engineer to implement quantized and sparse recipes in inference engines and to manage model export pipelines for correct serialization. You will build benchmarking harnesses and data analysis tools to improve developer productivity through infrastructure... 
    Senior

    NVIDIA AI

    Redmond, WA
    3 days ago
  •  ...only architectures, combining rigorous engineering with learning systems proven in globally deployed...  ...pipelines for training, evaluation, and inference on multimodal datasets. Build and...  ...experience in ML infrastructure or platform engineering. ~ Strong coding skills... 
    Senior
    Local area

    FieldAI

    Irvine, CA
    6 days ago
  •  ...Description Fetch is looking for a Senior Manager of Machine Learning Engineering to lead the team building and...  ...learning capabilities across our Ad Platform. You will partner with Product, Data...  ...understanding of experimentation, causal inference, and incrementality measurement.... 
    Senior
    Full time

    Fetch

    Remote
    a month ago
  •  ...-leading training and inference speeds; over 10 times...  ...software systems that power engineering workflows across...  ...coordinates complex work across machines, clusters, development...  ...and reusable software platforms that allow engineers...  ...through continuous learning, growth and support of... 
    Senior

    Cerebras Systems

    Sunnyvale, CA
    4 days ago
  • $292.5k - $409.5k

     ...information, visit . Who We Are: The Machine Learning Platform team at Reddit is a high-impact...  ...teams. What You’ll Do: As a Senior Staff Software Engineer, you will help define and lead...  ...knowledge of model serving, inference pipelines, monitoring, and observability... 
    Senior
    Full time
    For contractors
    Work experience placement
    Flexible hours

    Reddit

    United States
    2 days ago
  • $152k - $241.5k

     ...parallel computing. More recently, GPU deep learning ignited modern AI — the next era of...  ...looking for an AI & Deep Learning Compiler Engineer. NVIDIA is hiring software engineers for...  ...DLC has been the backbone of NVIDIA’s inference engine, spanning across data centers, personal... 
    Senior
    Full time
    Remote work

    Nvidia

    Austin, TX
    18 hours ago
  • $193.93k - $352.29k

     ...why we’re building a universal autonomy platform: self-driving for all roads and all...  ...deploy core infrastructure components in machine learning model life cycle, to push the autonomous...  ...road validation.Maintain an in-house ML inference platform to serve large language models... 
    Senior
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    2 days ago
  • Job Title: Senior Machine Learning Engineer100% Remote Full Time As a Senior Member of Technical...  ...data preparation, training, evaluation, inference, and iteration.Turn research ideas...  ...closely with research, product, and engineering to deliver real user impact.Mentor and... 
    Senior
    Full time
    Remote work

    Spectraforce Technologies

    Seattle, WA
    1 day ago
  • $209k - $313k

     ...themselves, live in the moment, learn about the world, and have fun...  ...other digital services.Snap Engineering teams build fun and...  ...forefront.We’re looking for a Machine Learning Engineer to join Snap...  ...Strong understanding of causal inference and modern approaches to estimating... 
    Full time
    Live in
    Work at office
    Local area

    Snap

    New York, NY
    4 days ago
  •  ...in TikTok E-commerce. We are currently looking for talented software engineers that have a deep understanding of machine learning (ML), operations research (OR), data mining and statistical inference. This position can be fulfilled in our San Jose and Seattle offices.Responsibilities... 
    Senior
    Work experience placement

    TikTok

    Seattle, WA
    2 days ago
  •  ...technology products.As a Senior Lead Software Engineer at JPMorgan Chase within...  ...Corporate Sector, Infrastructure Platforms team, you are an integral...  ...understanding of machine learning concepts, including transformer...  ..., ML training, and inference.Experience with Infrastructure... 
    Senior
    For contractors

    JP Morgan Chase

    Seattle, WA
    1 day ago
  • $174.72k - $295.68k

     ...transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.We are looking for a full-time Machine Learning Engineer - AI Foundation, with deep knowledge and...  ...model and accelerating model training/inference. Our mission is to solve the autonomous... 
    Senior
    Full time

    XPENG Motors

    Santa Clara, CA
    18 hours ago
  • Core ResponsibilitiesDesign, build, and maintain end-to-end machine learning pipelines from research through production deployment. Engineer scalable training, inference, and retraining workflows using AWS SageMaker. Develop and maintain feature engineering, feature storage... 
    Senior
    Full time
    Work experience placement

    Vanguard

    Malvern, PA
    1 day ago
  • $174.72k - $295.68k

     ...transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.Our mission is...  ...fine tuning, PTQ, QAT, on-vehicle inference and related fields.Key...  ...Strong Python programming and software engineering skills.Ability to work effectively across... 
    Senior
    Full time

    XPENG Motors

    Santa Clara, CA
    18 hours ago
  • $170k - $205k

     ...developed by a team of seasoned leaders, engineers, AI scientists, and clinicians spun...  ...the position:We're looking for a Senior Machine Learning Engineer with deep expertise in some...  ...multimodal systems, large-scale training and inference infrastructure, model evaluation and... 
    Senior
    Flexible hours

    Videa Health

    Boston, MA
    2 days ago
  • $195k - $230k

     ...NewsBreak is the Content Intelligence platform shaping the future content economy....  ...About the RoleWe are looking for a Senior Machine Learning Engineer to help evolve our large-scale recommendation...  ...from offline training online inference A/B experimentation metric analysis.... 
    Senior
    Full time
    Local area
    Work from home

    News Break

    Mountain View, CA
    1 day ago
  • $224k - $356.5k

    NVIDIA is looking for a Machine Learning Engineer to join the GPU accelerated Apache Spark team.Apache...  ..., SQL, and ML/DL model training and inference pipelines, spanning many domains and...  ...years) with large-scale data processing platforms, such as Apache Spark.Proven ability... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    15 hours ago
  •  ...the center of that shift is our ML platform and the "Next Best Action" framework...  ...-level data (the DT User Card).As a Senior Machine Learning Engineer, you'll design and scale the ML systems...  ...features.Optimize training and inference for compute efficiency (CPU/GPU utilization... 
    Senior
    Full time
    Local area
    Shift work

    Fyber

    New York, NY
    18 hours ago
  •  ...webAI.About the Role:We are seeking a Senior Machine Learning Engineer to support our Public Sector...  ...systems using LoRA, PEFT, and on-device inference strategies, leveraging PyTorch, TensorFlow...  ...requirements, and the employment platform or entity through which the employee... 
    Senior
    Full time
    Live out
    Work at office
    Local area

    webAI

    Austin, TX
    2 days ago
  • $259.03k - $311.43k

     ...to explore, create, play, learn, and connect with friends in...  ...’re building the tools and platform that empower our community...  ...we are seeking experienced machine learning engineers who thrive on solving complex...  ...based model training, inference, and product integration are... 
    Senior
    Full time
    Work experience placement
    H1b
    Work at office
    Local area
    Visa sponsorship
    Monday to Friday

    Roblox

    San Mateo, CA
    18 hours ago
  • Apple Inc. seeks a Sr. Machine Learning Engineer for Foundation Models Inference in Cloud OS & Inference to advance private, scalable AI inference across Siri, Apple Intelligence, and app ecosystems. You will bridge research and production, optimizing inference, hardware... 
    Senior

    Apple

    Seattle, WA
    1 day ago
  • $230k - $265k

     ...ML and work alongside industry-veteran scientists and engineers. As a Senior Machine Learning Engineer, you’ll bring your strong software engineering...  ...implementation of training, fine-tuning, post-training, and inference strategies for large language and speech models using... 
    Senior
    Permanent employment

    Otter.ai

    Mountain View, CA
    18 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Machine Learning Engineer (Inference Platform). Be the first to apply!