Senior Machine Learning Engineer (Inference Platform)
$200k - $250kWizard
About Wizard AI
At Wizard AI, we’re building a high-performing AI Shopping Agent that helps people discover the best products across the web with speed, accuracy, and trust. Our ML systems sit at the core of that experience, and we’re looking for a Senior MLOps Engineer to help us run them reliably and efficiently in production.
The Role
As a Senior MLOps Engineer at Wizard, you’ll own the end-to-end lifecycle of our ML systems — from packaging and deployment to monitoring, performance, and scaling — across a custom-built inference platform powering a live conversational product.
This isn’t a typical “pipeline” role. Our platform runs multiple specialized inference engines (LLMs, embeddings, and extraction models), each with different performance and scaling characteristics. A big part of the role is thinking through tradeoffs — latency vs. cost, throughput vs. reliability — and helping us evolve the system as we grow.
You’ll work closely with ML, Data, and DevOps, and have real input into how the platform is designed — not just how it’s maintained.
What You’ll Do
- Build and improve production ML pipelines, making it easy to move models from experimentation to reliable production use
- Help own and evolve our multi-engine inference platform (LLMs, embeddings, and extraction), improving how different workloads are served and scaled
- Put strong foundations in place for model versioning, rollouts, and rollbacks so systems stay reproducible and safe to iterate on
- Define and monitor key system metrics like latency, availability, and GPU utilization, and set clear expectations around performance
- Improve overall system performance — whether that’s reducing latency, increasing throughput, or making better use of GPU resources
- Design systems that are resilient and cost-aware, with thoughtful approaches to autoscaling, failure isolation, and graceful degradation
- Bring solid engineering practices (testing, CI/CD, observability) into ML workflows to help the team move faster without sacrificing reliability
- Partner closely with ML, Data, Product, and DevOps to turn ideas into production-ready systems and help guide technical decisions
What We’re Looking For
- 5–8+ years of experience in software, ML, platform, or infrastructure engineering, with hands-on ownership of production ML systems
- Experience deploying and running LLMs or other deep learning models in real-world environments
- Strong Python skills and a solid foundation in software engineering
- Familiarity with cloud platforms (AWS, GCP, Azure) and common ML tooling (model registries, experiment tracking, etc.)
- A good understanding of inference performance — batching, memory usage, quantization, and how systems behave across CPU and GPU
- Experience working with (or curiosity about) systems that serve different types of models with different constraints
- Ability to think through tradeoffs between speed, cost, and reliability in a practical way
- Comfort working in a fast-moving environment where things evolve quickly
What Success Looks Like
Reliable, Scalable Systems
Our ML systems run smoothly with clear visibility into performance, and can scale as demand grows without constant firefighting.
End-to-End Ownership
You’re able to take a model from idea to production and keep it running well, while making it easier for others to do the same.
Real Impact
You help shape how our ML platform evolves — improving performance, reducing costs, and making the overall system stronger over time.
Compensation & Benefits
The expected base salary range for this role is $200,000 – $250,000 USD, and will vary based on skills, experience, role level, and geographic location. Final compensation will be determined by considering these factors alongside overall role scope and responsibilities.
In addition to base salary, Wizard offers:
- Equity in the form of stock options
- Medical, dental, and vision coverage
- 401(k) plan
- Flexible PTO and company holidays
- Fully remote work within the United States
- Periodic company offsites and team gatherings
Wizard is committed to fair, transparent, and competitive compensation practices.
$200k - $250k
...experience, and we’re looking for a Senior MLOps Engineer to help us run them reliably and efficiently... ...and scaling — across a custom-built inference platform powering a live conversational... ...and running LLMs or other deep learning models in real-world environments...SeniorFull timeRemote workFlexible hours- ...growth and superior returns, as we deliver rare value and impact across our businesses. The Role As a Senior ML Engineer for AWS and Real-Time Inference, you'll own the fast path: ingesting live trading data and scoring it in near real time. It's a systems-heavy...SeniorFull time
- NVIDIA is building the software stack for fast LLM inference on edge AI hardware in Westford, MA. This role focuses on evaluating open-source inference frameworks and mapping architectures to NVIDIA GPUs to maximize throughput and minimize latency. You will own validation...Senior
$150k - $210k
...actionable recommendations. Our AI platform is central to this mission... ...day. WHOOP is hiring a Senior AI/ML Engineer to help scale the... ...of experience in applied machine learning, AI engineering, or ML-focused... ...deployments with inference optimization, observability...SeniorFull timeWork at officeRelocation$228.7k - $306.7k
...is a global organization of engineers, product developers,... ...building the products and platforms that will power our media,... ...As our team's Sr Principal Machine Learning Engineer (IC leadership role... ...librariesModel optimization and inference (TensorRT, ONNX, DeepSpeed)...Senior$228.7k - $306.7k
Job Posting Title:Senior Principal Machine Learning Engineer, Ad PlatformsReq ID:10150390Job Description:Technology... ...and building the products and platforms that will power our media, advertising... ...librariesModel optimization and inference (TensorRT, ONNX, DeepSpeed)Ad Tech...SeniorFull time$295k - $405.5k
...a technology wholesale platform built on the belief that... ...of tech, data, and machine learning to connect this thriving... ....About this roleAs the Senior Staff Machine Learning Platform Engineer, you will own the technical... ...including training, inference, feature management, governanceEstablish...SeniorWork experience placementWork at officeLocal areaRemote workMonday to FridayFlexible hours3 days per week- Haus Analytics in Seattle is seeking a senior ML engineer to drive high-impact projects on the cMMM space, blending ML, causal inference, and scalable production code. You will collaborate with applied scientists, data engineers, and cross-functional teams to deliver trustworthy...SeniorFlexible hours
- NVIDIA AI is seeking an engineer to implement quantized and sparse recipes in inference engines and to manage model export pipelines for correct serialization. You will build benchmarking harnesses and data analysis tools to improve developer productivity through infrastructure...Senior
- ...only architectures, combining rigorous engineering with learning systems proven in globally deployed... ...pipelines for training, evaluation, and inference on multimodal datasets. Build and... ...experience in ML infrastructure or platform engineering. ~ Strong coding skills...SeniorLocal area
- ...Description Fetch is looking for a Senior Manager of Machine Learning Engineering to lead the team building and... ...learning capabilities across our Ad Platform. You will partner with Product, Data... ...understanding of experimentation, causal inference, and incrementality measurement....SeniorFull time
- ...-leading training and inference speeds; over 10 times... ...software systems that power engineering workflows across... ...coordinates complex work across machines, clusters, development... ...and reusable software platforms that allow engineers... ...through continuous learning, growth and support of...Senior
$292.5k - $409.5k
...information, visit . Who We Are: The Machine Learning Platform team at Reddit is a high-impact... ...teams. What You’ll Do: As a Senior Staff Software Engineer, you will help define and lead... ...knowledge of model serving, inference pipelines, monitoring, and observability...SeniorFull timeFor contractorsWork experience placementFlexible hours$152k - $241.5k
...parallel computing. More recently, GPU deep learning ignited modern AI — the next era of... ...looking for an AI & Deep Learning Compiler Engineer. NVIDIA is hiring software engineers for... ...DLC has been the backbone of NVIDIA’s inference engine, spanning across data centers, personal...SeniorFull timeRemote work- Job Title: Senior Machine Learning Engineer100% Remote Full Time As a Senior Member of Technical... ...data preparation, training, evaluation, inference, and iteration.Turn research ideas... ...closely with research, product, and engineering to deliver real user impact.Mentor and...SeniorFull timeRemote work
$193.93k - $352.29k
...why we’re building a universal autonomy platform: self-driving for all roads and all... ...deploy core infrastructure components in machine learning model life cycle, to push the autonomous... ...road validation.Maintain an in-house ML inference platform to serve large language models...SeniorImmediate startFlexible hours$209k - $313k
...themselves, live in the moment, learn about the world, and have fun... ...other digital services.Snap Engineering teams build fun and... ...forefront.We’re looking for a Machine Learning Engineer to join Snap... ...Strong understanding of causal inference and modern approaches to estimating...Full timeLive inWork at officeLocal area- ...in TikTok E-commerce. We are currently looking for talented software engineers that have a deep understanding of machine learning (ML), operations research (OR), data mining and statistical inference. This position can be fulfilled in our San Jose and Seattle offices.Responsibilities...SeniorWork experience placement
- ...technology products.As a Senior Lead Software Engineer at JPMorgan Chase within... ...Corporate Sector, Infrastructure Platforms team, you are an integral... ...understanding of machine learning concepts, including transformer... ..., ML training, and inference.Experience with Infrastructure...SeniorFor contractors
- Core ResponsibilitiesDesign, build, and maintain end-to-end machine learning pipelines from research through production deployment. Engineer scalable training, inference, and retraining workflows using AWS SageMaker. Develop and maintain feature engineering, feature storage...SeniorFull timeWork experience placement
$174.72k - $295.68k
...transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.We are looking for a full-time Machine Learning Engineer - AI Foundation, with deep knowledge and... ...model and accelerating model training/inference. Our mission is to solve the autonomous...SeniorFull time$174.72k - $295.68k
...transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.Our mission is... ...fine tuning, PTQ, QAT, on-vehicle inference and related fields.Key... ...Strong Python programming and software engineering skills.Ability to work effectively across...SeniorFull time$195k - $230k
...NewsBreak is the Content Intelligence platform shaping the future content economy.... ...About the RoleWe are looking for a Senior Machine Learning Engineer to help evolve our large-scale recommendation... ...from offline training online inference A/B experimentation metric analysis....SeniorFull timeLocal areaWork from home$170k - $205k
...developed by a team of seasoned leaders, engineers, AI scientists, and clinicians spun... ...the position:We're looking for a Senior Machine Learning Engineer with deep expertise in some... ...multimodal systems, large-scale training and inference infrastructure, model evaluation and...SeniorFlexible hours$230k - $265k
...ML and work alongside industry-veteran scientists and engineers. As a Senior Machine Learning Engineer, you’ll bring your strong software engineering... ...implementation of training, fine-tuning, post-training, and inference strategies for large language and speech models using...SeniorPermanent employment- Apple Inc. seeks a Sr. Machine Learning Engineer for Foundation Models Inference in Cloud OS & Inference to advance private, scalable AI inference across Siri, Apple Intelligence, and app ecosystems. You will bridge research and production, optimizing inference, hardware...Senior
$224k - $356.5k
NVIDIA is looking for a Machine Learning Engineer to join the GPU accelerated Apache Spark team.Apache... ..., SQL, and ML/DL model training and inference pipelines, spanning many domains and... ...years) with large-scale data processing platforms, such as Apache Spark.Proven ability...SeniorFull time- ...webAI.About the Role:We are seeking a Senior Machine Learning Engineer to support our Public Sector... ...systems using LoRA, PEFT, and on-device inference strategies, leveraging PyTorch, TensorFlow... ...requirements, and the employment platform or entity through which the employee...SeniorFull timeLive outWork at officeLocal area
$254k - $350k
...optimize high-performance deep learning models that generate dense,... ...on sparse sensor modalities .Engineer temporal processing modules to... ...for real-time on-vehicle inference, balancing high-fidelity range... ...Computer Science, Robotics, Machine Learning, or related field with...SeniorFull timeTemporary workRelocation package- ...the center of that shift is our ML platform and the "Next Best Action" framework... ...-level data (the DT User Card).As a Senior Machine Learning Engineer, you'll design and scale the ML systems... ...features.Optimize training and inference for compute efficiency (CPU/GPU utilization...SeniorFull timeLocal areaShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Machine Learning Engineer (Inference Platform). Be the first to apply!
- ai ml engineer United States
- graduate machine learning engineer United States
- staff machine learning engineer United States
- junior machine learning research engineer United States
- junior machine learning engineer United States
- senior ml engineer United States
- machine learning ai engineer United States
- lead machine learning engineer United States
- entry level machine learning engineer United States
- computer vision machine learning engineer United States

