Staff ML Systems Engineer - Reliable Inference & Serving
Jobleads-US
Human Intuition Inc. is building the autonomous company and seeking an experienced engineer to ensure dependable, efficient model serving behind agents and training rollouts. You will work on serving, routing, and evaluating performance metrics that matter for task completion.
The role emphasizes reliability, scalability, and cost-aware design, with collaboration across research and infrastructure teams to support evaluation and post-training workloads.
#J-18808-Ljbffr Jobleads-US$190k - $260k
...enterprises who are building AI systems. We believe that our... ...team of researchers, engineers, designers, and more,... ..., scalable and reliable machine learning systems... ...to join the Model Serving team at Cohere. The team... ...latency and throughput of inference.Strong understanding...SuggestedFull timeWork experience placementWork at officeLocal areaRemote workHome office$295k
...enterprises who are building AI systems. We believe that our... ...a team of researchers, engineers, designers, and more,... ...that enable fast, reliable, and scalable model training... ...the full stack of ML systems, this role gives... ...Familiarity with evaluation and serving frameworks (vLLM,...SuggestedFull timeWork at officeLocal areaRemote workHome office- ...Conversational AI system that integrates seamlessly... ....About the AI & ML Platform TeamOur... ...build productive, reliable tools that empower... ...a Senior ML System Engineer on the AI & ML Platform’s Inference team, you will design... ...large-scale model serving systems end-to-end....SuggestedWork at officeLocal area
$180k - $275k
...intelligence layer for any system that makes a... ...from the ground up. Engineers here own major surface... ...Engineer focused on Inference and Serving at Yobi , you’ll design... ...models into performant, reliable, and continuously... ...This is an applied ML systems role—equal parts...SuggestedRemote job$120 per hour
...Jack Dorsey . Position: MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) Type: Contract... ...profiling , debugging , and inference serving . Write accurate, well-structured... ...improve AI model performance on ML systems and training infrastructure...SuggestedRemote jobContract workSummer work- ...searchable securely and reliably through an... ...technical lead for Search Serving, you will set its... ...where new ML capabilities can meaningfully... ...agentic search systems, addressing tail... ..., and ML inference. Turn advances in... ...systems, mentor senior engineers, and align technical...Work at officeLocal area
$209k - $313k
...other digital services.Snap Engineering teams build fun and technically... ...understanding of causal inference and modern approaches to estimating... ...tests) and leveraging causal ML in production... ...faster, reinforce our values, and serve our community, customers and...Full timeLive inWork at officeLocal area$250k - $350k
...s mission is to develop reliable AI systems for the world’s most important... ..., paired with applied ML research, design, and... ...like.About the RoleAs a Staff Machine Learning Research Engineer, you will operate across... ...— training/fine-tuning, inference, memory and retrieval,...Full time- AI/ML Ops EngineerLocation: Remote / Hybrid... ...experienced AI/ML Engineer to design, deploy,... ...production systems that drive critical... ...stakeholders to deliver reliable and scalable AI... ...scalable model-serving architectures, including... ...-time APIs, batch inference pipelines, and...Full timeContract workLocal areaRemote workFlexible hours
- ...distributed, and we serve a large, active... ...three advanced engineering programs for... ...kind of call a Staff or Principal... ...and defend real systems. Your role is to... ..., platform, ML, and infrastructure... ..., ADK), agent reliability and guardrails,... ...serving and inference cost. AI Systems...Hourly payWork at officeRemote workAfternoon shift
$200k - $250k
...accurate, fast, and reliable, because they sit... ...measures it Own model serving in production: low-latency inference, batching,... ...data Set the bar for ML infrastructure as the... ...years of software engineering experience, including... ...infrastructure for ML or LLM systems in production Hands...Work at office- ...Yobi is seeking a Machine Learning Engineer focused on Inference and Serving. You will design, optimize, and operate systems that bring Behavioral AI models to life in real time across production environments. You’ll package, version, roll out, and observe models to...
- ...AI Engineer New York City, NY (in person, relocation... ...Engineer. You’ll build the ML systems that make the... ...most efficient, most reliable model for the job,... ...class imbalance, and serving latency as they do... ...training, hosting, and inference platform. ~ Design and...Full timeWork at officeRelocation
$61k - $101k
...Computer Science or a related Engineering field with 10+ years of... ...for accelerating LLM inference on specific GPU... ...models and the challenges of serving large transformer-based systems. We require solid knowledge... ...on performance and reliability. We will implement quantization...Full time- ...Lead ML Engineer Location: Remote, Nationwide Our... ...leadership role in the systems that move sophisticated... ..., evaluation, inference, and deployment while... ...training through evaluation, serving, deployment, and continuous... ...time, scalability, reliability, and infrastructure...Remote work
$130k - $250k
The Core Engineering The Core Engineering builds... ...leveraging cutting-edge AI/ML techniques,... ...scalable and reliable end-to-end AI/ML solutions... ...for real-time inference such as... ...design for distributed systems.Experience with data... ...the communities we serve to grow. Founded in...Full timeTemporary workPart time$200k - $265k
...value. We are building systems that model... ...stakeholders into reliable behavioral predictions... ...building training and inference pipelines, integrating... ...across applied ML research and production engineering, developing new models... ...organizations we serve, this work is mission...Full time$229.5k - $360k
...on a robust, flexible ML platform built for experimentation... ..., ensuring our systems remain fast, reliable, and at the forefront... ...blends innovation, engineering excellence, and a deep... ...store, real-time inference services, vector DBs, etc., that serve millions of transactions...Work at officeLocal areaRemote workMonday to ThursdayFlexible hours$209k - $313k
...digital services.Snap Engineering teams build fun and technically... ..., build, and deploy ML systems for personalization,... ...ML investmentsBuild reliable, observable, scalable production ML systems serving Snapchatters at... ...SWEsExperience with causal inference, uplift modeling,...Full timeLive inWork at officeLocal area- ...Decisioning & Optimization engineering team owns the systems that determine which ad wins... ...spans three platform areas:ML infrastructure for model serving: real-time inference at 1M+ QPS, multi-model... ...excellence for ML systems: reliability, observability, capacity planning...Hourly payFull timeImmediate startFlexible hoursShift work
$130k - $160k
...team building the systems these markets need... ...The Role An AI Engineer at Octaura, is a software... ...models are reliable, explainable, and... ...help shape Octaura’s ML architecture and... ...support ML training and inference. Collaborate... ..., and model serving. Implement MLOps...Full timeWork at officeRemote workFlexible hours- ...eventually any AI system making decisions about... .... Our ML-driven programmatic... ...learning recommendation engine. We're seed-... ...the models, the serving path, and the... ...training, online inference, and the parity between... ...raw event data into reliable, queryable,...
- ...Site Reliability Engineer Baseten powers mission-critical inference for the world's most dynamic AI companies... ...day 2 operations for our ML infrastructure platform... ...and build robust systems, processes, automations... ...models are deployed and served at scale will serve you...Flexible hours
$170k - $225k
...we’re looking for a Staff Machine Learning Engineer to take technical ownership... ...of our core ranking system. Every job request on... ...most consequential ML systems we run.This... ...around you. You’ll also serve as the primary... ...strategy, and production reliability of the core ranking system...Work at officeImmediate startFlexible hours$192.5k - $357.5k
...Machine Learning Engineer to join our Foundation... ...LLMs) and agentic systems, enabling them to... ...tasks. Build reliable interfaces between... ...sources.Scalable ML Systems & Productionization... ...training and inference systems for foundation... ...models. Serve as a technical authority...Full timeLocal areaWorldwideRelocation package$282.1k - $414.8k
...causal decisioning systems for New Verticals: grocery... ...Machine Learning Engineer to lead the Causal ML pod and establish... ...causal estimates remain reliable as policies,... ...0+ years) in causal inference, econometrics, experimentation... ...frameworks, serving patterns, and monitoring...Hourly payWork at officeLocal areaImmediate startRemote workFlexible hours$188k - $250k
...demanding training, inference, and high-... ...teams build the systems needed to... ...infrastructure reliably at scale.About... ...production software engineering to help engineers... ...the underlying ML systems,... ...About the roleAs a Staff Machine Learning... ...workflows, model-serving paths, and...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$207k - $300k
...automation, and alerting systems to ensure model... ...learning model training and inference performance, and programming... ...’s degree or PhD in Engineering, Computer Science, or a... ...models using AI and ML techniques to predict user... ...Ads business, which serves users and generates business...$145.68k - $178.05k
Job Description:Reliability Engineer (Tonawanda, NY)Collaborate with Innovative 3Mers Around the WorldChoosing... ...and keeping our deployed technology systems current. Perform failure mode and... ...develop & implement improvement plans.Serve as the equipment reliability Subject...Full timeH1bFlexible hours- # Senior ML Operations (MLOps) EngineerEight... ...as a Sr MLOps Engineer to help us bring... ...pipelines that ensure reliable delivery of models... ...to ensure ML inference operates reliably... ...high-performance ML systems by optimizing compute... ...platforms for serving and monitoring ML...Daily paidFull timeImmediate startRemote workWorldwideFlexible hoursNight shift
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff ML Systems Engineer - Reliable Inference & Serving. Be the first to apply!
- assistant engineer New York, NY
- senior staff systems engineer New York, NY
- technology administrator New York, NY
- engineering aide New York, NY
- senior staff engineer New York, NY
- staff design engineer New York, NY
- assistant engineering manager New York, NY
- staff data engineer New York, NY
- software engineer staff New York, NY
- staff engineer New York, NY





