Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff ML Systems Engineer - Reliable Inference & Serving

Jobleads-US

Human Intuition Inc. is building the autonomous company and seeking an experienced engineer to ensure dependable, efficient model serving behind agents and training rollouts. You will work on serving, routing, and evaluating performance metrics that matter for task completion.

The role emphasizes reliability, scalability, and cost-aware design, with collaboration across research and infrastructure teams to support evaluation and post-training workloads.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff ML Systems Engineer - Reliable Inference & Serving in New York, NY vacancy
  • $190k - $260k

     ...enterprises who are building AI systems. We believe that our...  ...team of researchers, engineers, designers, and more,...  ..., scalable and reliable machine learning systems...  ...to join the Model Serving team at Cohere. The team...  ...latency and throughput of inference.Strong understanding... 
    Suggested
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    1 day ago
  • $295k

     ...enterprises who are building AI systems. We believe that our...  ...a team of researchers, engineers, designers, and more,...  ...that enable fast, reliable, and scalable model training...  ...the full stack of ML systems, this role gives...  ...Familiarity with evaluation and serving frameworks (vLLM,... 
    Suggested
    Full time
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    1 day ago
  •  ...Conversational AI system that integrates seamlessly...  ....About the AI & ML Platform TeamOur...  ...build productive, reliable tools that empower...  ...a Senior ML System Engineer on the AI & ML Platform’s Inference team, you will design...  ...large-scale model serving systems end-to-end.... 
    Suggested
    Work at office
    Local area

    Atlassian

    New York, NY
    23 hours ago
  • $180k - $275k

     ...intelligence layer for any system that makes a...  ...from the ground up. Engineers here own major surface...  ...Engineer focused on Inference and Serving at Yobi , you’ll design...  ...models into performant, reliable, and continuously...  ...This is an applied ML systems role—equal parts... 
    Suggested
    Remote job

    Jobleads-US

    New York, NY
    2 days ago
  • $120 per hour

     ...Jack Dorsey . Position: MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) Type: Contract...  ...profiling , debugging , and inference serving . Write accurate, well-structured...  ...improve AI model performance on ML systems and training infrastructure... 
    Suggested
    Remote job
    Contract work
    Summer work

    Mercor

    New York, NY
    24 days ago
  •  ...searchable securely and reliably through an...  ...technical lead for Search Serving, you will set its...  ...where new ML capabilities can meaningfully...  ...agentic search systems, addressing tail...  ..., and ML inference. Turn advances in...  ...systems, mentor senior engineers, and align technical... 
    Work at office
    Local area

    Atlassian

    New York, NY
    4 days ago
  • $209k - $313k

     ...other digital services.Snap Engineering teams build fun and technically...  ...understanding of causal inference and modern approaches to estimating...  ...tests) and leveraging causal ML in production...  ...faster, reinforce our values, and serve our community, customers and... 
    Full time
    Live in
    Work at office
    Local area

    Snap

    New York, NY
    1 day ago
  • $250k - $350k

     ...s mission is to develop reliable AI systems for the world’s most important...  ..., paired with applied ML research, design, and...  ...like.About the RoleAs a Staff Machine Learning Research Engineer, you will operate across...  ...— training/fine-tuning, inference, memory and retrieval,... 
    Full time

    Scale AI

    New York, NY
    1 day ago
  • AI/ML Ops EngineerLocation: Remote / Hybrid...  ...experienced AI/ML Engineer to design, deploy,...  ...production systems that drive critical...  ...stakeholders to deliver reliable and scalable AI...  ...scalable model-serving architectures, including...  ...-time APIs, batch inference pipelines, and... 
    Full time
    Contract work
    Local area
    Remote work
    Flexible hours

    Slalom

    New York, NY
    1 day ago
  •  ...distributed, and we serve a large, active...  ...three advanced engineering programs for...  ...kind of call a Staff or Principal...  ...and defend real systems. Your role is to...  ..., platform, ML, and infrastructure...  ..., ADK), agent reliability and guardrails,...  ...serving and inference cost. AI Systems... 
    Hourly pay
    Work at office
    Remote work
    Afternoon shift

    TripleTen

    New York, NY
    1 day ago
  • $200k - $250k

     ...accurate, fast, and reliable, because they sit...  ...measures it Own model serving in production: low-latency inference, batching,...  ...data Set the bar for ML infrastructure as the...  ...years of software engineering experience, including...  ...infrastructure for ML or LLM systems in production Hands... 
    Work at office

    ZeroDrift, Inc.

    New York, NY
    23 hours ago
  •  ...Yobi is seeking a Machine Learning Engineer focused on Inference and Serving. You will design, optimize, and operate systems that bring Behavioral AI models to life in real time across production environments. You’ll package, version, roll out, and observe models to... 

    Jobleads-US

    New York, NY
    2 days ago
  •  ...AI Engineer New York City, NY (in person, relocation...  ...Engineer. You’ll build the ML systems that make the...  ...most efficient, most reliable model for the job,...  ...class imbalance, and serving latency as they do...  ...training, hosting, and inference platform. ~ Design and... 
    Full time
    Work at office
    Relocation

    Sharpe Recruiting Ventures

    New York, NY
    1 day ago
  • $61k - $101k

     ...Computer Science or a related Engineering field with 10+ years of...  ...for accelerating LLM inference on specific GPU...  ...models and the challenges of serving large transformer-based systems. We require solid knowledge...  ...on performance and reliability. We will implement quantization... 
    Full time

    J.P. Morgan

    New York, NY
    2 days ago
  •  ...Lead ML Engineer Location: Remote, Nationwide Our...  ...leadership role in the systems that move sophisticated...  ..., evaluation, inference, and deployment while...  ...training through evaluation, serving, deployment, and continuous...  ...time, scalability, reliability, and infrastructure... 
    Remote work

    LinkedIn

    New York, NY
    1 day ago
  • $130k - $250k

    The Core Engineering The Core Engineering builds...  ...leveraging cutting-edge AI/ML techniques,...  ...scalable and reliable end-to-end AI/ML solutions...  ...for real-time inference such as...  ...design for distributed systems.Experience with data...  ...the communities we serve to grow. Founded in... 
    Full time
    Temporary work
    Part time

    Goldman Sachs

    New York, NY
    1 day ago
  • $200k - $265k

     ...value. We are building systems that model...  ...stakeholders into reliable behavioral predictions...  ...building training and inference pipelines, integrating...  ...across applied ML research and production engineering, developing new models...  ...organizations we serve, this work is mission... 
    Full time

    Triangle Analytics

    New York, NY
    3 days ago
  • $229.5k - $360k

     ...on a robust, flexible ML platform built for experimentation...  ..., ensuring our systems remain fast, reliable, and at the forefront...  ...blends innovation, engineering excellence, and a deep...  ...store, real-time inference services, vector DBs, etc., that serve millions of transactions... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    New York, NY
    23 hours ago
  • $209k - $313k

     ...digital services.Snap Engineering teams build fun and technically...  ..., build, and deploy ML systems for personalization,...  ...ML investmentsBuild reliable, observable, scalable production ML systems serving Snapchatters at...  ...SWEsExperience with causal inference, uplift modeling,... 
    Full time
    Live in
    Work at office
    Local area

    Snap

    New York, NY
    14 hours ago
  •  ...Decisioning & Optimization engineering team owns the systems that determine which ad wins...  ...spans three platform areas:ML infrastructure for model serving: real-time inference at 1M+ QPS, multi-model...  ...excellence for ML systems: reliability, observability, capacity planning... 
    Hourly pay
    Full time
    Immediate start
    Flexible hours
    Shift work

    Netflix

    New York, NY
    3 days ago
  • $130k - $160k

     ...team building the systems these markets need...  ...The Role An AI Engineer at Octaura, is a software...  ...models are reliable, explainable, and...  ...help shape Octaura’s ML architecture and...  ...support ML training and inference. Collaborate...  ..., and model serving. Implement MLOps... 
    Full time
    Work at office
    Remote work
    Flexible hours

    Octaura

    New York, NY
    23 hours ago
  •  ...eventually any AI system making decisions about...  .... Our ML-driven programmatic...  ...learning recommendation engine. We're seed-...  ...the models, the serving path, and the...  ...training, online inference, and the parity between...  ...raw event data into reliable, queryable,... 

    PatternIQ

    New York, NY
    -28
  •  ...Site Reliability Engineer Baseten powers mission-critical inference for the world's most dynamic AI companies...  ...day 2 operations for our ML infrastructure platform...  ...and build robust systems, processes, automations...  ...models are deployed and served at scale will serve you... 
    Flexible hours

    Baseten

    New York, NY
    23 hours ago
  • $170k - $225k

     ...we’re looking for a Staff Machine Learning Engineer to take technical ownership...  ...of our core ranking system. Every job request on...  ...most consequential ML systems we run.This...  ...around you. You’ll also serve as the primary...  ...strategy, and production reliability of the core ranking system... 
    Work at office
    Immediate start
    Flexible hours

    Taskrabbit

    New York, NY
    1 day ago
  • $192.5k - $357.5k

     ...Machine Learning Engineer to join our Foundation...  ...LLMs) and agentic systems, enabling them to...  ...tasks. Build reliable interfaces between...  ...sources.Scalable ML Systems & Productionization...  ...training and inference systems for foundation...  ...models. Serve as a technical authority... 
    Full time
    Local area
    Worldwide
    Relocation package

    Genentech

    New York, NY
    4 days ago
  • $282.1k - $414.8k

     ...causal decisioning systems for New Verticals: grocery...  ...Machine Learning Engineer to lead the Causal ML pod and establish...  ...causal estimates remain reliable as policies,...  ...0+ years) in causal inference, econometrics, experimentation...  ...frameworks, serving patterns, and monitoring... 
    Hourly pay
    Work at office
    Local area
    Immediate start
    Remote work
    Flexible hours

    Doordash

    New York, NY
    2 days ago
  • $188k - $250k

     ...demanding training, inference, and high-...  ...teams build the systems needed to...  ...infrastructure reliably at scale.About...  ...production software engineering to help engineers...  ...the underlying ML systems,...  ...About the roleAs a Staff Machine Learning...  ...workflows, model-serving paths, and... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    New York, NY
    3 days ago
  • $207k - $300k

     ...automation, and alerting systems to ensure model...  ...learning model training and inference performance, and programming...  ...’s degree or PhD in Engineering, Computer Science, or a...  ...models using AI and ML techniques to predict user...  ...Ads business, which serves users and generates business... 

    Google

    New York, NY
    2 days ago
  • $145.68k - $178.05k

    Job Description:Reliability Engineer (Tonawanda, NY)Collaborate with Innovative 3Mers Around the WorldChoosing...  ...and keeping our deployed technology systems current. Perform failure mode and...  ...develop & implement improvement plans.Serve as the equipment reliability Subject... 
    Full time
    H1b
    Flexible hours

    3M

    New York, NY
    3 days ago
  • # Senior ML Operations (MLOps) EngineerEight...  ...as a Sr MLOps Engineer to help us bring...  ...pipelines that ensure reliable delivery of models...  ...to ensure ML inference operates reliably...  ...high-performance ML systems by optimizing compute...  ...platforms for serving and monitoring ML... 
    Daily paid
    Full time
    Immediate start
    Remote work
    Worldwide
    Flexible hours
    Night shift

    Coinscapture

    New York, NY
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff ML Systems Engineer - Reliable Inference & Serving. Be the first to apply!