Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Systems Engineer

$145k - $165k

Bright Vision Technologies

ML Systems Engineer

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. Location: 100% Remote (U.S.) Position Type: Full-time Salary Range: $145,000–$165,000 Annually Experience Required: 6+ years Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary: We are seeking a ML Systems Engineer to design, build, and operate high-performance, highly reliable inference platforms for serving large machine learning models in production. The role focuses on the systems engineering side of AI deployment, including request routing, batching, caching, autoscaling, GPU utilization, and end-to-end observability across diverse model workloads. The ideal candidate brings strong distributed systems and performance engineering expertise, has shipped serving systems at scale, and understands the trade-offs between latency, throughput, cost, and quality in ML serving.

Required Qualifications

  • Bachelor's or Master's degree in Computer Science or a related field.
  • Six or more years of experience in distributed systems, infrastructure, or ML platform engineering.
  • Strong proficiency in Python and a systems language such as Go, Rust, or C++.
  • Deep experience operating high-throughput, low-latency services in production.
  • Hands-on experience with LLM or large model inference frameworks such as vLLM or TensorRT-LLM.
  • Strong understanding of GPU architecture, memory hierarchies, and accelerator utilization.
  • Familiarity with Kubernetes, autoscaling, and modern cloud platforms.
  • Experience with observability stacks including metrics, tracing, and structured logging.
  • Solid grounding in performance engineering and capacity planning.
  • Strong communication and incident response skills.

Preferred Qualifications

  • Open-source contributions to model serving infrastructure.
  • Experience with multi-region or globally distributed AI serving.
  • Familiarity with model quantization, distillation, and compression techniques.
  • Exposure to FinOps for AI workloads and cost-efficient serving design.
  • Experience supporting external-facing AI APIs at scale.
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the ML Systems Engineer in United States vacancy
  • $174.9k - $261.3k

     ...and understand the world!The Data Labeling Engineering team designs, builds, and operates hybrid...  ...engineering, data engineering, and AI/ML, defining the strategies, tooling, and quality...  ...leadership, and direct impact on systems that unblock the next generation of AV capabilities... 
    Suggested
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, TX
    13 hours ago
  • $90.1k - $191.8k

     ...and understand the world!The Data Labeling Engineering team designs, builds, and operates high‑...  ...engineering, data engineering, and ML, defining labeling strategies, tooling, and...  ...technical leadership, and work directly on systems that unblock the next generation of AV models... 
    Suggested
    Full time
    Work experience placement
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, TX
    4 days ago
  • $224k - $356.5k

     ...the next phase, we are building agentic systems that can reason about, build, evaluate, and...  ...about creating the meta-layer of modern ML: the agents, tooling, pipelines, and feedback...  .... We are looking for exceptional engineers who are passionate about the idea of AI-native... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    19 hours ago
  •  ...models for enterprises who are building AI systems. We believe that our work is instrumental...  .... Cohere is a team of researchers, engineers, designers, and more, who are all passionate...  ...you enjoy working across the full stack of ML systems, this role gives you the... 
    Suggested
    Full time
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    4 days ago
  •  ...work sits at the intersection of distributed systems, GPU performance, model training frameworks, RL pipelines, and production engineering. Your responsibilities Build and maintain...  ...with distributed model training, large-scale ML systems, or GPU cluster workloads.... 
    Suggested

    Nebius B.V.

    Palo Alto, CA
    19 hours ago
  • $300k - $400k

     ...possible. About the Role You will own the systems layer that makes our frontier model...  ...operations Profiling and benchmarking distributed ML systems to identify and eliminate...  ...team of the world’s best — the scientists, engineers, and problem-solvers who don’t just follow... 
    Visa sponsorship
    Flexible hours
    Shift work

    Periodic Labs

    Menlo Park, CA
    4 days ago
  •  ...ML Systems Engineer, ML Acceleration We are looking for a Machine Learning Systems Engineer to join our ML Acceleration team. In this role, you will be responsible for the core systems that enable our researchers to train frontier models at scale, focusing obsessively... 

    Motional

    Boston, MA
    3 days ago
  • $145k - $165k

     ...opportunity to join an established and well-respected organization offering tremendous career growth potential. Job Title: ML Systems Engineer Location: 100% Remote (U.S.) Position Type: Full-time, Direct W2 Salary Range: $145,000–$165,000 Annually... 
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    Bedford, TX
    19 hours ago
  • $110 per hour

     ...and Jack Dorsey . Position: MLOps Engineer (JAX, PyTorch, Pallas/Triton) Type:...  ...MLOps , training infrastructure, and ML framework-level topics . Design challenging...  ...-structured solutions to MLOps and ML systems problems . Evaluate MLOps tasks and... 
    Remote job
    Contract work
    Summer work
    Weekday work

    Mercor

    San Francisco, CA
    10 days ago
  • $110 per hour

     ...Role Responsibilities Guide research and engineering teams to close knowledge gaps and improve...  ...MLOps , training infrastructure, and ML framework-level topics . Design...  ...-structured solutions to MLOps and ML systems problems . Evaluate MLOps tasks and... 
    Contract work
    Summer work
    Remote work
    Weekday work

    Mercor

    New York, NY
    2 days ago
  •  ...Job Responsibilities: Engineer, design, implement, and improve highly-scalable machine learning systems and tools for enabling research Apply knowledge of relevant research...  ...Machine Learning Distributed training for ML models Experience with Machine Learning... 
    Work experience placement

    SGS Consulting

    Remote
    more than 2 months ago
  •  ...Job Responsibilities: Engineer, design, implement, and improve highly scalable machine learning systems and tools for enabling research Apply knowledge of relevant research...  ...experience ~0-2 years of Distributed ML Training (FSDP/DDP) experience ~5+ years of... 
    Work experience placement

    SGS Consulting

    Remote
    more than 2 months ago
  • $200.8k - $251k

     ...member to build and optimize a machine learning framework for large language models. Candidates should have system optimization experience and solid software engineering skills, particularly in tools like CUDA and Pytorch. This full-time position offers a competitive salary... 
    Full time

    Scale AI

    San Francisco, CA
    19 hours ago
  • General Motors' Data Labeling Engineering team builds and operates hybrid human/machine labeling tools powering autonomous vehicle ML models. We work across software, data, and ML to create scalable training data, with a modern full‑stack including TypeScript, React, GraphQL... 

    General Motors

    Seattle, WA
    1 day ago
  • Rhoda AI is hiring a Senior/Staff-level Research Engineer to ensure our robot-learning pipeline is reliable from data collection through...  ..., and real-robot evaluation. You will build validation systems, observability, and robust operating practices to distinguish model... 

    Socket.dev

    Mountain View, CA
    19 hours ago
  •  ..., Inc. is looking for a Member of Technical Staff focused on ML systems and inference in San Francisco. You will design and build inference...  .... Candidates should have strong foundations in software engineering, experience with ML inference systems, and performance tuning... 

    Gimlet Labs, Inc.

    San Francisco, CA
    1 day ago
  • NVIDIA Gruppe is seeking a Senior Engineer in Santa Clara, CA, to join the Cosmos team. This role focuses on creating AI-native systems that enhance the efficiency of machine learning workflows. Candidates should have extensive Python and PyTorch experience, along with... 

    NVIDIA Gruppe

    Santa Clara, CA
    3 days ago
  •  ...and production-grade workflows. You will work at the intersection of distributed systems, GPU performance, and ML framework integration. The role requires strong Python and PyTorch engineering skills, hands-on experience with distributed model training, and the ability to... 

    Nebius B.V.

    Palo Alto, CA
    1 day ago
  •  ...Member of Technical Staff to design and optimize inference systems. The role involves managing KV cache allocation and...  ...components. Ideal candidates should have strong software engineering skills and experience with ML inference systems, particularly in Python and C++. This... 

    Gimlet Labs

    San Francisco, CA
    4 days ago
  •  ...purpose-built AI products for life sciences. Looking for one strong systems engineer to own the distributed stack that keeps our training,...  ...Ray Track record of shipping or maintaining high-performance ML infrastructure High ownership, fast iteration, zero tolerance... 

    EM DASH LABS

    Wisconsin
    3 days ago
  • $170k - $220k

     ...of multidisciplinary Research Scientists and Engineers working on building a cutting-edge offline perception and auto-labelling system leveraging computer vision, and machine...  ...distributed computing frameworks. - Collaborate with ML researchers and engineers to seamlessly... 
    Full time
    Work at office
    Work from home
    Flexible hours

    Waabi

    Dallas, TX
    2 days ago
  • OpenAI in San Francisco seeks an experienced Software Engineer to help bring inference workloads to AWS Trainium and build the software stack to run frontier models efficiently on the platform. This deeply technical, cross‑stack role covers kernels, compilers, and model... 

    Slope

    San Francisco, CA
    3 days ago
  • Recruiting From Scratch is seeking a Machine Learning Systems Engineer to design and operate large-scale ML training and inference infrastructure in Palo Alto. The role focuses on building high-performance, GPU-accelerated systems for model serving and deployment across... 

    Recruiting from Scratch

    Palo Alto, CA
    2 days ago
  • Arch Systems is looking for a talented individual to design and implement optimal algorithms for Wi-Fi network performance, leveraging...  ...at least three years of experience in software or systems engineering. Key skills include Wi-Fi products development, WLAN management... 
    Remote work

    Arch Systems

    Palo Alto, CA
    3 days ago
  • A leading AI research firm located in San Francisco is seeking a Senior ML Systems Engineer to build and maintain the training framework for large-scale language models. The role involves designing distributed training solutions and improving training throughput across... 
    Flexible hours

    Cohere

    San Francisco, CA
    4 days ago
  • General Motors’ Data Labeling Engineering team is building cutting‑edge labeling tools and pipelines that power autonomous vehicle ML models. The role sits at the intersection of software...  ...and ML, focusing on scalable labeling systems and foundations for foundation‑model... 

    General Motors

    Mountain View, CA
    19 hours ago
  • Bright Vision Technologies is seeking an ML Systems Engineer to design, build, and operate high-performance inference platforms for serving large machine learning models in production. The role emphasizes systems engineering for AI deployment: request routing, batching,... 
    Remote job

    Bright Vision Technologies

    Bedford, TX
    2 days ago
  • Google is seeking software engineers for the Labs division, focusing on multimodal content generation across audio, video, and code. You...  ...emphasizes prompt engineering, model safety, and high-performance systems, collaborating with UX, research, and engineering to deliver... 

    Google

    Cambridge, MA
    2 days ago
  • A leading company in the financial technology sector is seeking a Senior Software Engineer to enhance trading systems through machine learning. The ideal candidate will have extensive software development experience and a strong skill set in Python, managing data pipelines... 

    The Hagen Ricci Group

    New York, NY
    19 hours ago
  • Jobzhr, a leading global trading firm, is expanding its ML infrastructure in New York. We are hiring engineers to build distributed training and low-latency inference systems that move models from research into production. You will work closely with researchers and traders... 

    Jobzhr

    New York, NY
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Systems Engineer. Be the first to apply!