ML Systems Engineer
$145k - $165kBright Vision Technologies
ML Systems Engineer
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. Location: 100% Remote (U.S.) Position Type: Full-time Salary Range: $145,000–$165,000 Annually Experience Required: 6+ years Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.
Job Summary: We are seeking a ML Systems Engineer to design, build, and operate high-performance, highly reliable inference platforms for serving large machine learning models in production. The role focuses on the systems engineering side of AI deployment, including request routing, batching, caching, autoscaling, GPU utilization, and end-to-end observability across diverse model workloads. The ideal candidate brings strong distributed systems and performance engineering expertise, has shipped serving systems at scale, and understands the trade-offs between latency, throughput, cost, and quality in ML serving.
Required Qualifications
- Bachelor's or Master's degree in Computer Science or a related field.
- Six or more years of experience in distributed systems, infrastructure, or ML platform engineering.
- Strong proficiency in Python and a systems language such as Go, Rust, or C++.
- Deep experience operating high-throughput, low-latency services in production.
- Hands-on experience with LLM or large model inference frameworks such as vLLM or TensorRT-LLM.
- Strong understanding of GPU architecture, memory hierarchies, and accelerator utilization.
- Familiarity with Kubernetes, autoscaling, and modern cloud platforms.
- Experience with observability stacks including metrics, tracing, and structured logging.
- Solid grounding in performance engineering and capacity planning.
- Strong communication and incident response skills.
Preferred Qualifications
- Open-source contributions to model serving infrastructure.
- Experience with multi-region or globally distributed AI serving.
- Familiarity with model quantization, distillation, and compression techniques.
- Exposure to FinOps for AI workloads and cost-efficient serving design.
- Experience supporting external-facing AI APIs at scale.
$174.9k - $261.3k
...and understand the world!The Data Labeling Engineering team designs, builds, and operates hybrid... ...engineering, data engineering, and AI/ML, defining the strategies, tooling, and quality... ...leadership, and direct impact on systems that unblock the next generation of AV capabilities...SuggestedFull timeLocal areaRemote workWork from homeRelocation packageFlexible hours$90.1k - $191.8k
...and understand the world!The Data Labeling Engineering team designs, builds, and operates high‑... ...engineering, data engineering, and ML, defining labeling strategies, tooling, and... ...technical leadership, and work directly on systems that unblock the next generation of AV models...SuggestedFull timeWork experience placementLocal areaRemote workWork from homeRelocation packageFlexible hours$224k - $356.5k
...the next phase, we are building agentic systems that can reason about, build, evaluate, and... ...about creating the meta-layer of modern ML: the agents, tooling, pipelines, and feedback... .... We are looking for exceptional engineers who are passionate about the idea of AI-native...SuggestedFull time- ...models for enterprises who are building AI systems. We believe that our work is instrumental... .... Cohere is a team of researchers, engineers, designers, and more, who are all passionate... ...you enjoy working across the full stack of ML systems, this role gives you the...SuggestedFull timeWork at officeLocal areaRemote workHome office
- ...work sits at the intersection of distributed systems, GPU performance, model training frameworks, RL pipelines, and production engineering. Your responsibilities Build and maintain... ...with distributed model training, large-scale ML systems, or GPU cluster workloads....Suggested
$300k - $400k
...possible. About the Role You will own the systems layer that makes our frontier model... ...operations Profiling and benchmarking distributed ML systems to identify and eliminate... ...team of the world’s best — the scientists, engineers, and problem-solvers who don’t just follow...Visa sponsorshipFlexible hoursShift work- ...ML Systems Engineer, ML Acceleration We are looking for a Machine Learning Systems Engineer to join our ML Acceleration team. In this role, you will be responsible for the core systems that enable our researchers to train frontier models at scale, focusing obsessively...
$145k - $165k
...opportunity to join an established and well-respected organization offering tremendous career growth potential. Job Title: ML Systems Engineer Location: 100% Remote (U.S.) Position Type: Full-time, Direct W2 Salary Range: $145,000–$165,000 Annually...Full timeH1bLocal areaImmediate startRemote workVisa sponsorship$110 per hour
...and Jack Dorsey . Position: MLOps Engineer (JAX, PyTorch, Pallas/Triton) Type:... ...MLOps , training infrastructure, and ML framework-level topics . Design challenging... ...-structured solutions to MLOps and ML systems problems . Evaluate MLOps tasks and...Remote jobContract workSummer workWeekday work$110 per hour
...Role Responsibilities Guide research and engineering teams to close knowledge gaps and improve... ...MLOps , training infrastructure, and ML framework-level topics . Design... ...-structured solutions to MLOps and ML systems problems . Evaluate MLOps tasks and...Contract workSummer workRemote workWeekday work- ...Job Responsibilities: Engineer, design, implement, and improve highly-scalable machine learning systems and tools for enabling research Apply knowledge of relevant research... ...Machine Learning Distributed training for ML models Experience with Machine Learning...Work experience placement
- ...Job Responsibilities: Engineer, design, implement, and improve highly scalable machine learning systems and tools for enabling research Apply knowledge of relevant research... ...experience ~0-2 years of Distributed ML Training (FSDP/DDP) experience ~5+ years of...Work experience placement
$200.8k - $251k
...member to build and optimize a machine learning framework for large language models. Candidates should have system optimization experience and solid software engineering skills, particularly in tools like CUDA and Pytorch. This full-time position offers a competitive salary...Full time- General Motors' Data Labeling Engineering team builds and operates hybrid human/machine labeling tools powering autonomous vehicle ML models. We work across software, data, and ML to create scalable training data, with a modern full‑stack including TypeScript, React, GraphQL...
- Rhoda AI is hiring a Senior/Staff-level Research Engineer to ensure our robot-learning pipeline is reliable from data collection through... ..., and real-robot evaluation. You will build validation systems, observability, and robust operating practices to distinguish model...
- ..., Inc. is looking for a Member of Technical Staff focused on ML systems and inference in San Francisco. You will design and build inference... .... Candidates should have strong foundations in software engineering, experience with ML inference systems, and performance tuning...
- NVIDIA Gruppe is seeking a Senior Engineer in Santa Clara, CA, to join the Cosmos team. This role focuses on creating AI-native systems that enhance the efficiency of machine learning workflows. Candidates should have extensive Python and PyTorch experience, along with...
- ...and production-grade workflows. You will work at the intersection of distributed systems, GPU performance, and ML framework integration. The role requires strong Python and PyTorch engineering skills, hands-on experience with distributed model training, and the ability to...
- ...Member of Technical Staff to design and optimize inference systems. The role involves managing KV cache allocation and... ...components. Ideal candidates should have strong software engineering skills and experience with ML inference systems, particularly in Python and C++. This...
- ...purpose-built AI products for life sciences. Looking for one strong systems engineer to own the distributed stack that keeps our training,... ...Ray Track record of shipping or maintaining high-performance ML infrastructure High ownership, fast iteration, zero tolerance...
$170k - $220k
...of multidisciplinary Research Scientists and Engineers working on building a cutting-edge offline perception and auto-labelling system leveraging computer vision, and machine... ...distributed computing frameworks. - Collaborate with ML researchers and engineers to seamlessly...Full timeWork at officeWork from homeFlexible hours- OpenAI in San Francisco seeks an experienced Software Engineer to help bring inference workloads to AWS Trainium and build the software stack to run frontier models efficiently on the platform. This deeply technical, cross‑stack role covers kernels, compilers, and model...
- Recruiting From Scratch is seeking a Machine Learning Systems Engineer to design and operate large-scale ML training and inference infrastructure in Palo Alto. The role focuses on building high-performance, GPU-accelerated systems for model serving and deployment across...
- Arch Systems is looking for a talented individual to design and implement optimal algorithms for Wi-Fi network performance, leveraging... ...at least three years of experience in software or systems engineering. Key skills include Wi-Fi products development, WLAN management...Remote work
- A leading AI research firm located in San Francisco is seeking a Senior ML Systems Engineer to build and maintain the training framework for large-scale language models. The role involves designing distributed training solutions and improving training throughput across...Flexible hours
- General Motors’ Data Labeling Engineering team is building cutting‑edge labeling tools and pipelines that power autonomous vehicle ML models. The role sits at the intersection of software... ...and ML, focusing on scalable labeling systems and foundations for foundation‑model...
- Bright Vision Technologies is seeking an ML Systems Engineer to design, build, and operate high-performance inference platforms for serving large machine learning models in production. The role emphasizes systems engineering for AI deployment: request routing, batching,...Remote job
- Google is seeking software engineers for the Labs division, focusing on multimodal content generation across audio, video, and code. You... ...emphasizes prompt engineering, model safety, and high-performance systems, collaborating with UX, research, and engineering to deliver...
- A leading company in the financial technology sector is seeking a Senior Software Engineer to enhance trading systems through machine learning. The ideal candidate will have extensive software development experience and a strong skill set in Python, managing data pipelines...
- Jobzhr, a leading global trading firm, is expanding its ML infrastructure in New York. We are hiring engineers to build distributed training and low-latency inference systems that move models from research into production. You will work closely with researchers and traders...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to ML Systems Engineer. Be the first to apply!
- data scientist machine learning engineer United States
- machine learning ai engineer United States
- computer vision machine learning engineer United States
- machine learning engineer United States
- ai ml engineer United States
- graduate machine learning engineer United States
- machine learning software engineer United States
- entry level machine learning engineer United States
- junior machine learning engineer United States
- staff machine learning engineer United States



