Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Platform Engineer

$100k - $160k
Full-time

Bright Vision Technologies

ML Platform Engineer - Remote

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title: ML Platform Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $100,000–$160,000 Annually
Experience Required: 10+ years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary:
We are seeking a ML Platform Engineer to design, build, and operate high-performance, highly reliable inference platforms for serving large machine learning models in production. The role focuses on the systems engineering side of AI deployment, including request routing, batching, caching, autoscaling, GPU utilization, and end-to-end observability across diverse model workloads. The ideal candidate brings strong distributed systems and performance engineering expertise, has shipped serving systems at scale, and understands the trade-offs between latency, throughput, cost, and quality in ML serving.

Key Responsibilities
  • Design and operate model serving platforms supporting diverse workloads including LLMs, vision models, and recommendation systems.
  • Optimize inference performance using continuous batching, paged attention, speculative decoding, and request multiplexing.
  • Implement multi-tenant routing, rate limiting, and quality-of-service policies across model endpoints.
  • Build autoscaling and capacity management systems that balance latency, throughput, and cost.
  • Tune GPU utilization, memory management, and KV cache strategies for LLM serving workloads.
  • Integrate model serving with API gateways, identity systems, and observability platforms.
  • Implement caching, prompt deduplication, and response reuse strategies where appropriate.
  • Drive end-to-end observability including latency histograms, queue dynamics, GPU utilization, and error tracking.
  • Develop deployment workflows including canary releases, shadow testing, and automated rollback.
  • Operate incident response for high-availability AI services and drive durable reliability improvements.
  • Collaborate with ML and product teams to support new model releases and capability rollouts.
  • Implement security controls including request signing, content filtering, and abuse detection at the serving layer.
  • Document operational procedures, performance characteristics, and tuning guidance for internal teams.
  • Stay current with AI serving research and translate advances into production capabilities.
Required Qualifications
  • Bachelor’s or Master’s degree in Computer Science or a related field.
  • 10 or more years of experience in distributed systems, infrastructure, or ML platform engineering.
  • Strong proficiency in Python and a systems language such as Go, Rust, or C++.
  • Deep experience operating high-throughput, low-latency services in production.
  • Hands-on experience with LLM or large model inference frameworks such as vcLLM or TensorRT-LLM.
  • Strong understanding of GPU architecture, memory hierarchies, and accelerator utilization.
  • Familiarity with Kubernetes, autoscaling, and modern cloud platforms.
  • Experience with observability stacks including metrics, tracing, and structured logging.
  • Solid grounding in performance engineering and capacity planning.
  • Strong communication and incident response skills.
Preferred Qualifications
  • Open-source contributions to model serving infrastructure.
  • Experience with multi-region or globally distributed AI serving.
  • Familiarity with model quantization, distillation, and compression techniques.
  • Exposure to FinOps for AI workloads and cost-efficient serving design.
  • Experience supporting external-facing AI APIs at scale.
How to Apply
Would you like to know more about this opportunity? For immediate consideration, please send your resume to View email address on aiapply.co or contact us at View phone number on aiapply.co. Learn more about Bright Vision Technologies at
Bright Vision Technologies is an Equal Opportunity Employer.

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Vacancy posted 17 hours ago
Similar jobs that could be interesting for youBased on the ML Platform Engineer in Remote vacancy
  •  ...AI/ML Infrastructure Engineer A hands-on engineer who can build and scale AI/ML infrastructure, ensuring reliable pipelines, versioning, monitoring...  ...and data infrastructure. 5+ years in DevOps/MLOps or AI platform engineering Expert in containerized workloads (... 
    Suggested
    Remote work

    ClifyX

    United States
    4 days ago
  • $205k - $235k

     ...step of the way - enabling you to shape your future with confidence.Within the EY-Parthenon service line, the EY Growth Platforms AI ML Engineering Director will collaborate with Business Leaders, Data Scientists, Project Managers, and other team members to design, build... 
    Suggested
    Full time
    For contractors
    Work experience placement
    Summer holiday
    Flexible hours

    EY (Ernst & Young)

    New York, NY
    2 days ago
  • $136k - $167k

    Role Description As an ML Platform Engineer at Stitch Fix, you will play a key role in building and maintaining the critical infrastructure that powers machine learning and AI across our organization. You will design, develop, and support scalable, resilient services and... 
    Suggested
    Full time

    Stitch Fix

    Remote
    3 days ago
  • Role Description Whisker Labs is seeking a Senior ML Platform Engineer to join our fully remote Data Science team. As part of the team, you will be responsible for advancing Whisker Labs’ technology to detect early warning signs of electrical fires in homes, preventing... 
    Suggested
    Full time
    Remote work

    Whisker Labs

    Remote
    4 days ago
  • $170k - $300k

     ..., while scaling our Factory-as-a-Service platform to transform how critical products are built...  ...looking for. The Role This is an ML infrastructure role at the core of...  ...without requiring close support from platform engineers.   What We’re Looking For Track record... 
    Suggested
    Permanent employment
    Full time
    Contract work
    Local area
    Relocation package
    Flexible hours

    Hadrian Automation

    Los Angeles, CA
    14 days ago
  • $100k - $160k

     ...ML Platform Engineer - Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected... 
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    North Andover, MA
    14 days ago
  •  ...PrizePicks is seeking an ML Platform Engineer to build and scale the ML infrastructure that supports real-time inference across our DFS ecosystem. You will work on a production-grade platform that powers dynamic oddsmaking, risk analysis, and smart deposit defaults.... 
    Remote work

    PrizePicks

    Atlanta, GA
    21 hours ago
  •  ...Faire is seeking a Staff Machine Learning Platform Engineer to design, improve, and operate a scalable ML platform that accelerates model training, deployment, and governance. You will bridge data science and production engineering, joining a small but critical team supporting... 
    Local area
    Remote work

    Faire

    San Francisco, CA
    1 day ago
  • $143k - $197k

    Role Description We are looking for an experienced Senior AI/ML Platform Engineer who is eager to build the platforms, infrastructure and services necessary to deliver ML and generative AI based products that make a significant impact within the organization. The ideal... 
    Full time

    Lyra Health

    Remote
    16 hours ago
  • £85k - £110k per year

     ...micro-savings portfolios, Monzo continuously engineers human-centric financial products that...  ...backend engineers with an uncompromised platform canvas to leverage cutting-edge multi-cloud...  ...connection-obsessed, and systems-minded ML/AI Platform Engineer to join our centralized... 
    Full time
    Work at office
    Remote work
    Work from home
    Visa sponsorship
    Relocation package
    Flexible hours
    Shift work

    Monzo

    United Kingdom
    a month ago
  •  ...Overview Employer:    Southern Glazer’s Wine and Spirits LLC Job Title:      Principal AI/ML Operations Platform Engineer  Locations:    2300 SW 145th Avenue, Miramar, FL 33027  14911 Quorum Drive, Dallas, TX 75254  Duties Design, build, and deploy machine... 
    Work at office
    Local area
    Remote work

    Southern Glazer’s Wine and Spirits, LLC

    Dallas, TX
    a month ago
  •  ...training software, bridging research and production while writing production-grade code alongside researchers. You’ll work with Python, ML frameworks, and distributed training tools. Join a world-class team with extensive compute resources, remote-friendly policies, and... 
    Remote work
    Worldwide

    Cohere

    San Francisco, CA
    16 hours ago
  • $130k - $147k

     ...put customers at the center of every decision. Our AI-first platform transforms proprietary data, advanced analytics and deep...  ...United States. Job Description We are seeking a Senior ML/AI Platform Engineer to help build and operate Curinos' Databricks-native AI platform... 
    Part time
    Work at office
    Remote work
    Work from home
    Flexible hours

    Curinos Inc

    New York, NY
    6 days ago
  • A leading open source software company is seeking a Python and Kubernetes Software Engineer focused on Data, AI/ML & Analytics. The role involves developing solutions for public cloud and private infrastructure in a remote setting. Ideal candidates have strong skills in... 
    Remote work

    Canonical

    Boston, MA
    2 days ago
  •  ...listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Staff ML Software Engineer (L6) — Platform Systems, AIMS Engineering based in the United States. This is a high-impact staff-level engineering role focused... 
    Full time
    Immediate start
    Remote work
    Flexible hours

    jobgether

    United States
    4 days ago
  • $180k - $230k

     ...Raft is seeking a Lead AI/ML Software Engineer to serve as a senior technical leader for [R]AIMS in Colorado Springs. This role involves driving...  ...decisions, mentoring engineers, and optimizing platform performance across diverse operational environments. The... 
    Remote work

    Raft

    Colorado Springs, CO
    21 hours ago
  • $205k - $235k

     ...of the way - enabling you to shape your future with confidence. Within the EY-Parthenon service line, the EY Growth Platforms AI ML Engineering Director will collaborate with Business Leaders, Data Scientists, Project Managers, and other team members to design, build... 
    Full time
    For contractors
    Work experience placement
    Summer holiday
    Flexible hours

    Ernst & Young Oman

    Hoboken, NJ
    21 hours ago
  • $100k

     ...audiences that are deeply engaged. Our Team The Ads Platform Engineering teams build advertising systems and integrations that powers...  ...state-of-art realtime inventory forecasting solution leveraging ML models and high performance ad server simulations. The team... 
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Remote
    16 hours ago
  •  ...remote company with a strong open-source focus, building data-driven identification and smart signals products. As Engineering Manager, Data Platform & ML Ops, you will lead 4-6 engineers, own the data warehouse and ML lifecycle, and partner with data scientists, PMs and... 
    Remote job

    Fingerprint

    Chicago, IL
    16 hours ago
  • $130k - $185k

     ...step of the way – enabling you to shape your future with confidence. Within the EY-Parthenon service line, the EY Growth Platforms AI ML Engineering Director will collaborate with Business Leaders, Data Scientists, Project Managers, and other team members to design,... 
    For contractors
    Work experience placement
    Summer holiday
    Flexible hours

    Ernst & Young

    Atlanta, GA
    15 days ago
  •  ...be part of Gen. About the Role: Engine by Gen is a leader in financial...  ...finance super app, a premier embedded finance platform for enterprise businesses and a world-class...  ..., and maintain infrastructure supporting ML training, deployment, and inference workloads... 
    Full time
    Flexible hours

    Gen Digital

    New York, NY
    16 hours ago
  • $150k - $210k

     ...data into clear insights and actionable recommendations. Our AI platform is central to this mission, turning raw physiological signals...  ...members can act on every day. WHOOP is hiring a Senior AI/ML Engineer to help scale the intelligence layer behind WHOOP’s AI-powered... 
    Full time
    Work at office
    Relocation

    Whoop

    Boston, MA
    16 hours ago
  • $195k - $298k

     ...expected to report to the GM Global Technical Center - Cole Engineering Center Podium or Mountain View Technical Center , CA at...  ...is eligible for relocation assistance. About the Team:The ML Compute Platform is part of the AI Compute Platform organization within Infrastructure... 
    Full time
    Local area
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Warren, MI
    16 hours ago
  • $295k - $405.5k

     ...About FaireFaire is a technology wholesale platform built on the belief that the future is local. Independent...  ...the Senior Staff Machine Learning Platform Engineer, you will own the technical vision and evolution of Faire’s ML platform. You will set standards, influence org... 
    Work experience placement
    Work at office
    Local area
    Remote work
    Monday to Friday
    Flexible hours
    3 days per week

    Faire

    San Francisco, CA
    4 days ago
  • $246.5k - $339k

    About FaireFaire is a technology wholesale platform built on the belief that the future is local. Independent...  ...this roleAs a Staff Machine Learning Platform Engineer, you will help design, improve, and operate a scalable ML platform to accelerate model training,... 
    Work experience placement
    Work at office
    Local area
    Remote work
    Monday to Friday
    Flexible hours
    3 days per week

    Faire

    San Francisco, CA
    2 days ago
  •  ...constantly pushing the envelope to enhance, build, and deliver top-notch technology products.As a Senior Lead Software Engineer - ML Engineer for Agent Platform at JPMorgan Chase within the Commercial and Investment Banking - Data Analytics Payments Team, you are a senior... 

    JP Morgan Chase

    Jersey City, NJ
    1 day ago
  • $120.1k - $214.5k

     ...start Caring. Connecting. Growing together.We are seeking a Senior AI/ML Software Engineer to build AI-powered data products and modern digital experiences across our analytics and healthcare platforms. This is a hands-on full-stack engineering role focused on developing... 
    Minimum wage
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work

    UnitedHealth Group

    Eden Prairie, MN
    3 days ago
  •  ...leave a clean trail of tests and dashboards behind. Required Qualifications: Expert-level PyTorch. Proven software engineer who loves ML; comfortable writing production code across the stack. Hands-on experience training or fine-tuning large language or other... 
    Full time
    Contract work
    Flexible hours
    Shift work

    Sesame, L.l.c.

    San Francisco, CA
    16 hours ago
  • $160k - $220k

     ...at scale. This is a hands-on role where you'll see direct impact on business metrics daily. What You'll Do Build and deploy ML models serving 100M+ predictions daily Develop ranking algorithms that balance relevance, diversity, and revenue Implement real... 
    Remote job
    Full time

    Launch Potato

    Remote
    16 hours ago
  • $200k - $300k

     ...building that model: the Food Foundation Model. As a Senior ML Engineer, Foundation Models, you will work at the frontier of large-...  ...just offline benchmark accuracy Collaborate with the data and platform teams on training data requirements, augmentation strategies,... 
    Full time
    Flexible hours

    Chef Robotics

    Remote
    16 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Platform Engineer. Be the first to apply!