MLOps Engineer
$100k - $150kBright Vision Technologies
Role Description
We are seeking a MLOps Engineer to design, build, and operate high-performance, highly reliable inference platforms for serving large machine learning models in production. The role focuses on the systems engineering side of AI deployment, including:
- Request routing
- Batching
- Caching
- Autoscaling
- GPU utilization
- End-to-end observability across diverse model workloads
The ideal candidate brings strong distributed systems and performance engineering expertise, has shipped serving systems at scale, and understands the trade-offs between latency, throughput, cost, and quality in ML serving.
Qualifications
- Bachelor’s or Master’s degree in Computer Science or a related field
- Six or more years of experience in distributed systems, infrastructure, or ML platform engineering
- Strong proficiency in Python and a systems language such as Go, Rust, or C++
- Deep experience operating high-throughput, low-latency services in production
- Hands-on experience with LLM or large model inference frameworks such as vLLM or TensorRT-LLM
- Strong understanding of GPU architecture, memory hierarchies, and accelerator utilization
- Familiarity with Kubernetes, autoscaling, and modern cloud platforms
- Experience with observability stacks including metrics, tracing, and structured logging
- Solid grounding in performance engineering and capacity planning
- Strong communication and incident response skills
Requirements
- Design and operate model serving platforms supporting diverse workloads including LLMs, vision models, and recommendation systems
- Optimize inference performance using continuous batching, paged attention, speculative decoding, and request multiplexing
- Implement multi-tenant routing, rate limiting, and quality-of-service policies across model endpoints
- Build autoscaling and capacity management systems that balance latency, throughput, and cost
- Tune GPU utilization, memory management, and KV cache strategies for LLM serving workloads
- Integrate model serving with API gateways, identity systems, and observability platforms
- Implement caching, prompt deduplication, and response reuse strategies where appropriate
- Drive end-to-end observability including latency histograms, queue dynamics, GPU utilization, and error tracking
- Develop deployment workflows including canary releases, shadow testing, and automated rollback
- Operate incident response for high-availability AI services and drive durable reliability improvements
- Collaborate with ML and product teams to support new model releases and capability rollouts
- Implement security controls including request signing, content filtering, and abuse detection at the serving layer
- Document operational procedures, performance characteristics, and tuning guidance for internal teams
- Stay current with AI serving research and translate advances into production capabilities
Benefits
- 100% Remote work
- Full-time, Direct W2 position
- Salary Range: $100,000–$150,000 Annually
- Career growth potential
- ...) EngineerLocationSunnyvale, CA (Work from Office)Primary SkillsMLOps, Ray , GPU , Python ScriptingMLOps (Ray Engineer)Experience: 3-5 Years .Job Role - MLOPS Engineer (RAY ENGINEER)Objective : MLOPS Engineer with experience in Ray & GPU configuration for engineering &...SuggestedFull timeWork at officeRemote work
- About the RoleAs Senior MLOps Engineer, you will focus on supporting cross-functional teams in designing, deploying, and operating machine learning solutions while building scalable infrastructure, tools, and best practices across the Machine Learning Engineering (MLE)...SuggestedRemote work
$230k - $250k
New York City, New York100% RemoteFull Time$230k - $250k A rapidly growing health technology company is seeking a Senior MLOps Engineer to help scale the next generation of machine learning-powered products. This is a full-time remote opportunity focused on AWS, Python...SuggestedFull timeRemote workFlexible hours- ...To build and scale the inference infrastructure for generative audio models, the full-time MLOps Engineer will design and deploy high-performance systems for low-latency model serving while working remotely. Key responsibilities Design and maintain inference infrastructure...SuggestedFull timeRemote work
- ...Working fully remote in a full-time capacity, the Senior MLOps Engineer I will transform models from ML Scientists and Data Scientists into reliable production services, focusing on infrastructure, deployment pipelines, and monitoring across various industry verticals...SuggestedFull timeRemote work
- ...influencing outcomes? We know that cybersecurity and its technologies evolve at a rapid pace. The defense community needs an engineering partner who can not only keep up, but bring the technical expertise and passion necessary to solve the new hardest problems — and...Full timeVisa sponsorshipWork visaFlexible hours
$94k - $141k
MLOps Engineer / MLOps SpecialistLead I - ML EngineeringWho We Are:Born digital, UST transforms lives through the power of technology. We walk alongside our clients and partners, embedding innovation and agility into everything they do. We help them create transformative...Full timeTemporary workPart timeWork at officeLocal areaRemote workFlexible hours$147.9k - $203k
Role Description We are looking for a Senior MLOps Engineer to join our Data Engineering & Analytics team. In this role, your primary focus will be leading the design and evolution of the platforms, workflows, and governance practices that enable machine learning teams...Full timeFlexible hours$171.5k - $201k
Role Description We are seeking a highly motivated and skilled Senior II MLOps Engineer. In this role, you will bridge the critical gap between machine learning model development and core system operations. You will be responsible for designing, building, and scaling the...Full timeRemote workFlexible hours- ...of data Qualifications ~Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field (or equivalent practical experience) ~5+ years of experience as MLOps engineer or in DevOps roles, working with MLOps platforms (MLflow, WandB etc.) and...Full timeWork at officeRelocation package
- ...make an impact, and work with people who care, we'd love to meet you! ABOUT THE ROLE: We are looking for a Middle/Senior MLOps Engineer to own the complete lifecycle transition from AI/ML experimentation to reliable production deployment, building and maintaining...Local areaRemote workVisa sponsorshipWork visaFlexible hours
- ...absolute mission to protect, optimize, and transform biomaterials engineering for industrial manufacturing at scale. Backed by elite venture... ...biomaterial data arrays, and deploy robust, production-grade MLOps pipelines globally. Position Overview We are seeking a highly...Full timeRemote workWork from homeHome officeShift work
- ...Job Title: MLOPS Engineer Location: Remote (travel to Santa Monica, CA if required) Technical Skills: Experience in MLOps, DevOps, or ML Engineering, with hands-on GCP experience Strong Python skills; comfortable with Bash and scripting for automation...Remote work
- ...environment. Join our multicultural team of visionaries and industry rebels in disrupting the traditional finance industry! As an MLOps Engineer, you will help design, build, and operate the next generation of our machine learning platform and infrastructure, enabling...Full timeWork at officeFlexible hours
$100k - $125k
...to join our research organization and contribute to machine learning-driven drug discovery efforts. This role will focus on data engineering, statistical modeling, bioinformatics, and scientific software development supporting internal therapeutic design programs. The ideal...Remote work$180k - $230k
SVP, MLOps / DevOps Engineer - Full Time - Hybrid We’re partnering with our client, a fast-growing fintech firm, on a senior-level MLOps-focused hire to help build and scale the infrastructure behind their AI and machine learning platforms. This is a highly visible, hands...Full timeRemote work$91.42k - $152.38k
City/StateVirginia Beach, VAWork ShiftMultiple shifts availableOverview:Sentara is hiring a Senior MLOps & Generative AI Engineer!This position is fully remote!Selected candidates would be required to be onsite for final round of team interviewCandidates must reside in...Full timeTemporary workRemote workShift work$90 - $140 per hour
Mercor is seeking an MLOps Engineer Expert to work remotely and help bridge knowledge gaps in AI model performance. The role involves guiding teams, designing tasks, and optimizing ML infrastructure solutions. You will need at least 2 years of hands-on experience with...Remote jobHourly pay- ...Sentara Health is seeking a Senior MLOps & Generative AI Engineer to accelerate production AI across healthcare services. You will design scalable ML infrastructure, CI/CD pipelines, feature stores, and model registries, while leading GenAI architectures with LLMs and...Remote work
$70 - $110 per hour
...Location: Remote Commitment: 40 hours/week Role Responsibilities Guide research and engineering teams to close knowledge gaps and improve AI model performance in MLOps , training infrastructure, and ML framework-level topics . Design challenging, domain...Contract workSummer workRemote workWeekday work$93k - $189k
DescriptionSummary: The MLOps Automation Engineering Senior Lead will lead a team responsible for building and deploying MLOps Automation for some of Huntington’s most valuable and most challenging data-driven projects.Duties and Responsibilities: Streamline the data,...Full timeWork at officeRemote workWork from homeFlexible hours$91.42k - $152.38k
Role Description Sentara is hiring a Senior MLOps & Generative AI Engineer! This position is fully remote! We are seeking a highly skilled and experienced Senior MLOps & Generative AI Engineer to join our growing AI organization and help advance current and future initiatives...Full timeTemporary workRemote work- A leading AI solutions provider is seeking an experienced MLOps / AI Ops Engineer for a remote 12-month contract. The role involves building and automating CI/CD pipelines for machine learning models, establishing monitoring frameworks, and managing deployment strategies...Contract workTemporary workRemote work
- ...Your Role We're seeking a Staff / Principal MLOps Engineer to join our team. Our ML footprint has grown quickly alongside the business: batch models that process records in the backend, real-time models that serve recommendations to job seekers, and daily pipelines...Full timeContract workTemporary workWork at officeRemote work
- ...Sentara Health Plans is seeking a Senior MLOps & Generative AI Engineer to architect and deploy enterprise-grade AI platforms in a healthcare setting. The role focuses on building scalable ML infrastructure, CI/CD for AI workloads, and GenAI applications using LLMs, RAG...Remote work
$170k - $260k
Lead Data Scientist / MLOps Engineer (Hybrid) Location: Annapolis Junction, MD (Partial Telework up to 16 hours per week) Compensation Base compensation ranges from $170,000 to $260,000. Responsibilities Lead analytic efforts and provide technical expertise for mission...16 hoursRemote work- Job Title: MLOps Platform Engineer (SageMaker) Duration: 12 Months Location: Plano, TX Pay Rate: $90/hr - $102/hr on W2 What you’ll be doing Set up SageMaker Unified Studio platform — domain configuration, project provisioning, persona-based roles, and multi-environment...Temporary workWork experience placement
$93k - $133k
...Requirements: We require a BS in computer science or a related engineering discipline. We look for 5+ years of experience building and... ...: We design, implement, and operate a unified MLOps platform that supports both on-premises Kubernetes clusters and...Full timeRemote workVisa sponsorshipFlexible hours- ...an IT services consultancy placing a Senior Machine Learning Engineer with one of our end clients — a growing healthcare technology... ...availability, performance, and reliability. Develop and maintain MLOps pipelines including CI/CD, model registry, feature stores,...Full timeContract work
$184.05k - $262.93k
...fulfillment and session generation Collaborate with cross-functional partners across user research, design, data science, product, and engineering Prototype new ML approaches and bring them into production at global scale Build and improve systems that connect artists...Full timeWork from homeWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to MLOps Engineer. Be the first to apply!







