MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling)
$90 - $120 per hourGridnaut Recruiting
Gridnaut Recruiting is hiring a remote MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) contractor (pay $90–$120/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: Have 2+ years of hands-on professional experience in ML systems, ML infrastructure, model serving, or GPU/accelerator performance engineering.; Have practical experience in GPU kernels, performance profiling/trace analysis, debugging distributed workloads, or serving large language models at scale.; Have production experience with JAX and/or PyTorch; framework-level depth is a strong plus.; Be familiar with modern accelerators such as A100, H100, B200 or TPU and reason about throughput, latency and memory trade-offs.; Engage reliably for at least 40 hours per week during weekdays.; Have strong written communication and explain complex technical decisions clearly..
$90 - $120 per hour
...MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit...SuggestedRemote jobFor contractors10 hours per week- ...organization, apply now.We are currently seeking a On-Premise LLM Inference & GPU Systems Engineer to join our team in Charlotte, North Carolina (US-NC),... .../decode optimization and KV cache management. Inference Serving: Deploy and manage inference engines including vLLM and...SuggestedWork at officeRemote workFlexible hours
- ...build the platform engineers turn to to ship AI... ...the global operating system for distributed,... ...We believe that as LLM and multi-modal workloads... ...to lead our GPU Networking efforts,... ...for Disaggregated Serving, Wide Expert Parallelism... .... Optimize Kernels: You will work with...SuggestedFull timeFlexible hours
$150k - $230k
...advanced AI, recommendation systems, and adtech.... ...hands-on Machine Learning Engineer to drive the post-training... ...on mid-to-large GPU clusters , applying distributed... ..., and ensure training–serving consistency. Stay... ...Requirements Hands-on LLM post-training...SuggestedFull timeLocal areaWork from home$200k - $280k
...AnalyticsJob Number: 192105Eligible for remote: YesApply: JobThe GPU- MLOps Engineer will own the Azure platform layer end-to-end — deploying AI/... ..., partnering closely with data scientists, ML engineers, and LLM engineers to keep everything running smoothly. What You Will...SuggestedRemote workFlexible hours- ...We’re hiring a Senior MLOps Engineer with deep machine learning... ...platform powering ML/LLM-driven healthcare workflows... ...secure, and compliant systems for model development,... ...(distributed training, GPU scheduling, cost... ...and inference (training-serving skew prevention). Establish...Remote workFlexible hours
$119.25k - $150.85k
...and we’re hiring engineers to help deliver... ...sister teams (kernels, compiler,... ...inference benchmarking/profiling infrastructure.... ..., operating systems, computer... ...ML compilers, GPU programming (CUDA... ...distributed training/serving infrastructure.... ...agentic or LLM-powered tools or...Full timeInternshipLocal areaWork from homeRelocation packageFlexible hours$295k
...who are building AI systems. We believe that... ...team of researchers, engineers, designers, and... ...responsible for large-scale LLM training.Design... ..., or custom kernels/fused ops.... ...with evaluation and serving frameworks (vLLM,... ...performance engineering, profiling, or low-level...Full timeWork at officeLocal areaRemote workHome office- ...junior Machine Learning Engineer. In this role, you... ...machine learning systems in real customer environments... ...for multi-GPU LLM training Profiling and optimizing training... ...pipelines for serving LLMs at scale Requirements... ...) Experience with MLOps and experiment...Full timeInternshipWork at office
- ...Machine Learning Engineer (Llama AI... ...connected business systems. We are building... ...and evaluate LLM performance for... ...on model serving, monitoring, logging... ...and MLOps practices.... ...infrastructure and GPU environments.... ...letter GitHub profile (if available)...Full timeRemote work
$213k - $263k
...We are looking for engineers with ML software & systems expertise to help b... ...performance ML runtime and serving system tailored for... ...robust tooling for profiling, benchmarking, and... ...building or scaling LLM serving systems,... ...Experience with custom kernel development (e.g., CUDA...Full timeRemote work$170.1k - $258.3k
...capable fully self-driving systems, to move us toward... ...accessible mobility. For the AI Kernels & Compilers team, that... ..., and performance engineering so that every cycle on... ...high‑performance GPU kernels and custom libraries... ...that make it easier to profile, debug, and validate CUDA...Full timeLocal areaRemote workWork from homeRelocation packageFlexible hours$193.3k - $261.5k
...Trainium.The Acceleration Kernel Library team is at the... ...software boundary, our engineers craft high-performance... ..., and machine learning systems, you'll bring expertise... ...analysis using profiling tools to identify and resolve... ...architectures- Experience with GPU kernel optimization and...InternshipLocal areaWork from homeFlexible hours- About the RoleAs Senior MLOps Engineer, you will focus on supporting cross... ..., including model serving (real-time and batch), training... ...environments, and orchestration systems, with a focus on performance,... ...orchestration of autonomous workflows or LLM-driven agentsJob SummaryJob...Remote work
- ...services and tools including LLM fine-tuning, alignment and... ...principal machine learning engineer, you will be responsible... ...compression, on-device inference, GPU inference, pytorch, kernel development Preferred... ...). Customer Support Systems : Experience with AI...Remote jobFull timeCasual workLive inWork at office
$190.2k - $345.65k
...Staff Machine Learning Engineer to architect and... ...intelligence — the systems that turn massive... ...retrieval stack that serves it, and the tool... ...correctness, and cost profile enterprise scale... ...ANN index tuning, GPU-accelerated enrichment... ...building retrieval for LLM and agentic systems...Full timeTemporary workLocal areaWorldwide- ...Overview We are seeking a Senior GPU Systems & Fabric Engineer to serve as the critical bridge between our physical... ...requires deep expertise in Linux kernel internals, GPU architectures, and... ...maximize cluster utilization. Profile and tune kernel-level parameters, device...Remote jobFull timeLocal area
- ...implement robust model serving infrastructure using platforms... ...# Improve GPU usage, enable autoscaling... ...degree in Computer Science, Engineering, or a related field (or... ...years of experience as MLOps engineer or DevOps... ...~ Experience in AI/ML systems security, compliance, and...Work at officeRemote workRelocation package
$198k - $326k
...power AI across LinkedIn. The LLM Serving team builds the critical... ...for a Senior Staff Software Engineer with deep expertise at the intersection of systems, machine learning, GPU infrastructure, and large-scale... ..., runtime, compiler, kernel, and hardware layersDrive model...For contractorsWork at officeFlexible hours- ...Intelligence team is building an NPC system that can (1) play any Roblox... ...across Discovery, Safety, Engine, and more. We are seeking... ...maintain core platform components: Serving Layer, Model Registry,... ...quantization) and systems on GPU architectures to maintain peak...Full time
$100k - $150k
...ML Performance Engineer - Remote Bright... ...neural network systems. The role spans the... ...from low-level kernel optimization to distributed... ...understanding of GPU architecture,... ...Profile and optimize end-to... ...speculative decoding for LLM serving. Drive compiler...Full timeH1bLocal areaImmediate startRemote workVisa sponsorship- ...to converse with all of their business systems through natural language to quickly find... ...workflow automation with Moveworks’ Reasoning Engine and natural language capabilities, we... ...edge ML infrastructure for building and serving LLM’s at Moveworks. This role will be...Full timeWork at officeRemote workFlexible hours
$397.46k
...scalable machine learning systems that power... ...looking for a Distinguished Engineer/Technical Director to lead... ...generative modeling, and LLM-powered personalization... ...from data ingestion to GPU training to live inference... ...for training, serving, and evaluating both traditional...Full timeWork experience placementH1bWork at officeLocal areaVisa sponsorshipMonday to Friday$130k - $220k
.... We’re hiring a Machine Learning Engineer (Recommendation Systems) to build the personalization engine... ...gives you the chance to work on systems serving 100M+ predictions daily, directly... ...distributed computing (Spark, Ray) and LLM/AI Agent frameworks ~ Track record...Full timeRemote work$135.2k - $306.4k
Oracle hardware platform development engineering is seeking a highly driven GPU/CPU Platform System Engineer at the Principal Engineer level. The GPU System Engineer... ...process out to production. You will also serve as the last level of engineering technical support...Temporary workWork experience placementRemote workFlexible hours- ...who builds with it every day. An engineer whose default instinct is to reach for an LLM, an agent framework, or a durable... .... You'll design and ship agentic systems that touch real users and real money... ...tools shop. Our platform serves thousands of traders. The things...Full timeLive inImmediate startWeekend work1 day per week
$145k - $165k
...potential. Job Title: ML Systems Engineer Location: 100% Remote (U.... ...reliable inference platforms for serving large machine learning models... ..., caching, autoscaling, GPU utilization, and end-to-end observability... ...Hands-on experience with LLM or large model inference...Full timeH1bLocal areaImmediate startRemote workVisa sponsorship$174.96k - $213.84k
...Machine Learning Engineer (L3) . About the... ...edge products that serve developers, builders... ...-latency, ML-based systems for real-time applications... ...orchestration and MLOps tooling.... ...Metaflow, SageMaker; LLM/agent frameworks such... ...; familiarity with GPU-based implementation...Full timeLocal areaRemote work- ...experienced Machine Learning Engineer to join our AI & Threat... ...for attackers, and the ML systems you build will serve as a first line of defense... ...engineering techniques and LLM frameworks ~ Experience building... ...who submit your profile References (with your consent...Remote jobFull timeTemporary work
$145k - $165k
..., Trust & Safety, Profile, Chat, Growth, and... ...Machine Learning Engineers (this role) who focus... ...training, serving, and feature management... ...machine learning systems that improve product... ...Understanding of MLOps practices including... ...systems Exposure to LLM-related use cases...Full timeWork experience placementCasual workWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling). Be the first to apply!
- application system engineer Remote
- system test engineer Remote
- sr systems engineer Remote
- software system engineer Remote
- distributed systems engineer Remote
- ground systems engineer Remote
- healthcare systems engineer Remote
- space systems engineer Remote
- senior windows systems engineer Remote
- digital communications systems engineer Remote



