MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling)
$90 - $120 per hourGridnaut Recruiting
Gridnaut Recruiting is hiring a remote MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) contractor (pay $90–$120/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: Have 2+ years of hands-on professional experience in ML systems, ML infrastructure, model serving, or GPU/accelerator performance engineering.; Have practical experience in GPU kernels, performance profiling/trace analysis, debugging distributed workloads, or serving large language models at scale.; Have production experience with JAX and/or PyTorch; framework-level depth is a strong plus.; Be familiar with modern accelerators such as A100, H100, B200 or TPU and reason about throughput, latency and memory trade-offs.; Engage reliably for at least 40 hours per week during weekdays.; Have strong written communication and explain complex technical decisions clearly..
- ...up. We're seeking MLOps Engineers with hands-on experience... ...of four areas: GPU kernel programming, performance profiling and trace analysis... ...inference serving. This role involves... ...assessing MLOps and ML systems tasks and... ...SGLang, TensorRT-LLM, Ray Serve, KV cache...SuggestedFull timeContract workTemporary workWeekday work
$90 - $120 per hour
...MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit...SuggestedRemote jobFor contractors10 hours per week- ...organization, apply now.We are currently seeking a On-Premise LLM Inference & GPU Systems Engineer to join our team in Charlotte, North Carolina (US-NC),... .../decode optimization and KV cache management. Inference Serving: Deploy and manage inference engines including vLLM and...SuggestedWork at officeRemote workFlexible hours
$200k - $400k
...Machine Learning Engineer (Reinforcement... .../Quantization/GPU), Multiple roles... ...advanced AI systems that can operate... ...high-performance kernels, compression... ...Language Models, LLM, Foundation Models... ..., Model Serving, Parallel Computing... ...Model Evaluation, MLOps Disclaimer...SuggestedRemote work$150k - $230k
...advanced AI, recommendation systems, and adtech.Recognized... ...-on Machine Learning Engineer to drive the post-... ...training on mid-to-large GPU clusters, applying distributed... ..., and ensure training-serving consistency.Stay... ...code.RequirementsHands-on LLM post-training...SuggestedFull timeLocal areaWork from home$200k - $280k
...AnalyticsJob Number: 192105Eligible for remote: YesApply: JobThe GPU- MLOps Engineer will own the Azure platform layer end-to-end — deploying AI/... ..., partnering closely with data scientists, ML engineers, and LLM engineers to keep everything running smoothly. What You Will...Remote workFlexible hours$100k - $150k
...MLOps Engineer -Remote Bright Vision Technologies is a technology... ...platforms for serving large machine learning... ...The role focuses on the systems engineering side of AI... ...KV cache strategies for LLM serving workloads. Integrate... ...understanding of GPU architecture, memory hierarchies...Full timeH1bLocal areaRemote workVisa sponsorship- ...observability, and the AI serving and routing layer.... ...treat that as an engineering responsibility,... ...agentic delivery system (the loops that... ...before you document MLOps experience: deploying and operating LLM or ML systems in... ...Jobot candidate profile, and any job...Temporary workLocal areaWork from homeHome officeFlexible hoursNight shift
$295k
...who are building AI systems. We believe that... ...team of researchers, engineers, designers, and... ...Implementing performance profiling across the ML... ...for large-scale LLM training.Design distributed... ..., or custom kernels/fused ops.... ...with evaluation and serving frameworks (vLLM,...Full timeWork at officeLocal areaRemote workHome office$135 per hour
...We believe great AI comes from exceptional, human-generated data. This remote, contract role focuses on ML inference and GPU kernel engineering. Competitive hourly pay ($135/hr) based on experience, with a broader range ($100–$170/hr) and a flexible, asynchronous...Remote jobHourly payContract workFlexible hours- ...senior ML researchers to define correct and excellent performance on hard, real-world problems, focusing on inference and GPU kernel engineering. This is research-and-evaluation work, not production engineering. The role is remote, flexible, and asynchronous, with...Remote jobFlexible hours
- ...AfterQuery seeks an ML Research Engineer focusing on inference and GPU kernels for remote, contract work. You will architect challenging evaluation problems, craft reference solutions, and judge AI outputs with rigor and domain insight. The role emphasizes research...Contract workRemote work
- ...AfterQuery is building a research lab focused on ML inference and GPU kernel engineering, covering inference serving systems, kernel optimization, and deployment infrastructure for large models. This is research-and-evaluation work, not production engineering, defining...Remote jobHourly pay
- ...frontiers through novel datasets and experimentation. We are building a cohort of senior ML researchers focused on inference and GPU kernel engineering to define criteria for correctness and excellence on real-world problems. This fully remote, contract role invites...Contract workRemote workFlexible hours
- ...AfterQuery is seeking an ML Research Engineer focusing on inference and GPU kernels. This remote contract role involves designing research scenarios, creating reference solutions, and grading AI outputs. You’ll contribute to frontier AI evaluation rather than production...Remote jobHourly payContract workFlexible hours
$170.1k - $258.3k
...capable fully self-driving systems, to move us toward... ...accessible mobility. For the AI Kernels & Compilers team, that... ..., and performance engineering so that every cycle on... ...high‑performance GPU kernels and custom libraries... ...that make it easier to profile, debug, and validate CUDA...Full timeLocal areaRemote workWork from homeRelocation packageFlexible hours$127k - $235k
...DescriptionAs a Senior Machine Learning Engineer, you will build and deploy production machine... ...learning and large language model (LLM) systems that extract actionable insights from complex... ...or data-sensitive industries.Experience serving or self-hosting large language models...Full timeWork at officeLocal areaRemote workFlexible hours$193.3k - $261.5k
...Trainium.The Acceleration Kernel Library team is at the... ...software boundary, our engineers craft high-performance... ..., and machine learning systems, you'll bring expertise... ...analysis using profiling tools to identify and resolve... ...architectures- Experience with GPU kernel optimization and...InternshipLocal areaWork from homeFlexible hours$200k - $230k
SVP, Lead AI/MLOps Infrastructure Engineer - Full Time - HybridWe’re partnering... ...to model serving, reliability, and cost... ..., compute (including GPU), storage, and model... ...deploymentProductionize AI/ML and GenAI (LLM) workloads in... ...similar)Solid Linux, systems, and troubleshooting...Full timeRemote work- ...AfterQuery is building a research cohort focused on ML inference and GPU kernel engineering, spanning inference serving systems to deployment infrastructure. This is research-and-evaluation work, not production engineering, defining what is considered correct and excellent...Remote job
$122.8k - $184.2k
...Job Area: Engineering Group, Engineering... ...machine learning systems that power next-generation... ...Benchmark and profile models across diverse... ...Integrate ML/LLM models into APIs,... ...and maintain model-serving infrastructure (e.... ...Distributed computing and GPU/accelerator...Work at officeWork from home$300k - $400k
...You will own the systems layer that makes our... ...: scheduling, kernels, RDMA, weight synchronization... ...and offline profilers that surface... ...communication and GPU kernels to extract... ..., scheduling, and serving architecture at production... ...— the scientists, engineers, and problem-...Visa sponsorshipFlexible hoursShift work$150k - $220k
...looking for a ML Systems Engineer, Inference. We want... ...the world to run LLM inference, meaning... ...effort. You'll own LLM serving performance end to... ...and repeatable. - Profile and diagnose... ...management down to kernels and interconnect.... ...node and multi-node GPU deployments. - Turn...Full timeRemote work- As a **Senior Machine Learning Engineer**, you will build and deploy production machine learning and large language model (LLM) systems that extract actionable insights from complex legal... ...quantitative field.*** Experience serving or self-hosting large language models and...Work at officeRemote workFlexible hours
- ...Overview We are seeking a Senior GPU Systems & Fabric Engineer to serve as the critical bridge between our physical... ...requires deep expertise in Linux kernel internals, GPU architectures, and... ...maximize cluster utilization. Profile and tune kernel-level parameters, device...Remote jobFull timeLocal area
$135.2k - $306.4k
Oracle hardware platform development engineering is seeking a highly driven GPU/CPU Platform System Engineer at the Principal Engineer level. The GPU System Engineer... ...process out to production. You will also serve as the last level of engineering technical support...Temporary workWork experience placementRemote workFlexible hours$397.46k
...scalable machine learning systems that power... ...looking for a Distinguished Engineer/Technical Director to lead... ...generative modeling, and LLM-powered personalization... ...from data ingestion to GPU training to live inference... ...for training, serving, and evaluating both traditional...Full timeWork experience placementH1bWork at officeLocal areaVisa sponsorshipMonday to Friday- ...Job Description Role : MLOPS Engineer / Architect \n Location :... ...ML pipelines \n Model serving \n CI/CD \n Monitoring... ...Conduct continuous performance profiling, load testing, and optimization... ..., alerting, and logging systems to track model drift, data quality...
$119.25k - $150.85k
...and we’re hiring engineers to help deliver... ...sister teams (kernels, compiler,... ...inference benchmarking/profiling infrastructure.... ..., operating systems, computer... ...ML compilers, GPU programming (CUDA... ...distributed training/serving infrastructure.... ...agentic or LLM-powered tools or...Full timeInternshipLocal areaWork from homeRelocation packageFlexible hours- ...Role : MLOPS Engineer / Architect Location : Charlotte NC (Onsite)... ...Kubernetes ML pipelines Model serving CI/CD Monitoring... ...Conduct continuous performance profiling, load testing, and... ...monitoring, alerting, and logging systems to track model drift, data quality...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling). Be the first to apply!
- senior linux systems engineer Remote
- system engineer contract Remote
- data systems engineer Remote
- senior staff systems engineer Remote
- microsoft systems engineer Remote
- operations support system engineer Remote
- advanced systems engineer Remote
- software system engineer Remote
- space systems engineer Remote
- adas systems engineer Remote






