Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling)

$90 - $120 per hour
Temporary

Gridnaut Recruiting

Gridnaut Recruiting is hiring a remote MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) contractor (pay $90–$120/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: Have 2+ years of hands-on professional experience in ML systems, ML infrastructure, model serving, or GPU/accelerator performance engineering.; Have practical experience in GPU kernels, performance profiling/trace analysis, debugging distributed workloads, or serving large language models at scale.; Have production experience with JAX and/or PyTorch; framework-level depth is a strong plus.; Be familiar with modern accelerators such as A100, H100, B200 or TPU and reason about throughput, latency and memory trade-offs.; Engage reliably for at least 40 hours per week during weekdays.; Have strong written communication and explain complex technical decisions clearly..

Vacancy posted 23 days ago
Similar jobs that could be interesting for youBased on the MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) in Remote vacancy
  •  ...up. We're seeking MLOps Engineers with hands-on experience...  ...of four areas: GPU kernel programming, performance profiling and trace analysis...  ...inference serving. This role involves...  ...assessing MLOps and ML systems tasks and...  ...SGLang, TensorRT-LLM, Ray Serve, KV cache... 
    Suggested
    Full time
    Contract work
    Temporary work
    Weekday work

    Mercor

    Remote
    19 days ago
  • $90 - $120 per hour

     ...MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit... 
    Suggested
    Remote job
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    22 hours ago
  •  ...organization, apply now.We are currently seeking a On-Premise LLM Inference & GPU Systems Engineer to join our team in Charlotte, North Carolina (US-NC),...  .../decode optimization and KV cache management. Inference Serving: Deploy and manage inference engines including vLLM and... 
    Suggested
    Work at office
    Remote work
    Flexible hours

    NTT DATA

    Charlotte, NC
    2 days ago
  • $200k - $400k

     ...Machine Learning Engineer (Reinforcement...  .../Quantization/GPU), Multiple roles...  ...advanced AI systems that can operate...  ...high-performance kernels, compression...  ...Language Models, LLM, Foundation Models...  ..., Model Serving, Parallel Computing...  ...Model Evaluation, MLOps Disclaimer... 
    Suggested
    Remote work

    Attis

    United States
    2 days ago
  • $150k - $230k

     ...advanced AI, recommendation systems, and adtech.Recognized...  ...-on Machine Learning Engineer to drive the post-...  ...training on mid-to-large GPU clusters, applying distributed...  ..., and ensure training-serving consistency.Stay...  ...code.RequirementsHands-on LLM post-training... 
    Suggested
    Full time
    Local area
    Work from home

    News Break

    Mountain View, CA
    12 hours ago
  • $200k - $280k

     ...AnalyticsJob Number: 192105Eligible for remote: YesApply: JobThe GPU- MLOps Engineer will own the Azure platform layer end-to-end — deploying AI/...  ..., partnering closely with data scientists, ML engineers, and LLM engineers to keep everything running smoothly. What You Will... 
    Remote work
    Flexible hours

    Molex

    Fremont, CA
    3 days ago
  • $100k - $150k

     ...MLOps Engineer -Remote Bright Vision Technologies is a technology...  ...platforms for serving large machine learning...  ...The role focuses on the systems engineering side of AI...  ...KV cache strategies for LLM serving workloads. Integrate...  ...understanding of GPU architecture, memory hierarchies... 
    Full time
    H1b
    Local area
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    Minneapolis, MN
    5 days ago
  •  ...observability, and the AI serving and routing layer....  ...treat that as an engineering responsibility,...  ...agentic delivery system (the loops that...  ...before you document MLOps experience: deploying and operating LLM or ML systems in...  ...Jobot candidate profile, and any job... 
    Temporary work
    Local area
    Work from home
    Home office
    Flexible hours
    Night shift

    Jobot

    Atlanta, GA
    1 day ago
  • $295k

     ...who are building AI systems. We believe that...  ...team of researchers, engineers, designers, and...  ...Implementing performance profiling across the ML...  ...for large-scale LLM training.Design distributed...  ..., or custom kernels/fused ops....  ...with evaluation and serving frameworks (vLLM,... 
    Full time
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    12 hours ago
  • $170.1k - $258.3k

     ...capable fully self-driving systems, to move us toward...  ...accessible mobility. For the AI Kernels & Compilers team, that...  ..., and performance engineering so that every cycle on...  ...high‑performance GPU kernels and custom libraries...  ...that make it easier to profile, debug, and validate CUDA... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Austin, TX
    1 day ago
  • $127k - $235k

     ...DescriptionAs a Senior Machine Learning Engineer, you will build and deploy production machine...  ...learning and large language model (LLM) systems that extract actionable insights from complex...  ...or data-sensitive industries.Experience serving or self-hosting large language models... 
    Full time
    Work at office
    Local area
    Remote work
    Flexible hours

    Thomson Reuters

    Ann Arbor, MI
    1 day ago
  • $193.3k - $261.5k

     ...Trainium.The Acceleration Kernel Library team is at the...  ...software boundary, our engineers craft high-performance...  ..., and machine learning systems, you'll bring expertise...  ...analysis using profiling tools to identify and resolve...  ...architectures- Experience with GPU kernel optimization and... 
    Internship
    Local area
    Work from home
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  •  ...experiments. Join a team focused on advanced ML inference and GPU kernel challenges, working remotely with flexible asynchronous...  ...excellent results, and shaping evaluation methodologies in a research setting rather than production engineering.J-18808-Ljbffr AfterQuery
    Remote job
    Flexible hours

    AfterQuery

    Costa Mesa, CA
    4 days ago
  • $200k - $230k

    SVP, Lead AI/MLOps Infrastructure Engineer - Full Time - HybridWe’re partnering...  ...to model serving, reliability, and cost...  ..., compute (including GPU), storage, and model...  ...deploymentProductionize AI/ML and GenAI (LLM) workloads in...  ...similar)Solid Linux, systems, and troubleshooting... 
    Full time
    Remote work

    Benchmark IT

    New York, NY
    2 days ago
  • $122.8k - $184.2k

     ...Job Area: Engineering Group, Engineering...  ...machine learning systems that power next-generation...  ...Benchmark and profile models across diverse...  ...Integrate ML/LLM models into APIs,...  ...and maintain model-serving infrastructure (e....  ...Distributed computing and GPU/accelerator... 
    Work at office
    Work from home

    Qualcomm

    San Diego, CA
    2 days ago
  • $150k - $220k

     ...looking for a ML Systems Engineer, Inference. We want...  ...the world to run LLM inference, meaning...  ...effort. You'll own LLM serving performance end to...  ...and repeatable. - Profile and diagnose...  ...management down to kernels and interconnect....  ...node and multi-node GPU deployments. - Turn... 
    Full time
    Remote work

    Runpod

    Remote
    13 days ago
  •  ...Overview We are seeking a Senior GPU Systems & Fabric Engineer to serve as the critical bridge between our physical...  ...requires deep expertise in Linux kernel internals, GPU architectures, and...  ...maximize cluster utilization. Profile and tune kernel-level parameters, device... 
    Remote job
    Full time
    Local area

    Bitdeer Technologies Group

    Remote
    7 days ago
  • $135.2k - $306.4k

    Oracle hardware platform development engineering is seeking a highly driven GPU/CPU Platform System Engineer at the Principal Engineer level. The GPU System Engineer...  ...process out to production. You will also serve as the last level of engineering technical support... 
    Temporary work
    Work experience placement
    Remote work
    Flexible hours

    Oracle Corporation

    Seattle, WA
    1 day ago
  • $397.46k

     ...scalable machine learning systems that power...  ...looking for a Distinguished Engineer/Technical Director to lead...  ...generative modeling, and LLM-powered personalization...  ...from data ingestion to GPU training to live inference...  ...for training, serving, and evaluating both traditional... 
    Full time
    Work experience placement
    H1b
    Work at office
    Local area
    Visa sponsorship
    Monday to Friday

    Roblox

    San Mateo, CA
    12 hours ago
  • $119.25k - $150.85k

     ...and we’re hiring engineers to help deliver...  ...sister teams (kernels, compiler,...  ...inference benchmarking/profiling infrastructure....  ..., operating systems, computer...  ...ML compilers, GPU programming (CUDA...  ...distributed training/serving infrastructure....  ...agentic or LLM-powered tools or... 
    Full time
    Internship
    Local area
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    3 days ago
  •  ...Job Description Role : MLOPS Engineer / Architect \n Location :...  ...ML pipelines \n Model serving \n CI/CD \n Monitoring...  ...Conduct continuous performance profiling, load testing, and optimization...  ..., alerting, and logging systems to track model drift, data quality... 

    Arkhya Tech Inc.

    Charlotte, NC
    6 days ago
  •  ...Role : MLOPS Engineer / Architect Location : Charlotte NC (Onsite)...  ...Kubernetes ML pipelines Model serving CI/CD Monitoring...  ...Conduct continuous performance profiling, load testing, and...  ...monitoring, alerting, and logging systems to track model drift, data quality... 

    Arkhya Tech Inc.

    Charlotte, NC
    2 days ago
  •  ...to converse with all of their business systems through natural language to quickly find...  ...workflow automation with Moveworks’ Reasoning Engine and natural language capabilities, we...  ...edge ML infrastructure for building and serving LLM’s at Moveworks. This role will be... 
    Permanent employment
    Work at office
    Remote work
    Flexible hours

    Moveworks

    Mountain View, CA
    2 days ago
  • $145k - $165k

     ...potential. Job Title: ML Systems Engineer Location: 100% Remote (U.S...  ...reliable inference platforms for serving large machine learning models...  ..., caching, autoscaling, GPU utilization, and end-to-end observability...  ...Hands-on experience with LLM or large model inference... 
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    Austin, TX
    4 days ago
  • $100k - $150k

     ...MLOps Engineer - Remote    Bright Vision Technologies...  ...inference platforms for serving large machine learning...  ...The role focuses on the systems engineering side of AI...  ...caching, autoscaling, GPU utilization, and end-to...  ...KV cache strategies for LLM serving workloads.  Integrate... 
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    Plymouth, MN
    1 day ago
  • $147.6k - $274k

     ...enable researchers and engineers to move models from...  ...to our model-serving platform, and to the...  ...scope extends beyond LLM serving. You will...  ...inference workloads, GPU-backed services,...  ...Kubernetes, and distributed systems. Prior inference-...  ..., DevOps, MLOps, or a related area.... 
    Full time
    Local area
    Immediate start
    Worldwide
    Relocation package

    Genentech

    York, PA
    3 days ago
  • $144k - $233.1k

     ...industry's most essential operating system, serving owners and residents...  ...for a Senior Machine Learning Engineer to drive the next generation...  ...environments.Familiarity with modern LLM tooling, model serving, and...  ...with distributed training or GPU-based model workloads.... 
    Full time
    Local area
    Remote work
    Worldwide
    Flexible hours

    Entrata

    Lehi, UT
    3 days ago
  • $90 - $120 per hour

     ...General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) Type: Contract Compensation: $90–$120/hour Location: Remote Commitment: 40 hours/... 
    Contract work
    Summer work
    Remote work

    Mercor

    Chicago, IL
    23 days ago
  • $184k - $287.5k

    NVIDIA is seeking a Senior MLOps Engineer to join our DSX Enablement team...  ...and solve full-stack AI and ML system problems, and work on a team...  ...ML initiatives, including LLM performance evaluation and supporting...  ...compelling AI applications,Profile and tune large-scale training... 
    Full time
    Remote work

    Nvidia

    Seattle, WA
    4 days ago
  • $148k - $222k

     ...Platform and Observability Engineering (DPOE) team is building...  ..., and LangSmith for LLM/agentic tracing and...  ...expertise in distributed systems and big data while contributing...  ...and participate in profiling and resolving latency...  ...paths that will serve as the foundation for AI... 
    Full time
    Work at office
    Remote work
    Home office
    Flexible hours

    Workday

    Pleasanton, CA
    12 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling). Be the first to apply!