Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling)

$90 - $120 per hour
Temporary

Gridnaut Recruiting

Gridnaut Recruiting is hiring a remote MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) contractor (pay $90–$120/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: Have 2+ years of hands-on professional experience in ML systems, ML infrastructure, model serving, or GPU/accelerator performance engineering.; Have practical experience in GPU kernels, performance profiling/trace analysis, debugging distributed workloads, or serving large language models at scale.; Have production experience with JAX and/or PyTorch; framework-level depth is a strong plus.; Be familiar with modern accelerators such as A100, H100, B200 or TPU and reason about throughput, latency and memory trade-offs.; Engage reliably for at least 40 hours per week during weekdays.; Have strong written communication and explain complex technical decisions clearly..

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) in Remote vacancy
  • $90 - $120 per hour

     ...MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit... 
    Suggested
    Remote job
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    21 hours ago
  •  ...organization, apply now.We are currently seeking a On-Premise LLM Inference & GPU Systems Engineer to join our team in Charlotte, North Carolina (US-NC),...  .../decode optimization and KV cache management. Inference Serving: Deploy and manage inference engines including vLLM and... 
    Suggested
    Work at office
    Remote work
    Flexible hours

    NTT DATA

    Charlotte, NC
    1 day ago
  •  ...build the platform engineers turn to to ship AI...  ...the global operating system for distributed,...  ...We believe that as LLM and multi-modal workloads...  ...to lead our GPU Networking efforts,...  ...for Disaggregated Serving, Wide Expert Parallelism...  .... Optimize Kernels: You will work with... 
    Suggested
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    1 day ago
  • $150k - $230k

     ...advanced AI, recommendation systems, and adtech....  ...hands-on Machine Learning Engineer to drive the post-training...  ...on mid-to-large GPU clusters , applying distributed...  ..., and ensure training–serving consistency. Stay...  ...Requirements Hands-on LLM post-training... 
    Suggested
    Full time
    Local area
    Work from home

    News Break

    Remote
    1 day ago
  • $200k - $280k

     ...AnalyticsJob Number: 192105Eligible for remote: YesApply: JobThe GPU- MLOps Engineer will own the Azure platform layer end-to-end — deploying AI/...  ..., partnering closely with data scientists, ML engineers, and LLM engineers to keep everything running smoothly. What You Will... 
    Suggested
    Remote work
    Flexible hours

    Molex

    Lisle, IL
    2 days ago
  •  ...We’re hiring a Senior MLOps Engineer with deep machine learning...  ...platform powering ML/LLM-driven healthcare workflows...  ...secure, and compliant systems for model development,...  ...(distributed training, GPU scheduling, cost...  ...and inference (training-serving skew prevention). Establish... 
    Remote work
    Flexible hours

    C the Signs

    United States
    2 days ago
  • $119.25k - $150.85k

     ...and we’re hiring engineers to help deliver...  ...sister teams (kernels, compiler,...  ...inference benchmarking/profiling infrastructure....  ..., operating systems, computer...  ...ML compilers, GPU programming (CUDA...  ...distributed training/serving infrastructure....  ...agentic or LLM-powered tools or... 
    Full time
    Internship
    Local area
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    6 days ago
  • $295k

     ...who are building AI systems. We believe that...  ...team of researchers, engineers, designers, and...  ...responsible for large-scale LLM training.Design...  ..., or custom kernels/fused ops....  ...with evaluation and serving frameworks (vLLM,...  ...performance engineering, profiling, or low-level... 
    Full time
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    4 days ago
  •  ...junior Machine Learning Engineer. In this role, you...  ...machine learning systems in real customer environments...  ...for multi-GPU LLM training Profiling and optimizing training...  ...pipelines for serving LLMs at scale Requirements...  ...) Experience with MLOps and experiment... 
    Full time
    Internship
    Work at office

    Pangram Labs

    Remote
    1 day ago
  •  ...Machine Learning Engineer (Llama AI...  ...connected business systems. We are building...  ...and evaluate LLM performance for...  ...on model serving, monitoring, logging...  ...and MLOps practices....  ...infrastructure and GPU environments....  ...letter GitHub profile (if available)... 
    Full time
    Remote work

    Performacentric

    Remote
    1 day ago
  • $213k - $263k

     ...We are looking for engineers with ML software & systems expertise to help b...  ...performance ML runtime and serving system tailored for...  ...robust tooling for profiling, benchmarking, and...  ...building or scaling LLM serving systems,...  ...Experience with custom kernel development (e.g., CUDA... 
    Full time
    Remote work

    Waymo

    Remote
    1 day ago
  • $170.1k - $258.3k

     ...capable fully self-driving systems, to move us toward...  ...accessible mobility. For the AI Kernels & Compilers team, that...  ..., and performance engineering so that every cycle on...  ...high‑performance GPU kernels and custom libraries...  ...that make it easier to profile, debug, and validate CUDA... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Warren, MI
    5 days ago
  • $193.3k - $261.5k

     ...Trainium.The Acceleration Kernel Library team is at the...  ...software boundary, our engineers craft high-performance...  ..., and machine learning systems, you'll bring expertise...  ...analysis using profiling tools to identify and resolve...  ...architectures- Experience with GPU kernel optimization and... 
    Internship
    Local area
    Work from home
    Flexible hours

    Amazon

    Cupertino, CA
    4 days ago
  • About the RoleAs Senior MLOps Engineer, you will focus on supporting cross...  ..., including model serving (real-time and batch), training...  ...environments, and orchestration systems, with a focus on performance,...  ...orchestration of autonomous workflows or LLM-driven agentsJob SummaryJob... 
    Remote work

    Kohl's

    Menomonee Falls, WI
    5 days ago
  •  ...services and tools including LLM fine-tuning, alignment and...  ...principal machine learning engineer, you will be responsible...  ...compression, on-device inference, GPU inference, pytorch, kernel development Preferred...  ...). Customer Support Systems : Experience with AI... 
    Remote job
    Full time
    Casual work
    Live in
    Work at office

    Airbnb, Inc.

    United States
    1 day ago
  • $190.2k - $345.65k

     ...Staff Machine Learning Engineer to architect and...  ...intelligence — the systems that turn massive...  ...retrieval stack that serves it, and the tool...  ...correctness, and cost profile enterprise scale...  ...ANN index tuning, GPU-accelerated enrichment...  ...building retrieval for LLM and agentic systems... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    New York, NY
    5 days ago
  •  ...Overview We are seeking a Senior GPU Systems & Fabric Engineer to serve as the critical bridge between our physical...  ...requires deep expertise in Linux kernel internals, GPU architectures, and...  ...maximize cluster utilization. Profile and tune kernel-level parameters, device... 
    Remote job
    Full time
    Local area

    Bitdeer Technologies Group

    Remote
    16 days ago
  •  ...implement robust model serving infrastructure using platforms...  ...# Improve GPU usage, enable autoscaling...  ...degree in Computer Science, Engineering, or a related field (or...  ...years of experience as MLOps engineer or DevOps...  ...~ Experience in AI/ML systems security, compliance, and... 
    Work at office
    Remote work
    Relocation package

    Fundamental

    United States
    3 days ago
  • $198k - $326k

     ...power AI across LinkedIn. The LLM Serving team builds the critical...  ...for a Senior Staff Software Engineer with deep expertise at the intersection of systems, machine learning, GPU infrastructure, and large-scale...  ..., runtime, compiler, kernel, and hardware layersDrive model... 
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Mountain View, CA
    4 days ago
  •  ...Intelligence team is building an NPC system that can (1) play any Roblox...  ...across Discovery, Safety, Engine, and more. We are seeking...  ...maintain core platform components: Serving Layer, Model Registry,...  ...quantization) and systems on GPU architectures to maintain peak... 
    Full time

    Roblox

    Remote
    1 day ago
  • $100k - $150k

     ...ML Performance Engineer - Remote    Bright...  ...neural network systems. The role spans the...  ...from low-level kernel optimization to distributed...  ...understanding of GPU architecture,...  ...Profile and optimize end-to...  ...speculative decoding for LLM serving.  Drive compiler... 
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    Renton, WA
    8 days ago
  •  ...to converse with all of their business systems through natural language to quickly find...  ...workflow automation with Moveworks’ Reasoning Engine and natural language capabilities, we...  ...edge ML infrastructure for building and serving LLM’s at Moveworks. This role will be... 
    Full time
    Work at office
    Remote work
    Flexible hours

    Servicenow

    Remote
    1 day ago
  • $397.46k

     ...scalable machine learning systems that power...  ...looking for a Distinguished Engineer/Technical Director to lead...  ...generative modeling, and LLM-powered personalization...  ...from data ingestion to GPU training to live inference...  ...for training, serving, and evaluating both traditional... 
    Full time
    Work experience placement
    H1b
    Work at office
    Local area
    Visa sponsorship
    Monday to Friday

    Roblox

    San Mateo, CA
    4 days ago
  • $130k - $220k

     .... We’re hiring a Machine Learning Engineer (Recommendation Systems) to build the personalization engine...  ...gives you the chance to work on systems serving 100M+ predictions daily, directly...  ...distributed computing (Spark, Ray) and LLM/AI Agent frameworks ~ Track record... 
    Full time
    Remote work

    Launch Potato

    Fort Lauderdale, FL
    1 day ago
  • $135.2k - $306.4k

    Oracle hardware platform development engineering is seeking a highly driven GPU/CPU Platform System Engineer at the Principal Engineer level. The GPU System Engineer...  ...process out to production. You will also serve as the last level of engineering technical support... 
    Temporary work
    Work experience placement
    Remote work
    Flexible hours

    Oracle Corporation

    Seattle, WA
    5 days ago
  •  ...who builds with it every day. An engineer whose default instinct is to reach for an LLM, an agent framework, or a durable...  .... You'll design and ship agentic systems that touch real users and real money...  ...tools shop. Our platform serves thousands of traders. The things... 
    Full time
    Live in
    Immediate start
    Weekend work
    1 day per week

    Tradeify

    Remote
    23 days ago
  • $145k - $165k

     ...potential. Job Title: ML Systems Engineer Location: 100% Remote (U....  ...reliable inference platforms for serving large machine learning models...  ..., caching, autoscaling, GPU utilization, and end-to-end observability...  ...Hands-on experience with LLM or large model inference... 
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    United States
    4 days ago
  • $174.96k - $213.84k

     ...Machine Learning Engineer (L3) . About the...  ...edge products that serve developers, builders...  ...-latency, ML-based systems for real-time applications...  ...orchestration and MLOps tooling....  ...Metaflow, SageMaker; LLM/agent frameworks such...  ...; familiarity with GPU-based implementation... 
    Full time
    Local area
    Remote work

    Twilio

    United States
    1 day ago
  •  ...experienced Machine Learning Engineer to join our AI & Threat...  ...for attackers, and the ML systems you build will serve as a first line of defense...  ...engineering techniques and LLM frameworks ~ Experience building...  ...who submit your profile References (with your consent... 
    Remote job
    Full time
    Temporary work

    Keeper Security

    United States
    1 day ago
  • $145k - $165k

     ..., Trust & Safety, Profile, Chat, Growth, and...  ...Machine Learning Engineers (this role) who focus...  ...training, serving, and feature management...  ...machine learning systems that improve product...  ...Understanding of MLOps practices including...  ...systems Exposure to LLM-related use cases... 
    Full time
    Work experience placement
    Casual work
    Work at office

    Match Group

    Remote
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling). Be the first to apply!