Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Remote ML Research Engineer — Inference & GPU Kernel Expert

Jobleads-US

Draper, UT
  • Remote job

AfterQuery is building a research cohort focused on ML inference and GPU kernel engineering, spanning inference serving systems to deployment infrastructure. This is research-and-evaluation work, not production engineering, defining what is considered correct and excellent for problems you know deeply.

Responsibilities include designing realistic scenarios, authoring reference solutions and rubrics, and evaluating AI outputs for depth and domain judgment.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Remote ML Research Engineer — Inference & GPU Kernel Expert in Draper, UT vacancy
  •  ...AfterQuery is assembling a research cohort focused on ML inference and GPU kernel engineering, spanning inference serving systems, kernel optimization, and deployment...  ..., domain-specific problems. The role is fully remote, asynchronous, and contract-based, with emphasis... 
    Remote job
    Contract work
    Work at office

    Jobleads-US

    Boise, ID
    3 days ago
  •  ...AfterQuery seeks an ML Research Engineer focusing on inference and GPU kernels for remote, contract work. You will architect challenging evaluation problems, craft reference solutions, and judge AI outputs with rigor and domain insight. The role emphasizes research,... 
    Remote work
    Contract work

    Jobleads-US

    Tempe, AZ
    3 days ago
  •  ...AfterQuery is building a research lab focused on ML inference and GPU kernel engineering, covering inference serving systems, kernel optimization, and deployment...  ...10 years of experience or equivalent. It’s fully remote, asynchronous, and offers competitive hourly compensation... 
    Remote job
    Hourly pay

    Jobleads-US

    Laurel, MD
    3 days ago
  •  ...AfterQuery is a research lab exploring the boundaries of artificial...  ...build a cohort of senior ML researchers to define...  ...world problems, focusing on inference and GPU kernel engineering. This is research-and-...  ...engineering. The role is remote, flexible, and asynchronous... 
    Remote job
    Flexible hours

    Jobleads-US

    Ann Arbor, MI
    3 days ago
  •  ...AfterQuery is a research lab exploring AI frontiers through...  ...building a cohort of senior ML researchers focused on inference and GPU kernel engineering to define criteria for correctness...  ...problems. This fully remote, contract role invites experts to design scenarios,... 
    Remote work
    Contract work
    Flexible hours

    Jobleads-US

    Irvine, CA
    3 days ago
  •  ...AfterQuery is seeking an ML Research Engineer focusing on inference and GPU kernels. This remote contract role involves designing research scenarios, creating reference solutions, and grading AI outputs. You’ll contribute to frontier AI evaluation rather than production... 
    Remote job
    Hourly pay
    Contract work
    Flexible hours

    Jobleads-US

    Irving, TX
    3 days ago
  • $135 per hour

     ...AfterQuery is a research lab exploring the boundaries of artificial intelligence through novel datasets...  ...from exceptional, human-generated data. This remote, contract role focuses on ML inference and GPU kernel engineering. Competitive hourly pay ($135/hr) based on... 
    Remote job
    Hourly pay
    Contract work
    Flexible hours

    Jobleads-US

    Rockville, MD
    3 days ago
  •  ...AfterQuery is a research lab exploring AI boundaries through novel datasets and evaluation, focusing on ML inference and GPU kernel engineering. This remote, contract role invites senior experts to define correct and excellent solutions for hard, real-world problems.... 
    Remote job
    Contract work

    Jobleads-US

    Lansing, MI
    3 days ago
  •  ...AfterQuery is a research lab exploring the boundaries of artificial intelligence through novel...  ...datasets and experimentation. We recruit for an ML Research Engineer focusing on inference and GPU kernel engineering, working remotely on research-and-evaluation tasks rather... 
    Remote job

    Jobleads-US

    New York, NY
    2 days ago
  •  ...AfterQuery seeks an ML Research Engineer focused on inference and GPU kernels. The role is fully remote and contract-based, offering competitive hourly compensation and the chance to work on frontier AI research problems. You will define problems, develop reference... 
    Remote job
    Hourly pay
    Contract work

    Jobleads-US

    Bridgewater, MA
    3 days ago
  •  ...partner is looking for an AI Research Engineer (Kernel & Inference Optimization) based in...  ...inference strategies and GPU kernels. You will work...  ...teams in a highly technical, remote environment focused on pushing...  ...pipeline parallelism, and expert parallelism for large-... 
    Remote work
    Full time

    Jobgether

    Remote
    7 days ago
  • $70 - $90 per hour

    Gridnaut Recruiting is hiring a remote GPU Kernel Expert contractor (pay $70-$90/hr). Contribute to frontier AI research and evaluation work. Ideal candidates: 3+ years GPU/accelerator kernel development; Proficient in at least two: CUDA, Triton, NKI, Pallas; Experience... 
    Remote work
    Temporary work
    For contractors

    Gridnaut Recruiting

    Remote
    23 days ago
  • $170.1k - $258.3k

     ...mobility. For the AI Kernels & Compilers team,...  ..., and planning research into production‑grade...  ..., and performance engineering so that every...  ...high‑performance GPU kernels and custom...  ...of our on‑vehicle ML inference for ADAS and autonomous...  ...of America; Remote - Washington; Austin... 
    Remote work
    Full time
    Local area
    Work from home
    Relocation package
    Flexible hours

    General Motors

    San Francisco, CA
    5 days ago
  •  ...partner is looking for an AI Research Engineer (Kernel & Inference Optimization) based in...  ...inference strategies and GPU kernels. You will work...  ...teams in a highly technical, remote environment focused on pushing...  ...pipeline parallelism, and expert parallelism for large-... 
    Remote job
    Full time

    jobgether

    United States
    6 days ago
  •  ...ElevenLabs is an AI research and product...  ...researchers, engineers, and operators....  ...Optimizing inference performance across...  ...strategies, and custom kernels. Building...  ...and serving ML models in production...  ...skills in GPU programming and...  ...This role is remote and can be executed... 
    Remote work
    Full time
    Immediate start

    Valor Capital Group

    United Kingdom
    9 days ago
  •  ...At Atlassian, the ML System Engineer will design and optimize large-scale model serving systems...  ...distributed infrastructure to low-level GPU kernel optimizations. You’ll own end-to-end...  ...serving systems, benchmarking and tuning inference engines, and partnering with senior ML... 
    Remote job

    Jobleads-US

    Kentucky
    3 days ago
  •  ...Member of Technical Staff, ML Systems to accelerate model training and inference across image, video, and world...  ...with a founding team on kernels, runtimes, and distributed engines that power production-scale...  ...stacks. You’ll optimize GPU performance, profile bottlenecks... 

    Jobleads-US

    Menlo Park, CA
    2 days ago
  • $90 - $120 per hour

    Gridnaut Recruiting is hiring a remote MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) contractor (pay $90-$1...  ...0/hr). Contribute to frontier AI research and evaluation work. Ideal candidates...  ...-on professional experience in ML systems, ML infrastructure, model... 
    Remote work
    Temporary work
    For contractors
    Weekday work

    Gridnaut Recruiting

    Remote
    22 days ago
  • $150k - $220k

     ...developers, from indie researchers to teams running...  ...than 20 billion inference requests. We...  ...We're a small, remote-first team. We take...  ...re looking for a ML Systems Engineer, Inference. We want...  ...down to kernels and interconnect....  ...node and multi-node GPU deployments. Turn... 
    Remote work
    Full time
    Home office
    Visa sponsorship
    Work visa
    Flexible hours

    Runpod

    Remote
    6 days ago
  • $150k

     ...Models We are a dedicated research lab for building, understanding...  ..., data scientists, and engineers, tackling the most fundamental...  ...pioneers. The Role The GPU Kernel Engineer will play a role at...  ...especially at training and inference, and support the team to... 
    Full time
    Visa sponsorship

    Institute of Foundation Models

    Sunnyvale, CA
    2 days ago
  • $193.3k - $261.5k

     ....The Acceleration Kernel Library team is at...  ...for AWS's custom ML accelerators. Working...  ...boundary, our engineers craft high-performance...  ...unparalleled ML inference and training performance...  ...cutting-edge research, and mentor a...  ...- Experience with GPU kernel optimization... 
    Internship
    Local area
    Work from home
    Flexible hours

    Amazon

    Cupertino, CA
    5 days ago
  •  ...Avride is seeking a software engineer with leadership experience to drive the ML infrastructure layer across the company. You will lead GPU inference optimization for onboard and high-throughput offboard scenarios, while guiding broader ML infrastructure across pipelines... 

    Jobleads-US

    Austin, TX
    3 days ago
  • $250k - $350k

     ...seeking Senior/Staff level Inference Engineers to accelerate the performance...  ...edge inference acceleration, GPU parallelism, advanced model...  ...techniques, and work closely with researchers and engineers across the...  ...high-performance computing kernels and distributed workloads... 
    Work at office
    3 days per week

    Pika

    Palo Alto, CA
    5 days ago
  •  ...'re seeking MLOps Engineers with hands-on experience...  ...of four areas: GPU kernel programming,...  ...and high-throughput inference serving. This role...  ...assessing MLOps and ML systems tasks and...  ...them. ~ Guide research and engineering teams...  ...subject matter experts to keep training data... 
    Full time
    Contract work
    Temporary work
    Weekday work

    Mercor

    Remote
    18 days ago
  • $90 - $120 per hour

     ...with leading AI research labs....  ...Position: MLOps Engineer, LLM Systems (Serving, GPU Kernels, Profiling) Type...  ...Location: Remote Commitment:...  ...debugging , and inference serving . Write...  ...performance on ML systems and training...  ...subject matter experts to ensure... 
    Remote work
    Contract work
    Summer work

    Mercor

    San Francisco, CA
    7 days ago
  •  ...professors, SAIL researchers, Olympiad medalists...  ...state‑of‑the‑art CUDA kernels to power AI...  ...scale model training, inference, and reinforcement...  ...the limits of GPU utilization for compute...  ...researchers and engineers, you’ll help make...  ...and semiconductor experts to translate domain... 

    Voltai

    Palo Alto, CA
    5 days ago
  •  ...Evaluate the quality, correctness, and completeness of GPU/accelerator kernel development tasks used to train and evaluate a frontier AI lab...  ...(NKI/Pallas/TPU) ecosystems • Background in compiler engineering, MLIR, or intermediate-representation lowering • Understanding... 
    Temporary work

    Mercor

    Remote
    26 days ago
  •  ...California. The Role: As a Research Engineer - AI Performance & Kernel Optimization , you will...  ...model training and inference stacks. You will work closely...  ...optimization for large-scale ML workloads, using any level...  ..., HIP, Triton, or other GPU DSLs Performance tuning... 
    Work at office
    Relocation package

    Zyphra

    San Francisco, CA
    6 days ago
  •  ...Pluralis Research works on Protocol Learning: training...  ...Protocol Learning]( Our inference pipeline generates the...  ...transport, the serving engine, and failure handling....  ...a high base salary. Remote-First Culture : Flexible...  ...deeply technical team of ML researchers. Pluralis... 
    Remote work
    Full time
    Immediate start
    Visa sponsorship
    Relocation package
    Flexible hours

    Pluralis Research

    California
    16 days ago
  • $236k - $330k

     ...systems developers and researchers to join the Snowflake...  ...of the art in LLM inference systems and optimization...  ...runtime systems to GPU kernels and model-system co-design...  ...We embrace AI-native engineering, using AI not only as...  ...pipeline, data, and expert parallelism.Develop... 

    Snowflake

    Bellevue, WA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Remote ML Research Engineer — Inference & GPU Kernel Expert. Be the first to apply!