Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Research Engineer: Inference & GPU Kernels (Remote)

Jobleads-US

Bridgewater, MA
  • Remote job

AfterQuery seeks an ML Research Engineer focused on inference and GPU kernels. The role is fully remote and contract-based, offering competitive hourly compensation and the chance to work on frontier AI research problems.

You will define problems, develop reference solutions, and evaluate outputs in a rigorous, research-driven setting. The ideal candidate has 6–10 years of hands-on ML research experience, a PhD or a Master's with strong research credentials, and the ability to work

#J-18808-Ljbffr Jobleads-US
Vacancy posted 17 hours ago
Similar jobs that could be interesting for youBased on the AI Research Engineer: Inference & GPU Kernels (Remote) in Bridgewater, MA vacancy
  •  ...steps. Our partner is looking for an AI Research Engineer (Kernel & Inference Optimization) based in United States...  ...novel inference strategies and GPU kernels. You will work with complex...  ...functional teams in a highly technical, remote environment focused on pushing the boundaries... 
    Remote work
    Full time

    Jobgether

    Remote
    3 days ago
  •  ...steps. Our partner is looking for an AI Research Engineer (Kernel & Inference Optimization) based in Austria....  ...develop novel inference strategies and GPU kernels. You will work with complex...  ...teams in a highly technical, remote environment focused on pushing the boundaries... 
    Remote work
    Full time

    jobgether

    United States
    4 days ago
  •  ...AfterQuery is a research lab exploring AI boundaries through novel datasets and evaluation, focusing on ML inference and GPU kernel engineering. This remote, contract role invites senior experts to define correct and excellent solutions for hard, real-world problems.... 
    Remote job
    Contract work

    Jobleads-US

    Lansing, MI
    16 hours ago
  •  ...AfterQuery seeks an ML Research Engineer focusing on inference and GPU kernels for remote, contract work. You will architect challenging evaluation problems, craft reference solutions, and judge AI outputs with rigor and domain insight. The role emphasizes research... 
    Remote work
    Contract work

    Jobleads-US

    Tempe, AZ
    17 hours ago
  •  ...AfterQuery is a research lab exploring the boundaries of artificial intelligence...  ...hard, real-world problems, focusing on inference and GPU kernel engineering. This is research-and-evaluation...  ...production engineering. The role is remote, flexible, and asynchronous, with... 
    Remote job
    Flexible hours

    Jobleads-US

    Ann Arbor, MI
    16 hours ago
  •  ...AfterQuery is seeking an ML Research Engineer focusing on inference and GPU kernels. This remote contract role involves designing research scenarios, creating reference solutions, and grading AI outputs. You’ll contribute to frontier AI evaluation rather than production... 
    Remote job
    Hourly pay
    Contract work
    Flexible hours

    Jobleads-US

    Irving, TX
    16 hours ago
  •  ...AfterQuery is a research lab investigating the boundaries of artificial intelligence through...  ...and experimentation. We believe great AI comes from exceptional, human-generated data...  ...leading AI labs. This role offers fully remote, flexible, asynchronous work with competitive... 
    Remote work
    Hourly pay
    Flexible hours

    Jobleads-US

    Hartford, CT
    17 hours ago
  •  ...innovation in model serving and inference architectures, the full-time AI Research Engineer will optimize deployment...  ...AI systems in a fully remote environment, focusing on...  ...experience in low-level kernel optimizations and...  ...Strong expertise in writing GPU kernels for mobile... 
    Remote work
    Full time

    Virtual Vocations Inc

    United States
    5 days ago
  •  ...the human layer of AI. Our mission is...  ...through pioneering research in multimodal AI...  ...Scientist/Engineer with a focus on model...  ...understanding of inference performance and GPU/accelerator fundamentals...  ...Triton/CUDA kernels or low-level performance...  ...support offered. Remote candidates are... 
    Remote work
    Full time
    Relocation package
    Flexible hours

    Tavus

    San Francisco, CA
    4 days ago
  •  ...ElevenLabs is an AI research and product company...  ...are researchers, engineers, and operators. IOI...  ...infrastructure. Optimizing inference performance across...  ..., and custom kernels. Building and...  ...skills in GPU programming and inference...  ...This role is remote and can be executed... 
    Remote work
    Full time
    Immediate start

    Valor Capital Group

    United Kingdom
    6 days ago
  •  ...oprecruiting.comPhone: (***) ***-****Job Title: AI Research Engineer - Deep Learning InferenceLocation: New...  ...acceleration, building low-latency inference systems that process continuous global...  ...model execution, including custom kernel creation, data streaming pipelines, and... 

    Objective Paradigm

    Chicago, IL
    2 days ago
  •  ...Description This is an applied research role, not a machine learning...  ...in one person — systems engineering, inference-stack depth, and honest experimental...  ...when the numbers say so. GPU and serving-stack modeling....  ...learning. Flexible remote working environment.... 
    Remote work
    Immediate start
    Flexible hours

    Akka

    San Francisco, CA
    3 days ago
  • $182k - $242k

     ...is The Essential Cloud for AI™. Built for pioneers by pioneers...  ...cloud for high-performance GPU infrastructure across AI/ML...  ..., rendering, and real-time inference. Our stack is engineered for speed, scale, and cost-...  ...team, focused on kernel authoring and optimization.... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Bellevue, WA
    a month ago
  •  ...FMX is seeking a Senior AI Research Engineer to help design, build, and scale...  ...tuning, governed deployment, and inference optimization.Bachelor's...  ...TGI, ONNX, batching, caching, GPU utilization, and latency/...  ...LangChain, LlamaIndex, Semantic Kernel, AutoGen, CrewAI, or similar... 

    Cantor Fitzgerald

    New York, NY
    5 days ago
  • $70 - $90 per hour

     ...technical talent with leading AI research labs. Headquartered in San...  ...Jack Dorsey . Position: GPU Kernel Expert Type: Contract...  ...70–$90/hour Location: Remote Role Responsibilities...  .... Background in compiler engineering, MLIR , or intermediate-representation... 
    Remote work
    Contract work
    Summer work

    Mercor

    Chicago, IL
    4 days ago
  • $230k - $350k

     ...Technical Staff, Inference Systems, to build...  ...high-performance AI inference platform...  ...and strong systems engineering skills, with Rust...  ...workloads across multi-GPU and multi-node...  ...CUDA or Triton kernel development experience...  ...accelerator company, research lab, or similar... 
    Work at office

    Premier Global Links

    Palo Alto, CA
    14 days ago
  •  ...semiconductors where AI can design and...  ...professors, SAIL researchers, Olympiad medalists...  ...state‑of‑the‑art CUDA kernels to power AI...  ...scale model training, inference, and reinforcement...  ...the limits of GPU utilization for compute...  ...researchers and engineers, you’ll help make... 

    Voltai

    Palo Alto, CA
    2 days ago
  •  ...is to enable every engineering organization to...  ...workforce. It combines AI agents,...  ...semiconductor design, GPU-accelerated computing...  ...to turn advanced research into technology that...  ...the model training, inference, and...  ...optimize custom GPU kernels (CUDA, Triton) and... 
    Full time

    International Recruiting LLC

    Bellevue, WA
    6 days ago
  • GPU Kernel Engineer Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable... 
    Flexible hours

    Baseten

    New York, NY
    2 days ago
  • $224k - $356.5k

     ...a Senior Software Engineer to join our team!...  ...combining the best AI and graphics technology...  ...AI models from research to real-time,...  ...Optimize models and inference for latency, token...  ...model architecture to GPU kernels.Integrate models...  ...: US, CA, Remote; US, CA, Santa ClaraType... 
    Remote work
    Full time

    Nvidia

    Santa Clara, CA
    4 hours ago
  •  ...Senior GPU Systems / AI Infrastructure Engineer (NYC) Location: New York City (Hybrid /...  ...scale model training and inference. This role sits at the intersection...  ...of GPU systems, kernel optimisation, distributed...  ...closely with ML researchers and infra engineers to remove... 
    Full time
    New York, NY
    more than 2 months ago
  • $200k - $250k

     ...About the Role We're looking for an AI Inference Platform Engineer to build, operate, and optimize the...  ...platform engineering, with deep GPU literacy. You'll own the serving platform...  ...bottlenecks from individual GPU kernels through multi-node inference systems.... 
    Temporary work
    Flexible hours

    DRW

    Chicago, IL
    2 days ago
  • $70 - $90 per hour

     ...Review GPU and accelerator kernel development tasks that support the training and evaluation of advanced AI models. This role focuses on assessing task...  ...Background in compiler engineering, MLIR, or intermediate-representation...  ...calls. Work Terms Remote role open to candidates... 
    Remote work
    Hourly pay

    SaidGig

    Remote
    a month ago
  • $100k - $150k

     ...AI Performance Engineer – Remote Bright Vision Technologies is a technology...  ...end AI training and inference pipelines for...  ...current with AI systems research and translate...  ...profiling tools across CPU, GPU, and distributed...  ...Familiarity with custom kernel authoring in Triton... 
    Remote work
    Full time
    H1b
    Local area
    Immediate start
    Visa sponsorship

    Bright Vision Technologies

    Monroeville, PA
    9 days ago
  • $215k - $285k

     ...future of physical AI. Founded in 2017 and...  ...include occasional remote work, starting the...  ...for a performance engineer who specializes in...  ...high-throughput batch inference sweeping petabytes...  ...wastes 30% of its GPU-hours on stalled data...  ..., augmentation, kernel execution, gradient... 
    Remote work
    Full time
    For contractors
    For subcontractor
    Casual work
    Work at office
    Day shift

    Applied Intuition

    California
    22 days ago
  •  ...Francisco, California. The Role: As a Research Engineer - AI Performance & Kernel Optimization , you will improve...  ...-scale language model training and inference stacks. You will work closely with...  ...to CUDA, HIP, Triton, or other GPU DSLs Performance tuning for training... 
    Work at office
    Relocation package

    Zyphra

    San Francisco, CA
    3 days ago
  • $144.7k - $261.3k

     ...infrastructure, and ML/AI GPU platforms for AV research and development...  ...Senior Performance Engineer to join the AV Capacity...  ...scale ML training and inference environments....  ...Nsight Compute for kernel‑level performance tuning...  ...discretion. Ability to sit remote in Seattle, WA until... 
    Remote work
    Work at office
    Local area
    Work from home
    Flexible hours
    3 days per week

    General Motors LLC

    Sunnyvale, CA
    3 hours ago
  • $100k - $150k

     ...AI Research Engineer – Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data...  ...methodology. Build production-quality training and inference pipelines using modern ML frameworks and orchestration tools... 
    Remote work
    Full time
    H1b
    Local area
    Immediate start
    Visa sponsorship

    Bright Vision Technologies

    Remote
    3 days ago
  •  ...Pluralis Research works on Protocol Learning: training...  ...Protocol Learning]( Our inference pipeline generates the...  ...transport, the serving engine, and failure handling....  ...trustless, and sovereign AI. Nice to Have Familiarity...  ...a high base salary. Remote-First Culture :... 
    Remote work
    Full time
    Immediate start
    Visa sponsorship
    Relocation package
    Flexible hours

    Pluralis Research

    California
    13 days ago
  •  ...Pluralis Research works on Protocol Learning: training and serving large...  ...Learning]( Our training and inference network is trustless, and the...  ...collective, trustless, and sovereign AI. Nice to Have...  ...addition to a high base salary. Remote-First Culture : Flexible work... 
    Remote work
    Full time
    Visa sponsorship
    Relocation package
    Flexible hours

    Pluralis Research

    California
    13 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Research Engineer: Inference & GPU Kernels (Remote). Be the first to apply!