Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Trainium (NKI) Kernel Developer for AI Model Evaluation

$70 - $90 per hour

SaidGig

Role Overview

Help evaluate Neuron Kernel Interface development tasks that support the training and evaluation of advanced AI models. You will assess kernel quality, numerical correctness, CUDA-to-NKI migration fidelity, and whether implementations are well suited to AWS Trainium hardware, then deliver clear written feedback using defined evaluation criteria.

Key Responsibilities

  • Evaluate NKI kernel development tasks for quality, correctness, and hardware appropriateness.
  • Review CUDA-to-NKI migrations for fidelity and implementation quality.
  • Assess Trainium-specific performance optimization, including effective use of hardware resources.
  • Evaluate cross-platform numerical-correctness standards, including accumulation order, rounding behavior, and mixed-precision semantics across GPU and Trainium environments.
  • Provide clear, rubric-based written feedback.

Qualifications

  • At least 2 years of hands-on experience developing or optimizing NKI kernels for AWS Trainium or Inferentia2 hardware.
  • Strong knowledge of NKI development patterns, including tile-based computation, SBUF, PSUM, and HBM memory-hierarchy management, partition-dimension constraints, and DMA orchestration.
  • Experience evaluating CUDA-to-NKI migration quality.
  • Familiarity with Trainium performance profiling, including NeuronCore pipeline utilization, tensor-engine throughput, and memory-bandwidth bottlenecks.
  • Experience defining or evaluating numerical-correctness standards across platforms.

Preferred Qualifications

  • Experience with the AWS Neuron SDK, Neuron Compiler internals, or NKI kernel-library contributions.
  • Prior CUDA or Triton kernel-development experience.
  • Knowledge of Trainium hardware, including NeuronCore-v2 architecture, on-chip SRAM topology, and FP32, BF16, FP8, and INT8 data types.
  • Experience benchmarking machine-learning training workloads on Trn1 or Trn2 instances.

Work Terms

  • Remote role, open to candidates located in the United States.
  • Hourly engagement.

Compensation

  • $70 to $90 per hour.
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Trainium (NKI) Kernel Developer for AI Model Evaluation in Remote vacancy
  • $70 - $90 per hour

     ...Review GPU and accelerator kernel development tasks that support the training and evaluation of advanced AI models. This role focuses on assessing...  ...of hands-on experience developing, optimizing, or verifying GPU...  ...following: CUDA, Triton, NKI, or Pallas with JAX. Strong... 
    Suggested
    Hourly pay
    Remote work

    SaidGig

    Remote
    3 days ago
  •  ...Infrastructure Engineer to design and build data pipelines, evaluation harnesses, and annotation tooling powering AI systems. This fully remote hourly contract role focuses on production-grade Python systems for model evaluation at scale. You will collaborate with research... 
    Suggested
    Remote job
    Hourly pay
    Contract work

    Alignerr

    New York, NY
    17 hours ago
  •  ...remote Kotlin Engineer to review AI-generated responses and create...  .... Responsibilities include developing AI prompts, optimizing AI performance, and ensuring model accuracy. The ideal candidate has...  ...network and requires critical evaluation of technical concepts. #J-1880... 
    Suggested
    Remote job

    SME Careers

    New York, NY
    5 days ago
  •  ...The Opportunity Deepgram is looking for a Senior Software Engineer - Model Evaluation & AI Systems to join the team responsible for validating the quality of our speech, audio, and multilingual models before they reach customers. This team owns the evaluation and... 
    Suggested
    Full time

    Deepgram

    Remote
    a month ago
  • $85 per hour

     ...elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our...  ...Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-generated implementations involving cloud platforms... 
    Suggested
    Contract work
    Summer work
    Remote work

    Mercor

    New York, NY
    3 days ago
  • $165.2k - $223.6k

     ...accelerators, Inferentia and Trainium.The AWS Neuron SDK, developed by the Annapurna Labs team...  ...running a wide range of models and supporting novel...  ...and create high-performance kernels for ML functions, ensuring...  ...boundaries of what's possible in AI acceleration.As part of... 
    Work experience placement
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $400 per month

     ...Role Overview Evaluate and improve frontier AI coding models by using AI coding agents to complete and assess realistic infrastructure engineering workflows. Work will focus on reviewing model-generated infrastructure solutions across cloud platforms, container orchestration... 
    Hourly pay
    Remote work

    SaidGig

    United States
    a month ago
  •  ...Vice President – Technical AI Foundation Model Engineer Discover your opportunity...  ...AI architectures. Develop and maintain Retrieval-...  ...semantic search capabilities. Evaluate, benchmark, and recommend...  ...LangChain LlamaIndex Semantic Kernel Cloud & Platform... 
    Work at office
    Remote work

    MUFG

    Jersey City, NJ
    5 days ago
  •  ...AI Foundation Model Engineer NTT DATA strives to hire exceptional, innovative and passionate...  ...experimentation to deployment, monitoring, evaluation, rollback, and continuous improvement....  ...Face, LangChain, LlamaIndex, Semantic Kernel, or equivalent frameworks. ~... 
    Temporary work
    Work at office
    Remote work
    Flexible hours

    NTT DATA

    Jersey City, NJ
    4 days ago
  •  ...The Opportunity Deepgram is looking for a Senior Software Engineer - Model Evaluation & AI Systems to join the team responsible for validating the quality of our speech, audio, and multilingual models before they reach customers. This team owns the evaluation and... 
    Full time

    Deepgram

    Remote
    23 days ago
  • $85 per hour

     ...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our...  ...Use frontier AI coding agents to complete and evaluate complex engineering tasks. Review model-generated mobile application code for correctness, quality... 
    Contract work
    Summer work
    Remote work

    Mercor

    San Francisco, CA
    29 days ago
  • $70 - $90 per hour

     ...technical talent with leading AI research labs....  ...Dorsey . Position: GPU Kernel Expert Type: Contract...  ...Responsibilities Evaluate the quality and correctness...  ...and evaluating AI models. Assess numerical correctness...  ...CUDA , Triton , NKI , or Pallas (JAX) .... 
    Contract work
    Summer work
    Remote work

    Mercor

    Chicago, IL
    2 days ago
  •  ...Job Title: Senior SAS/Python Model Validation and Modernization Developer Company: BLN24 About Us: We find strength...  ..., and business stakeholders to evaluate model performance, conduct what-if...  ...technologies, particularly AI  Strategic mindset with the ability... 
    Full time
    Work at office
    Remote work

    BLN24

    McLean, VA
    more than 2 months ago
  • $145k - $200k

     ...our platforms empower our partners to develop lifesaving drugs, forecast supply...  ...team with expertise in enabling ML models in production. We deploy AI models to run in variety of environments...  ...and the ability to quickly evaluate and integrate new models and technologies... 
    Full time
    Work experience placement
    Work at office
    Remote work
    Work from home
    Relocation package

    Palantir Technologies

    Palo Alto, CA
    1 day ago
  • $86.8k - $198k

    Model and Simulation Software EngineerThe Opportunity: You will play a critical...  ...Simulation (M&S) capabilities by designing, developing, and integrating AI‑enabled models that support complex...  ...for next‑generation M&S systems by evaluating new frameworks, enhancing simulation... 
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Suffolk, VA
    1 day ago
  •  ...Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant... 
    Remote work
    Flexible hours

    Prolific

    Charlotte, NC
    2 days ago
  • $70 - $150 per hour

     ...Contribute frontend engineering expertise to AI-focused projects, including model training and evaluation, real-world task design, and specialized feedback that helps advance AI research. Role Overview This is an open application for future remote, hourly contract opportunities... 
    Hourly pay
    Contract work
    Remote work

    SaidGig

    United States
    29 days ago
  • $204k - $259k

     ...U.S. states. The core challenge within Model Lifecycle is accelerating Waymo's ML development...  ...ready for efficient model training and evaluation. Develop infrastructure to produce reliable, high...  .... Passionate about data‑centric AI and autonomous driving applications. We... 
    Full time
    Remote work

    Waymo

    Kirkland, WA
    5 days ago
  •  ...Solerity is seeking Mid to Senior Machine Learning Engineers and AI Model Developers to support an upcoming Federal program expected to begin in...  ...August/September 2026. This effort focuses on developing, evaluating, and integrating machine learning models within a secure... 
    For contractors
    Remote work
    Flexible hours

    Solerity

    Fort Belvoir, VA
    4 days ago
  • $117.7k - $221.4k

     ...learning, data infrastructure, and developer productivity, building...  ...iteration across perception and evaluation workflows.Our goal is to make...  ...cost efficient for embodied AI systems. We believe the next...  ...depends not only on stronger models, but also on better infrastructure... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    7 hours ago
  •  ...Evaluate AI-generated music and lyrics across a broad range of genres, applying your Bengali music expertise to help assess quality, originality, and natural expression. Key Responsibilities Compare AI-generated lyrics with published songs to identify similarities.... 
    Hourly pay
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    16 days ago
  • $35 - $62 per hour

     ...Apply your Japanese music expertise to evaluate AI-generated music and lyrics across a wide range of genres. You will assess outputs against detailed quality standards in both Japanese and English. Key Responsibilities Compare AI-generated lyrics with published songs... 
    Hourly pay
    For contractors
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    16 days ago
  • $11 - $19 per hour

     ...Evaluate AI-generated music and lyrics in Telugu and English, helping assess outputs across a broad range of genres against detailed quality standards. Key Responsibilities Compare AI-generated lyrics with published songs to identify similarities. Rate lyrics for... 
    Hourly pay
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    16 days ago
  • $14 - $42 per hour

     ...Evaluate AI-generated music and lyrics across a broad range of genres, applying detailed quality standards in both Hindi and English. Key Responsibilities Compare AI-generated lyrics with published songs to identify similarities. Rate lyrics for quality, creativity... 
    Hourly pay
    For contractors
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    5 days ago
  • $20 - $60 per hour

     ...Role Overview Help train and evaluate next-generation AI systems by creating rigorous, real-world assessments that test how advanced models learn, reason, and perform. This remote contract...  ...required. Key Responsibilities Develop original, challenging question-and-... 
    Hourly pay
    Contract work
    For contractors
    Remote work

    SaidGig

    United States
    5 days ago
  • $100 per hour

     ...expertise to improve the performance of large language models on finance tasks. You will work with AI researchers to identify model weaknesses in areas...  ...focused on advanced AI systems. Key Responsibilities Evaluate LLM performance in finance areas where models... 
    Hourly pay
    Contract work
    For contractors
    Freelance
    Remote work
    10 hours per week
    Flexible hours

    SaidGig

    United States
    more than 2 months ago
  •  ...Role Overview Work with a leading AI lab to evaluate outputs from generative music models in German and English. This role focuses on listening, scoring, and annotating AI-generated music and lyrics across genres, using music production and audio engineering vocabulary... 
    Hourly pay
    Part time
    Immediate start
    Remote work
    10 hours per week

    SaidGig

    United States
    a month ago
  • $46 per hour

     ...steps. Our partner is looking for a Legal Domain Expert (SME) – AI Model Evaluation based in the United States. This is a remote, flexible...  ..., and adherence to applicable legal standards. Develop and apply objective evaluation criteria, scoring frameworks,... 
    Full time
    Contract work
    Remote work
    Flexible hours

    jobgether

    United States
    2 days ago
  • $237.6k - $318.24k

     ...intelligence . As the only vertically integrated AI infrastructure company built from the...  ...Staff Software Engineer for the AI Model Lifecycle team will play a crucial role in...  ...experiment management: versioning, lineage, evaluation, and reproducible fine-tuning at scale.... 
    Full time
    Temporary work

    Crusoe

    Remote
    4 days ago
  • $28 - $60 per hour

     ...Evaluate AI-generated music and lyrics across a wide range of genres, applying your knowledge of the Dutch music scene and strong editorial judgment to detailed quality standards. Key Responsibilities Assess AI-generated music and rate it against detailed quality criteria... 
    Hourly pay
    For contractors
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Trainium (NKI) Kernel Developer for AI Model Evaluation. Be the first to apply!