Trainium (NKI) Kernel Developer for AI Model Evaluation
$70 - $90 per hourSaidGig
Role Overview
Help evaluate Neuron Kernel Interface development tasks that support the training and evaluation of advanced AI models. You will assess kernel quality, numerical correctness, CUDA-to-NKI migration fidelity, and whether implementations are well suited to AWS Trainium hardware, then deliver clear written feedback using defined evaluation criteria.
Key Responsibilities
- Evaluate NKI kernel development tasks for quality, correctness, and hardware appropriateness.
- Review CUDA-to-NKI migrations for fidelity and implementation quality.
- Assess Trainium-specific performance optimization, including effective use of hardware resources.
- Evaluate cross-platform numerical-correctness standards, including accumulation order, rounding behavior, and mixed-precision semantics across GPU and Trainium environments.
- Provide clear, rubric-based written feedback.
Qualifications
- At least 2 years of hands-on experience developing or optimizing NKI kernels for AWS Trainium or Inferentia2 hardware.
- Strong knowledge of NKI development patterns, including tile-based computation, SBUF, PSUM, and HBM memory-hierarchy management, partition-dimension constraints, and DMA orchestration.
- Experience evaluating CUDA-to-NKI migration quality.
- Familiarity with Trainium performance profiling, including NeuronCore pipeline utilization, tensor-engine throughput, and memory-bandwidth bottlenecks.
- Experience defining or evaluating numerical-correctness standards across platforms.
Preferred Qualifications
- Experience with the AWS Neuron SDK, Neuron Compiler internals, or NKI kernel-library contributions.
- Prior CUDA or Triton kernel-development experience.
- Knowledge of Trainium hardware, including NeuronCore-v2 architecture, on-chip SRAM topology, and FP32, BF16, FP8, and INT8 data types.
- Experience benchmarking machine-learning training workloads on Trn1 or Trn2 instances.
Work Terms
- Remote role, open to candidates located in the United States.
- Hourly engagement.
Compensation
- $70 to $90 per hour.
$70 - $90 per hour
...Review GPU and accelerator kernel development tasks that support the training and evaluation of advanced AI models. This role focuses on assessing... ...of hands-on experience developing, optimizing, or verifying GPU... ...following: CUDA, Triton, NKI, or Pallas with JAX. Strong...SuggestedHourly payRemote work- ...Infrastructure Engineer to design and build data pipelines, evaluation harnesses, and annotation tooling powering AI systems. This fully remote hourly contract role focuses on production-grade Python systems for model evaluation at scale. You will collaborate with research...SuggestedRemote jobHourly payContract work
- ...remote Kotlin Engineer to review AI-generated responses and create... .... Responsibilities include developing AI prompts, optimizing AI performance, and ensuring model accuracy. The ideal candidate has... ...network and requires critical evaluation of technical concepts. #J-1880...SuggestedRemote job
- ...The Opportunity Deepgram is looking for a Senior Software Engineer - Model Evaluation & AI Systems to join the team responsible for validating the quality of our speech, audio, and multilingual models before they reach customers. This team owns the evaluation and...SuggestedFull time
$85 per hour
...elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our... ...Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-generated implementations involving cloud platforms...SuggestedContract workSummer workRemote work$165.2k - $223.6k
...accelerators, Inferentia and Trainium.The AWS Neuron SDK, developed by the Annapurna Labs team... ...running a wide range of models and supporting novel... ...and create high-performance kernels for ML functions, ensuring... ...boundaries of what's possible in AI acceleration.As part of...Work experience placementInternshipLocal areaFlexible hours$400 per month
...Role Overview Evaluate and improve frontier AI coding models by using AI coding agents to complete and assess realistic infrastructure engineering workflows. Work will focus on reviewing model-generated infrastructure solutions across cloud platforms, container orchestration...Hourly payRemote work- ...Vice President – Technical AI Foundation Model Engineer Discover your opportunity... ...AI architectures. Develop and maintain Retrieval-... ...semantic search capabilities. Evaluate, benchmark, and recommend... ...LangChain LlamaIndex Semantic Kernel Cloud & Platform...Work at officeRemote work
- ...AI Foundation Model Engineer NTT DATA strives to hire exceptional, innovative and passionate... ...experimentation to deployment, monitoring, evaluation, rollback, and continuous improvement.... ...Face, LangChain, LlamaIndex, Semantic Kernel, or equivalent frameworks. ~...Temporary workWork at officeRemote workFlexible hours
- ...The Opportunity Deepgram is looking for a Senior Software Engineer - Model Evaluation & AI Systems to join the team responsible for validating the quality of our speech, audio, and multilingual models before they reach customers. This team owns the evaluation and...Full time
$85 per hour
...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our... ...Use frontier AI coding agents to complete and evaluate complex engineering tasks. Review model-generated mobile application code for correctness, quality...Contract workSummer workRemote work$70 - $90 per hour
...technical talent with leading AI research labs.... ...Dorsey . Position: GPU Kernel Expert Type: Contract... ...Responsibilities Evaluate the quality and correctness... ...and evaluating AI models. Assess numerical correctness... ...CUDA , Triton , NKI , or Pallas (JAX) ....Contract workSummer workRemote work- ...Job Title: Senior SAS/Python Model Validation and Modernization Developer Company: BLN24 About Us: We find strength... ..., and business stakeholders to evaluate model performance, conduct what-if... ...technologies, particularly AI Strategic mindset with the ability...Full timeWork at officeRemote work
$145k - $200k
...our platforms empower our partners to develop lifesaving drugs, forecast supply... ...team with expertise in enabling ML models in production. We deploy AI models to run in variety of environments... ...and the ability to quickly evaluate and integrate new models and technologies...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package$86.8k - $198k
Model and Simulation Software EngineerThe Opportunity: You will play a critical... ...Simulation (M&S) capabilities by designing, developing, and integrating AI‑enabled models that support complex... ...for next‑generation M&S systems by evaluating new frameworks, enhancing simulation...Full timeContract workPart timeWork at officeLocal areaRemote work- ...Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant...Remote workFlexible hours
$70 - $150 per hour
...Contribute frontend engineering expertise to AI-focused projects, including model training and evaluation, real-world task design, and specialized feedback that helps advance AI research. Role Overview This is an open application for future remote, hourly contract opportunities...Hourly payContract workRemote work$204k - $259k
...U.S. states. The core challenge within Model Lifecycle is accelerating Waymo's ML development... ...ready for efficient model training and evaluation. Develop infrastructure to produce reliable, high... .... Passionate about data‑centric AI and autonomous driving applications. We...Full timeRemote work- ...Solerity is seeking Mid to Senior Machine Learning Engineers and AI Model Developers to support an upcoming Federal program expected to begin in... ...August/September 2026. This effort focuses on developing, evaluating, and integrating machine learning models within a secure...For contractorsRemote workFlexible hours
$117.7k - $221.4k
...learning, data infrastructure, and developer productivity, building... ...iteration across perception and evaluation workflows.Our goal is to make... ...cost efficient for embodied AI systems. We believe the next... ...depends not only on stronger models, but also on better infrastructure...Full timeLocal areaRemote workWork from homeRelocation packageFlexible hours- ...Evaluate AI-generated music and lyrics across a broad range of genres, applying your Bengali music expertise to help assess quality, originality, and natural expression. Key Responsibilities Compare AI-generated lyrics with published songs to identify similarities....Hourly payImmediate startRemote workFlexible hours
$35 - $62 per hour
...Apply your Japanese music expertise to evaluate AI-generated music and lyrics across a wide range of genres. You will assess outputs against detailed quality standards in both Japanese and English. Key Responsibilities Compare AI-generated lyrics with published songs...Hourly payFor contractorsImmediate startRemote workFlexible hours$11 - $19 per hour
...Evaluate AI-generated music and lyrics in Telugu and English, helping assess outputs across a broad range of genres against detailed quality standards. Key Responsibilities Compare AI-generated lyrics with published songs to identify similarities. Rate lyrics for...Hourly payImmediate startRemote workFlexible hours$14 - $42 per hour
...Evaluate AI-generated music and lyrics across a broad range of genres, applying detailed quality standards in both Hindi and English. Key Responsibilities Compare AI-generated lyrics with published songs to identify similarities. Rate lyrics for quality, creativity...Hourly payFor contractorsImmediate startRemote workFlexible hours$20 - $60 per hour
...Role Overview Help train and evaluate next-generation AI systems by creating rigorous, real-world assessments that test how advanced models learn, reason, and perform. This remote contract... ...required. Key Responsibilities Develop original, challenging question-and-...Hourly payContract workFor contractorsRemote work$100 per hour
...expertise to improve the performance of large language models on finance tasks. You will work with AI researchers to identify model weaknesses in areas... ...focused on advanced AI systems. Key Responsibilities Evaluate LLM performance in finance areas where models...Hourly payContract workFor contractorsFreelanceRemote work10 hours per weekFlexible hours- ...Role Overview Work with a leading AI lab to evaluate outputs from generative music models in German and English. This role focuses on listening, scoring, and annotating AI-generated music and lyrics across genres, using music production and audio engineering vocabulary...Hourly payPart timeImmediate startRemote work10 hours per week
$46 per hour
...steps. Our partner is looking for a Legal Domain Expert (SME) – AI Model Evaluation based in the United States. This is a remote, flexible... ..., and adherence to applicable legal standards. Develop and apply objective evaluation criteria, scoring frameworks,...Full timeContract workRemote workFlexible hours$237.6k - $318.24k
...intelligence . As the only vertically integrated AI infrastructure company built from the... ...Staff Software Engineer for the AI Model Lifecycle team will play a crucial role in... ...experiment management: versioning, lineage, evaluation, and reproducible fine-tuning at scale....Full timeTemporary work$28 - $60 per hour
...Evaluate AI-generated music and lyrics across a wide range of genres, applying your knowledge of the Dutch music scene and strong editorial judgment to detailed quality standards. Key Responsibilities Assess AI-generated music and rate it against detailed quality criteria...Hourly payFor contractorsImmediate startRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Trainium (NKI) Kernel Developer for AI Model Evaluation. Be the first to apply!



