NKI Kernel Developer for AI Model Evaluation
$70 - $90 per hourSaidGig
Role Overview
Evaluate Neuron Kernel Interface, NKI, development tasks used to train and assess advanced AI models. You will review CUDA-to-NKI migrations, Trainium-focused performance optimizations, and cross-platform numerical correctness, then deliver clear written feedback using defined rubrics.
Key Responsibilities
- Assess NKI kernel-development tasks for quality, correctness, and suitability for AWS Trainium and Inferentia2 hardware.
- Evaluate the fidelity of CUDA-to-NKI migrations.
- Review Trainium-specific performance optimization quality.
- Evaluate numerical-correctness standards across GPU and Trainium platforms.
- Provide clear, rubric-based written feedback.
Qualifications
- At least 2 years of hands-on experience developing or optimizing NKI kernels for AWS Trainium or Inferentia2 hardware.
- Strong knowledge of tile-based computation, SBUF, PSUM, and HBM memory-hierarchy management, partition-dimension constraints, and DMA orchestration.
- Experience evaluating CUDA-to-NKI migration quality.
- Familiarity with Trainium performance profiling, including NeuronCore pipeline utilization, tensor-engine throughput, and memory-bandwidth bottlenecks.
- Experience establishing or assessing cross-platform numerical-correctness standards, including GPU versus Trainium accumulation order, rounding behavior, and mixed-precision semantics.
Preferred Qualifications
- Experience with the AWS Neuron SDK, Neuron Compiler internals, or NKI kernel-library contributions.
- Prior CUDA or Triton kernel-development experience.
- Familiarity with NeuronCore-v2 architecture, on-chip SRAM topology, and FP32, BF16, FP8, and INT8 data types.
- Experience benchmarking machine-learning training workloads on Trn1 or Trn2 instances.
Work Terms
- Remote role, open to candidates located in the United States.
- Hourly engagement.
Compensation
$70 to $90 per hour.
$70 - $90 per hour
...Review GPU and accelerator kernel development tasks that support the training and evaluation of advanced AI models. This role focuses on assessing... ...of hands-on experience developing, optimizing, or verifying GPU... ...following: CUDA, Triton, NKI, or Pallas with JAX. Strong...SuggestedHourly payRemote work$65 - $105 per hour
...deep engineering judgment to help frontier AI models reason more accurately about real-world... ...define high-quality engineering work, evaluate model performance, and turn expert practice... ...meaningful model improvement. Help develop engineering-specific skills and tools...SuggestedHourly payFull timeFreelanceInternshipLive inRelocationRelocation package$220k - $320k
...out of GPUs, diving deep into CUDA kernels, and turning optimization... ...trains and hosts specialized language models for companies that need frontier-quality AI at a fraction of the cost. The models... ...-to-end: distillation, training, evaluation, and planet-scale hosting. We...SuggestedFull timeWork at office$149k - $205k
...Vice President - Technical AI Foundation Model Engineer Role Summary... ...agentic AI architectures. Develop and maintain Retrieval-Augmented... ...search capabilities. Evaluate, benchmark, and recommend foundation... ...LlamaIndex Semantic Kernel Cloud & Platform...SuggestedWork at officeLocal areaRemote work1 day per week$193.3k - $261.5k
...Trainium.The Amazon Neuron SDK, developed by the Annapurna Labs team at... ...of running a wide range of models and supporting novel architecture... ...and create high-performance kernels for ML functions, ensuring every... ...of what's possible in AI acceleration.As part of the broader...SuggestedWork experience placementInternshipLocal areaFlexible hours$208.73k - $279.57k
...intelligence . As the only vertically integrated AI infrastructure company built from the... ...The Staff Software Engineer for the AI Model Lifecycle team will play a crucial role... ...management: versioning, lineage, evaluation, and reproducible fine-tuning at scale....Full timeTemporary work- ...Job Title: Senior SAS/Python Model Validation and Modernization Developer Company: BLN24 About Us: We find strength... ..., and business stakeholders to evaluate model performance, conduct what-if... ...technologies, particularly AI Strategic mindset with the ability...Full timeWork at officeRemote work
$60 - $90 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Jack Dorsey . Position: Machine Learning Engineer — Model Evaluation & Experimentation Type: Contract Compensation...Full timeContract workSummer workRemote work$70 - $150 per hour
...Role Overview Contribute Full-Stack Engineering expertise to AI-focused projects for leading labs and companies. This open application... ..., and project needs. Key Responsibilities Train and evaluate AI models in Full-Stack Engineering. Create tasks and deliverables...Hourly payContract workRemote work$70 - $150 per hour
...Contribute frontend engineering expertise to AI-focused projects, including model training and evaluation, real-world task design, and specialized feedback that helps advance AI research. Role Overview This is an open application for future remote, hourly contract opportunities...Hourly payContract workRemote work$70 - $150 per hour
...Contribute backend engineering expertise to AI-focused projects, including training and evaluating models, designing real-world technical tasks, and providing feedback that supports advanced AI research. This is an open application for future remote contract opportunities...Hourly payContract workRemote work$100 per hour
...expertise to help improve advanced AI systems through rigorous... ...review, prompt development, evaluation, and annotation. This remote,... ...real-world input that helps AI models learn, reason, and produce... ...Work contributes to a project developing and refining advanced AI systems...Hourly payContract workPart timeFor contractorsRemote work$180k - $245k
...power of innovation, Red Cell is developing powerful tools and solutions... ..., Trase Systems is AI, Uncomplicated. Trase empowers... ...you’ll own the core execution model and platform architecture of... ...structured logs, metrics, and evaluation hooks; build an “explainable...Full timeTemporary workShift work$70 - $80 per hour
...Role Overview Apply advanced drug safety expertise to help improve AI systems through rigorous, real-world evaluation of pharmacovigilance documentation and data. This remote contract role focuses on the quality, accuracy, and regulatory alignment of complex safety reports...Hourly payContract workRemote work$20 - $36 per hour
...Role Overview Evaluate generative music AI across a wide range of genres, applying your knowledge of Hungarian music and lyrics to detailed quality standards. You will work in both Hungarian and English to help assess the quality, originality, and naturalness of AI-generated...Hourly payFor contractorsImmediate startRemote workFlexible hours$100 per hour
...expertise to improve the performance of large language models on finance tasks. You will work with AI researchers to identify model weaknesses in areas... ...focused on advanced AI systems. Key Responsibilities Evaluate LLM performance in finance areas where models...Hourly payContract workFor contractorsFreelanceRemote work10 hours per weekFlexible hours- ...Role Overview Use your investment and finance expertise to evaluate and improve AI model performance on financial reasoning, valuation, markets, and real-world investment scenarios. Key Responsibilities Assess AI model outputs on valuation, financial modeling, markets...For contractorsRemote work
- ...us Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand,... ...from feature integration through training, evaluation, approval and promotion. Integrate the platform...Full timeWork at officeWork from home
$20 per hour
SupportFinity™ is seeking an Editorial Proofreader to evaluate AI models and improve their quality through expert writing and editing skills. This role can be part‑time or full‑time, allowing for a flexible schedule and project selection. Applicants must be fluent in English...Remote jobHourly payFull timePart timeFlexible hours$2,000 per month
...history. Job Summary Every model release presents a new opportunity to push the frontier on kernel engineering. Future performance breakthroughs will come from AI systems that can understand... ...software roadmap. Continuously evaluate new model releases and deploy...Full timeWork at officeRelocation package- ...The role We’re looking for a Model Release Engineer to build the platform that moves Wayve’s AI models from promising new features through training, simulation... ...process, from feature integration through training, evaluation, approval and promotion. Integrate the...Full timeWork at officeWork from home
- ...is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia. Responsibilities...Remote jobHourly payFlexible hours
$20 per hour
SupportFinity™ is looking for an Editorial Proofreader to join our team for AI model training. In this remote role, you'll evaluate AI chatbots and enhance model quality. Candidates should have fluency in English and strong editing skills. This position can be full‑time...Remote jobHourly payFull timeContract workPart time$100 - $150 per hour
...data scientists who will be considered for future projects evaluating how well AI systems perform real-world data science tasks. Members of this... ...criteria, assess AI or human-produced analyses and models, document decisions in writing, and iterate on evaluations with...Hourly payImmediate startRemote work- ...’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and... ...About the role We're looking for Research Engineers to build the evaluations that tell us — and the world — what Claude can actually do....Full time
- ...reasoning and computational problem solving to improve and evaluate large language models. You will design rigorous math problems, produce clear, logically... ...How this work supports customers Accelerate frontier AI research by contributing high quality data and evaluation...Contract workFor contractorsFreelanceRemote work
$50 - $100 per hour
...Apply your software engineering expertise to help train and evaluate next-generation AI systems through real-world coding tasks, technical feedback... ...contract role focuses on code generation workflows and model evaluation; prior AI experience is not required. Key Responsibilities...Hourly payContract workFor contractorsRemote work$60 - $80 per hour
...mathematics experts considered for future contract opportunities with AI labs and companies. This is an open application, not a posting... ...by applying mathematics knowledge to real-world tasks, model evaluation, and domain-specific feedback. Key Responsibilities...Hourly payContract workRemote work$110 per hour
...Apply to join a physician talent network supporting AI labs and companies with medical expertise. This is an open application... ...projects by bringing real-world clinical expertise to model development and evaluation. Key Responsibilities Train and evaluate AI models...Hourly payContract workRemote work$60 - $90 per hour
...engineering judgment to improve how advanced AI models reason through real-world engineering... ...correct solutions, and create rigorous evaluations grounded in industry practice. Key... ...Write clear instruction specifications, develop worked reference solutions, and create tasks...Hourly payFull timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to NKI Kernel Developer for AI Model Evaluation. Be the first to apply!



