Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

$165.2k - $223.6k

Annapurna Labs (U.S.) Inc.

The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon's custom machine learning accelerators, Inferentia and Trainium.

The Acceleration Kernel Library team is at the forefront of maximizing performance for AWS's custom ML accelerators. Working at the hardware-software boundary, our engineers craft high-performance kernels for ML functions, ensuring every FLOP counts in delivering optimal performance for our customers' demanding workloads. We combine deep hardware knowledge with ML expertise to push the boundaries of what's possible in AI acceleration.

The AWS Neuron SDK, developed by the Annapurna Labs team at AWS, is the backbone for accelerating deep learning and GenAI workloads on Amazon's Inferentia and Trainium ML accelerators. This comprehensive toolkit includes an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch, enabling unparalleled ML inference and training performance.

As part of the broader Neuron Compiler organization, our team works across multiple technology layers - from frameworks and compilers to runtime and collectives. We not only optimize current performance but also contribute to future architecture designs, working closely with customers to enable their models and ensure optimal performance. This role offers a unique opportunity to work at the intersection of machine learning, high-performance computing, and distributed architectures, where you'll help shape the future of AI acceleration technology

This is an opportunity to work on cutting-edge products at the intersection of machine-learning, high-performance computing, and distributed architectures. You will architect and implement business-critical features, publish cutting-edge research, and mentor a brilliant team of experienced engineers. We operate in spaces that are very large, yet our teams remain small and agile. There is no blueprint. We're inventing. We're experimenting. It is a very unique learning culture. The team works closely with customers on their model enablement, providing direct support and optimization expertise to ensure their machine learning workloads achieve optimal performance on AWS ML accelerators.

Explore the product and our history!

Key job responsibilities

Our kernel engineers collaborate across compiler, runtime, framework, and hardware teams to optimize machine learning workloads for our global customer base. Working at the intersection of software, hardware, and machine learning systems, you'll bring expertise in low-level optimization, system architecture, and ML model acceleration. In this role, you will:

  • Design and implement high-performance compute kernels for ML operations, leveraging the Neuron architecture and programming models
  • Analyze and optimize kernel-level performance across multiple generations of Neuron hardware
  • Conduct detailed performance analysis using profiling tools to identify and resolve bottlenecks
  • Implement compiler optimizations such as fusion, sharding, tiling, and scheduling
  • Work directly with customers to enable and optimize their ML models on AWS accelerators
  • Collaborate across teams to develop innovative kernel optimization techniques

A day in the life

As you design and code solutions to help our team drive efficiencies in software architecture, you'll create metrics, implement automation and other improvements, and resolve the root cause of software defects. You'll also:

Build high-impact solutions to deliver to our large customer base.

Participate in design discussions, code review, and communicate with internal and external stakeholders.

Work cross-functionally to help drive business decisions with your technical input.

Work in a startup-like development environment, where you're always working on the most important stuff.

About the team

Amazon Diverse Experiences

values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying.

Inclusive Team Culture

Here at Amazon, we embrace our differences. We are committed to furthering our culture of inclusion. We have ten employee-led affinity groups, reaching 40,000 employees in over 190 chapters globally. We have innovative benefit offerings, and host annual and ongoing learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (gender diversity) conferences. Amazon's culture of inclusion is reinforced within our 16 Leadership Principles, which remind team members to seek diverse perspectives, learn and be curious, and earn trust.

Work/Life Balance

Our team puts a high value on work-life balance. It isn't about how many hours you spend at home or at work; it's about the flow you establish that brings energy to both parts of your life. We believe striking the right balance between your personal and professional life is critical to life-long happiness and fulfillment. We offer flexibility in working hours and encourage you to find your own balance between your work and personal lives.

BASIC QUALIFICATIONS

  • 3+ years of non-internship professional software development experience
  • 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience
  • Experience programming with at least one software programming language

PREFERRED QUALIFICATIONS

  • 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience
  • Bachelor's degree in computer science or equivalent

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at

USA, CA, Cupertino - 165,200.00 - 223,600.00 USD annually

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs in Cupertino, CA vacancy
  • $183k - $247.6k

    AWS Utility Computing (UC) provides product innovations...  ...their cloud services.Annapurna Labs (our organization...  ...a Hardware Design Engineer with role in the definition...  ...AWS next generation ML Chips, Cards and server...  ...improve your products performance, quality and cost. We’... 
    Amazon Web Service
    Performance
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  • $165.2k - $223.6k

     ...Annapurna Labs is an integral part of AWS and develops hardware and software components that...  ...customer experience. The AWS Neuron Collectives team is seeking a Software Engineer to optimize collective...  ..., you'll push for maximum performance using C/C++, interfacing with... 
    Amazon Web Service
    Performance
    Local area
    Work from home
    Flexible hours

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    1 day ago
  • $165.2k - $223.6k

     ...The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit...  ...Inferentia and Trainium ML accelerators....  ...and training performance. The Inference...  ...software boundary, our engineers build systematic...  ...high-performance kernels for ML functions,... 
    Amazon Web Service
    Performance
    Work experience placement
    Internship
    Local area
    Flexible hours

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    1 day ago
  • $165.2k - $223.6k

     ...software and hardware solutions that make it possible. AWS Neuron is the SDK that optimizes the performance of complex ML models executed on AWS Inferentia and Trainium,...  ...workloads. This role is for a software engineer in the Compiler team for AWS Neuron. As part of... 
    Amazon Web Service
    Performance
    Internship
    Local area
    Flexible hours

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    1 day ago
  • $157.3k - $212.8k

     ...centers including technologies such as AWS Inferentia which is a machine learning inference...  ...product designed to deliver high performance at low cost. You'll provide leadership...  ...Bachelor's degree in Electrical Engineering or a related field Block Design using... 
    Amazon Web Service
    Performance
    Local area
    Flexible hours

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    7 days ago
  • $165.2k - $223.6k

     ...As a Neuron Collectives Software Developer, you...  ...for optimal training performance Use tools like...  ...day in the life Annapurna Labs, a crucial part of AWS, is responsible for...  ...for Machine Learning (ML) and High-Performance...  ...infrastructure experts, hardware engineers, RTL engineers,... 
    Amazon Web Service
    Performance
    Internship
    Local area
    Flexible hours

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    1 day ago
  • $165.2k - $223.6k

     ...AWS Neuron is the complete software stack for the AWS Inferentia and...  ...As the Software Development Engineer for the Neuron Runtime Team,...  ...develop and maintain high-performance runtime libraries and drivers...  ...behavior. Improving performance of ML Kernels and ML Frameworks. In... 
    Amazon Web Service
    Performance
    Internship
    Local area
    Work from home
    Flexible hours

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    1 day ago
  • $157.3k - $212.8k

     ...around the world.We are seeking an experienced Design Verification Engineers to build the next generation of our cloud server platforms....  ...bugs in architecture, algorithms, functionality, and performance with strong overall debugging skills- Experience verifying at... 
    Amazon Web Service
    Performance
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    12 hours ago
  • $213k - $287.5k

    Annapurna Labs designs silicon and software that accelerates innovation...  ...Chip) live at the heart of AWS Machine Learning servers. As...  ...for an ASIC Physical Design Engineer to help us trail-blaze new technologies...  ...studies, explore power-performance-area tradeoffs for physical... 
    Amazon Web Service
    Performance
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    12 hours ago
  • $165.2k - $223.6k

     ...AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale...  ...them. This position is for a Software Engineer that will lead the development of various...  ...monitoring solutions to track system performance. Identify bottlenecks and optimize system... 
    Amazon Web Service
    Performance
    Work experience placement
    Internship
    Local area
    Flexible hours

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    1 day ago
  • $157.3k - $212.8k

     ...AWS Utility Computing (UC) provides product innovations — from foundational services...  ...inference product designed to deliver high performance at low cost. You'll provide...  ...and teamwork with other physical design engineers as well as with the RTL/Arch. teams BASIC... 
    Amazon Web Service
    Performance
    Full time
    Local area
    Flexible hours

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    7 days ago
  • $183k - $247.6k

     ...deadline: Sep 27, 2026AWS Compute & ML Services owns the design,...  ..., and operation of all AWS global infrastructure. In other...  ...software, hardware, and network engineers, supply chain specialists, security...  ...are industry-leading in performance, frugality and operational excellence... 
    Amazon Web Service
    Performance
    Local area
    Flexible hours

    AmazonWebServices

    Cupertino, CA
    1 day ago
  • $180k - $220k

     ...are looking for a Senior DevOps Engineer to own the build, deployment, and...  ...systems that push software and ML models to a production fleet of...  ...including networking, storage, performance, and security. ~ Practical experience with AWS and container technologies including... 
    Amazon Web Service
    Performance
    Full time

    Knightscope

    Sunnyvale, CA
    a month ago
  •  ...collaborative team of researchers and engineers who thrive on pushing the...  .... Collaborate closely with ML and product teams to integrate...  ..., Java, C++) or Python in a performance-sensitive context. ~...  ...stack: cloud infrastructure (AWS/GCP), containerized deployments... 
    Amazon Web Service
    Performance

    Boson AI

    Santa Clara, CA
    8 days ago
  • $248k - $396.75k

    Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline...  ...their reliability and performance objectives while enabling developers...  ...Security, Networking, and AI/ML organizations to make high-...  ...cloud platforms such as AWS, Azure, or GCP.Strong proficiency... 
    Amazon Web Service
    Performance
    Full time

    NVIDIA

    Santa Clara, CA
    2 days ago
  • $193.3k - $261.5k

    The AWS BIOS Engineering team creates and maintains custom firmware solutions for Amazon's innovative...  ...are industry-leading in security, performance, and operational excellence, and are critical...  ...computingAccess to advanced hardware labs and debug equipmentOpportunity to... 
    Amazon Web Service
    Performance
    Internship
    Local area
    Worldwide
    Flexible hours

    Amazon

    Cupertino, CA
    4 days ago
  • $183k - $247.6k

     ...breakthrough innovation in AI/ML and HPC workloads. If you’re passionate...  ...about pushing the limits of performance, efficiency, and scalability...  ...that define what’s next for AWS — and for the entire AI...  ...software, hardware, and network engineers, supply chain specialists, security... 
    Amazon Web Service
    Performance
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    4 days ago
  • $185.5k - $265k

     ...Principal System and Solution Test Engineer to join our team. This is a...  ...understanding of AI/ML technologies and experience leveraging...  ..., TCP/IP stack internals, and performance testing methodologies using industry...  ...Public Cloud networking in AWS or Azure What Will Make... 
    Amazon Web Service
    Performance
    Full time
    Work at office
    Local area

    Zscaler

    Santa Clara, CA
    3 days ago
  •  ...are a dedicated research lab for building,...  ..., data scientists, and engineers, tackling the most fundamental...  ...a global hub for high-performance computing in deep learning...  ...Engineer focused on ML infrastructure and MLOps...  ...scalable ML infrastructure on AWS (e.g., compute, storage... 
    Amazon Web Service
    Performance
    Visa sponsorship

    Institute of Foundation Models

    Sunnyvale, CA
    a month ago
  • $215k - $250k

    Onehouse Data Infrastructure Engineer Onehouse is a mission-driven company...  ...analytics to real-time AI / ML). We are a team of self-driven...  ...including Uber, Snowflake, AWS, Linkedin, Confluent and many...  ...processing and analytical query performance. Design systems that help... 
    Amazon Web Service
    Performance
    Odd job
    Work at office
    Local area
    Remote work
    Relocation
    Relocation package

    OneHouse LLC

    Sunnyvale, CA
    4 days ago
  •  ...Site Reliability Engineer Onsite- Bay Area, CA Skills Relevant...  ...cloud infrastructure (GCP or AWS, on-prem). Build, maintain,...  ...Monitor system health and performance using Grafana and other observability...  ...with scalable GPU infrastructure for AI/ML #J-18808-Ljbffr
    Amazon Web Service
    Performance

    Amiri Recruiting

    Mountain View, CA
    4 days ago
  • $2,500 per month

     ...investors and staffed by leading engineers, Etched is redefining the...  ...software engineering, and high-performance computing (HPC), building systems...  ...and manage cloud resources (AWS, GCP) for scaling compute, storage...  ...Data pipelining for AI/ML workflows using Airflow, Prefect... 
    Amazon Web Service
    Performance
    Work at office
    Relocation package

    Etched

    San Jose, CA
    13 days ago
  • $165.2k - $223.6k

     ...AWS Networking builds the infrastructure that powers every customer workload in the...  ...to deliver industry-leading networking performance: dramatically lower tail latency, higher...  ...are looking for a Software Development Engineer to join the ENA Express team. You will design... 
    Amazon Web Service
    Performance
    Internship
    Local area
    Flexible hours

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    1 day ago
  •  ...Principal Machine Learning Systems Engineer (P60) to lead technical...  ...algorithms with reliable, high-performance infrastructure.Working at AtlassianAtlassians...  ...You'll Doð Design and Build ML SystemsArchitect and implement...  ...with cloud environments (AWS, GCP, Azure) and container/... 
    Amazon Web Service
    Performance
    Work at office
    Local area

    Atlassian

    Mountain View, CA
    1 day ago
  •  ...Shift Wand is building a high-performing global team who take full...  ...experienced Senior Staff SRE Engineer to act as a senior technical authority...  ...platform, product, data, and ML teams, helping us...  ...secure production environments (AWS preferred). Lead reliability... 
    Amazon Web Service
    Performance
    Shift work

    Wand AI

    Palo Alto, CA
    4 days ago
  • $212.7k - $287.7k

     ...network, and automation systems that keep AWS's global network infrastructure running....  ...means the tooling and platforms network engineers and automated systems use to deploy,...  ...roadmaps, and dependencies- Drive hiring, performance management, and career growth for a team... 
    Amazon Web Service
    Performance
    Local area
    Flexible hours
    Shift work
    Day shift

    Amazon

    Santa Clara, CA
    12 hours ago
  • $171.6k - $222.2k

     ...Description The AWS Neuron Science Team is looking for...  ...collaborate closely with our engineering teams to implement...  ...state-of-the-art ML systems. As part of a...  ...ML/RL approaches for kernel/code generation and optimization...  ...: Designing high-performance kernels optimized for... 
    Amazon Web Service
    Performance
    Local area
    Flexible hours

    Amazon

    Santa Clara, CA
    5 days ago
  • $180k - $270k

     ...technology - hyperscalers, AI labs, the AI hardware supply...  ...As a Machine Learning Engineer on the Digital...  ...inference. Data & ML Infrastructure: Design...  ...Build monitoring for model performance, data drift, and...  ...cloud-native environment (AWS, GCP, or Azure).Working... 
    Amazon Web Service
    Performance
    Full time
    Work at office
    Flexible hours
    Shift work

    Everpure

    Santa Clara, CA
    3 days ago
  •  ...cutting-edge Machine Learning (ML) and Artificial Intelligence (...  ...a skilled Site Reliability Engineer to join our growing team. The...  ...reliability, scalability, and performance of our hybrid-based (Cloud & On...  ...optimizing cloud infrastructure (AWS preferred, or GCP, Azure),... 
    Amazon Web Service
    Performance
    Work at office
    Weekend work

    FLUIX

    Palo Alto, CA
    4 days ago
  • $170k - $230k

     ...Site Reliability Engineer (SRE) Palo Alto / San Francisco...  ...to the stability and performance of Mithril's global GPU...  ...one major provider (AWS, GCP, or Azure), including...  ...workloads or AI/ML infrastructure. Exposure...  ....g., CoreWeave, Lambda Labs, Nebius). Familiarity... 
    Amazon Web Service
    Performance
    Work at office
    Local area
    1 day per week

    Mithril

    Palo Alto, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs. Be the first to apply!