Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Neuron Runtime Software Engineer

Full-time

Annapurna Labs (U.S.) Inc.

Salary: $143,400 - 165,600 per year Requirements:

  • We need 3+ years of professional software development experience outside of internships.
  • We need 2+ years of experience designing or architecting new and existing systems, including patterns for reliability and scaling.
  • We need hands-on programming experience in at least one software language.
  • We prefer 3+ years of full software development lifecycle experience, including coding standards, code reviews, source control, build processes, testing, and operations.
  • We prefer a bachelors degree in computer science or an equivalent background.
  • We value experience architecting, building, and operating distributed systems with strong availability and fault-tolerance characteristics.
  • We value production experience with AWS services such as EC2, ECS, CloudWatch, S3, and Lambda.
  • We value ownership of services end to end, including deployment, monitoring, alerting, on-call support, and post-incident review.
Responsibilities:
  • We develop and maintain high-performance runtime libraries and drivers for machine learning applications and AI accelerators.
  • We lead the design, development, and deployment of Neuron Runtime and related Neuron components.
  • We improve the profiler to help internal and external customers identify performance bottlenecks and better optimize AI workloads across Trainium and Inferentia hardware.
  • We enhance the performance of ML kernels and ML frameworks.
  • We manage the full development lifecycle of the Neuron Runtime with a focus on scalability, reliability, and usability.
  • We collaborate with cross-functional partners so our C++ compiler surfaces the right information for customers to understand and tune custom hardware performance.
  • We drive improvements that expand profiler support across multiple frameworks, including PyTorch, JAX, and XLA.
  • We partner with executive leadership and senior technical leaders to define product direction and deliver solutions to customers.
  • We contribute to massive-scale distributed training and inference systems across software, servers, and chips.
Technologies:
  • AI
  • AWS
  • Lambda
  • Cloud
  • CloudWatch
  • EC2
  • Hardware
  • Support
  • Machine Learning
  • PyTorch
  • DevOps
  • Flow

More:

We are AWS Neuron, the complete software stack for AWS Inferentia and Trainium cloud-scale machine learning accelerators and the servers that run them. Our team builds the full stack of software, servers, and chips to accelerate machine learning at the highest scale. We foster an inclusive culture with employee-led affinity groups, learning events, and strong support for work-life balance, mentorship, and career growth. We also offer comprehensive benefits, flexible working arrangements, and compensation that includes base salary, sign-on payments, restricted stock units, and additional benefits. The role is based in the USA, with listed locations in Cupertino, CA and Seattle, WA.

last updated 37 week of 2026

Vacancy posted more than 2 months ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Neuron Runtime Software Engineer. Be the first to apply!