Neuron Runtime Software Engineer
Annapurna Labs (U.S.) Inc.
Salary: $143,400 - 165,600 per year Requirements:
- We need 3+ years of professional software development experience outside of internships.
- We need 2+ years of experience designing or architecting new and existing systems, including patterns for reliability and scaling.
- We need hands-on programming experience in at least one software language.
- We prefer 3+ years of full software development lifecycle experience, including coding standards, code reviews, source control, build processes, testing, and operations.
- We prefer a bachelors degree in computer science or an equivalent background.
- We value experience architecting, building, and operating distributed systems with strong availability and fault-tolerance characteristics.
- We value production experience with AWS services such as EC2, ECS, CloudWatch, S3, and Lambda.
- We value ownership of services end to end, including deployment, monitoring, alerting, on-call support, and post-incident review.
- We develop and maintain high-performance runtime libraries and drivers for machine learning applications and AI accelerators.
- We lead the design, development, and deployment of Neuron Runtime and related Neuron components.
- We improve the profiler to help internal and external customers identify performance bottlenecks and better optimize AI workloads across Trainium and Inferentia hardware.
- We enhance the performance of ML kernels and ML frameworks.
- We manage the full development lifecycle of the Neuron Runtime with a focus on scalability, reliability, and usability.
- We collaborate with cross-functional partners so our C++ compiler surfaces the right information for customers to understand and tune custom hardware performance.
- We drive improvements that expand profiler support across multiple frameworks, including PyTorch, JAX, and XLA.
- We partner with executive leadership and senior technical leaders to define product direction and deliver solutions to customers.
- We contribute to massive-scale distributed training and inference systems across software, servers, and chips.
- AI
- AWS
- Lambda
- Cloud
- CloudWatch
- EC2
- Hardware
- Support
- Machine Learning
- PyTorch
- DevOps
- Flow
More:
We are AWS Neuron, the complete software stack for AWS Inferentia and Trainium cloud-scale machine learning accelerators and the servers that run them. Our team builds the full stack of software, servers, and chips to accelerate machine learning at the highest scale. We foster an inclusive culture with employee-led affinity groups, learning events, and strong support for work-life balance, mentorship, and career growth. We also offer comprehensive benefits, flexible working arrangements, and compensation that includes base salary, sign-on payments, restricted stock units, and additional benefits. The role is based in the USA, with listed locations in Cupertino, CA and Seattle, WA.
last updated 37 week of 2026
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Neuron Runtime Software Engineer. Be the first to apply!
- software engineer - web development Cupertino, CA
- software developer positions Cupertino, CA
- senior software engineer remote Cupertino, CA
- software engineer contract Cupertino, CA
- cybersecurity software engineer Cupertino, CA
- part time software developer remote Cupertino, CA
- junior software developer internship Cupertino, CA
- software engineer remote Cupertino, CA
- software engineer Cupertino, CA
- software engineer amazon Cupertino, CA
