Machine Learning - Compiler Engineer , AWS Neuron, Annapurna Labs
$165.2k - $223.6kAmazon Locker
Do you want to be part of AI revolution? At AWS our vision is to make deep learning pervasive for everyday developers and to democratize access to AI hardware and software infrastructure. In order to deliver on that vision, we’ve created innovative software and hardware solutions that make it possible. AWS Neuron is the SDK that optimizes the performance of complex ML models executed on AWS Inferentia and Trainium, our custom chips designed to accelerate deep-learning workloads.This role is for a software engineer in the Compiler team for AWS Neuron. As part of this role, you will be responsible for building next generation Neuron compiler which transforms ML models written in ML frameworks (e.g, PyTorch, TensorFlow, and JAX) to be deployed AWS Inferentia and Trainium based servers in the Amazon cloud. You will be responsible for solving hard compiler optimization problems to achieve optimum performance for variety of ML model families including massive scale large language models like Llama, Deepseek, and beyond as well as stable diffusion, vision transformers and multi-model models. You will be required to understand how these models work inside-out to make informed decisions on how to best coax the compiler to generate optimal implementation instruction. You will leverage your technical communications skill to partner with internal and external customers/stakeholders and will be involved in pre-silicon design, bringing new products/features to market, ultimately, making Neuron compiler highly performant and easy-to-use. Experience in object-oriented languages like C++/Java is a must, experience with compilers or building ML models using ML frameworks on accelerators (e.g., GPUs) is preferred but not required. Experience with technologies like OpenXLA, StableHLO, MLIR will be added bonus!Explore the product and our history! Utility Computing (UC) provides product innovations — from foundational services such as Amazon’s Simple Storage Service (S3) and Amazon Elastic Compute Cloud (EC2), to consistently released new product innovations that continue to set AWS’s services and features apart in the industry. As a member of the UC organization, you’ll support the development and management of Compute, Database, Storage, Internet of Things (Iot), Platform, and Productivity Apps services in AWS, including support for customers who require specialized security solutions for their cloud services.Key job responsibilitiesYou will design, implement, test, deploy and maintain innovative software solutions to transform Neuron compiler’s performance, stability and user-interface. You will work side by side with chip architects, runtime/OS engineers, scientists and ML Apps teams to seamlessly deploy state of the art ML models from our customers on AWS accelerators with optimal cost/performance benefits. You will have opportunity to work with open-source software (e.g., StableHLO, OpenXLA, MLIR) to pioneer optimizing advanced ML workloads on AWS software and hardware. You will also work on building innovative features that will deliver best possible experiences for our customers – developers across the globe.A day in the lifeAs you design and code solutions to help our team drive efficiencies in compiler architecture, you’ll create compiler optimization and verification passes, build features surface features and peculiarities of AWS accelerators to developers, implement tools to analyze numerical errors, and resolve the root cause of compiler defects. You’ll also participate in design discussions, code review, and communicate with internal (other Neuron SDK and Amazon wide teams) and external stakeholders (open-source communities). Lastly, work in a startup-like development environment, where you’re always working on the most important stuff.About the teamAbout the TeamOur team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we’re building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future.Diverse ExperiencesAWS values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying. About AWSAmazon Web Services (AWS) is the world’s most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating — that’s why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses.Inclusive Team CultureHere at AWS, it’s in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon conferences, inspire us to never stop embracing our uniqueness.Work/Life BalanceWe value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve in the cloud. Mentorship & Career GrowthWe’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.Basic qualifications- 3+ years of non-internship professional software development experience- 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience- Experience programming with at least one software programming languagePreferred qualification - Master's degree or PhD in Computer Science, or a related technical field.- 3+ years of experience writing production grade code in object-oriented languages such as C++/Java.- Experience in compiler design for CPU/GPU/Vector engines/ML-accelerators.- Experience with OpenSource compiler toolset like LLVM/MLIR.- Experience with the following technologies: PyTorch, OpenXLA, StableHLO, JAX, TVM, deep learning models, and algorithms.- Experience with modern build systems like Bazel/CMake.Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company’s reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at .USA, CA, Cupertino - 165,200.00 - 223,600.00 USD annually
$193.3k - $261.5k
The Product: AWS Machine Learning accelerators are at the forefront... ...stack, the AWS Neuron Software Development... ...includes an ML compiler, runtime and natively... ...a whole, the Amazon Annapurna Labs team is responsible... ...disciplines including silicon engineering, hardware design and...Amazon Web ServiceInternshipLocal areaWork from homeRelocationFlexible hours$193.3k - $261.5k
The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit... ...accelerate deep learning and GenAI... ...Amazon’s custom machine learning accelerators... ...software boundary, our engineers craft high-... ...includes an ML compiler, runtime, and application...Amazon Web ServiceInternshipLocal areaWork from homeFlexible hours$165.2k - $223.6k
The AWS Neuron Compiler team is actively seeking skilled compiler engineers to join our efforts in developing a... ...state-of-the-art deep learning compiler stack. This... ...our custom-built Machine Learning accelerators... ...-software co-design.Annapurna Labs (our organization within...Amazon Web ServiceInternshipLocal areaFlexible hours$165.6k
...prefer experience with compiler design for CPU, GPU, vector engines, or ML accelerators... ..., JAX, TVM, deep learning models, and... ...solutions that improve Neuron compiler performance... ...the-art ML models on AWS accelerators with... ...Java LLVM Machine Learning PyTorch...Amazon Web ServiceFull timeInternship$193.5k
...experience developing compiler features and optimizations... ..., PyTorch, or JAX deep learning models. We prefer... ...user experience of the Neuron compiler. We develop... ...JAX for deployment on AWS Inferentia and Trainium... ...architects, runtime and OS engineers, scientists, and ML...Amazon Web ServiceFull time$212.7k - $287.7k
The Product: AWS Machine Learning accelerators are at the forefront... ...enabled by the AWS Neuron Software Development... ...includes an ML compiler, the Neuron Kernel Interface... ...Team: The Amazon Annapurna Labs team is responsible... ...'s most talented engineers. Our team covers...Amazon Web ServiceLocal areaWork from homeRelocationFlexible hoursDay shift$165.6k
...performance compute kernels for machine learning operations using the Neuron architecture and... ...remove bottlenecks Apply compiler optimizations such as... ...learning models on AWS accelerators Partner... ...More: We are the Annapurna Labs team at Amazon Web Services...Amazon Web ServiceFull timeInternship$193.5k
...mentor, tech lead, or engineering team lead ~... ...using the Neuron architecture and... ...bottlenecks Develop compiler optimizations including... ...ML models on AWS accelerators... ...Hardware LLVM Machine Learning PyTorch TensorFlow... ...: We are the Annapurna Labs team at Amazon...Amazon Web ServiceFull timeInternshipFlexible hours$193.3k - $261.5k
Annapurna Labs is an integral part of AWS and develops hardware and software components... ...experience.The AWS Neuron Collectives team is... ...seeking a Software Engineer to optimize... ...annual and ongoing learning experiences, including... ...to revolutionize machine learning at Amazon...Amazon Web ServiceLocal areaWork from homeFlexible hours$93k - $128k
...least 2 years of engineering team management experience... ...understanding of compilers, including... ...Technologies: AWS Cloud... ...Hardware Support Machine Learning Flow PyTorch... ...We are Amazon Annapurna Labs, building innovative... ...cloud scale. Our AWS Neuron and Trainium...Amazon Web ServiceFull timeRelocation$193.3k - $261.5k
Annapurna Labs designs silicon and software that accelerates... ...change the world. AWS Neuron is the complete... ...3), our cloud scale Machine Learning accelerators and we... ...seeking a Senior Software Engineer to join our ML... ...with chip architects, compiler engineers, runtime engineers...Amazon Web ServiceInternshipLocal areaFlexible hours- ...Amazon Annapurna Labs is seeking a Sr. Software Engineer for AI/ML distributed training to design and optimize large-scale ML training on Trainium instances.... ...expertise. On-site in Cupertino, you will collaborate with AWS solution architects and customers to deploy #J-18808...Amazon Web Service
$212.7k - $287.7k
...need 2+ years of engineering team management experience... ...understanding of compilers, including... ...for custom AWS hardware. Translate... ...Hardware Machine Learning PyTorch Flow... ...More: We are AWS Annapurna Labs, building innovative... ...works on AWS Neuron and Trainium, enabling...Amazon Web ServiceFull timeRelocationFlexible hours$206.9k - $279.9k
AWS Neuron is looking for an experienced Technical... ...Interface (NKI), a compiler library enabling custom... ...innovation in machine learning acceleration software... ...contribute to and influence engineering discussions around... ....About Amazon Annapurna Labs:Amazon Annapurna Labs...Amazon Web ServiceFlexible hours$165.2k - $223.6k
The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit... ...accelerate deep learning and GenAI... ...Amazon’s custom machine learning accelerators... ...includes an ML compiler, runtime, and application... ...boundary, our engineers build systematic...Amazon Web ServiceWork experience placementInternshipLocal areaFlexible hours$176.6k - $239k
Annapurna Labs was a startup company acquired by AWS in 2015, and is now fully integrated... ...silicon engineering, hardware design... ...2 Instances, AWS Neuron, Inferentia and Trainium... ...The Product: AWS Machine Learning accelerators are... ...which includes an ML compiler, runtime and...Amazon Web ServiceLocal areaFlexible hours$175k - $236.8k
AWS Trainium is deployed at scale... ...models. AWS Neuron is the... ...to run deep learning and generative... ...partner with engineering teams building... ...responsible for compiler, runtime, NKI... ...About Amazon Annapurna LabsAmazon Annapurna Labs team (our organization... ...performance machine learning with...Amazon Web ServiceLocal areaFlexible hours$212.7k - $287.7k
The Product: AWS Machine Learning accelerators are at the forefront... ...stack, the AWS Neuron Software Development... ...includes an ML compiler, runtime and natively... ...The Team: The Amazon Annapurna Labs team is a responsible... ...world’s most talented engineers. Our team covers...Amazon Web ServiceLocal areaWork from homeRelocationFlexible hours$165.2k - $223.6k
...Solid grounding in machine learning and large... ...apply sound software engineering practices in large... ...PyTorch within the Neuron SDK Tune... ...and efficiency on AWS Trainium and Inferentia... ...across compiler, runtime, framework... ...More: We are the Annapurna Labs team at Amazon Web...Amazon Web ServiceFull timeInternship$208.3k - $281.8k
AWS Trainium is deployed at scale... ...models. AWS Neuron is the software... ...customers to run deep learning and generative... ...partner with engineering teams building... ...for Neuron compiler, runtime, and... ...About Amazon Annapurna Labs Amazon... ...high performance machine learning with...Amazon Web ServiceLocal areaFlexible hours$127.1k - $185k
Annapurna Labs was a startup acquired by AWS in 2015 and is now fully integrated.... ...org spans silicon engineering, hardware design and... ...2 Instances, AWS Neuron, Inferentia and... ...Trainium cloud-scale machine learning accelerators and... ...profiling.- Exposure to compiler toolchains, code...Amazon Web ServiceInternshipLocal areaFlexible hours$212.7k - $287.7k
...We develop AWS Neuron, the complete software stack for Trainium,... ...Amazon's custom cloud-scale machine learning accelerators. Join us to... ...lead a team of expert AI/ML engineers to onboard and optimize... ...inference library, Neuron compiler, runtime, and collectives....Amazon Web ServiceLocal areaFlexible hours$193.3k - $261.5k
...tech lead, or leading an engineering team ~ Masters... ...background ~ Experience with machine learning and large language... ...across our Neuron ecosystem Mentor team... ...across the stack with our compiler and runtime teams Coach... ...Technologies: AWS C# Cloud Java...Amazon Web ServiceFull timeInternship$151.3k - $261.5k
...Fundamental knowledge of machine learning and large language... ...best software engineering practices in large-scale... ...for PyTorch within the Neuron SDK. ~ I am responsible... ...on customer AWS Trainium and Inferentia... ...More: Our team at Annapurna Labs within Amazon Web Services...Amazon Web ServiceFull time$129.3k - $223.6k
...Fundamental knowledge of machine learning and large language... ...in software engineering for large-scale systems... ...operations, utilizing the Neuron architecture and programming... ...their ML models on AWS accelerators.... ...PyTorch More: Our Annapurna Labs team at Amazon Web Services...Amazon Web ServiceFull timeInternship$143.7k - $223.6k
...hands-on experience with AWS services in production... ...We work alongside engineering peers to develop and maintain... ...and drivers for machine learning applications and AI accelerators... ..., deploy, and evolve Neuron Runtime and related... ...teams so our C++ compiler produces the...Amazon Web ServiceFull timeInternshipFlexible hours$165.6k
...We look for knowledge of engineering practices across the... ...We use tools such as Neuron Explorer to pinpoint bottlenecks... ...Technologies: AI AWS EC2 Embedded... ...Network Cloud Flow Machine Learning More: We are Annapurna Labs, part of AWS, and we design...Amazon Web ServiceFull time$193.5k
...maximize training performance Use Neuron Explorer and similar tools to... ...solutions Technologies: AI AWS EC2 Firmware Hardware Support Cloud Flow Machine Learning More: We are Annapurna Labs, part of AWS, and we build the hardware...Amazon Web ServiceFull timeFlexible hours$143.4k - $165.6k
...value production experience with AWS services such as EC2, ECS,... ...libraries and drivers for machine learning applications and AI accelerators... ..., and deployment of Neuron Runtime and related Neuron components... ...functional partners so our C++ compiler surfaces the right...Amazon Web ServiceFull timeInternshipFlexible hours$193.3k - $261.5k
...We develop AWS Neuron, the complete software stack for Trainium, Amazon's custom cloud-scale machine learning accelerators. Join us to optimize the... ...Sr. Software Development Engineer on the Inference Model Enablement... ...collaboration with the compiler and runtime teams...Amazon Web ServiceInternshipLocal areaFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Machine Learning - Compiler Engineer , AWS Neuron, Annapurna Labs. Be the first to apply!
- machine learning engineer Cupertino, CA
- computer vision machine learning engineer Cupertino, CA
- senior ml engineer Cupertino, CA
- senior aws cloud engineer Cupertino, CA
- aws cloud architect Cupertino, CA
- aws developer Cupertino, CA
- junior aws engineer Cupertino, CA
- aws cloud Cupertino, CA
- aws cloud security engineer Cupertino, CA
- aws devops Cupertino, CA



