Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Development Engineer - AI/ML Networking Disaggregated Inference, Annapurna Labs , Elastic Collectives

$165.2k - $223.6k
Full-time

Amazon

Every token a large language model generates depends on data reaching the right accelerator at the right moment. As AI models outgrow any single chip, the network between accelerators becomes the bottleneck that decides how fast — and how affordably — the world's largest models can serve real users. That network layer is what our team builds.We're looking for an engineer to work at the frontier of disaggregated inference: splitting LLM serving into separate prefill and decode pools and moving the model's KV cache between them at the limit of what the hardware allows. Get it right and users get answers in milliseconds; get it wrong and the fastest accelerators in the world sit idle waiting on data. You'll build components of the high-speed transfer path that make that difference, and you'll learn to measure success in how close we run to the theoretical peak of the machine.In this role you will:Build and optimize the low-level data-movement software that transfers KV cache and activations across accelerators, servers, and heterogeneous memory — over AWS's highest-performance network fabric.Profile real workloads, find the true bottleneck, and close the gap between "it works" and "it runs fast" — pushing components toward the hardware's limit.Work across the stack — from network transport up to the inference frameworks — learning from the teams building the chips, runtime, and models.Deliver features that ship to our largest clusters, for our largest customers, serving the largest AI models in production.What we're looking for:Strong C/C++ and a genuine interest in low-level, performance-critical systems — solid command of Linux, memory, and writing fast code.The instinct to ask "how fast could this go?" and the discipline to measure it.Exposure to high-speed networking, HPC interconnects, or GPU/accelerator systems (RDMA, InfiniBand, libfabric, UCX, NCCL, MPI) is a strong plus; embedded-systems experience is welcome.Prior AI/ML experience is not required — if you're a strong systems engineer eager to learn, we'll teach you the ML side.If you like solving genuinely hard problems, working alongside HPC and ML customers, iterating fast, and shipping at a scale few places can offer, come join us. You'll work alongside senior engineers and Principal Engineers who've built this layer from the ground up, with real room to grow your scope and technical depth — on a team at the leading edge of AI/ML infrastructure.About the team: You'd be joining Annapurna Labs, an integral part of AWS. Annapurna designs the hardware and software building blocks behind EC2 — every EC2 instance runs on hardware we designed. We specialize in the chips, systems, and software that optimize the AWS customer experience, and this team sits where AI meets the silicon and the network underneath it.A day in the lifeAnnapurna Labs, a crucial part of AWS, is responsible for developing hardware and software components for EC2 infrastructure. Our team focuses on building networking solutions that for Machine Learning (ML) and High-Performance Computing (HPC) workloads on AWS.We have mixed discipline orgs, you’d be working side by side with infrastructure experts, hardware engineers, RTL engineers, scientists & architects. Our workforce spans the globe and is truly international, you’ll find yourself working side by side with individuals from numerous countries. We take mentorship seriously, you can both expect senior mentorship and will be expected to mentor new and junior engineers. The pace is fast as we work on the latest advancements of AI/ML, but we take the time to bond as a team and enjoy the successes. We offer flexibility in working hours, and respect WLB as a core org tenet. The team enjoys working with numerous principal-level engineers and closely with directors, career growth opportunities are certainly available. This is a role where you will always be encouraged to keep learning, the AI/ML field is fast moving and constantly evolving.About the teamOur team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we’re building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future.Diverse ExperiencesAWS values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying.About AWSAmazon Web Services (AWS) is the world’s most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating — that’s why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses.Inclusive Team CultureHere at AWS, it’s in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (diversity) conferences, inspire us to never stop embracing our uniqueness.Work/Life BalanceWe value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve in the cloud.Mentorship & Career GrowthWe’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.Basic qualifications- 3+ years of non-internship professional software development experience- 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience- Experience programming with at least one software programming language- Experience with C/C++Preferred qualification - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience- Bachelor's degree in computer science or equivalentAmazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company’s reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at .USA, CA, Cupertino - 165,200.00 - 223,600.00 USD annually

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Software Development Engineer - AI/ML Networking Disaggregated Inference, Annapurna Labs , Elastic Collectives in Cupertino, CA vacancy
  • $165.2k - $223.6k

     ...an experienced engineer to work on distributed AI/ML systems. This...  ...involves working on collective operations -...  ...high-speed networking or HPC...  ...be joining is Annapurna Labs, an integral part...  ...hardware and software components that...  ...professional software development experience- 2+... 
    Network
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  • $193.3k - $261.5k

     ...moment. As AI models outgrow...  ...chip, the network between...  ...looking for an engineer to work at...  ...of disaggregated inference: splitting...  ...data-movement software across accelerators...  ....Prior AI/ML experience...  ...be joining Annapurna Labs, an...  ...professional software development experience-... 
    Network
    Full time
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $193.3k - $261.5k

     ...an experienced engineer to work on distributed AI/ML systems. This...  ...involves working on collective operations -...  ...high-speed networking or HPC...  ...be joining is Annapurna Labs, an integral part...  ...hardware and software components that...  ...professional software development experience- 5+... 
    Network
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $193.3k - $261.5k

    The Annapurna Labs team at Amazon builds Amazon...  ...Neuron, the software development kit used to...  ...and Trainium ML accelerators....  ...unparalleled ML inference and training performance...  ...boundary, our engineers build...  ...s possible in AI acceleration.As...  ...runtime and collectives. We not only optimize... 
    Suggested
    Work experience placement
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  • $165.2k - $223.6k

    Annapurna Labs is an integral part of AWS and develops...  ...hardware and software components that...  ...experience.The AWS Neuron Collectives team is seeking a Software Engineer to optimize...  ...the frontier AI models being trained...  .../hardware/networks development life cycle, including... 
    Network
    Local area
    Work from home
    Flexible hours

    Amazon

    Cupertino, CA
    4 days ago
  • $127.1k - $185k

     ...talented early-career engineer to join our team that owns the network stack for EC2 distributed AI/ML systems. You'll work on software that enables the world...  ...Familiarity with Linux development environments and...  ...programming (sockets, MPI, collective communication patterns... 
    Network
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  • $165.2k - $223.6k

     ...professional software development experience....  ...high-speed networking, HPC interconnects...  .... Prior AI/ML experience...  ...systems engineering background and...  ...transport through inference frameworks,...  ...We focus on disaggregated inference,...  ...: We are Annapurna Labs, an integral... 
    Network
    Full time
    Internship

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    5 days ago
  • $165.2k - $223.6k

     ...professional software development experience...  ...training and inference lifecycles, plus...  ...sound software engineering practices in...  ...on custom ML accelerators...  ...Technologies: AI AWS...  ...We are the Annapurna Labs team at Amazon...  ...runtime, and collectives, to deliver high... 
    Full time
    Internship

    Annapurna Labs Inc.

    Cupertino, CA
    5 days ago
  • $165.2k - $223.6k

     ...building complex software systems that...  ...need knowledge of engineering practices and patterns...  ..., hardware, and networks development life cycle,...  ...Familiarity with collective communication...  ...operations to scale AI compute across...  ...More: We are Annapurna Labs, part of AWS, and... 
    Network
    Full time
    Flexible hours

    Annapurna Labs Inc.

    Cupertino, CA
    5 days ago
  • $165.2k - $223.6k

    As a Neuron Collectives Software Developer, you will...  ...to scale AI compute across...  ...device driver development* Work closely...  ...lifeAnnapurna Labs, a crucial part...  ...focuses on building networking solutions that...  ...Learning (ML) and High-Performance...  ..., hardware engineers, RTL engineers... 
    Network
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  • $193.3k - $261.5k

     ...We develop AWS Neuron, the complete software stack for Trainium, Amazon's custom cloud-scale machine learning accelerators...  ...fast on the Trainium hardware. As a Sr. Software Development Engineer on the Inference Model Enablement team, you will onboard and optimize state... 
    Internship
    Local area
    Flexible hours

    Annapurna Labs (U.S.)

    Cupertino, CA
    10 days ago
  •  ...'s largest AI chip, 56 times...  ...and inference speeds; over...  ...leading model labs, global enterprises...  ...of disaggregated AI inference...  ...-Scale Engine.We are hiring a Software Engineer to...  ...GPU nodes, networking, and rack-scale...  ...ROCm/HIP, collective communication...  ...-source ML systems project... 
    Network

    Cerebras Systems

    Sunnyvale, CA
    3 days ago
  • $129.3k - $223.6k

     ...of professional software development experience (non-internship...  ..., training, and inference processes, along...  ...in software engineering for large-scale...  ...deployment on custom ML hardware...  ...Technologies: AI AWS Hardware...  ...More: Our Annapurna Labs team at Amazon Web... 
    Full time
    Internship

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    6 days ago
  •  ...potential of generative AI to power the...  ...the forefront of software and hardware...  ...System Software Engineer, AI Inference ExecutionWhat you...  ...responsible for the development, enhancement, and...  ...other software (ML and compilers) and...  ...distributed systems collectives such as NCCL and... 
    3 days per week

    d-Matrix

    Santa Clara, CA
    3 days ago
  • $165.2k - $223.6k

     ...internship professional software development experience....  ...We will enhance collective algorithms and...  ...to scale AI compute across the...  ...More: We are Annapurna Labs, part of AWS, and...  ...Our team develops networking solutions for machine...  ..., hardware engineers, RTL engineers,... 
    Network
    Full time
    Internship

    Annapurna Labs Inc.

    Cupertino, CA
    5 days ago
  • $165.6k

     ...internship professional software development experience....  ...: We enhance collective algorithms and...  ...to scale AI compute across the...  ...More: We are Annapurna Labs, part of AWS, and...  ...Our team builds networking solutions for machine...  ..., hardware engineers, RTL engineers,... 
    Network
    Full time
    Internship

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    6 days ago
  • $193.3k - $261.5k

     ...seeking an experienced engineer and technical...  ...team that owns the network stack for EC2 distributed AI/ML systems. The team...  ...would be joining is Annapurna Labs, an integral part...  ...hardware and software components that are...  ...of full software development life cycle, including... 
    Network
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $206.9k - $279.9k

     ...custom kernel development and...  ...acceleration software. AWS Neuron...  ...-in-class ML...  ...influence engineering discussions...  ...-class ML inference performance...  ...About Amazon Annapurna Labs:Amazon Annapurna...  ...chips, in networking and...  ...ENA), and Elastic Fabric Adapter...  ...generative AI services and... 
    Network
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  • $175k - $236.8k

     ...training and inference of...  ...Neuron is the software stack for...  ...AI workloads...  ...training AI/ML ecosystem...  ...partner with engineering teams building...  ...Amazon Annapurna...  ...Annapurna Labs team (our...  ...chips, in networking and security...  ...ENA), and Elastic Fabric Adapter...  ...software development, network... 
    Network
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  • $193.3k - $261.5k

     ...training and inference clusters....  ...the low-level software stack that brings...  ..., and collective communication...  ...together across a network.We're...  ...Systems Software Engineer who wants to...  ...software development, and enable...  ...part of the ML accelerator...  ...teammatesAnnapurna Labs, our... 
    Network
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    11 hours ago
  • $151.3k - $261.5k

     ...professional experience in software development ~5+ years of...  ...architecture, training, and inference lifecycles, with hands...  ...practices in software engineering in large-scale systems...  ...deployment on custom ML hardware accelerators...  ...Technologies: AI AWS Hardware... 
    Full time
    Internship

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    6 days ago
  • $127.1k - $185k

    Annapurna Labs was a startup acquired by AWS in 201...  ...org spans silicon engineering, hardware design and verification, software, and operations. We...  ...and Trainium ML Accelerators, and...  ...looking for a Software Development Engineer to help build...  ...on custom AI accelerators. You'... 
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $193.3k - $261.5k

     ...seeking an experienced engineer and technical...  ...team that owns the network stack for EC2 distributed AI/ML systems. The team...  ...would be joining is Annapurna Labs, an integral part...  ...hardware and software components that are...  ...of full software development life cycle, including... 
    Network
    Full time
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $184k - $287.5k

     ...skilled and motivated software engineers to join us and build AI inference systems that serve large...  ...parallelism, prefill-decode disaggregation.Develop, optimize, and...  ...for the field of ML Systems; survey recent...  ...advance AI research and development to create groundbreaking... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...performance engineering background who...  ...Physical AI workloads using...  ...on the development and test of...  ...hardware and software. If you are...  ...NVLINK) and network protocols (InfiniBand...  ...validated ML/DL...  ...interconnects, collective communication...  ...training and inference.Effective verbal... 
    Network
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    3 days ago
  • $165.6k

     ...has built complex software systems that...  ...for knowledge of engineering practices across...  ..., hardware, and network development lifecycle, including...  ...with collective communication algorithms...  ...collective operations so AI compute can...  ...More: We are Annapurna Labs, part of AWS,... 
    Network
    Full time

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    6 days ago
  • $200k - $420k

     ...At River AI, our mission is to...  ...hardware for local inference, bespoke...  ...are scientists, engineers, and builders...  ...companies and AI labs. We bring a...  ...concurrency, and networking....  ...speculative decoding or disaggregated prefill and decode...  ..., GPU collectives, and quantized... 
    Network
    Full time
    Local area
    Visa sponsorship
    Relocation package

    River AI Inc.

    Palo Alto, CA
    4 days ago
  • $92k - $135k

     ...Essential Cloud for AI™. Built for...  ...Trusted by leading AI labs, startups, and global...  ...You'll Do: Join the Inference team to ship production...  ...from experienced engineers. About the role:...  ..., algorithms, and networked services....  ...a microservice or ML inference demo.... 
    Network
    Permanent employment
    Full time
    Temporary work
    Casual work
    Internship
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    27 days ago
  •  ...of physical AI. Founded in 2...  ...performance engineer who specializes...  ...throughput batch inference sweeping...  ...accelerators, ML frameworks, and...  ...users to collect feedback, and...  ...storage and network I/O, scheduling...  ...GPU kernel development experience: CUDA...  ...tolerance and elastic training for... 
    Network
    Full time
    For contractors
    For subcontractor
    Casual work
    Work at office
    Remote work
    Day shift

    Decisive Point

    Sunnyvale, CA
    4 days ago
  • $151.3k - $261.5k

     ...of professional software development experience. ~ At...  ...architecture, training, and inference lifecycle,...  ...best software engineering practices in...  ...deployment on specialized ML hardware...  ...Technologies: AI AWS Hardware...  ...: Our team at Annapurna Labs within Amazon Web... 
    Full time

    Annapurna Labs Inc.

    Cupertino, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Development Engineer - AI/ML Networking Disaggregated Inference, Annapurna Labs , Elastic Collectives. Be the first to apply!