Sr Cloud Hardware Dev Engineer, AWS Generative AI & ML Servers
$183k - $247.6kAmazon Locker
AWS operates the world's largest fleet of GPU-accelerated servers powering AI/ML training and inference at cloud scale. Our team defines the server architectures, drives the hardware designs, and owns the fleet quality for these platforms — from component selection through datacenter operations. If you want to shape the physical hardware that frontier models train on, this is the role.We are seeking a Cloud Hardware Development Engineer to define server architectures based on workload demand, translate them into detailed component specifications, and drive validation from PCBA bring-up through rack integration. You will lead ODM design partners through development and production, triage hardware issues across manufacturing and datacenters, and own fleet quality metrics post-launch.What You Will DoYou will define the hardware that runs the world's largest AI training workloads. Your designs span thermal, mechanical, power, and signal integrity across GPU-accelerated platforms. You will drive validation from first silicon through fleet-scale deployment, triage failures correlating across PCIe, power delivery, memory, and accelerator interconnects, and feed root cause findings back into design improvements. When a new server platform launches at a large scale, the architecture, component choices, and quality gates are yours.Why You Will Love ItThe world's most advanced frontier models train on the hardware you design. You will see your architecture decisions scale to a large fleet of servers. The team is deeply technical and high-trust — you own platforms end to end from architecture definition through fleet operations.The Ideal CandidateYou think across the full hardware stack — from silicon packaging and power delivery to rack-level thermal and mechanical design. You are as comfortable reviewing a schematic as you are analyzing fleet failure data. You drive quality through data, not assumption, and you hold design partners to the same standard you hold yourself. You mentor and develop junior engineers, contribute to hiring, and share your expertise to make the team stronger.Key job responsibilitiesArchitecture & Design* Define server architectures based on workload demand and customer requirements, translating them into detailed designs and component specifications that enable high-performance AI training and inference at scale* Work with interdisciplinary teams of component, firmware, test, qualification, and integration engineers to deliver cohesive designs* Drive design reviews with ODM/JDM partners covering schematic, layout, BOM, and manufacturing DFx (Design for Test, Design for Manufacturing)Validation & Bring-up* Define and execute validation strategies from PCBA bring-up through server and rack integration — covering power sequencing, signal integrity, thermal characterization, and accelerator interconnect performance* Own hardware debug during EVT/DVT/PVT builds, correlating failures across PCIe, power rails, memory channels, and GPU subsystems* Triage hardware issues at both ODM facilities and datacenters, conduct root cause analysis, and implement corrective actionsFleet Quality & Continuous Improvement* Own fleet quality metrics post-launch: server-level annualized failure rates and component-level failure modes* Monitor operational telemetry to identify systemic issues and drive design or process changes for current and future platforms* Partner with test and automation teams to improve manufacturing yield and reduce test dwell timesCross-Team Collaboration* Work with EC2 architecture teams to align on instance definitions, workload requirements, and platform trade-offs* Drive ODM/JDM design partners through development milestones and production ramp* Collaborate with firmware, software, and operations teams to ensure designs are debuggable, serviceable, and automation-readyMay require occasional (<10%) regional and international travel to Design and Manufacturing Partner sites.A day in the lifeYou start the day reviewing thermal and power validation data from an EVT build at your ODM partner. Mid-morning, you join a design review to close signal integrity findings on a high-speed accelerator interconnect. In the afternoon, you triage a fleet quality signal — correlating component-level failure data with manufacturing lot information to identify a systemic issue. You end the day aligning with architecture teams on requirements for the next-generation platform.About the teamThe Hardware Engineering AI/ML UltraServer platform team is a group of engineers and technical program managers directly responsible for launching GPU-accelerated servers into the AWS fleet. Located in Seattle, Austin, and Cupertino, we collaborate with global development teams and ODM partners to deliver next-generation AI/ML infrastructure deployed in datacenters worldwide. We move fast with small, empowered teams delivering end-to-end — from server conception through fleet-scale operations.Basic qualifications- Bachelor's degree in electrical engineering, computer engineering, or equivalent- Experience in developing functional specifications, design verification plans and functional test procedures- 7+ years of hardware design and development experience for server, compute, or large-scale infrastructure platforms- Experience in one or more server technologies: thermal/mechanical design, power delivery, high-speed signal integrity, or accelerator subsystems- Experience leading hardware development through full product lifecycle (concept through production ramp)Preferred qualification - Master's degree or above in electrical engineering, computer engineering, or equivalent- Experience working in data centers or critical infrastructure- 5+ years of experience working with ODMs through the product development and manufacturing lifecycle (EVT, DVT, PVT)- In-depth expertise in high-speed bus design, signal integrity analysis, or power delivery for GPU/accelerator platforms- 5+ years of experience with hardware bring-up, debug, and root cause analysis across PCIe, NVMe, memory, and accelerator interconnects- Experience owning fleet quality metrics and driving design improvements based on operational failure data- Experience with thermal/mechanical design for high-power-density compute platforms (liquid cooling, air cooling, or hybrid)- Track record of defining engineering standards and design best practices adopted across teams or partner organizationsAmazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company’s reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at .USA, CA, Cupertino - 183,000.00 - 247,600.00 USD annuallyUSA, TX, Austin - 159,200.00 - 215,300.00 USD annuallyUSA, WA, Seattle - 159,200.00 - 215,300.00 USD annually
$183k - $247.6k
...to build the backbone of Generative AI cloud at AWS? Do you want to build the... ...performance and scalability in AI/ML and HPC workloads.Utility... ...team of software, hardware, and network engineers, supply chain specialists... ...AWS Generative AI & ML Servers team, you will partner with...Amazon Web ServiceCloudLocal areaFlexible hours$171k - $231.4k
AWS Infrastructure Services owns the... ...people who keep the cloud running. We... ...and all of the servers, networking, power... ...team of software, hardware, and network engineers, supply chain... ...offerings to power AI/ML workloads across... ...for future generations.You will work directly...Amazon Web ServiceCloudSeniorLocal areaWorldwideFlexible hours$173.9k - $235.2k
AWS runs the world's largest fleet of AI/ML accelerator servers. When a model with billions of parameters... ...intersection of hardware, software, and... ...Development Engineer to build automation... ...across servers in the cloud. When you ship,... ...to deliver next-generation AI/ML...Amazon Web ServiceCloudSeniorPermanent employmentInternshipLocal areaWorldwideFlexible hoursNight shiftDay shift$148.7k - $201.2k
...you want to build the backbone of Generative AI cloud at AWS? Do you want to build the future... ...performance and scalability in AI/ML and HPC workloads.You are... ...looking for builders like you. The AWS Hardware Engineering team creates server designs for Amazon’s innovative web...Amazon Web ServiceCloudInternshipLocal areaFlexible hours$168.1k - $227.4k
...building the future of AI-powered data... ...to create next-generation data quality and... ...foundational ML infrastructure.... ...visibility with AWS leadership and the... ...We're seeking Sr. SDE who thrives... ...Leadership: Mentor engineers, drive design... ...broadly adopted cloud platform. We pioneered...Amazon Web ServiceCloudSeniorInternshipFlexible hours$168.1k - $227.4k
Lead our pioneering AI initiative at AWS and define the future... ...will power the next generation of AI agents from small... ...projects requiring multiple engineers, balancing business... ...and broadly adopted cloud platform. We... ...and implementing AI/ML systems, including working...Amazon Web ServiceCloudSeniorInternshipFlexible hoursDay shift$168.1k - $227.4k
...mentor, tech lead, or engineering team leader.... ...experience building ML pipelines, large-scale... ...evaluation frameworks for AI agents and models,... ..., and ground-truth generation. We mentor... ...AI AI Agents AWS Fine-tuning Support... ...Learning Cloud Architect Web...Amazon Web ServiceCloudSeniorFull timeInternship$173.9k - $235.2k
...position is part of the AWS Specialist and... ...challenges.Within ASP, the AI & Strategic Partner Engineering team is seeking a... ...SAP applications and cloud computing, is conversant... ...infrastructure, and generate documentation, treating... ...building with AI/ML technologies including...Amazon Web ServiceCloudSeniorInternshipWorldwideFlexible hours- ...production infrastructure for Generative AI at DoorDash, leading... ...and inference engines, fine-tuning and... ...technical bar for — ML engineers, product engineers... ...with Kubernetes, cloud infrastructure (AWS/GCP), GPUs, serverless... ...deploying AI agents or MCP servers in...Amazon Web ServiceCloudSeniorHourly payWork at officeLocal areaRemote workFlexible hours
$161.9k - $218.6k
...translate complex cloud solutions into compelling... ...customer value? At AWS, we're seeking a... ...the future of AI-powered developer tools... ...tools, AI/ML technologies, or cloud... ...developer tools, code generation, AI agents, or AI-assisted... ...0 USD annually #J-18808-Ljbffr Socket.devAmazon Web ServiceCloudSeniorWork experience placementLocal areaFlexible hours$175k - $236.8k
AWS Trainium is deployed at scale,... ...deep learning and generative AI workloads with... ...training AI/ML ecosystem and what... ...partner with engineering teams building... ...achieve in the cloud. About Amazon... ...silicon engineering, hardware design,... ...annually #J-18808-Ljbffr Socket.devAmazon Web ServiceCloudSeniorLocal areaFlexible hours$168.1k - $227.4k
As part of the AWS Applied AI Solutions organization, we have a vision... ...globally. We're building next-generation services that combine Amazon... ...Senior Software Development Engineer in the AWS Healthcare AI... ...comprehensive and broadly adopted cloud platform. We pioneered cloud...Amazon Web ServiceCloudSeniorInternshipWorldwideFlexible hours$168.1k - $227.4k
As part of the AWS Applied AI Solutions organization, we have... .... We're building next-generation services that combine... ...Software Development Engineer in the AWS Healthcare... ...comprehensive and broadly adopted cloud platform. We pioneered... .... Familiarity with AI/ML frameworks and...Amazon Web ServiceCloudSeniorInternshipWorldwideFlexible hours$147.9k - $200.1k
...changes with Agentic AI? The Next Gen... ...services that enable AWS customers to leverage... ...Join the team as a Sr. Technical Business... ...and broadly adopted cloud platform. We pioneered... ...not limited to, AI/ML (Artificial Intelligence... ...Machine Learning), GenAI (Generative AI), Analytics,...Amazon Web ServiceCloudSeniorImmediate startWorldwideFlexible hours$153.6k - $207.8k
...Solutions Architect at AWS, you will help... ...to build AI-powered migration... ...Windows/.NET, SQL Server, and virtualized... ...-deployed engineering mindset. You will... ...capabilities and Generative AI to accelerate... ...technical skills across cloud migration,... ...modernization, and AI/ML, plus the...Amazon Web ServiceCloudSeniorWork experience placementFlexible hours$162.7k - $220.2k
The AWS Worldwide Data & AI Strategy Team is focused on working directly with... ...Database, Analytics, AI/ML, and Generative AI services.As a member of... ...as diving deep with data engineers and data scientists. You bring... ...AWS as the leader in cloud-based data and AI. You will...Amazon Web ServiceCloudSeniorLocal areaWorldwideFlexible hours$162.7k - $220.2k
...agentic workflows with AWS customers and evolution... ...optimize price/performance for AI Inference at scale? AWS... ...GenAI and ML workloads with specific... ...the teamWe are a team of Generative AI Go-to-Market specialists... ...enterprise software or cloud-based applications- Experience...Amazon Web ServiceCloudSeniorLocal areaWorldwideFlexible hours$182.8k - $247.3k
Amazon Web Services (AWS) is seeking an... ...deeply technical Generative AI Solutions Architect... ...will guide telco engineering and product teams... ...architectures with ML engineers, designing... ...companies to accelerate cloud transformation and... ...AI acceleration hardware, or experience...Amazon Web ServiceCloudFlexible hoursShift work$176.6k - $239k
...position is part of the AWS Specialist and... ...ADAS), and physical AI applications? Then... ...us define the next generation of autonomous operations... ...management, engineering, business development... ...and broadly adopted cloud platform. We... ...data-intensive AI/ML workloads in industrial...Amazon Web ServiceCloudSeniorLocal areaFlexible hoursShift work$162.7k - $220.2k
...Description The AWS Data and AI Strategic Partner GTM team supports the... ...trends • Work with a strategic Generative AI partner to define and... ...comprehensive and broadly adopted cloud platform. We pioneered cloud... ..., but not limited to, AI/ML, GenAI, Analytics, Database,...Amazon Web ServiceCloudSeniorLocal areaWorldwideFlexible hours$125k - $150k
...development, and engineering to harness the... ...Consulting hardware and robotics team... ...perception and applied AI. While the... ...and own CV/ML pipeline architecture... ...synthetic data generation from simulation... ...with hardware, cloud, and software... ...such as AWS, GCP or Azure...Amazon Web ServiceCloudSenior- ...Infosys Limited is seeking an AI Architect to design, implement, and lead enterprise AI/ML and Generative AI solutions across industries. You... ...and deliver end-to-end results on cloud-native and hybrid environments with platforms like AWS, Azure, and GCP. You will work...Amazon Web ServiceCloudSenior
$153.6k - $207.8k
...you a power systems engineer who wants to use AI to transform how utilities... ...electric grid? The AWS Energy & Utilities AI... ...how they evaluate generation queue projects. You bring... ...and broadly adopted cloud platform. We... ...- Experience with AI/ML applications in power...Amazon Web ServiceCloudSeniorFlexible hoursDay shift$152.2k - $205.9k
As part of the AWS - Applied AI Solutions organization, we have a vision... ...team and build world-class Generative AI products for our healthcare... ...service roadmap and work with engineering and scientists to execute on... ...and broadly adopted cloud platform. We pioneered cloud...Amazon Web ServiceCloudSeniorWorldwideFlexible hours- ...shape the future of AI-driven reliability engineering? At JPMorganChase, we... ...JPMorganChase within the AI/ML Data Platforms... ...application stacks, and cloud infrastructure —... ...developing production generative AI applications at scale... ...experience (e.g., AWS) and infrastructure-as...Amazon Web ServiceCloudSenior
$147.9k - $200.1k
AWS Infrastructure Services (AIS) owns the design, planning, delivery, and operation... ...people who keep the cloud running. We support all... ...centers and all of the servers, storage, networking,... ...diverse team of software, hardware, and network engineers, supply chain...Amazon Web ServiceCloudSeniorFlexible hours$168.1k - $227.4k
...choose us to create cloud solutions that... ...customers change the world.AWS Neuron is the... ...Trn1 and Inf2/Inf1 servers that use them. This... ...is for a software engineer in the Distributed... ...a wide variety of ML model families, including... ...as multi-modal generation models such as Stable...Amazon Web ServiceCloudSeniorInternshipFlexible hours- ...Senior AI Engineer – Privacy Location: Bellevue WA... ...), retrieval-augmented generation (RAG), multi-agent orchestration... ...workflows. Data & ML Engineering Build... ...pipelines. Cloud & MLOps Deploy and... ...workloads on Azure or AWS, including serverless inference...Amazon Web ServiceCloudSenior
$142k - $220.5k
Job DescriptionA Senior Engineer on the AI Enablement team... ...Nordstrom Technology use Generative AI safely and at scale... ...with developer tools, cloud services, LLM providers... ...using Kubernetes and AWS.Mentorship & CollaborationMentor... ...as code.AI/ML understanding: understanding...Amazon Web ServiceCloudSeniorFull timeTemporary work$148.7k - $201.2k
AWS Infrastructure Services owns the design... ...who keep the cloud running. We support... ...centers and all of the servers, storage, networking... ...managers, software engineers, data specialists,... ...Infrastructure Services (AIS) designs, delivers,... ...rapidly growing AI/ML server footprint....Amazon Web ServiceCloudSeniorInterim roleFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr Cloud Hardware Dev Engineer, AWS Generative AI & ML Servers. Be the first to apply!
- senior aws cloud engineer Seattle, WA
- aws cloud architect Seattle, WA
- remote cloud architect Seattle, WA
- informatica cloud developer Seattle, WA
- senior principal cloud computing engineer Seattle, WA
- cloud network engineer Seattle, WA
- google cloud architect Seattle, WA
- cloud engineer Seattle, WA
- senior cloud solutions architect Seattle, WA
- entry level cloud engineer Seattle, WA



