Software Engineer- GPU/AI/ML
Advanced Micro Devices Inc
WHAT YOU DO AT AMD CHANGES EVERYTHING At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career. THE ROLE:At AMD, we build products that accelerate next-generation computing—from AI and data centers to PCs, gaming, and embedded systems. Progress here comes from bold ideas, strong execution, and people who care about hard problems.This role sits at the center of that mission: making AMD the platform of choice for the most demanding AI workloads by improving how models train, align, and run on our GPUs.THE PERSON:We're looking for a senior software engineer who combines deep systems performance work with modern AI—someone who can shape software from GPU kernels through distributed training and inference.You'll join a core team of specialists working on the latest AMD hardware and software. Your work will directly influence the ROCm ecosystem and how foundation models and agentic systems perform on AMD GPUs.The challenge: Help train and run AI systems that make AI itself more efficient on GPUs—tuning stacks, kernels, and workflows in ways that can materially shift what's possible on our hardware.This is a high-impact, hands-on role. You'll own hard technical problems, influence direction across teams, and mentor others as we scale AMD's AI software strategy.WHAT YOU'LL DOOwn the AI software stack: Establish best practices and drive performance from low-level GPU kernels to large-scale distributed systems. Use modern LLMs and agent-based tooling where it accelerates development and tuning of the ROCm ecosystem.Accelerate foundation models and agents: Improve training, post-training, and inference for LLMs and autonomous AI workloads so AMD is the default platform for the most demanding use cases.Co-design hardware and software: Partner on the full lifecycle—from GPU architecture input to software for new accelerators—and engage with the broader AI community to keep AMD at the forefront.WHAT WE'RE LOOKING FORWe need someone who can go deep in the areas below and collaborate effectively.SYSTEMS & GPU PERFORMANCE (KERNEL ENGINEERING)Expert-level modern C++ and design of large, performance-critical systems.Strong grasp of GPU architecture, memory hierarchy, and kernel optimization (HIP/CUDA).Hands-on delivery on large-scale C++/HIP/CUDA codebases, such as ROCm (rocBLAS, hipDNN, Composable Kernel, AITemplate), the CUDA ecosystem (cuBLAS, cuDNN, CUTLASS, Thrust, CUB, NCCL), and ML framework cores such as PyTorch, TensorFlow, or JAX (C++/HIP/CUDA paths).Comfort diagnosing bottlenecks with profilers (for example, ROCm Profiler and Nsight) in multi-GPU, distributed settings.AI POST-TRAINING & LLM SYSTEMSDeep understanding of transformers, attention, and the full model lifecycle.Hands-on work in alignment and post-training—for example, SFT, RLHF, and GRPO.Awareness of current LLM trends, including MoE, quantization, speculative decoding, and agentic systems.Experience optimizing post-training and inference pipelines at scale.PREFERRED BACKGROUNDSubstantial professional experience in software development within performance-critical environments.Extensive HIP/CUDA experience optimizing deep learning and OSS LLM inference/training kernels and operators.Strong technical ownership and a track record of shipping complex systems.Clear communication and influence across teams.Plus: Deep familiarity with the AMD ROCm/HIP ecosystem.Plus: Working knowledge of RTL design and Verilog/SystemVerilog for hardware–software co-design.EDUCATIONBachelor's in Computer Science, Computer Engineering, Electrical Engineering, or equivalent experience.Master's preferred; PhD a plus.Publications in AI/ML, GPU computing, or systems optimization are valued.#LI-BW1 #LI-HYBRIDBenefits offered are described: AMD benefits at a glance.AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.This posting is for an existing vacancy.
$224k - $356.5k
...unlimited potential of AI to define the next era... ...computing. An era in which our GPU acts as the brains of... ...team is building the software stack that makes large... ...Science, Computer Engineering, Electrical Engineering... ...depth in GPU computing, ML systems, or high-performance...SuggestedFull timeLocal area- ...healthcare and scientific discovery to powering AI and the technologies people rely on... ...ROLE :AMD is looking for an experienced software engineer to help build Fleet Manager, a secure... ...control plane for operating large-scale AMD GPU infrastructure.Fleet Manager provides...Suggested
$61k - $101k
...training or certification in software engineering concepts, plus 3+ years of applied... ..., with a strong focus on ML systems. We need hands-on... ...using enterprise-authorized AI-assisted development tools in... ...exposure to deploying or operating GPU workloads in Kubernetes...SuggestedFull time$152k - $241.5k
...computing model focused on visual and AI computing. For two decades,... ..., with our invention of the GPU. The GPU has also shown to be... ...is looking for Architects, Software Engineers, and AI application developers... ...Proficiency in C++, Python and ML frameworks like LangChain, LangSmith...SuggestedFull time- ...opportunity for you to take your software engineering career to the next level. As... ...enterprise-authorized AI coding assist tools within the... ...experience, with emphasis on ML systems.Hands-on experience using... ...to deploying or operating GPU workloads in Kubernetes environmentsExposure...Suggested
- ...generation computing experiences—from AI and data centers, to PCs,... ...is looking for an influential software engineer who is passionate about... ...performance from the lowest-level GPU kernels to large-scale distributed... ..., or the C++/HIP/CUDA core of ML frameworks like PyTorch,...
$152k - $241.5k
NVIDIA’s invention of the GPU in 1999 sparked the growth of the... ...deep learning ignited modern AI — the next era of computing —... ...advancement.Are you a motivated system software engineer with a deep understanding of... ...improve the scheduling of AI/ML workloads on our GPUS to be...Full time$124k - $195.5k
...Visualization. Our invention—the GPU—functions as the visual cortex... ...groundbreaking applications from generative AI to autonomous vehicles. We are now looking for a Software Engineer to help accelerate the next era... ..., and deploy the most advanced ML models on some of the world’s...Full timeRemote work- ...computing experiences—from AI and data centers, to... ...critical in enhancing GPU kernels, deep learning... ...technologies and advanced engineering principles to drive... ...integrating graph compilers. Software Engineering Best... ...accelerated compute into ML frameworks (e.g., PyTorch...
- ...computing experiences—from AI and data centers, to... ...for high-performance GPU kernels, powering major... ...strategically critical to AMD’s AI software roadmap.AMD GPUs are an... .../Gluon kernels for ML kernels powering the... ..., and performance engineering. You are comfortable working...
$184k - $287.5k
...seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-... ...performance inference stacks, optimize GPU kernels and compilers, drive... ...the pareto frontier for the field of ML Systems; survey recent publications...Full time$90k - $180k
...WorkflowsBuild agentic AI services (planning, tool... ...Data PipelinesDevelop GPU‑accelerated pipelines using... ...to design reviews and engineering best practices. Mentor... ...(preferably in AI/ML contexts). Proficiency... ...Tech. We’re a team of software engineers, data scientists...Full timeTemporary workPart time$127.1k - $185k
We're looking for a talented early-career engineer to join our team that owns the network stack for EC2 distributed AI/ML systems. You'll work on software that enables the world's largest AI models to train across massive GPU clusters, developing support for communication...InternshipLocal areaFlexible hours$92k - $135k
...CoreWeave is The Essential Cloud for AI™. Built for pioneers by... ...cost for model serving on our GPU platform. As an IC1, you'll implement... ...mentorship from experienced engineers. About the role: Implement... ...deployed a microservice or ML inference demo. Coursework/research...Permanent employmentFull timeTemporary workCasual workInternshipWork at officeFlexible hours$140k - $224.25k
...driven tools to improve software quality, and ensuring customers... ..., and hands-on software engineer with a test to failure... ...experience with AI technologies for automation... ...the testing workflows in GPU domain.Write... ...experience with building ML and DL based applications...Full time- ...the Kubernetes-native AI infrastructure company,... ...Mirantis empowers platform engineering teams to deliver... ...delivers the automation, GPU orchestration, and policy... ...topology with a customer's ML infra lead, and closed... ...training clusters. Software and platform layer:...Full time
$182k - $242k
...CoreWeave is The Essential Cloud for AI™. Built for pioneers by... ...We're looking for a Senior Engineer to be a driving force on CoreWeave... ..., in near real time, whether a GPU fleet, a fabric, or a training... ...ideally for infrastructure or ML systems rather than general-purpose...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$108k - $172.5k
...organization is looking for passionate software support engineers to partner closely with our... ...Experience with MLOps workflows or ML infrastructure Familiarity with GPU workloads or distributed training... ...an existing vacancy. NVIDIA uses AI tools in its recruiting processes...Full timeRemote work$165.6k
...of non-internship professional software development experience. We... ...networking, HPC interconnects, or GPU/accelerator systems such as... ...is welcome. Prior AI/ML experience is not required; we... ...the ML side to strong systems engineers. Responsibilities: We build...Full timeInternship$120.75k - $161k
...building the next generation of Agentic AI products that help millions of people and... ...talent decisions. We’re looking for a Software Engineer to design and build highly scalable, user... ...with product managers, designers, and AI/ML engineers to bring intelligent agents to...Full timeWork at officeRemote workFlexible hours- ...: We are seeking highly skilled Applied AI Engineer (Software Engineer) to build intelligent systems that... ..., building predictive and deploying ML/AI solutions for complex PCB Design workflows... ..., Gymnasium or custom RL environments. GPU training, distributed experimentation,...
$152k - $241.5k
...Silicon Co-Design Group is seeking an Applied AI Engineer to innovate, develop, and integrate... ...infrastructure that powers our chips. Every CPU, GPU, and Tegra SoC NVIDIA has shipped in the... ...-on experience building and deploying ML/AI systems or data-intensive backend...Full timeRemote work- ...computing experiences—from AI and data centers, to PCs,... ...are hiring Applied AI Engineers to work directly with hardware and software engineering teams on high... ...Experience building applied AI, ML, agentic, automation, or... ....Experience with GPU/CPU performance engineering...
$143k - $286k
...Home OfficeRole summary: The (USA) Staff, Software Engineer plays a critical role in delivering... ...architecture principles, and integrates AI/ML solutions to enhance platform functionality... ...quality solutions. Background in compute and GPU architecture understanding and...Full timeTemporary workPart time$200k - $250k
Thanks for your interest in Oklo! We are searching for a Software Engineer to join our team. Join us in pioneering the next generation of nuclear... ....Grow and manage a team of software engineers to support your AI initiatives.Present roadmaps, product demos, and progress...Remote workFlexible hours- ...performance computing, cloud, and AI. Whether you’re designing next... ...compiler for high-performance GPU kernels, powering major AI... ...closely with GPU architecture and software teams to help establish AMD... ..., cross-functional engineering environmentsStrong problem-solving...Remote work
$147k - $211k
...and deploy scalable and agentic AI solutions for high-value, real-... ....2 years of experience with software development in Python or C++.1 year of experience with ML infrastructure (e.g., model deployment... ...statistics.Google's software engineers develop the next-generation...$174k - $252k
Write and test production software for agentic validation platforms,... ...efficiency, code quality, and engineering best practices.Contribute to technical... ...solutions in specialized ML areas, leveraging ML... ...iOS, Espresso), or closed-loop AI agent toolchains.Experience developing...Shift work$174k - $252k
...implement GenAI solutions, leverage ML infrastructure, and evaluate... ..., maintaining, or launching software products, and 1 year of... ...technologies.Google's software engineers develop the next-generation technologies... ...push technology forward.The AI and Infrastructure team is...Worldwide$147k - $210k
...Artificial Intelligence (GenAI) solutions, utilize ML infrastructure, and contribute to data preparation,... ...optimization.Passionate about pushing the boundaries of AI adoption in the enterprise world.Our mission is to engineer the platform for building and running AI agents,...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Software Engineer- GPU/AI/ML. Be the first to apply!
- software developer intern Santa Clara, CA
- software developer fintech Santa Clara, CA
- ngo software engineer Santa Clara, CA
- software engineer - web development Santa Clara, CA
- intel software engineer Santa Clara, CA
- senior software engineer remote Santa Clara, CA
- junior software developer remote Santa Clara, CA
- software engineer full time Santa Clara, CA
- software engineer staff Santa Clara, CA
- senior software engineer Santa Clara, CA


