Fellow, AI Workload Optimization
Advanced Micro Devices Inc
WHAT YOU DO AT AMD CHANGES EVERYTHING
At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world's most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career. Fellow, AI Software (Workload Optimization) We are looking for a visionary technical leader to join the AI Software group. As a Fellow, you will be accountable for defining and driving the end-to-end software optimization strategy to achieve industry-leading performance for our top-tier customers. You will sit at the intersection of architecture, customer engagement, and software engineering, ensuring that AMD’s software stack—from ROCm and compilers to high-level AI frameworks—is tuned to extract maximum performance for the world's most demanding AI workloads.THE PERSON
The ideal candidate is a technical powerhouse with a proven track record of solving complex performance bottlenecks at scale. You possess deep knowledge of AI hardware architecture and software optimization, with the ability to map emerging model architectures to low-level software. You are an exceptional communicator who can translate complex technical challenges into strategic roadmaps and influence senior leadership, while also engaging deeply with key customers to solve their most critical performance needs.KEY RESPONSIBILITIES
Strategic Leadership: Set the technical vision and roadmap for workload optimization across the AI software stack, ensuring AMD remains the platform of choice for top-tier AI customers. Workload Performance Engineering: Lead the profiling, analysis, and tuning of large-scale models (LLMs, Diffusion, Multimodal, and MoE) to ensure “out-of-the-box” performance excellence on AMD hardware. Customer Engagement: Partner with top customers and hyperscalers to understand their unique workload requirements and deliver tailored architectural wins and software optimizations. Hardware-Software Co-design: Collaborate across hardware architecture, compiler, and framework teams to influence future silicon features based on evolving AI workload trends. Ecosystem Innovation: Drive the development of advanced tools and frameworks for performance estimation, modeling, and automated reporting. Community & Mentorship: Act as a technical ambassador in industry forums and open-source communities. Mentor and inspire the next generation of AMD's technical leaders and engineers.PREFERRED EXPERIENCE
15+ years of software development experience with at least 5 years in a high-level technical leadership role (Fellow or equivalent). Deep expertise in AI Frameworks (PyTorch, JAX, vLLM, SGLang) and the ROCm software stack. Proven history of optimizing distributed inference and training at scale across multi-node/multi-GPU environments. Mastery of performance profiling tools (e.g., TorchProfiler, ROCm Profiler, Nsight) and hardware-level performance modeling. Strong understanding of modern model architectures (Transformer, Attention, KV Cache) and optimization techniques like quantization, speculative decoding, and FlashAttention. Demonstrated ability to drive cross-functional initiatives in fast-paced, ambiguous environments.ACADEMIC CREDENTIALS
PhD or Master's degree in Computer Science, Electrical Engineering, or a related field, or equivalent experience. Demonstrated research or applied experience in AI/ML, including areas such as deep learning, model training/inference optimization, large language models, or computer vision. Benefits offered are described: AMD benefits at a glance. AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants' needs under the respective laws throughout all stages of the recruitment and selection process. AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD's “Responsible AI Policy” is available here. This posting is for an existing vacancy. #J-18808-Ljbffr Advanced Micro Devices, Inc.Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Fellow, AI Workload Optimization in Bellevue, WA vacancy
- NVIDIA seeks a Senior Performance Engineer to profile and optimize large-scale AI workloads within the DGX Cloud AI Efficiency Team in Redmond, WA. You will characterize workloads, establish baselines, diagnose bottlenecks, and drive end-to-end improvements from investigation...Suggested
- ...performance computing, cloud, and AI. Whether you’re designing next... ...for a strong, Principal or Fellow level software engineer to... ...efficiency, and reliability of AI workloads across both model training and... ...performance bottlenecks, optimize workloads, and ensure that models...Suggested
$182k - $242k
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave... ...Performance team, focused on kernel authoring and optimization. You will write, profile, and tune the... ...(and Training) runs, including workload setup, cluster configuration, runbooks,...SuggestedPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$236k - $330k
...enterprise. To usher in this new era, we seek AI-native thinkers across every function... ...of the art in LLM inference systems and optimization.Our mission is to build the next... ...new models, architectures, hardware, and workloads.Our work spans the full inference stack—...Suggested$168.1k - $227.4k
...collective communications and memory utilization to compiler optimizations and kernel performance.Key job responsibilitiesThis role will... ...compiler and runtime to enable and tune large-scale training workloads on the latest Trainium instances.About the teamOur team is dedicated...SuggestedInternshipFlexible hours- ...enterprise. To usher in this new era, we seek AI-native thinkers across every function who... ...applications that scale to enterprise workloads without ever leaving their data warehouse... ...isolation requirements Design and optimize APIs for performance, reliability, and developer...Full timeWorldwide
$99.5k - $160k
Every time an Amazon Customer makes a purchase, the Fulfillment Optimization (FO) Team determines how to fulfill that order in the most cost... ...data pipelines, develop automated reporting systems, and apply AI-powered tooling to accelerate insight generation and decision-...Full timeTemporary workSeasonal workWorldwideFlexible hoursNight shiftDay shift$184.9k - $250.2k
...landscape through industry-leading generative AI technologies, revolutionizing how... ...advertising lifecycle — from ad creation and optimization to performance analysis and customer... ..., handling high-throughput, low-latency workloads at Amazon scale.- Bring GenAI to production...InternshipWorldwideFlexible hours$182k - $242k
...Description Job Description CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform... ...the Role As a Senior Software Engineer II (IC4) on the AI Workload Orchestration Platform team, you will help build and operate...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$99.5k - $160k
...solutions that work across our vendors, warehouses and carriers to optimize both time & cost of getting the packages delivered. Our services... ...S3, and Redshift- Experience developing, deploying and managing AI products at scalePreferred qualification - Master's degree in BI...WorldwideFlexible hours$152.2k - $205.9k
...and the hardest unsolved problem in it is AI cost. Customers are moving from... ...second to deliver the cost, usage, and optimization insights that power financial decisions... ...attribution, and control for generative AI workloads are still being defined across the industry...Flexible hours$154.56k - $193.2k
...hyperscaler for the edge, delivering modular AI infrastructure from first deployment to... ...preparation through model development, optimization, evaluation, deployment, observability,... ...deployment. Deploy containerized AI workloads across Kubernetes, cloud, on-premises,...Work at officeRemote workFlexible hours$184.9k - $250.2k
The AWS Insights and Optimizations (AIO) team owns cloud insight and optimization products, where customers can understand, control and optimize... ...; and 3) enhancing our recommendation products with generative AI. If you love building high-performance software, stay current...Flexible hours$200k - $300k
Member of Technical Staff — Model Optimization and Inference (New Grad) Seattle, Washington About... ...is building photorealistic, real-time AI avatars with emotional intelligence: a... ...etc.) and extend them for our specific workloads Profile and benchmark end-to-end latency...InternshipH1bWork at officeVisa sponsorship- ...the deployment and operation of cutting-edge AI models. Our work spans system software,... ...architecture, fleet-level monitoring, and performance optimization. About the Role We're hiring an SW Engineer to enable production workloads and end-to-end testing on new platforms....
- ...insight, experimentation, attribution, and AI-driven decision-making across the company... ...and increasingly complex training workloads.You will play a key role in shaping how model... ...preparation, or ML feature pipelinesExperience optimizing big data pipelines and infrastructure for...Full timeWork at officeWorldwideRelocation package
$207k - $275k
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave... ...performance, reliability, and efficiency for GPU workloads at hyperscale. What You'll Do... ..., and AI infrastructure performance optimization. Architect and develop backend...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours- ...near real time pipelines Integrate Databricks workloads with downstream reporting, analytics, and AI/ML use cases Collaborate with cross functional... ...frameworks Familiarity with performance tuning and cost optimization on Databricks Experience supporting analytics...
$145k - $298.9k
...that wants you to grow and succeed. The context engine that makes AI enterprise ready. Anyone can build an AI agent. What makes SAP'... ...detection, causal inference, operations research, mathematical optimization, probabilistic modeling, or simulation and scenario search....Permanent employmentFull timeWorldwideFlexible hours$157.7k - $213.8k
...building and running the world's best data and AI infrastructure platform so our customers... ...abstractions to support diverse workloads ranging from ETL to data science.Below are... ...Engineering: Build the next generation query optimizer and execution engine that's fast, tuning...Local areaWorldwide$162.7k - $220.2k
...leadership? Are you passionate about helping customers optimize price/performance for AI Inference at scale? AWS is seeking a Go-To-Market (GTM)... ...build, customize, deploy and operationalize GenAI and ML workloads with specific focus on GenAI Inference. You will take your...Local areaWorldwideFlexible hours$184.9k - $250.2k
Amazon’s Supply Chain Optimization Technologies (SCOT) is looking for experienced Software Development Managers who want to build Amazon’... ...to solve highly complex supply chain challenges. You and your fellow engineers are responsible for designing the architecture, building...Flexible hours$254k - $350k
...autonomous system intelligence.As a Machine Learning and System Optimization Engineer, you will orchestrate and allocate overall system... ...FP8, FP4, BF16/FP16).Proficiency in low-level programming for AI accelerators, specifically developing and optimizing custom ML...Full timeTemporary workRelocation package- ...science stack; proficiency with Spark/PySpark for distributed workloads. We use the right tech for the task and often work within... ...client needs, develop impactful advanced analytics and AI solutions, optimize code, and solve complex business challenges across industries...ApprenticeshipWork at officeLocal areaEasy work
$254k - $350k
...possible. The OpportunityAre you excited to drive our ML Performance Optimization initiatives and make our ML models that enable autonomous... ...backgrounds, experiences, and skills.We may use artificial intelligence (AI) tools to support parts of the hiring process, such as...Full timeRemote work- ...Senior Data Engineer to design, build, and optimize production-grade data solutions... ...PySpark, Python, and ADF. Optimize Spark workloads including large-scale joins, partitioning... ...or customer data. Exposure to Generative AI, LLMs, RAG, or Agentic AI. Telecommunications...Temporary work
$168.1k - $227.4k
...wondered how it got to you so fast? What if you could build the AI-powered systems that make that happen — and make it faster,... ...video to learn more about our organization, SCOT: Supply Chain Optimization Technologies (SCOT) organization is responsible for the end-to-...InternshipFlexible hours$169k - $338k
...technical leadership and long‑term vision for next‑generation AI systems and platforms. This senior individual contributor will... ...scale. The role combines deep technical expertise (ML, simulation, optimization, agentic AI) with cross‑functional influence, mentoring of...Full timeTemporary workPart timeHome office$200k - $287.5k
...enterprise. To usher in this new era, we seek AI-native thinkers across every function who... ..., forecasts requirements, and delivers optimal CPU and GPU capacity on schedule. We... ...procurement and reservation lifecycle for AI/ML workloads (training, fine-tuning, and model serving...$110k - $220k
...business decisions. This role involves developing and deploying optimization and machine learning models, performing statistical analysis,... ...team: The Inventory Decision Intelligence team, part of Applied AI, develops advanced AI and mathematical optimization solutions to...Full timeTemporary workPart time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Fellow, AI Workload Optimization. Be the first to apply!



