Principal AI Performance Modeling Architect (Santa Clara)
AMD
WHAT YOU DO AT AMD CHANGES EVERYTHING At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career. THE ROLE:As a Principal Engineer, you will spearhead the next generation of AI infrastructure by defining GPU architecture specifications that enable massive model training at scale. Your expertise will drive 2-3x performance gains in both training and inference pipelines through innovative system design and optimization. You will champion the adoption of cutting-edge techniques across the engineering organization, from efficient attention mechanisms to advanced parallelization strategies. By establishing comprehensive best practices for distributed ML systems, you will create a framework that enables seamless scaling from single-GPU to thousand-GPU deployments.THE PERSON:You have a deep understanding of GPU microarchitecture, memory hierarchies, and their impact on large-scale ML workloads You are passionate about software engineering and possess leadership skills to drive sophisticated issues to resolution. You are able to communicate effectively and work optimally with different teams across AMD. KEY RESPONSIBILITIES:Lead performance modeling and optimization for multi-trillion parameter LLM training/inference including Dense, Mixture of Experts (MoE) with multiple modalities (text, vision, speech)Model/optimize novel parallelization strategies across tensor, pipeline, context, expert and data parallel dimensionsArchitect memory-efficient training systems utilizing techniques like structured pruning, quantization (MX formats), continuous batching/chunked prefill, speculative decodingIncorporate and extend SOTA models such as GPT-4, Reasoning models (Deepseek-R1), and multi-modal architecturesCollaborate with internal and external stakeholders/ML researchers to disseminate results and iterate at rapid pace.REQUIRED EXPERIENCE:Extensive and Senior experience optimizing large-scale ML systems and GPU architecturesDeep expertise in CUDA programming, GPU memory hierarchies, and hardware-specific optimizationsProven track record architecting distributed training systems handling large scale systemsExpert knowledge of transformer architectures, attention mechanisms, and model parallelism techniquesPREFERRED EXPERIENCE:PyTorch, CUDA, TensorRT, OpenAI TritonDistributed systems: Ray, Megatron-LMPerformance analysis tools: NSight Compute, nvprof, PyTorch ProfilerKV cache optimization, Flash Attention, Mixture of ExpertsHigh-speed networking: InfiniBand, RDMA, NVLinkACADEMIC CREDENTIALS:Bachelors, MS/PhD in Computer Science/Engineering or equivalent industry experienceLOCATION: Austin, Tx or Santa Clara, Ca strongly preferred; Remote is a possibility for the right candidateThis role is not eligible for visa sponsorship.#LI-RL1Benefits offered are described: AMD benefits at a glance.AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.This posting is for an existing vacancy.
$200k - $300k
...Job Title: Principal ASIC Architect - Memory Systems & AI InterconnectsJob Location: Santa Clara, CA or Boston, MACompensation: $200K - $300K plus bonus and equityRequirements... ...UALink, RDMA, ARM CHI, UCIe, PCIe), and performance modeling to define high-bandwidth, low-latency...PrincipalPerformancePart time$195k - $285k
...the potential of generative AI to power the transformation... ...AI. Working onsite at our Santa Clara, CA headquarters 3 days per week Hybrid.The role: Principal Architect- Performance Analysis and Modelingd-Matrix... ...-modal LLMs, CoT reasoning models, video/audio-generation)...PrincipalPerformancePart time3 days per week$272k - $431.25k
...group is solving some of AI’s hardest... ...computing interconnects.This Principal Architect role leads the research... ...prefill/decode, model parallelism).Integrating... ...deep expertise in high-performance networking (... ...SummaryLocation: US, CA, Santa Clara; US, TX, Austin; US,...PrincipalPerformanceFull timePart timeRemote work$272k - $431.25k
...used for artificial intelligence (AI), agentic workloads, deep learning (DL), high-performance computing (HPC), cloud service... ...stack, enabling faster AI model training, agentic use-cases, efficient... ...law.SummaryLocation: US, CA, Santa Clara; US, TX, Austin; US, OR,...PrincipalPerformanceFull timePart time$272k - $431.25k
...We are looking for a Principal Architect for SoC and System performance modelling.The NVIDIA Architecture Modeling group is looking... ...an existing vacancy. NVIDIA uses AI tools in its recruiting processes... ...protected by law.SummaryLocation: US, CA, Santa ClaraType: Full time...PrincipalPerformanceFull timePart timeWork experience placement$208k - $327.75k
...accelerated computing, AI, and autonomous machines... ...state-of-the-art AI, high-performance compute, and scalable... ...looking for a Senior AI Architect to help define the next generation of AI model paradigms for autonomous... ....SummaryLocation: US, CA, Santa ClaraType: Full time...PerformanceFull timePart timeWorldwide$184k - $287.5k
...Imagine shaping the future of AI and computing at one of... ...advancements in High-Performance Computing, Artificial Intelligence... ...visionary computer architects to design and develop models for the next generation... ...SummaryLocation: US, CA, Santa Clara; US, TX, Austin; US, OR,...PerformanceFull timePart timeRemote work$272k - $431.25k
...seeking a world-class Principal Memory Simulation Architect to design and develop... ...on-chip interconnect performance and functional models. Ideal candidates will... ...architectural simulation, AI-assisted model... ...SummaryLocation: US, CA, Santa Clara; US, CA, RemoteType: Full...PrincipalPerformanceFull timePart timeNight shift$272k - $431.25k
...the unlimited potential of AI to define the next era of... ...vehicles. Join our CPU Performance Architecture team in Santa Clara, CA, and be part of a company... ...microarchitectural ideas.Modeling potential improvements to... ...of collaboration with architects and microarchitects to...PrincipalPerformanceFull timePart time$232k - $368k
...at scale.We are looking for a Principal Performance and Manufacturing Architect who has built the models, defined the specs, and seen them... ...exceptional hire also uses AI deliberately — with proven workflow... ...law.SummaryLocation: US, CA, Santa Clara; US, TX, Remote; US, OR,...PrincipalPerformanceFull timePart timeRemote workShift work$272k - $431.25k
...unlimited potential of AI to define the next era... ...seamless scaling of models to clusters comprising... ...application developers to architect and implement... ...industry experience in high-performance computing (HPC) or distributed... ...: US, CA, Santa Clara; US, TX, Austin; US, RemoteType...PrincipalPerformanceFull timePart time$191.53k - $286.9k
...enterprise, cloud and AI, and carrier... ...hyperscalers, system architects, and Marvell’s... ....As a Senior Principal Memory Architect,... ...ExpectArchitect high-performance HBM/DDR memory... ...specifications, performance models, and product... ...in-office at our Santa Clara, Irvine or Boise...PrincipalPerformancePermanent employmentPart timeInternshipWork at officeRemote workWork from home$100k
...industry on cutting-edge AI technology, revolutionizing performance expectations, ease of... ...innovations in software models, compilers, platforms, networking... ...Performance Modeling Architect to help shape the next... ...ishybrid, based out of Santa Clara, CA or Austin, TX.We...PerformancePermanent employmentPart time$272k - $431.25k
...innovate how we architect and develop our GPU... ...for the changing AI and accelerated... ...are looking for a Principal System Architect... ...to fruition high-performance, high-volume System... ...architectural modeling to identify optimal... ...SummaryLocation: US, CA, Santa Clara; US, OR,...PrincipalPerformanceFull timePart timeRemote work$184k - $287.5k
...NVIDIA is hiring an AI Hardware Architect to analyze and architect the next... ...Study the applications and models running on Nvidia hardware... ..., analyze, and explain the performance and power advantages of Nvidia... ....SummaryLocation: US, CA, Santa Clara; US, TX, AustinType: Full time...PerformanceFull timePart time- ...computing experiences—from AI and data centers, to PCs... ...your career. The Role:Architect, Analyze and optimize high-performance SoCs for Cloud computing... ..., performance modeling, and analysis for next-generation... ...preferredLocation:Austin TX preferred; Santa Clara, CAWe will ensure that...PrincipalPerformancePart time
$272k - $431.25k
...into the unlimited potential of AI to define the next era of... ...libraries. Moreover, you will build performance analysis tools and strategies... ...workloads and deep learning models aimed at large-scale LLM... ...law.SummaryLocation: US, CA, Santa Clara; US, TX, Remote; US, CO, Remote...PrincipalPerformanceFull timePart timeRemote work$272k - $431.25k
...seeking an exceptional Principal Perception... ...advanced 3D perception models using multi-camera... ...perception performance; analyze large-scale... ...Hands-on experience architecting and deploying DNN-... ...vacancy. NVIDIA uses AI tools in its... ...SummaryLocation: US, CA, Santa ClaraType: Full...PrincipalPerformanceFull timePart time$196k - $300k
...the potential of generative AI to power the transformation... ...Hybrid, working onsite at our Santa Clara, Ca headquarters 3-5 days per week.The Role: Principal Technical Program Manager,... ...design verification, emulation, modeling (power, performance), Design for Test (DFT),...PrincipalPerformancePart time3 days per week$272k - $431.25k
...components, driver/platform layers, and performance counter/trace providers.Establish profiling models that integrate with existing ML/AI workflows (e.g., PyTorch/XLA) to turn low... ...protected by law.SummaryLocation: US, CA, Santa Clara; US, TX, AustinType: Full time...PrincipalPerformanceFull timePart time$138.84k - $208k
...enterprise, cloud and AI, and carrier... ...experienced FP&A Systems Principal Professional / Solution Architect to lead the design,... ..., and scenario modeling processes• Drive architecture... ...enhance enterprise performance management (EPM)... ...1SummaryLocation: Santa Clara, CA; Irvine, CAType...PrincipalPerformancePermanent employmentFull timePart timeInternshipWork from home$190k - $300k
...potential of generative AI to power the... ...of AI. Location:Santa Clara, CA, headquarters... ...Canada.The role: Principal Software Engineer... ...you will do:Help architect and develop the... ...kernel authoring and modeling lowering flows.... ...optimizing high-performance kernels targeting...PrincipalPerformancePart timeWork experience placement$248k - $379.5k
...Diagnostic, Prescriptive and AI-augmented Analytics solutions... ...detection, root cause and predictive modeling for the benefit of millions... ...cloud computing/streaming performance and experience. Our... ...#deeplearningSummaryLocation: US, CA, Santa Clara; US, CA, RemoteType: Full...PrincipalPerformanceFull timePart time$272k - $431.25k
...serving generative AI and reasoning models across multi-node distributed... .... Built in Rust for performance and Python for... ....We are seeking a Principal Systems Engineer to... ...scale LLM inference.Architect and implement deep integrations... ...: US, CA, Santa Clara; US, WA, Remote; US,...PrincipalPerformanceFull timePart timeLocal areaRemote work- ...potential of generative AI to power the... ...working onsite at our Santa Clara, CA, headquarters 3+ days... ...days per week.The Role: Principal Software Engineer, KernelsWhat... ...or TensorFlow) and ML models for CV, NLP, or... ...individual and company performance. This is in addition to...PrincipalPerformancePart timeWork experience placement3 days per week
$248k - $391k
...unlimited potential of AI to define the next... ...a highly skilled Principal Software Engineer... ...optimizing the performance of our infrastructure... ...initiatives to architect and transform our... ...to frontier-class models. You will develop... ...SummaryLocation: US, CA, Santa Clara; US, RemoteType:...PrincipalPerformanceFull timePart time$208k - $260k
...educational organizations. As a Principal Security Detections... ...engineering, and high-performance systems software to... ...is based out of our Santa Clara, CA headquarters,... ...working with security data models, telemetry... ...automated tools, including AI-based systems, to help...PrincipalPerformancePart timeLocal areaWorldwide3 days per week$272k - $431.25k
...a world-class computer architect to contribute to the development... ...of future high-performance computing systems, with... ...architecture prototype models for power and noise... ...existing vacancy. NVIDIA uses AI tools in its recruiting... ...: US, CA, Santa ClaraType: Full time...PrincipalPerformanceFull timePart time$182.5k - $260.5k
...employees spread across offices in Santa Clara, St. Louis, Bangalore, London... ...layer that makes AI in agentic workflows fast, efficient... .... You fine-tune and evaluate models, push latency and throughput... ...new product that changes the performance and economics of agentic AI.Cutting...PrincipalPerformancePart timeWork at office$272k - $431.25k
...unlimited potential of AI to define the next era... ...world.At NVIDIA, as a Principal Rack Scale Systems... ...silicon, or other high-performance computing systems.Expertise... ...multiple adoption models — internal services, CSP... ...: US, CA, Santa Clara; US, NC, Remote; US, TX...PrincipalPerformanceFull timePart timeRemote workShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal AI Performance Modeling Architect (Santa Clara). Be the first to apply!
- principal architect Santa Clara, CA
- senior principal cloud computing engineer Santa Clara, CA
- senior principal scientist Santa Clara, CA
- principal cloud computing engineer Santa Clara, CA
- principal Santa Clara, CA
- system performance engineer Santa Clara, CA
- performance test engineer Santa Clara, CA
- senior performance tester Santa Clara, CA
- IT performance management Santa Clara, CA
- senior performance engineer Santa Clara, CA








