Senior Staff AI Accelerator Performance Architect
$175k - $275kCerebras Systems
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.Senior Staff AI Accelerator Performance ArchitectWafer-scale computing creates a distinctive architecture space in which compute placement, memory capacity and bandwidth, communication, kernel execution and system-level behavior must be understood together.We are looking for a performance architect to guide the evolution of our next-generation AI systems. You will connect real workloads to architectural behavior, identify the bottlenecks that matter, quantify potential improvements and influence hardware and software roadmaps through rigorous performance analysis.This role is ideal for someone with deep knowledge of hardware architecture, developed through hardware, compiler, kernel or system-performance work, who enjoys operating at the intersection of applications, kernels, architecture and system performance.What You’ll DoOwn and evolve performance models and modeling methodologies for next-generation accelerator and system architectures.Build and extend analytical, simulation-based or trace-driven models across workloads, architectural features and product generations.Analyze important AI workloads, from individual kernels through end-to-end inference and training execution, to determine where time, bandwidth, compute and capacity are spent.Identify hardware and software bottlenecks and quantify opportunities to improve latency, throughput, utilization and energy efficiency.Evaluate proposed architectural features and determine their expected performance return across representative workloads.Study how models and kernels map onto the underlying compute, memory and communication architecture.Partner with architecture, compiler, kernel, runtime and systems teams to evaluate alternative mappings and optimizations.Develop workload projections and competitive performance analyses grounded in transparent assumptions.Create concise recommendations that translate complex performance results into architectural and product decisions.Improve modeling methodology, validation and correlation with RTL, emulation and silicon measurements.Help define representative workloads, performance targets and success criteria for future products.What We’re Looking For7+ years of experience in performance analysis, performance modeling or architecture exploration for CPUs, GPUs, AI accelerators or other high-performance computing systems.Strong understanding of hardware architecture developed through hardware, compiler, kernel, runtime or system-performance work.Experience developing analytical, simulation-based or trace-driven performance models using Python, C++ or similar environments.Solid understanding of processor architecture, memory systems, interconnects, parallel execution and hardware resource constraints.Ability to move between kernel-level behavior and end-to-end application or system performance.Experience profiling workloads, forming performance hypotheses and validating them with quantitative evidence.Understanding of how software mapping and programmability affect realized hardware performance.Ability to communicate modeling assumptions, uncertainty, bottlenecks and recommendations clearly.MS or PhD in Electrical Engineering, Computer Engineering, Computer Science or equivalent practical experience.Particularly Relevant ExperiencePerformance analysis of transformer inference or training workloads.Attention, GEMM/GEMV, collective communication, mixture-of-experts, quantization or memory-capacity-constrained execution.Kernel optimization, compiler performance, runtime scheduling or distributed accelerator systems.Model validation using RTL simulation, emulation, FPGA prototypes or silicon measurements.Competitive analysis of AI accelerators and large-scale AI systems.Role FocusThis is a performance and architecture role, not a production RTL-design position. You should be comfortable reasoning about microarchitecture and working with architecture, RTL and physical-design teams, but you will not be expected to own detailed microarchitecture specifications, production RTL implementation, synthesis closure or physical design.This role evaluates architectural features and recommends improvements; the AI Accelerator Architect owns the detailed feature definition and implementation-ready microarchitecture specification.Your primary deliverables are trusted models, workload insights, feature ROI and architectural recommendations.The base salary range for this position is $175,000 to $275,000 annually. Actual compensation may include bonus and equity, and will be determined based on factors such as experience, skills, and qualifications.Why Join CerebrasPeople who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:Build a breakthrough AI platform beyond the constraints of the GPU.Publish and open source their cutting-edge AI research.Work on one of the fastest AI supercomputers in the world.Enjoy job stability with startup vitality.Our simple, non-corporate work culture that respects individual beliefs.Find out more about what it's like to work at Cerebras here! Apply today and become part of the forefront of groundbreaking advancements in AI!Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.LocationSunnyvale, CAEmployment TypeFull timeLocation TypeOn-siteDepartmentHardware
$200k - $300k
...Systems builds the world's largest AI chip, 56 times larger than... ...speed inference.Principal AI Accelerator ArchitectOur architecture... ...are looking for an experienced architect to conceive and drive new... ...AI workloads, quantify their performance and efficiency value, develop...Performance$184k - $287.5k
We are now looking for a Senior AI Training Performance ArchitectNVIDIA is seeking a senior engineer who is obsessed with performance analysis... ...has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of...SeniorPerformanceFull time$184k - $287.5k
...Infrastructure, and Agentic AI - the biggest technology breakthroughs... ...data moves, connects, and accelerates workloads at scale, and we’re seeking a visionary Product Architect with strong expertise in... ...product architectures, including performance, scalability,...SeniorPerformanceFull time$165k - $265k
...unleashing the potential of generative AI to power the transformation of... ...Role Overviewd-Matrix is looking for a Senior Staff Power Performance Architect to own pre-silicon power estimation across our next-generation AI accelerators. In this role, you will drive both RTL...SeniorPerformance- ...recognized globally for innovation, performance and quality. Sandisk has two facilities... ...moving forward. Job Description An AI Interconnect Architect defines and engineers high-speed... ...Architecture: Familiarity with GPU/accelerator clusters and data center infrastructure...SeniorPerformance
$224k - $356.5k
...used for artificial intelligence (AI) / deep learning (DL), high-performance computing (HPC), cloud service providers... ....Work with CPU and interconnect architects to improve future CPU and system... ...experience.Knowledge of GPU-accelerated workloads and modeling performance...SeniorPerformanceFull time$168k - $258.75k
...are building the next generation of AI-powered simulation tools to accelerate hardware and silicon development.... ...the core of hardware workflows. As a Senior Technical Program Manager, you will... ...over technical direction, not just performing tasks or tracking metrics. You...SeniorPerformanceFull time$224k - $356.5k
...Machine Learning Engineer to join the GPU accelerated Apache Spark team.Apache Spark is the... ...with GPUs. You will apply the latest ML/AI methods to empower enterprises to migrate... ...implement machine learning solutions for performance prediction and optimization of GPU...SeniorPerformanceFull time$200k - $275k
...Intelligent Edge. ADI combines analog, digital, AI, and software technologies into... ...at and on LinkedIn and X.Principal AI Accelerator Architect Lead Boston, MA; San Jose, CA Team:... ...position qualifies for a discretionary performance-based bonus which is based on personal...PerformancePermanent employmentFull timeWork at officeDay shift- EngineersOfAI is seeking a highly accomplished GPU Architect to lead the next generation of AI accelerators and multi-GPU cluster architecture. This... ...GPUs or AI accelerators and exceptional skills in performance modeling, manufacturing techniques, and reliability...Performance
$210.87k - $329.86k
...San Jose Summary Celestica is accelerating the adoption of Artificial Intelligence... ...right execution. We are seeking a senior AI Architect – Hardware Engineering to lead this... ...quality improvement, first-time-right performance, adoption, and return on investment....SeniorPerformanceTemporary workLocal areaWorldwideShift work$184k - $287.5k
...computer graphics, PC gaming, and accelerated computing for more than 25... ...the unlimited potential of AI to define the next era of computing... ...lives. We’re searching for a Senior Systems Software Engineer... ..., containers, and systems performance and scalability. The ideal candidate...SeniorPerformanceFull timeRemote work$180k - $225k
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers... ...the future of work is Human + AI and are building an AI-... ...Zscaler.RoleWe are looking for a Senior Staff Rust Developer to join our... ...orchestration layersOptimize system performance through profiling tools...SeniorPerformanceFull timeWork at officeLocal area- ...AWS operates the world's largest fleet of GPU-accelerated servers powering AI/ML training and inference at cloud scale. Our team defines the server... ...designs and component specifications that enable high-performance AI training and inference at scale* Work with...SeniorPerformance
- ...strong software fundamentals with practical AI-assisted development habits to move... ...journeys Use AI-assisted workflows to accelerate implementation, code review, testing, debugging... ...architectures with an emphasis on performance, reliability, maintainability, and cost...SeniorPerformanceLocal areaWork from homeRelocation packageFlexible hours
$272k - $431.25k
...group is solving some of AI’s hardest... ...interconnects. This Principal Architect role leads the... ...bodies, and mentoring senior engineers across the organization... ...expertise in high-performance networking (InfiniBand... ...MPI, NVSHMEM), and GPU accelerated systems, with track...Performance$193.3k - $261.5k
...these custom-designed accelerator SoCs for use by AWS internal... .... We’re looking for a Senior SoC Modeling Engineer... ...infrastructure performance improvements to help our... ...growing suite of generative AI services and other cutting... ..., supervisors, and staff; adhere to standards of...SeniorPerformanceInternshipLocal areaFlexible hours$171k - $231.4k
Amazon Web Services (AWS) is seeking a Senior AI Acceleration Lead to join the Customer Success... ...GrowthWe’re continuously raising our performance bar as we strive to become Earth’s Best... ...with other employees, supervisors, and staff; adhere to standards of excellence despite...SeniorPerformanceWork at officeLocal areaFlexible hours$224k - $356.5k
...tapping into the unlimited potential of AI to define the next era of computing. An... ...diagnostics software to ensure quality and performance at scale across ODM and partner... ...partners.What we need to see:Proven experience architecting diagnostics for complex server systems,...SeniorPerformanceFull time$224k - $356.5k
...transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It... ...into the unlimited potential of AI to define the next era of computing.... ...management, thermal regulation, and performance optimization. This senior technical leadership role requires...SeniorPerformanceFull time$184k - $287.5k
...transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a... ...tapping into the unlimited potential of AI to define the next era of computing. An... ...improve product yield while maintaining performance and architectural simplicity.Analyze how...SeniorPerformanceFull time- ...integration, empowering the creation of high-performance silicon chips and software content. Join... ..., partners, and internal teams. Accelerate customer outcomes by delivering sustainable... ..., and drive customer success. As a senior technical product manager, you will play...SeniorPerformanceShift work
$152k - $241.5k
We are looking for a highly skilled Performance Modeling Architect to lead the architectural definition and... ...own internal tools or frameworks to accelerate architectural exploration rather than... ...for an existing vacancy. NVIDIA uses AI tools in its recruiting processes....SeniorPerformanceFull timeNight shift$184k - $287.5k
...seeking a highly motivated power architect to own and advance a critical... ..., system architecture, performance modeling, and silicon analysis... ...memory systems, interconnects, accelerators, and power-management... ...for complex SoCs, CPUs, GPUs, AI accelerators, or automotive platforms...SeniorPerformanceFull time$152k - $241.5k
We are now looking for a Senior Hardware SoC Architect! Do you want to be a part of Artificial... ...that are at the forefront of accelerating machine learning, automotive and high-performance computing applications. We... ...vacancy. NVIDIA uses AI tools in its recruiting processes...SeniorPerformanceFull timeWork experience placement$165.5k - $289.6k
...meaningful work. Today, ServiceNow is the AI control tower for business reinvention.... ...is seeking a highly experienced Senior Staff Cloud FinOps Analyst to lead enterprise... ...efficiency, unit economics, commitment performance, and gross margin. What you get to do...SeniorPerformanceFull timeWork at officeImmediate startRemote workFlexible hours$134.9k - $237.3k
Be the one building AI-powered experiences where they matter most. At Genesys, we... ...world enterprise environments every day. Senior AI Architect, Presales United States Role Overview:... ...and freshness while balancing latency, performance, security, and governance requirements...SeniorPerformanceRemote workWork from homeWorldwideFlexible hours$183k - $247.6k
...want to shape the future of AI? Join the team building the foundation... ...about pushing the limits of performance, efficiency, and scalability... ...performance server and/or accelerator server and rack system... ...employees, supervisors, and staff; adhere to standards of excellence...SeniorPerformanceLocal areaFlexible hours- ...empowering the creation of high-performance silicon chips and software... ...and visionary thinking in AI and infrastructure-native architectures... ...audiences. Mentoring senior architects and engineers comes... ...standards and governance. Accelerate time-to-market via reusable,...Performance
$120k - $275k
...tailored for the world’s best AI models. Our hardware will... ...largest models. MatX is seeking an Architect to join our team as we create... ...-in-class silicon for high-performance and sustainable GenAI. The... ...-performance CPU, GPU, or AI accelerator architecture and hardware/software...PerformanceDaily paidFull timeWork experience placementWork at officeLocal areaRemote workMonday to FridayFlexible hours3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Staff AI Accelerator Performance Architect. Be the first to apply!
- senior computer engineer Sunnyvale, CA
- senior manager customer operations Sunnyvale, CA
- senior development engineer Sunnyvale, CA
- senior software engineer ruby on rails Sunnyvale, CA
- sr finance manager Sunnyvale, CA
- senior customer service Sunnyvale, CA
- senior business manager Sunnyvale, CA
- senior compensation manager Sunnyvale, CA
- senior cloud data engineer Sunnyvale, CA
- senior data integration developer Sunnyvale, CA


