Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Staff AI Accelerator Performance Architect

$175k - $275k

Cerebras Systems

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.Senior Staff AI Accelerator Performance ArchitectWafer-scale computing creates a distinctive architecture space in which compute placement, memory capacity and bandwidth, communication, kernel execution and system-level behavior must be understood together.We are looking for a performance architect to guide the evolution of our next-generation AI systems. You will connect real workloads to architectural behavior, identify the bottlenecks that matter, quantify potential improvements and influence hardware and software roadmaps through rigorous performance analysis.This role is ideal for someone with deep knowledge of hardware architecture, developed through hardware, compiler, kernel or system-performance work, who enjoys operating at the intersection of applications, kernels, architecture and system performance.What You’ll DoOwn and evolve performance models and modeling methodologies for next-generation accelerator and system architectures.Build and extend analytical, simulation-based or trace-driven models across workloads, architectural features and product generations.Analyze important AI workloads, from individual kernels through end-to-end inference and training execution, to determine where time, bandwidth, compute and capacity are spent.Identify hardware and software bottlenecks and quantify opportunities to improve latency, throughput, utilization and energy efficiency.Evaluate proposed architectural features and determine their expected performance return across representative workloads.Study how models and kernels map onto the underlying compute, memory and communication architecture.Partner with architecture, compiler, kernel, runtime and systems teams to evaluate alternative mappings and optimizations.Develop workload projections and competitive performance analyses grounded in transparent assumptions.Create concise recommendations that translate complex performance results into architectural and product decisions.Improve modeling methodology, validation and correlation with RTL, emulation and silicon measurements.Help define representative workloads, performance targets and success criteria for future products.What We’re Looking For7+ years of experience in performance analysis, performance modeling or architecture exploration for CPUs, GPUs, AI accelerators or other high-performance computing systems.Strong understanding of hardware architecture developed through hardware, compiler, kernel, runtime or system-performance work.Experience developing analytical, simulation-based or trace-driven performance models using Python, C++ or similar environments.Solid understanding of processor architecture, memory systems, interconnects, parallel execution and hardware resource constraints.Ability to move between kernel-level behavior and end-to-end application or system performance.Experience profiling workloads, forming performance hypotheses and validating them with quantitative evidence.Understanding of how software mapping and programmability affect realized hardware performance.Ability to communicate modeling assumptions, uncertainty, bottlenecks and recommendations clearly.MS or PhD in Electrical Engineering, Computer Engineering, Computer Science or equivalent practical experience.Particularly Relevant ExperiencePerformance analysis of transformer inference or training workloads.Attention, GEMM/GEMV, collective communication, mixture-of-experts, quantization or memory-capacity-constrained execution.Kernel optimization, compiler performance, runtime scheduling or distributed accelerator systems.Model validation using RTL simulation, emulation, FPGA prototypes or silicon measurements.Competitive analysis of AI accelerators and large-scale AI systems.Role FocusThis is a performance and architecture role, not a production RTL-design position. You should be comfortable reasoning about microarchitecture and working with architecture, RTL and physical-design teams, but you will not be expected to own detailed microarchitecture specifications, production RTL implementation, synthesis closure or physical design.This role evaluates architectural features and recommends improvements; the AI Accelerator Architect owns the detailed feature definition and implementation-ready microarchitecture specification.Your primary deliverables are trusted models, workload insights, feature ROI and architectural recommendations.The base salary range for this position is $175,000 to $275,000 annually. Actual compensation may include bonus and equity, and will be determined based on factors such as experience, skills, and qualifications.Why Join CerebrasPeople who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:Build a breakthrough AI platform beyond the constraints of the GPU.Publish and open source their cutting-edge AI research.Work on one of the fastest AI supercomputers in the world.Enjoy job stability with startup vitality.Our simple, non-corporate work culture that respects individual beliefs.Find out more about what it's like to work at Cerebras here! Apply today and become part of the forefront of groundbreaking advancements in AI!Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.LocationSunnyvale, CAEmployment TypeFull timeLocation TypeOn-siteDepartmentHardware

Vacancy posted 9 days ago
Similar jobs that could be interesting for youBased on the Senior Staff AI Accelerator Performance Architect in Sunnyvale, CA vacancy
  • $200k - $300k

     ...Systems builds the world's largest AI chip, 56 times larger than...  ...speed inference.Principal AI Accelerator ArchitectOur architecture...  ...are looking for an experienced architect to conceive and drive new...  ...AI workloads, quantify their performance and efficiency value, develop... 
    Performance

    CEREBRAS SYSTEMS INC.

    Sunnyvale, CA
    3 days ago
  • $184k - $287.5k

    We are now looking for a Senior AI Training Performance ArchitectNVIDIA is seeking a senior engineer who is obsessed with performance analysis...  ...has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  • $184k - $287.5k

     ...Infrastructure, and Agentic AI - the biggest technology breakthroughs...  ...data moves, connects, and accelerates workloads at scale, and we’re seeking a visionary Product Architect with strong expertise in...  ...product architectures, including performance, scalability,... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    22 days ago
  • $165k - $265k

     ...unleashing the potential of generative AI to power the transformation of...  ...Role Overviewd-Matrix is looking for a Senior Staff Power Performance Architect to own pre-silicon power estimation across our next-generation AI accelerators. In this role, you will drive both RTL... 
    Senior
    Performance

    d-Matrix

    Santa Clara, CA
    a month ago
  •  ...recognized globally for innovation, performance and quality. Sandisk has two facilities...  ...moving forward. Job Description An AI Interconnect Architect defines and engineers high-speed...  ...Architecture: Familiarity with GPU/accelerator clusters and data center infrastructure... 
    Senior
    Performance

    Sandisk

    Milpitas, CA
    2 days ago
  • $224k - $356.5k

     ...used for artificial intelligence (AI) / deep learning (DL), high-performance computing (HPC), cloud service providers...  ....Work with CPU and interconnect architects to improve future CPU and system...  ...experience.Knowledge of GPU-accelerated workloads and modeling performance... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  • $168k - $258.75k

     ...are building the next generation of AI-powered simulation tools to accelerate hardware and silicon development....  ...the core of hardware workflows. As a Senior Technical Program Manager, you will...  ...over technical direction, not just performing tasks or tracking metrics. You... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  • $224k - $356.5k

     ...Machine Learning Engineer to join the GPU accelerated Apache Spark team.Apache Spark is the...  ...with GPUs. You will apply the latest ML/AI methods to empower enterprises to migrate...  ...implement machine learning solutions for performance prediction and optimization of GPU... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    12 days ago
  • $200k - $275k

     ...Intelligent Edge. ADI combines analog, digital, AI, and software technologies into...  ...at and on LinkedIn and X.Principal AI Accelerator Architect Lead Boston, MA; San Jose, CA Team:...  ...position qualifies for a discretionary performance-based bonus which is based on personal... 
    Performance
    Permanent employment
    Full time
    Work at office
    Day shift

    Analog Devices

    San Jose, CA
    a month ago
  • EngineersOfAI is seeking a highly accomplished GPU Architect to lead the next generation of AI accelerators and multi-GPU cluster architecture. This...  ...GPUs or AI accelerators and exceptional skills in performance modeling, manufacturing techniques, and reliability... 
    Performance

    EngineersOfAI

    Milpitas, CA
    5 days ago
  • $210.87k - $329.86k

     ...San Jose  Summary Celestica is accelerating the adoption of Artificial Intelligence...  ...right execution. We are seeking a senior AI Architect – Hardware Engineering to lead this...  ...quality improvement, first-time-right performance, adoption, and return on investment.... 
    Senior
    Performance
    Temporary work
    Local area
    Worldwide
    Shift work

    Celestica International LP

    San Jose, CA
    a month ago
  • $184k - $287.5k

     ...computer graphics, PC gaming, and accelerated computing for more than 25...  ...the unlimited potential of AI to define the next era of computing...  ...lives. We’re searching for a Senior Systems Software Engineer...  ..., containers, and systems performance and scalability. The ideal candidate... 
    Senior
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    a month ago
  • $180k - $225k

    Zscaler (NASDAQ: ZS) accelerates digital transformation so customers...  ...the future of work is Human + AI and are building an AI-...  ...Zscaler.RoleWe are looking for a Senior Staff Rust Developer to join our...  ...orchestration layersOptimize system performance through profiling tools... 
    Senior
    Performance
    Full time
    Work at office
    Local area

    Zscaler

    San Jose, CA
    21 days ago
  •  ...AWS operates the world's largest fleet of GPU-accelerated servers powering AI/ML training and inference at cloud scale. Our team defines the server...  ...designs and component specifications that enable high-performance AI training and inference at scale* Work with... 
    Senior
    Performance

    Amazon

    Cupertino, CA
    3 days ago
  •  ...strong software fundamentals with practical AI-assisted development habits to move...  ...journeys Use AI-assisted workflows to accelerate implementation, code review, testing, debugging...  ...architectures with an emphasis on performance, reliability, maintainability, and cost... 
    Senior
    Performance
    Local area
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    3 days ago
  • $272k - $431.25k

     ...group is solving some of AI’s hardest...  ...interconnects. This Principal Architect role leads the...  ...bodies, and mentoring senior engineers across the organization...  ...expertise in high-performance networking (InfiniBand...  ...MPI, NVSHMEM), and GPU accelerated systems, with track... 
    Performance

    NVIDIA Gruppe

    Santa Clara, CA
    2 days ago
  • $193.3k - $261.5k

     ...these custom-designed accelerator SoCs for use by AWS internal...  .... We’re looking for a Senior SoC Modeling Engineer...  ...infrastructure performance improvements to help our...  ...growing suite of generative AI services and other cutting...  ..., supervisors, and staff; adhere to standards of... 
    Senior
    Performance
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    a month ago
  • $171k - $231.4k

    Amazon Web Services (AWS) is seeking a Senior AI Acceleration Lead to join the Customer Success...  ...GrowthWe’re continuously raising our performance bar as we strive to become Earth’s Best...  ...with other employees, supervisors, and staff; adhere to standards of excellence despite... 
    Senior
    Performance
    Work at office
    Local area
    Flexible hours

    AmazonWebServices

    Mountain View, CA
    15 days ago
  • $224k - $356.5k

     ...tapping into the unlimited potential of AI to define the next era of computing. An...  ...diagnostics software to ensure quality and performance at scale across ODM and partner...  ...partners.What we need to see:Proven experience architecting diagnostics for complex server systems,... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  • $224k - $356.5k

     ...transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It...  ...into the unlimited potential of AI to define the next era of computing....  ...management, thermal regulation, and performance optimization. This senior technical leadership role requires... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a...  ...tapping into the unlimited potential of AI to define the next era of computing. An...  ...improve product yield while maintaining performance and architectural simplicity.Analyze how... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    12 days ago
  •  ...integration, empowering the creation of high-performance silicon chips and software content. Join...  ..., partners, and internal teams. Accelerate customer outcomes by delivering sustainable...  ..., and drive customer success. As a senior technical product manager, you will play... 
    Senior
    Performance
    Shift work

    Synopsys Inc

    Sunnyvale, CA
    4 days ago
  • $152k - $241.5k

    We are looking for a highly skilled Performance Modeling Architect to lead the architectural definition and...  ...own internal tools or frameworks to accelerate architectural exploration rather than...  ...for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.... 
    Senior
    Performance
    Full time
    Night shift

    Nvidia

    Santa Clara, CA
    a month ago
  • $184k - $287.5k

     ...seeking a highly motivated power architect to own and advance a critical...  ..., system architecture, performance modeling, and silicon analysis...  ...memory systems, interconnects, accelerators, and power-management...  ...for complex SoCs, CPUs, GPUs, AI accelerators, or automotive platforms... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  • $152k - $241.5k

    We are now looking for a Senior Hardware SoC Architect! Do you want to be a part of Artificial...  ...that are at the forefront of accelerating machine learning, automotive and high-performance computing applications. We...  ...vacancy. NVIDIA uses AI tools in its recruiting processes... 
    Senior
    Performance
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    a month ago
  • $165.5k - $289.6k

     ...meaningful work. Today, ServiceNow is the AI control tower for business reinvention....  ...is seeking a highly experienced Senior Staff Cloud FinOps Analyst to lead enterprise...  ...efficiency, unit economics, commitment performance, and gross margin. What you get to do... 
    Senior
    Performance
    Full time
    Work at office
    Immediate start
    Remote work
    Flexible hours

    ServiceNow

    Santa Clara, CA
    12 days ago
  • $134.9k - $237.3k

    Be the one building AI-powered experiences where they matter most. At Genesys, we...  ...world enterprise environments every day. Senior AI Architect, Presales United States Role Overview:...  ...and freshness while balancing latency, performance, security, and governance requirements... 
    Senior
    Performance
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Genesys Cloud Services, Inc.

    Menlo Park, CA
    2 days ago
  • $183k - $247.6k

     ...want to shape the future of AI? Join the team building the foundation...  ...about pushing the limits of performance, efficiency, and scalability...  ...performance server and/or accelerator server and rack system...  ...employees, supervisors, and staff; adhere to standards of excellence... 
    Senior
    Performance
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  •  ...empowering the creation of high-performance silicon chips and software...  ...and visionary thinking in AI and infrastructure-native architectures...  ...audiences. Mentoring senior architects and engineers comes...  ...standards and governance. Accelerate time-to-market via reusable,... 
    Performance

    Synopsys Inc

    Sunnyvale, CA
    a month ago
  • $120k - $275k

     ...tailored for the world’s best AI models. Our hardware will...  ...largest models. MatX is seeking an Architect to join our team as we create...  ...-in-class silicon for high-performance and sustainable GenAI. The...  ...-performance CPU, GPU, or AI accelerator architecture and hardware/software... 
    Performance
    Daily paid
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work
    Monday to Friday
    Flexible hours
    3 days per week

    MatX

    Mountain View, CA
    14 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Staff AI Accelerator Performance Architect. Be the first to apply!