Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff TPU Performance Architect for Large-Scale ML

Socket.dev

Google's software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one another. Our products handle information at massive scale and extend beyond web search. We're seeking engineers who bring ideas from AI, distributed computing, and large-scale system design, with flexibility to switch teams as needed. As a software engineer, you will design, develop, test, deploy, maintain, and enhance software solutions, #J-18808-Ljbffr Socket.dev

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff TPU Performance Architect for Large-Scale ML in Sunnyvale, CA vacancy
  •  ...leading tech firm in Sunnyvale is seeking a Senior Staff Architect to drive cutting-edge TPU technology for AI/ML applications. This role will focus on developing...  ...collaborating with domain experts to enhance performance and efficiency. Candidates must have significant... 
    Performance
    Worldwide

    Google

    Sunnyvale, CA
    1 day ago
  • Google in Sunnyvale, CA is seeking a Senior Staff Co-Design Engineer on the TPU Chip Architecture team to shape AI/ML hardware acceleration. You will drive TPU architecture...  ..., software, and hardware teams to push performance and efficiency. This role requires deep expertise... 
    Performance

    Socket.dev

    Sunnyvale, CA
    1 day ago
  •  ...to shape the PCIe subsystem for next-generation TPU accelerators. You will architect and implement SoC-level RTL, drive cross-team...  ...collaboration with software and hardware, and ensure high-performance PCIe integration within AI/ML workloads. You will lead RTL lifecycle,... 
    Performance

    Socket.dev

    Sunnyvale, CA
    1 day ago
  • Google Cloud's TPU Chip Architecture and Performance team analyzes ML workloads to balance performance and cost on next-generation TPU systems. You will work with model researchers, compiler developers, and systems engineers to drive architecture decisions and performance... 
    Performance

    Socket

    Sunnyvale, CA
    3 days ago
  • A leading autonomous technology firm in California seeks a Senior IC to advance AI/ML infrastructure for large models. This role involves collaborating with global teams to enhance simulation realism and designing distributed systems for ML lifecycles. Candidates should... 
    Suggested

    Waymo

    Mountain View, CA
    1 day ago
  •  ...massive model training at scale. Your expertise will drive 2-3x performance gains in both training...  ...practices for distributed ML systems, you will create...  ..., and their impact on large-scale ML workloads You are...  ...track record architecting distributed training systems... 
    Performance
    Remote work

    AMD

    Santa Clara, CA
    23 hours ago
  • $208k - $327.75k

     ...through state-of-the-art AI, high-performance compute, and scalable...  ...are looking for a Senior AI Architect to help define the next generation...  ...2+ years of experience in AI/ML systems, deep learning...  ...modern AI architectures and large-scale model systemsExperience mapping... 
    Performance
    Full time
    Worldwide

    Nvidia

    Santa Clara, CA
    4 days ago
  • $184k - $287.5k

     ...are looking for an outstanding hands-on architect/engineer for a Senior HPC architect role to support deployment and bringup of large-scale GPU compute clusters. Be a key player...  ...architect, develop and bring up large scale performance platforms.What you’ll be doing:Provide... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $272k - $431.25k

     ...interconnects.This Principal Architect role leads the research...  ...systems communicate at scale—across GPUs, DPUs, NICs...  ...runtimes for large-scale AI workloads (KV...  ...deep expertise in high-performance networking (InfiniBand,...  ...networking.Understanding of ML systems concepts—transformer... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    3 days ago
  • Google DeepMind is seeking a Senior hardware architect to design and implement accelerators for...  ..., collaborating with application and performance teams to optimize efficiency and...  ...verification, and physical design with ML model design and numerics. Google DeepMind... 
    Performance

    Google DeepMind

    Mountain View, CA
    1 day ago
  • $227k - $320k

     ...shape the future of AI/ML hardware acceleration....  ...to drive cutting-edge TPU (Tensor Processing...  ...systems. As a Senior Staff Architect, you will collaborate...  ...generational improvements in performance, power, and cost while...  ...at unparalleled scale, efficiency, reliability... 
    Performance
    Full time
    Worldwide

    Google

    Sunnyvale, CA
    1 day ago
  •  ...empowering the creation of high-performance silicon chips and software...  ...audiences. Mentoring senior architects and engineers comes naturally...  ...Infrastructure solutions. Architecting large-scale GenAI and agentic AI...  ...hands-on experience with AI/ML platforms and GenAI systems.... 
    Performance

    Synopsys

    Sunnyvale, CA
    4 days ago
  •  ...seeking a seasoned software/solutions architect to mentor teams and lead the architecture...  ...distributed systems. You will identify performance bottlenecks, design elastic systems,...  ...security, and engineering excellence across large-scale data plane platforms, with... 
    Performance

    Oracle

    Santa Clara, CA
    4 days ago
  • $192k - $278k

     ...with SoC design, IC packaging, Power, Performance, and Area (PPA) constraints, thermal-aware...  ...of the Tensor Silicon Engineering team scaling mobile AI capabilities, you will guide...  ...envelopes. Your focus will be architecting the SoC Thermal Budget and mitigation policies... 
    Performance
    Worldwide
    Shift work

    Google

    Mountain View, CA
    3 days ago
  •  ...times larger than GPUs. Our novel wafer‑scale architecture provides the AI compute...  ...machine learning users to effortlessly run large‑scale ML applications, without the hassle of...  ...requirements Bring up new features in the performance/power model Perform comprehensive PPA... 
    Performance

    Cerebras

    Sunnyvale, CA
    2 days ago
  •  ...Description In this AI/ML ASIC Architecture...  ...product. As an AI/ML ASIC Architect you will help drive...  ...changes and assess the performance, power, area, and endurance...  ...AI Storage with GPU/TPU/xPU accelerators, with...  ...experience optimizing large-scale ML systems, GPU... 
    Performance
    Temporary work
    Remote work
    Flexible hours
    Shift work
    Night shift

    Sandisk

    Milpitas, CA
    3 days ago
  • Waymo is seeking a Senior Tech Lead for ML Infrastructure to drive efficient deployment of large-scale models across diverse hardware. You will work across data...  ...efforts for billions-parameter models, develop performance tooling, and coordinate multiple teams toward high... 
    Performance

    Neura Market

    Mountain View, CA
    1 day ago
  • Rhoda AI in Mountain View is seeking a Staff / Principal ML Training Systems Engineer to lead the performance of large-scale multimodal training systems. This role involves improving training efficiency and collaborating closely with research teams to accelerate model iteration... 
    Performance

    Rhoda AI

    Mountain View, CA
    4 days ago
  •  .... Every day, our work helps care teams perform with greater precision and patients recover...  ...-generation robotic platforms. As a Staff AI/ML Architect, you will own the end-to-end...  ...models and data and designing systems that scale.Roles and ResponsibilitiesDefine and own... 
    Contract work
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    Sunnyvale, CA
    4 days ago
  •  ...the Fortune 100, 10,000 large enterprises, and...  ...an experienced Senior Architect to lead the design and...  ...evolution of enterprise-scale distributed systems supporting...  ...reliability, security, and performance.The ideal candidate...  ...ISO, GDPR)Exposure to AI/ML data pipelines and... 
    Performance
    Full time
    Remote work
    Flexible hours

    Proofpoint

    Sunnyvale, CA
    4 days ago
  •  ...connectivity provides payments performance. Key products and services...  .... About the Role: The Staff Architect is a senior engineering leader...  ...Responsibilities: Architect for Scale and Reliability: Design and...  ...payments, ad tech, telecom, large-scale SaaS). Deep understanding... 
    Performance
    Remote work
    Work from home
    Worldwide
    Flexible hours

    GrabJobs

    San Jose, CA
    4 days ago
  • $272k - $431.25k

     ...environments. Built in Rust for performance and Python for extensibility...  ...single system at datacenter scale. As large language models rapidly...  ...support large-scale LLM inference.Architect and implement deep...  ...high-performance storage, or ML systems infrastructure in C/... 
    Performance
    Full time
    Local area
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $168k - $264.5k

     ...a Senior P&R Methodology Architect to define and own the next...  ...(3nm and below) and high performance GPU, CPU and SoC designs....  ...Python‑based analytics and ML/GenAI techniques to mine large QoR datasets, recommend flow...  ...automating A/B tests and large‑scale regressions, and analyzing... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $200k - $300k

    Job Title: Principal ASIC Architect - Memory Systems & AI InterconnectsJob...  ..., ARM CHI, UCIe, PCIe), and performance modeling to define high-...  ..., low-latency solutions that scale across accelerator, host, and...  ...high-performance solutions for ML hardware and AI workloads.Key... 
    Performance

    CyberCoders

    Santa Clara, CA
    23 hours ago
  • $183k - $247.6k

     ...seeking an experienced Power Architect to join our Machine...  ...constraints for ML accelerator designs in...  ...solutions that maximize performance while maintaining power...  ...Voltage and Frequency Scaling) flows and power control...  ...employees, supervisors, and staff; adhere to standards of... 
    Performance
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $206.4k - $379.1k

     ...whether individuals or large organizations, to effortlessly...  ...drives creativity at scale in design, imaging,...  ...for a Principal Architect to build and implement...  ...Express, merging strong ML skills with proficiency...  ...frameworks.Develop high-performance runtime services for inference... 
    Performance
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours

    Adobe Systems

    San Jose, CA
    1 day ago
  • $162.7k - $284.7k

     ...Preferred QualificationsThe Principal HPC Architect designs, builds, optimizes, and supports large scale compute environments used for scientific computing, AI/ML workloads, simulation, and data...  ...role blends systems engineering, performance tuning, cluster architecture, and... 
    Performance
    Minimum wage
    Full time
    Work experience placement
    Flexible hours

    KLA-Tencor

    Milpitas, CA
    4 days ago
  • $142.8k - $274.8k

     ...planning process, quality, delivery, scale and sustainability related to...  ...for a Principal AI Silicon Architect to join the team. ResponsibilitiesDrive High-Performance AI Silicon Architecture direction...  ...knowledge and expertise in large matrix multiplication, data formats... 
    Performance
    Ongoing contract
    Permanent employment
    Work at office
    Local area
    Worldwide
    3 days per week

    Microsoft

    Mountain View, CA
    3 days ago
  • $100k - $220k

     ...infrastructure. The role focuses on optimizing domain decomposition and scalable preconditioning while ensuring GPU-accelerated performance. Candidates should have experience shipping robust solutions for real hardware workflows and a strong understanding of parallel numerical... 
    Performance

    Vinci4D.ai

    Palo Alto, CA
    3 days ago
  • Advanced Micro Devices seeks a Staff Software Developer to advance AI software on AMD GPUs, spanning low-level kernel work to large-scale distributed systems. You will shape the ROCm ecosystem and drive performance improvements across AI workloads. The role emphasizes deep... 
    Performance

    Advanced Micro Devices , Inc.

    Santa Clara, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff TPU Performance Architect for Large-Scale ML. Be the first to apply!