Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

TPU Performance Engineer — ML Efficiency & Scale

Google

Google is seeking software engineers for the AI and Infrastructure teams to push performance across large-scale systems and accelerator hardware. You will work across JAX and PyTorch to optimize production and research workloads, including Gemini models, across TPU fleets. Responsibilities include designing and delivering high-performance software, collaborating with cross-functional teams, and driving scale, efficiency, and reliability in a fast-paced environment with opportunities to switch #J-18808-Ljbffr Google

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the TPU Performance Engineer — ML Efficiency & Scale in Sunnyvale, CA vacancy
  • $163k - $236k

     ...to ensure robust, high-performance silicon delivery....  ...degree in Electrical Engineering, Computer Engineering,...  ...shape the future of AI/ML hardware acceleration....  ...opportunity to drive TPU (Tensor Processing Unit...  ...Infrastructure at unparalleled scale, efficiency, reliability and... 
    Performance
    Worldwide

    Google

    Sunnyvale, CA
    5 days ago
  • $138k - $197k

    Define TPU microarchitecture and develop high-quality, performant, and power-efficient SystemVerilog RTL code for complex...  ...degree in Electrical Engineering, Computer Engineering...  ...shape the future of AI/ML hardware acceleration...  ...at unparalleled scale, efficiency, reliability... 
    Performance
    Worldwide

    Google

    Sunnyvale, CA
    1 day ago
  • Rhoda AI in Mountain View is seeking a Staff / Principal ML Training Systems Engineer to lead the performance of large-scale multimodal training systems. This role involves improving training efficiency and collaborating closely with research teams to accelerate model iteration... 
    Performance

    Rhoda AI

    Mountain View, CA
    4 days ago
  • $184k - $287.5k

     ...application is built. We are seeking a Sr. HPC Performance engineer to join our team of scientists and...  ...of scientific machine learning (ML) frameworks. Starting with digital biology...  ...computationally performant features for large scale, CUDA-backed ML training frameworks,... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $163k - $236k

     ...degree in Electrical Engineering, Computer...  ...power architecture, performance, or SoC design).Experience...  ...and Frequency Scaling (DVFS), turboing,...  ...the future of AI/ML hardware acceleration...  ...cutting-edge TPU (Tensor Processing...  ...improving the power efficiency of our TPUs. You will... 
    Performance
    Worldwide

    Google

    Sunnyvale, CA
    3 days ago
  • $163k - $236k

     ...Science, Electrical Engineering, Computer...  ...Teradyne (e.g., EXA scale, UltraFlex+).In this...  ...the future of AI/ML hardware acceleration...  ...opportunity to drive TPU (Tensor Processing...  ...to validate performance and screen bad devices...  ...unparalleled scale, efficiency, reliability and... 
    Performance
    Worldwide

    Google

    Sunnyvale, CA
    2 days ago
  • $320k

     ...looking for a Distinguished Engineer to join NVIDIA's architecture...  ...accelerated computing systems scale from a single processor to multi...  ...rely on you to set long-term performance strategy across applications,...  ...improvements in performance, efficiency, or scaling for priority... 
    Performance
    Full time
    Local area
    Shift work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

     ...are seeking an AI Compiler Engineer with deep expertise in compiler...  ...measurable improvements in performance and efficiency, and advancing LLM-enabled...  ...capabilities that power large-scale, high-impact products.Design...  ...software engineering and AI/ML experience, preferably in... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $320k

     ...a team of innovative engineers dedicated to solving some...  ...global strategy for scaled-out AI inferencing. You...  ...to run with peak efficiency on our accelerated computing...  ...-optimization, drive performance tuning at the kernel...  ...systems to support AI/ML workloads. Industry Expertise... 
    Performance
    Full time
    Worldwide

    Nvidia

    Santa Clara, CA
    3 days ago
  • $138k - $197k

     ...solutions for new High Performance Computing (HPC)...  ...degree in Electrical Engineering, Computer Engineering,...  ...shape the future of AI/ML hardware acceleration....  ...to drive cutting-edge TPU (Tensor Processing Unit...  ...Infrastructure at unparalleled scale, efficiency, reliability and... 
    Performance
    Worldwide

    Google

    Sunnyvale, CA
    23 hours ago
  • $138k - $198k

     ...degree in Electrical Engineering, Computer Engineering,...  ...in bringing up ASICs, performing functional and performance...  ...the future of AI/ML hardware acceleration....  ...to drive cutting-edge TPU (Tensor Processing Unit...  ...Infrastructure at unparalleled scale, efficiency, reliability and... 
    Performance
    Worldwide

    Google

    Sunnyvale, CA
    5 days ago
  • $138k - $197k

     ...and debugging tests and performing functional validation...  ...degree in Electrical Engineering, Computer Engineering,...  ...shape the future of AI/ML hardware acceleration....  ...to drive cutting-edge TPU (Tensor Processing...  ...Infrastructure at unparalleled scale, efficiency, reliability and... 
    Performance
    Worldwide

    Google

    Sunnyvale, CA
    5 days ago
  • $138k - $198k

     ...interacting with design engineers to identify...  ...methodologies to improve efficiency.Experience with...  ...shape the future of AI/ML hardware...  ...opportunity to drive TPU (Tensor Processing...  ...active projects and perform direct verification...  ...Infrastructure at unparalleled scale, efficiency,... 
    Worldwide

    Google

    Sunnyvale, CA
    4 days ago
  •  ...generalist robots and is seeking a Staff / Principal ML Training Systems Engineer to own training systems performance end-to-end. You will optimize large-scale multimodal training, define parallelism strategies, and drive efficiency across GPUs, memory, and compute. This is a... 
    Performance

    Rhoda AI

    Palo Alto, CA
    1 day ago
  • $183k - $247.6k

     ...world. As a member of the Cloud-Scale Machine Learning Acceleration...  ...in silicon yield & performance - it’s still Day One here at...  ...experienced Design Verification Engineers to build the next generation...  ...verifying complex CPU, GPU, or ML accelerator designsAmazon is... 
    Performance
    Local area
    Work from home
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  • $174k - $258k

    Staff Hardware Systems Design Engineer, AI Infrastructure, TPU Join to apply for the Staff Hardware Systems...  ..., delivering unparalleled performance, efficiency, and integration. Our team is on...  ...mission to redefine the future of at-scale computing by delivering a groundbreaking... 
    Performance
    Full time
    Worldwide

    Google

    Sunnyvale, CA
    1 day ago
  • $240k - $333k

     ..., compute, power, and performance.Execute custom silicon...  ...degree in Electrical Engineering, Computer Engineering,...  ...shape the future of AI/ML hardware acceleration....  ...to drive cutting-edge TPU (Tensor Processing Unit...  ...at unparalleled scale, efficiency, reliability and velocity... 
    Performance
    Contract work
    Worldwide

    Google

    Sunnyvale, CA
    23 hours ago
  •  ...production environments and scale solutions that elevate creative...  ..., increasing operational efficiency, and strengthening competitive...  ...stakeholders. Present pilot results, performance benchmarks, and scaling...  .... Familiarity with AI/ML frameworks (TensorFlow, PyTorch... 
    Performance

    Advanced Micro Devices , Inc.

    Santa Clara, CA
    5 days ago
  • $163k - $236k

     ...Silicon validation of TPU (Tensor Processing Unit - Google’s custom ML accelerator) design in...  ...framework.Develop test plan, perform pre-silicon validation...  ...s degree in Electrical Engineering, Computer Engineering,...  ...performance, efficiency, and integration.Google... 
    Performance
    Worldwide

    Google

    Mountain View, CA
    5 days ago
  •  ...GPUs. Our novel wafer‑scale architecture provides...  ...effortlessly run large‑scale ML applications, without...  ...are hiring a Senior Performance Analyst to join our...  ...Collaborate with Product and Engineering to identify where...  ...) with a systems or efficiency focus. Contributions... 
    Performance
    Contract work
    Shift work

    Cerebras

    Sunnyvale, CA
    5 days ago
  • $180k - $250k

    A leading AI infrastructure firm is seeking a TPU Systems Engineer to develop high-performance systems using JAX, XLA, and Pallas. This role involves pushing...  ...should have at least 3 years of experience in production ML systems and a strong foundation in Python. Competitive... 
    Performance

    RadixArk

    Palo Alto, CA
    4 days ago
  • $117k - $234k

     ...organization, designs, engineers, and operates the next...  ...at multi-petabyte scale.We are hiring a Staff...  ...improving reliability and performance, and enabling self-service...  ...infrastructure efficiency at massive scale.You will...  ...frameworks.Knowledge of AI/ML, AIOps, or advanced... 
    Performance
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    2 days ago
  • $174.9k - $261.3k

     ...world!The Data Labeling Engineering team designs, builds,...  ...engineering, and AI/ML, defining the strategies...  ...training data at scale. Our tools and platform...  ...and test scalable, high‑performance user experiences and services...  ...data quality (e.g., efficiency dashboards, auto‑QA,... 
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    1 day ago
  • $170.6k - $261.3k

     ...transportation on a global scale.As a Senior Machine Learning Engineer on the State Estimation...  ...develop and improve the ML perception model that powers...  ...to improve model performance against those metrics. Analyze...  ...sensor conditions. Implement efficient training and inference... 
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    5 days ago
  • $124k - $195.5k

     ...As the complexity and scale of artificial intelligence...  ...maximum hardware performance for emerging AI workloads...  ...network communication efficiency and programmability....  ...Computer Science, Computer Engineering, Electrical...  ...learning compilers and ML systems, including graph... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $184k - $287.5k

     ...skilled and motivated software engineers to join us and build AI...  ...inference systems that serve large-scale models with extreme efficiency. You’ll architect and implement high-performance inference stacks, optimize...  ...frontier for the field of ML Systems; survey recent... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $90.1k - $191.8k

     ...world!The Data Labeling Engineering team designs, builds,...  ...data engineering, and ML, defining labeling...  ...reliable training data at scale.Our team builds the...  ...and test scalable, high‑performance user experiences and services...  ...data quality (e.g., efficiency dashboards, auto‑QA,... 
    Performance
    Full time
    Work experience placement
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    1 day ago
  • $163k - $237k

     ...large form-factor package for ML high-performance computers (HPCs).Develop...  ...methodology and CAD flow for efficient substrate design and...  ...), thermal, and mechanical engineering teams to refine and optimize...  ...Infrastructure at unparalleled scale, efficiency, reliability and... 
    Performance
    Worldwide

    Google

    Sunnyvale, CA
    4 days ago
  • $120k - $275k

     ...train and run the largest ML workloads for AGI.About...  ...for a Hardware Systems Engineer to join our Hardware...  ...interconnects for our rack-scale AI data center platform...  ...role in driving power efficiency, scalability, reliability, and system performance at scaleWhat You’ll Do... 
    Performance
    Full time
    Work experience placement
    Local area
    Remote work
    Monday to Friday
    Flexible hours

    MatX

    Mountain View, CA
    2 days ago
  • $144.7k - $261.3k

    Job DescriptionThe Senior ML Validation Research Engineer will lead applied machine learning...  ...prototypes that enhance the efficiency, accuracy, and coverage of...  ...research concepts into performant tools integrated into CI/CD and large-scale validation pipelines.Advance... 
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Flexible hours

    General Motors

    Sunnyvale, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to TPU Performance Engineer — ML Efficiency & Scale. Be the first to apply!