Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Datacenter & Agentic AI Workload Performance Optimization Engineer

$100k

Tenstorrent

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities.Tenstorrent is looking for a Workload Performance Optimization Engineer to help optimize the software workloads that run on our next-generation RISC-V platforms. You’ll work across modern datacenter and agentic AI workloads—including Java, Python, PHP, Node.js, Lua, Go, and Rust—to identify performance bottlenecks and develop software, compiler, runtime, and hardware-aware optimizations that improve throughput, latency, and efficiency. This role sits at the intersection of software runtimes, compilers, CPU microarchitecture, and RISC-V silicon. You’ll bring up and tune major runtimes, profile real-world applications, investigate memory and concurrency behavior, and explore optimizations using RISC-V Vector/Matrix capabilities and custom instructions. You’ll also work with AI-assisted development and automated optimization workflows to accelerate the performance engineering process. Your work will directly influence both the RISC-V software ecosystem and the architecture of future Tenstorrent CPUs.This role isremote, based out of North America.We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting.Who You AreYou’re a performance engineer who enjoys getting deep into runtimes, compilers, applications, and CPU microarchitecture to understand why software is fast—or slow.You have hands-on experience optimizing software on RISC-V or another modern CPU architecture, with a strong understanding of the hardware/software boundary.You’re comfortable profiling complex systems, finding bottlenecks, forming hypotheses, and iterating through optimizations using data.You’re excited about emerging agentic AI development workflows and using AI tools to automate profiling, coding, benchmarking, and optimization.You’re a strong technical collaborator who can work across compiler, runtime, systems software, hardware, and performance modeling teams.What We NeedMaster’s or PhD in Computer Engineering, Electrical Engineering, Computer Science, or a related field, with strong experience in performance optimization, computer architecture, compilers, or systems software.Hands-on experience with runtime or compiler optimization, such as OpenJDK/JIT, LLVM, GCC, V8, Python, or equivalent systems.Strong understanding of CPU performance, memory hierarchies, concurrency, garbage collection, vector/SIMD optimization, and RISC-V architecture.Expertise with performance profiling and analysis tools such as Linux perf, runtime profilers, QEMU, tracing tools, and performance modeling environments.Strong programming skills in Java, Python, C/C++, and RISC-V assembly, with the ability to work effectively across multiple software layers.What You Will LearnHow to optimize modern software stacks from application and runtime all the way down to CPU microarchitecture and silicon.How RISC-V Vector, Matrix, and custom ISA capabilities can be used to accelerate real-world datacenter and AI workloads.How runtime, compiler, memory, and concurrency decisions impact performance at datacenter scale.How to build automated and AI-assisted performance optimization workflows that continuously profile, analyze, modify, and benchmark software.How to influence future CPU architecture by connecting real workload behavior and software optimization opportunities to hardware design decisions.Compensation for all engineers at Tenstorrent ranges from $100k - $500k including base and variable compensation targets. Experience, skills, education, background and location all impact the actual offer made.Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer.This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology. Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2). These requirements apply to persons located in the U.S. and all countries outside the U.S. As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency. If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.

Vacancy posted 22 hours ago
Similar jobs that could be interesting for youBased on the Datacenter & Agentic AI Workload Performance Optimization Engineer in Santa Clara, CA vacancy
  •  ...Technologist, Private Cloud AI - Applied & Agentic AIThis role has been...  ..., ensuring performant, reliable, and trustworthy...  ...and partner with engineering to take POCs into...  ...services. Troubleshoot and optimize AI systems at scale...  ...integrating AI workloads into production‑grade... 
    Performance
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start
    2 days per week

    Hewlett Packard Enterprise

    San Jose, CA
    3 days ago
  • $138k - $206k

     ...partners, and communities.Senior Engineer, Architecture & Performance Research Engineer for Data Center and Agentic AI CPUWhat You’ll...  ...performance for emerging computing workloads. We explore new...  ...performance bottlenecks and optimization opportunities across the front... 
    Performance
    Work at office
    Flexible hours
    Shift work

    Samsung Semiconductor

    San Jose, CA
    5 days ago
  •  ...computing experiences—from AI and data centers, to PCs...  ...looking for a Principal Engineer to serve as a hands-on...  ...team lead driving the performance and scalability of frontier AI workloads on AMD GPUs, including large...  ...challenges, from kernel optimization to framework-level... 
    Performance

    AMD

    San Jose, CA
    3 days ago
  • $152k - $241.5k

     ...unlimited potential of AI to define the next...  ...Developer Technology Engineer you will be at the forefront...  ...RTX & DGX SoC performance for agentic use. This role offers...  ...ecosystem partners, optimizing their applications to...  ...AI, and 3D graphics workloads, working with domain... 
    Performance
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    3 days ago
  • $138k - $206k

     ...Senior Physical Design AI/ML Engineer, Logic Pathfinding...  ...in predicting design performance, power, and area (PPA). Experience in optimization of design-technology...  ...explorationExpertise with agentic AI frameworks (...  ...state-of-the-art AI workloads and their compute and... 
    Performance
    Work at office
    Local area
    Flexible hours

    Samsung Semiconductor

    San Jose, CA
    4 days ago
  • $230k - $316.25k

     ...combines analog, digital, AI, and software technologies...  ...Role We’re looking for Deep Agentic Reasoning Engineer (open rank) to design and...  ...modifications, and optimization techniques Strong expertise...  ...qualifies for a discretionary performance-based bonus which is based... 
    Performance
    Permanent employment
    Full time
    Work at office
    Day shift

    Analog Devices

    San Jose, CA
    4 days ago
  • $100k

     ...the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use,...  ...is looking for a Workload Performance Analysis Engineer to help shape the performance...  ...RISC-V CPUs across modern datacenter and agentic AI workloads. In this role,... 
    Performance
    Permanent employment

    Tenstorrent

    Santa Clara, CA
    3 days ago
  •  ...discovery to powering AI and the technologies people...  ...seeking an AI Systems Engineer to join our AMD IT...  ...of High-Performance Computing (HPC) infrastructure...  ...GPU clusters, and AI workload schedulers. THE PERSON...  ...based clusters, ensuring optimal performance Administer... 
    Performance

    AMD

    San Jose, CA
    2 days ago
  • $152k - $241.5k

     ...unlimited potential of AI to define the next era...  ...a Developer Technology Engineer, you will be at the forefront...  ...to enable professional agentic AI workflows at the...  ...in suboptimal runtime performance.Conduct hands-on trainings...  ...deployment targeting optimal runtime performance.... 
    Performance
    Full time
    Local area

    Nvidia

    Santa Clara, CA
    5 days ago
  • $169k - $338k

     ...Home OfficeAI/ML & Agentic Systems Technical...  ...advanced agentic AI systems that can...  ...reliability engineering workflows, predictive...  ...analysis, and self-optimization across all...  ...capacity planning, and performance optimization...  ...supporting diverse workloads with varying reliability... 
    Performance
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    5 days ago
  •  ...the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use,...  ...is looking for a Workload Performance Analysis Engineer to help shape the performance...  ...RISC-V CPUs across modern datacenter and agentic AI workloads. In this role,... 
    Performance
    Full time
    Remote work

    Tenstorrent

    Santa Clara, CA
    9 days ago
  • $152k - $241.5k

     ...outstanding Senior Compiler Engineer to help build the...  ...of compilers, agentic systems, numerical...  ...about, generate, optimize, and validate code...  ..., system performance, and confidence in...  ...across GPU-centric workloads.Build the data, environments...  ...in compilers, AI systems, numerical... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ...NVIDIA Corporation in Santa Clara, CA, seeks a Sr. Inference Engineer to accelerate LLM inference through GPU kernel optimization. You will lead kernel benchmarking, model-level performance analysis, and AI-driven optimization workflows across silicon and software stacks... 
    Performance

    Jobleads-US

    Santa Clara, CA
    2 days ago
  • $195.2k - $315.49k

     ...highly technical AI Systems &...  ...generation of Intel datacenter platforms for AI workloads, including Generative AI and Agentic AI.Working...  ...customers and engineering teams, you will...  ...validation to evaluate performance, scalability,...  ..., workload optimizations, and platform enhancements... 
    Performance
    Full time
    Local area
    Immediate start
    Shift work

    Intel

    Santa Clara, CA
    22 hours ago
  • Staff Technologist-AI...  ...accelerate computational workloads, integrate across...  ...how work gets done. Engineers define intent, author...  ...characterization, performance modeling, and full...  ...models and autonomous agentic systems. You will...  ...model workloads to optimal hardware.Experience... 
    Performance

    Dell Technologies

    Santa Clara, CA
    1 day ago
  • $90k - $180k

     ...ResponsibilitiesAI Systems & Agentic WorkflowsBuild agentic AI services (planning,...  .../cuML/cuGraph) and optimize end‑to‑end performance. Use Ray (or similar)...  ...to design reviews and engineering best practices. Mentor...  ...accelerating data science workloads with RAPIDS, and... 
    Performance
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    3 days ago
  •  ...discovery to powering AI and the technologies people...  ...AI Training Systems & Performance Engineering PhD Intern/Co-Op to...  ...accelerate the adoption and optimization of cutting-edge AI training workloads on AMD Instinct™ GPUs...  ...LLM-powered agents and agentic AI techniques. You... 
    Performance
    Full time
    Summer work
    Internship
    Summer internship
    Worldwide

    AMD

    Santa Clara, CA
    5 hours ago
  •  ...Clara-based MTIA Software Team is hiring a Software Engineer, Systems ML specializing in compilers and kernels. You will help develop the AI compiler stack, contribute to PyTorch core components, and optimize high-performance kernels for next-generation hardware... 
    Performance

    Jobleads-US

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

     ...looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to push...  ...operation to its performance ceiling? Our LLM Inference...  ...performance projection tooling, and agentic optimization systems that...  ...optimization: applying AI-driven analysis to diagnose... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  • $184k - $287.5k

     ...the unlimited potential of AI to define the next era of computing...  ...an experienced Compiler Optimization Engineer for an exciting role in our...  ...range of computational workloads, ranging from deep learning...  ...to achieve the best performance of their applications. If this... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $152k - $241.5k

     ...the unlimited potential of AI to define the next era of...  ...computer graphics. As an engineer in our EDA Workflow Optimization team, you will partner...  ...chips in the world. You will perform investigations to...  ...experience running GPU-based workloads in a batch computing environment... 
    Performance
    Full time
    Worldwide

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

    We are looking for software engineers to contribute to the design and...  ...are revolutionizing AI, data analytics, and scientific...  ...centers powered by GPUs and high-performance linear algebra libraries. Applications...  ...in developing, debugging and optimizing high-performance software... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ...Tenstorrent is seeking a Datacenter & Agentic AI Workload Performance Analysis Engineer to shape CPU performance across datacenter and AI workloads. You will bridge hardware and software, profiling real-world workloads, and guiding architectural improvements with engineers... 
    Performance
    Remote work

    Jobleads-US

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

     ...unlimited potential of AI to define the next era...  ...an AI Compiler Engineer with deep expertise in...  ...measurable improvements in performance and efficiency, and advancing...  ...end-to-end compiler optimization workflows, from...  ...impact on representative workloads and benchmark suites.Lead... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...team is looking for a senior engineer to join our development efforts...  ...of kernel generation for AI and HPC, specifically targeting...  ...implementing high quality and performance numerical dense linear...  ...maintenance, and performance optimization of HPC software using C++.Strong... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    3 days ago
  • $320k

     ...As a Distinguished Engineer, Power...  ...architecture. This includes datacenter GPUs, GeForce and...  ...the broader edge-AI portfolio. This role...  ...class platforms.Drive performance-versus-power...  ...end.Drive ML- and workload-aware modeling and post-silicon optimization.Influence across the... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $136k - $218.5k

     ...Senior Power Architecture & Optimization Engineer to push the limits of...  ...using advanced analytics and AI, including LLMs trained specifically...  .... As AI and graphics workloads scale explosively, we believe...  ...closely with Architects, Performance, Software, ASIC Design, and... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...Senior Developer Technology Engineer, Artificial Intelligence!...  ...parallel algorithms to accelerate AI workloads on advanced computer...  ...achieve the best possible performance of computer hardware? Could...  ...perform in-depth analysis and optimization of complex AI and HPC... 
    Performance
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    3 days ago
  • $266.05k - $396k

     ...potential of their data, from AI to multicloud....  ...Summary Distinguished Engineer - AI Infrastructure...  ...experience with high-performance inference engines (...  ...Runtime, Triton), model optimization (quantization, pruning...  ...engines optimized for AI workloads (checkpoints, model... 
    Performance
    Work at office
    Local area

    NetApp

    San Jose, CA
    1 day ago
  • $184k - $287.5k

     ...globally. We seek a Senior Engineer to lead technical...  ...in deploying advanced AI agent frameworks and local...  ...efforts to optimize the agent runtimes for...  ...AI orchestration and agentic frameworks (e.g., OpenClaw...  ...particularly C++ (for performance-critical systems/OS integration... 
    Performance
    Full time
    Local area
    Shift work

    Nvidia

    Santa Clara, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Datacenter & Agentic AI Workload Performance Optimization Engineer. Be the first to apply!