Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Performance Engineer, GPU

$315k

Anthropic

Performance Engineer, GPU Join to apply for the Performance Engineer, GPU role at Anthropic . About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About The Role Pioneering the next generation of AI requires breakthrough innovations in GPU performance and systems engineering. As a GPU Performance Engineer, you'll architect and implement the foundational systems that power Claude and push the frontiers of what's possible with large language models. You'll be responsible for maximizing GPU utilization and performance at unprecedented scale, developing cutting‑edge optimizations that directly enable new model capabilities and dramatically improve inference efficiency. Working at the intersection of hardware and software, you'll implement state‑of‑the‑art techniques from custom kernel development to distributed system architectures. Your work will span the entire stack—from low‑level tensor core optimizations to orchestrating thousands of GPUs in perfect synchronization. Strong candidates will have a track record of delivering transformative GPU performance improvements in production ML systems and will be excited to shape the future of AI infrastructure alongside world‑class researchers and engineers. You Might Be a Good Fit If You Have deep experience with GPU programming and optimization at scale Are impact‑driven, passionate about delivering measurable performance breakthroughs Can navigate complex systems from hardware interfaces to high‑level ML frameworks Enjoy collaborative problem‑solving and pair programming Want to work on state‑of‑the‑art language models with real‑world impact Care about the societal impacts of your work Thrive in ambiguous environments where you define the path forward Strong Candidates May Also Have Experience With GPU Kernel Development: CUDA, Triton, CUTLASS, Flash Attention, tensor core optimization ML Compilers & Frameworks: PyTorch/JAX internals, torch.compile, XLA, custom operators Performance Engineering: Kernel fusion, memory bandwidth optimization, profiling with Nsight Distributed Systems: NCCL, NVLink, collective communication, model parallelism Low‑Precision: INT8/FP8 quantization, mixed‑precision techniques Production Systems: Large‑scale training infrastructure, fault tolerance, cluster orchestration Representative Projects Co‑design attention mechanisms and algorithms for next‑generation hardware architectures Develop custom kernels for emerging quantization formats and mixed‑precision techniques Design distributed communication strategies for multi‑node GPU clusters Optimize end‑to‑end training and inference pipelines for frontier language models Build performance modeling frameworks to predict and optimize GPU utilization Implement kernel fusion strategies to minimize memory bandwidth bottlenecks Create resilient systems for planet‑scale distributed training infrastructure Profile and eliminate performance bottlenecks in production serving infrastructure Partner with hardware vendors to influence future accelerator capabilities and software stacks Deadline to apply None. Applications will be reviewed on a rolling basis. The Expected Salary Range For This Position Is The expected base compensation for this position is below. Our total compensation package for full‑time employees includes equity, benefits, and may include incentive compensation. Annual Salary

$315,000—$560,000 USD

Logistics Education requirements: We require at least a Bachelor's degree in a related field or equivalent experience. Location‑based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren’t able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How We're Different We believe that the highest‑impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large‑scale research efforts. And we value impact — advancing our long‑term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest‑impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT‑3, Circuit‑Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process. Seniority Level Mid‑Senior level Employment Type Full‑time Job Function Engineering and Information Technology Industries Research Services #J-18808-Ljbffr Anthropic

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Performance Engineer, GPU in San Francisco, CA vacancy
  •  ...AGI. Join us in shaping the future of AI and pushing the boundaries of what's possible in video generation.We're seeking a GPU Performance Engineer to squeeze every last FLOP from our H100 infrastructure and optimize our model serving stack to its absolute limits.The... 
    Performance

    Genmo

    San Francisco, CA
    3 days ago
  •  ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by one...  .... Every day, our work helps care teams perform with greater precision and patients...  ...Function of PositionAs a Senior Systems GPU Engineer - AI & Robotics, you will be responsible... 
    Performance
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    San Francisco, CA
    3 days ago
  • $315k

    A leading AI research company in San Francisco is seeking a mid-senior GPU Performance Engineer. In this role, you'll architect systems that enhance GPU performance for groundbreaking AI models. Responsibilities include developing optimizations, collaborating with teams... 
    Performance

    Anthropic

    San Francisco, CA
    4 days ago
  • A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU kernels and using advanced techniques to boost computation efficiency. Ideal... 
    Performance

    Baseten

    San Francisco, CA
    5 days ago
  • $220k - $320k

    inference.net, a growing company in San Francisco, seeks an experienced engineer to optimize AI inference performance. The ideal candidate will have over 2 years of experience in ML systems and GPU programming. Key responsibilities include implementing optimization techniques... 
    Performance

    inference.net

    San Francisco, CA
    5 days ago
  • MakerMaker.AI in San Francisco is seeking a skilled Software Engineer to write and optimize GPU kernels. You will work on deep low-level tasks that directly impact the performance of machine learning models. The ideal candidate has over 4 years of experience with GPU kernels... 
    Performance

    MakerMaker.AI

    San Francisco, CA
    2 days ago
  • $285k - $315k

    SF Tensor is looking for a Founding GPU Kernel Engineer in San Francisco, specializing in GPU architecture and kernel optimization for machine...  ...has deep expertise, proven capabilities in hand-optimizing performance-critical kernels, and strong programming skills in C++ and... 
    Performance
    Full time
    Relocation package

    SF Tensor

    San Francisco, CA
    1 day ago
  • San Francisco Tensor Company is seeking a Founding GPU Kernel Engineer to enhance GPU performance for AI applications. You will optimize and write kernels while collaborating with compiler teams to improve efficiencies across architectures. The ideal candidate has deep... 
    Performance
    Work at office
    Relocation package

    San Francisco Tensor Company

    San Francisco, CA
    2 days ago
  •  ...seeks a Member of Technical Staff, Infrastructure to own the end-to-end cloud infrastructure for a high-performance compression API. You will build global, low-latency GPU ML inference systems in the critical path of customer traffic, driving reliability, scalability, and... 
    Performance

    Jack & Jill

    San Francisco, CA
    2 days ago
  • Adobe is seeking a Graphics/Engine Software Engineer to help rebuild Photoshop's GPU-first core engine (NGE) in California. You will design, build and optimize...  ...workflows. You will also write shaders, profile performance, and contribute to testing, reviews and documentation... 
    Performance

    Adobe Inc.

    San Francisco, CA
    5 days ago
  • Mirai Labs in San Francisco seeks engineers to join a senior team building the full on-device stack for real-time local intelligence...  ...modern language models work, and experience in writing high-performance GPU kernels or Rust systems programming. We welcome applications... 
    Performance
    Local area

    Mirai Labs

    San Francisco, CA
    4 days ago
  • $100k - $120k

     ...cheaper and faster. Responsibilities Lead a team of kernel and system engineers focused on performance-critical code Design, implement, and optimize custom compute kernels for CPU (AVX/ARM NEON), GPU (CUDA/ROCm), and hardware accelerators Find bottlenecks in memory hierarchy... 
    Performance

    Coda Robotics

    San Francisco, CA
    2 days ago
  •  ...reality. You would collaborate with software engineers, AI researchers, and hardware specialists to develop high-performance solutions that meet the stringent requirements...  ...mobility. Key Responsibilities Optimize end-to-end GPU performance for real-time autonomous driving... 
    Performance

    Bot Auto

    San Francisco, CA
    4 days ago
  • $285k - $315k

    About The Role We're looking for a Founding GPU Kernel Engineer who lives right at the boundary between hardware and software. Someone who...  ...workloads (matmuls, attention, normalization, etc.) to set the performance ceilings Profile at the microarchitectural level: look... 
    Performance
    Full time
    Work at office
    Relocation package

    SF Tensor

    San Francisco, CA
    1 day ago
  • $180k - $280k

     ...tier investors. Since mid-2024, we've been engineering the foundation for what comes after the...  .... About the role We're looking for a GPU kernel engineer with deep, low-level CUDA...  ...include: Write, optimize, and maintain high-performance GPU kernels (e.g., in CUDA / CuTe DSL)... 
    Performance
    Work at office
    Visa sponsorship
    Shift work

    TypeSafe AI

    San Francisco, CA
    2 days ago
  • $285k - $315k

     ...Company, we believe the future of AI and high-performance computing depends on rethinking the...  .... We are partnering with researchers, engineers, and organizations who share our belief...  ...About the Role We're hiring a Founding GPU Compiler Engineer to build the core compilation... 
    Performance
    Full time
    Work at office
    Relocation package

    San Francisco Tensor Company

    San Francisco, CA
    2 days ago
  • $167.2k - $209k

     ...world. DigitalOcean is seeking a Senior Engineer 2 to play a key technical role in our AI...  ...we can offer the industry-leading performance for our inference services. You will be...  ...optimizations at the inference engine and GPU kernel layers, ensuring our infrastructure... 
    Performance
    Local area
    Remote work
    Worldwide
    Flexible hours

    DigitalOcean

    San Francisco, CA
    1 day ago
  •  ...infrastructure. You will design, deploy, and operate large-scale GPU clusters powering training, evaluation, and serving for the...  ...ensuring reliability with observability. You'll work with researchers to optimize performance and placement. #J-18808-Ljbffr Linuxcareers
    Performance

    Linuxcareers

    San Francisco, CA
    1 day ago
  • Vast.ai Inc. is seeking a systems engineer with HPC or parallel programming experience to...  ...inference. You will design and optimize GPU kernels and tensor libraries, leveraging...  ...frameworks to push the bleeding edge of AI performance. This role is based on-site in San... 
    Performance

    Vast.ai Inc.

    San Francisco, CA
    3 days ago
  • Vast.ai is seeking a systems engineer to scale AI inference and optimize GPU performance at our San Francisco or Los Angeles offices. You will leverage your HPC background to push the bleeding edge of AI, working with CUDA/C++ and a modern tech stack. Ideal candidates have... 
    Performance
    Full time

    Vast.ai

    San Francisco, CA
    3 days ago
  •  ...Conviction. Join us and help build the platform engineers turn to to ship AI products. At...  ...for foundational engineers to lead our GPU Networking efforts, making RDMA a first-...  ...characterize and validate networking performance on bleeding-edge clusters (H100/H200, B2... 
    Performance
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    1 day ago
  •  ...and leadership is earned by shipping excellence. We seek engineers with strong intrinsic drive, a true passion for...  ...scale AI inference. You’ll leverage your knowledge of high-performance systems to optimize GPU performance at the bleeding edge of AI. Full-Time On-site... 
    Performance
    Full time
    Work at office

    Vast.ai Inc.

    San Francisco, CA
    3 days ago
  • CoreWeave is seeking a Bare Metal Support Engineer in San Francisco, CA, to ensure high performance and reliability of our GPU infrastructure. You will engage directly with customers and collaborate with engineering teams to resolve issues and improve our cloud services... 
    Performance

    CoreWeave

    San Francisco, CA
    5 days ago
  • $135.2k - $306.4k

    Job Overview Oracle hardware platform development engineering is seeking a highly driven GPU/CPU Platform System Engineer at the Principal Engineer level...  ...development, design reviews, system integration, performance testing and characterization. You will interact closely... 
    Performance
    Temporary work
    Work experience placement
    Remote work
    Flexible hours

    Oracle

    San Francisco, CA
    1 day ago
  • A technology startup is seeking a Founding Engineer (Systems + ML) to develop GPU-accelerated engines and build end-to-end pipelines. The ideal candidate...  ...GPU code, alongside a deep understanding of systems performance. You'll join as the first technical hire, expecting... 
    Performance
    Full time

    Partcl

    San Francisco, CA
    3 days ago
  • $170k - $250k

     ...over $500K in revenue within six months and is scaling rapidly with a small, high-performing team. This company is looking for an engineer to work directly with the CTO on complex GPU virtualisation challenges. The role offers hands-on involvement with a production... 
    Performance
    Full time
    Visa sponsorship
    Flexible hours
    San Francisco, CA
    8 days ago
  • A leading consulting firm is seeking a Software Engineer (C++ Systems) in San Francisco to optimize microsecond-level performance in GPU virtualization software. Ideal candidates will have elite C++ expertise, with at least 2 years of experience in low-level systems engineering... 
    Performance

    SK HR Consultants.com

    San Francisco, CA
    2 days ago
  • An innovative company is seeking a talented software engineer to join their dynamic Inference team. This role involves designing and...  ...infrastructure for large-scale multimodal models, focusing on high-performance delivery of audio and image inputs. You'll collaborate closely... 
    Performance

    Jobleads-US

    San Francisco, CA
    5 days ago
  • $250k

     ...infrastructure provider building a next-generation GPU platform designed for AI training,...  ...for a Senior / Staff Site Reliability Engineer to support and scale large-scale HPC and...  ...the reliability, scalability, and performance of HPC and cloud infrastructure environments... 
    Performance
    Full time
    Remote work
    San Francisco, CA
    more than 2 months ago
  • nineDots.io is hiring a Senior Software Engineer to build secure GPU sandbox environments and scalable GPU compute platforms. You will help define...  ...with opportunities to influence software foundations and performance-critical paths in a fast-growing AI infrastructure... 
    Performance

    nineDots.io

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Performance Engineer, GPU. Be the first to apply!