Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Engineer: GPU Kernels & AI Performance

Gimlet Labs

A cutting-edge technology company in San Francisco is seeking a Member of Technical Staff focused on kernels and GPU performance. This role involves optimizing GPU and accelerator kernels for AI workloads by analyzing performance across various hardware. Ideal candidates have strong software engineering foundations and experience with performance-critical systems. Familiarity with tools like CUDA and performance profiling is preferred. This position offers a dynamic environment focused on real-world performance optimizations. #J-18808-Ljbffr Gimlet Labs

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Staff Engineer: GPU Kernels & AI Performance in San Francisco, CA vacancy
  • $150k - $350k

     ...is seeking a Member of Technical Staff focused on optimizing GPU and accelerator kernels for AI workloads. This role involves analyzing and tuning performance across diverse execution platforms...  ...a strong foundation in software engineering and experience with performance-critical... 
    Performance

    Gimlet Labs, Inc.

    San Francisco, CA
    4 days ago
  • A leading data and AI company in San Francisco seeks a Staff Software Engineer to lead kernel-level performance engineering for GenAI workloads. The role involves designing and optimizing high-performance GPU kernels, mentoring engineers, and driving performance roadmaps... 
    Performance

    Databricks

    San Francisco, CA
    17 hours ago
  • $250k - $300k

     ...vertically integrated AI infrastructure company...  ...and be part of a high-performing team that believes in each...  ...direction for Crusoe's Linux kernel team, owning the...  ...mentoring and growing the engineers around you to deliver...  ...technologies tailored for GPU-accelerated AI and HPC... 
    Performance
    Temporary work

    Crusoe

    San Francisco, CA
    1 day ago
  •  ...processing down to the lowest layers of the stack, optimize kernel performance, develop new request scheduling and parallelism strategies, and...  ...and implement exotic parallelism schemes, write custom GPU kernels for regimes like cascade attention, and understand every... 
    Performance

    Sail

    San Francisco, CA
    17 hours ago
  • B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role involves designing and operating high-performance systems for model inference, synthetic data generation...  ...a passion for working in cutting-edge AI. Benefits include top-tier compensation... 
    Performance

    B Capital

    San Francisco, CA
    2 days ago
  • $220k - $280k

     ...the Role Together AI is building the best...  ...We're looking for a Staff ML Engineer to drive the model serving...  .... You'll profile GPU utilization, design batching...  ...-in-class inference performance — architect and implement...  ...analysis from GPU kernel behavior to framework-... 
    Performance
    Full time

    Together Ai

    San Francisco, CA
    17 hours ago
  •  ...the next generation of AI-driven game experiences...  ...Senior Machine Learning Engineer for On-Device & Mobile...  ...export, quantization, and kernel-level tuning, to a...  ...tuning across NPU, mobile GPU, and desktop/laptop GPU...  ...bars.Do low-level performance work: write and tune WebGPU... 
    Performance
    Full time
    Work at office
    Remote work
    Worldwide

    Unity Technologies

    San Francisco, CA
    17 hours ago
  •  ...intelligence open and accessible. We design, build, and operate GPU-heavy infrastructure for high-throughput model inference...  .... Expect collaboration with research teams and a focus on performance tuning, kernel optimization, and scalable distributed systems. #J-18808-... 
    Performance

    Visa Hunt

    San Francisco, CA
    2 days ago
  •  ...Tensor in San Francisco is building the fastest GPU compiler and a suite of AI-driven tools. We’re hiring a Member of Technical Staff for AI-Driven Compilation to own the search,...  ...systems, integrate with compiler and kernel teams, and ship learned components into production... 
    Relocation package

    SF Tensor

    San Francisco, CA
    3 days ago
  • A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU kernels and using advanced techniques to boost computation efficiency. Ideal... 
    Performance

    Baseten

    San Francisco, CA
    1 day ago
  • TypeSafe AI in San Francisco is seeking a GPU kernel engineer with deep CUDA expertise to accelerate our training and inference workloads. You will write, optimize, and maintain high-performance kernels close to the metal. Work alongside research and platform engineers... 
    Performance
    Work at office

    TypeSafe AI

    San Francisco, CA
    3 days ago
  •  ...for the world's most dynamic AI companies, like Cursor, Notion...  ...and help build the platform engineers turn to to ship AI products....  ...THE ROLE We’re seeking a GPU Kernel Engineer to join our team at...  ...your code directly impacts the performance of state-of-the-art machine... 
    Performance
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    17 hours ago
  • Sciforium is seeking a GPU Kernel Engineer to push performance on modern accelerators. You will design and optimize custom GPU kernels, from low-level development to integrating ops in ML frameworks used for large-scale training and inference. Ideal candidates have 5+ years... 
    Performance

    Sciforium

    San Francisco, CA
    1 day ago
  • $150k - $350k

     ...heterogeneous neocloud for AI workloads. As AI systems...  ...hardware that best fits its performance and efficiency needs. This...  ...a Member of Technical Staff focused on kernels and GPU performance. In this role,...  .... This role is ideal for engineers who enjoy deep performance... 
    Performance

    Gimlet Labs, Inc.

    San Francisco, CA
    4 days ago
  • $173.5k - $331.05k

     ...on a mission to build the modern, AI-powered video rendering pipeline behind...  ...looking for a senior, hands-on engineer to own and evolve the cross-platform GPU rendering platform at the heart of...  ...on Windows and Mac Own rendering performance, stability, and device... 
    Performance
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    17 hours ago
  • $250k - $300k

     ...vertically integrated AI infrastructure...  ...be part of a high-performing team that believes...  ...the Role:As a Senior Staff/Principal Deployment Automation Engineer for the Compute Team...  ...-scale, multi-node GPU clusters. You will...  ...Knowledge of Linux kernel internals, specifically... 
    Performance
    Temporary work

    Crusoe

    San Francisco, CA
    4 days ago
  • $155k - $269k

     ...Description Waabi, founded by AI visionary Raquel Urtasun, is...  ...team of Research Scientists and Engineers building the content backbone...  ...designing, launching, and debugging GPU jobs in the cloud (AWS, GCP,...  ...awards and an annual performance bonus. Perks/Benefits: -... 
    Performance
    Full time
    Work at office
    Work from home
    Flexible hours

    Waabi

    San Francisco, CA
    9 days ago
  • $225k - $275k

     ...the future of high-performance compute We firmly believe...  ...that the future of AI depends on the...  ...are building our Kernel Optimizer, which takes...  ...for researchers, engineers and organizations who...  ...build the fastest GPU compiler in the...  ...Member of Technical Staff for Product Engineering... 
    Performance
    16 hours
    Full time
    Live in
    Work at office
    Relocation package
    Night shift

    San Francisco Tensor Company

    San Francisco, CA
    1 day ago
  • $172.5k - $306.63k

     ...is the new family of creative generative AI models coming to Adobe products that offers...  ...frameworks leveraging GPUs to improve performance and scalability. Improve resiliency, elasticity...  ..., orchestration, and management of GPU resources  Experience with machine learning... 
    Performance
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    3 days ago
  • $197.3k - $313.7k

     ...SalesforceSalesforce is the #1 AI CRM, where humans with...  ...OPPORTUNITIES*Slack is looking for a Staff Machine Learning Engineer with deep expertise in...  ...training pipelines on GPU infrastructure.Brainstorm with...  ..., assessment of job performance, discipline, termination, and... 
    Performance
    Full time

    Salesforce

    San Francisco, CA
    3 days ago
  •  ...San Francisco is hiring for a Member of Technical Staff for Sandbox Infrastructure to build a serverless GPU container service across multiple vendors. You...  ...migration with socket preservation. This is a high-performance compute role requiring deep systems expertise. You... 
    Performance

    SF Tensor

    San Francisco, CA
    3 days ago
  • Causal Labs is building a Large Physics foundation Model and GPU-driven compute environment to enable rapid research iteration at...  ...provisioning to observability, collaborating with researchers to optimize performance and placement. A strong systems background and experience with... 
    Performance

    Causal Labs

    San Francisco, CA
    3 days ago
  • $190.2k - $345.65k

     ...custom multimedia generative AI — deep-tuned image, video,...  ....We are hiring a Senior Staff Machine Learning Engineer to architect and lead the data...  ...production.Own the performance and cost envelope of the platform...  ...throughput SLAs, ANN index tuning, GPU-accelerated enrichment (... 
    Performance
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    4 days ago
  • VSCO is seeking a Senior Staff Engineer to own Reflex, our production GPU image-processing engine, and guide its path into Studio Pro on iOS and macOS. You will shape a high-performance imaging stack, drive GPU and memory budgets, and collaborate with Product/Design to... 
    Performance

    VSCO

    San Francisco, CA
    17 hours ago
  •  ...Francisco is hiring Members of Technical Staff to build systems that accelerate LLM...  ...end to end. You will work on high-performance kernels, inference engine internals, and production infrastructure...  ...collaborating with a fast-growing AI inference company. #J-18808-Ljbffr Simplify
    Performance

    Simplify

    San Francisco, CA
    4 days ago
  • $286.2k - $326.7k

     ...Overview Senior Staff Engineer, AI Compute (Remote Eligible) At Capital One, we are creating...  ...product experiences and scalable, high-performance AI infrastructure. At Capital One, you...  ...compute infrastructure on top of CPU and GPU substrates. Your contributions will... 
    Performance
    Full time
    Part time
    Local area
    Remote work

    Capital One

    San Francisco, CA
    6 days ago
  • $220k

    We build and run the inference engine behind every Perplexity query and deploy dozens...  ...management to support in API Gateway. GPU kernels migration to CuTe DSL. Port our in-house...  ...keep up with rapidly growing traffic. Performance optimisation. Profile and fix... 
    Performance

    Perplexity

    San Francisco, CA
    1 day ago
  • $285k - $315k

     ...the future of high-performance compute We firmly believe...  ...that the future of AI depends on the...  ...are building our Kernel Optimizer, which takes...  ...for researchers, engineers and organizations who...  ...build the fastest GPU compiler in the...  ...Member of Technical Staff for GPU Kernel Engineering... 
    Performance
    Full time
    Work at office
    Immediate start
    Relocation package

    San Francisco Tensor Company

    San Francisco, CA
    1 day ago
  • $350k

     ...Join a rapidly growing AI infrastructure...  ...across global cloud and GPU environments. The organization...  ...build reliable, high-performance platforms supporting...  ...opportunity is for a Staff Site Reliability Engineer to lead the...  ...expertise, including kernel tuning, CUDA lifecycle... 
    Performance
    Full time
    San Francisco, CA
    a month ago
  • Magic AI, Inc. is seeking a engineer for the Supercomputing Platform & Infrastructure to design, build, and operate large-scale GPU infrastructure powering model training and inference. You will implement Terraform-driven IaC across cloud and hybrid environments, manage... 
    Visa sponsorship
    Relocation package

    Magic AI Corp.

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Engineer: GPU Kernels & AI Performance. Be the first to apply!