Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Engineer: GPU Kernels & AI Performance

Gimlet Labs

A cutting-edge technology company in San Francisco is seeking a Member of Technical Staff focused on kernels and GPU performance. This role involves optimizing GPU and accelerator kernels for AI workloads by analyzing performance across various hardware. Ideal candidates have strong software engineering foundations and experience with performance-critical systems. Familiarity with tools like CUDA and performance profiling is preferred. This position offers a dynamic environment focused on real-world performance optimizations. #J-18808-Ljbffr Gimlet Labs

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff Engineer: GPU Kernels & AI Performance in San Francisco, CA vacancy
  • $250k - $300k

     ...vertically integrated AI infrastructure company...  ...and be part of a high-performing team that believes in each...  ...direction for Crusoe's Linux kernel team, owning the...  ...mentoring and growing the engineers around you to deliver...  ...technologies tailored for GPU-accelerated AI and HPC... 
    Performance
    Temporary work

    Crusoe

    San Francisco, CA
    1 day ago
  • $150k - $350k

     ...is seeking a Member of Technical Staff focused on optimizing GPU and accelerator kernels for AI workloads. This role involves analyzing and tuning performance across diverse execution platforms...  ...a strong foundation in software engineering and experience with performance-critical... 
    Performance

    Gimlet Labs, Inc.

    San Francisco, CA
    2 days ago
  •  ...processing down to the lowest layers of the stack, optimize kernel performance, develop new request scheduling and parallelism strategies, and...  ...and implement exotic parallelism schemes, write custom GPU kernels for regimes like cascade attention, and understand every... 
    Performance

    SAIL

    San Francisco, CA
    3 days ago
  • B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role involves designing and operating high-performance systems for model inference, synthetic data generation...  ...a passion for working in cutting-edge AI. Benefits include top-tier compensation... 
    Performance

    B Capital

    San Francisco, CA
    15 hours ago
  • $220k - $280k

     ...the Role Together AI is building the best...  ...We're looking for a Staff ML Engineer to drive the model serving...  .... You'll profile GPU utilization, design batching...  ...-in-class inference performance — architect and implement...  ...analysis from GPU kernel behavior to framework-... 
    Performance
    Full time

    Together Ai

    San Francisco, CA
    15 hours ago
  •  ...the next generation of AI-driven game experiences...  ...Senior Machine Learning Engineer for On-Device & Mobile...  ...export, quantization, and kernel-level tuning, to a...  ...tuning across NPU, mobile GPU, and desktop/laptop GPU...  ...bars.Do low-level performance work: write and tune WebGPU... 
    Performance
    Full time
    Work at office
    Remote work
    Worldwide

    Unity Technologies

    San Francisco, CA
    3 days ago
  • A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU kernels and using advanced techniques to boost computation efficiency. Ideal... 
    Performance

    Baseten

    San Francisco, CA
    4 days ago
  • TypeSafe AI in San Francisco is seeking a GPU kernel engineer with deep CUDA expertise to accelerate our training and inference workloads. You will write, optimize, and maintain high-performance kernels close to the metal. Work alongside research and platform engineers... 
    Performance
    Work at office

    TypeSafe AI

    San Francisco, CA
    1 day ago
  •  ...for the world's most dynamic AI companies, like Cursor, Notion...  ...and help build the platform engineers turn to to ship AI products....  ...THE ROLE We’re seeking a GPU Kernel Engineer to join our team at...  ...your code directly impacts the performance of state-of-the-art machine... 
    Performance
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    15 hours ago
  •  ...heterogeneous neocloud for AI workloads. As AI systems...  ...hardware that best fits its performance and efficiency needs. This...  ...a Member of Technical Staff focused on kernels and GPU performance. In this role,...  ...hardware. This role is ideal for engineers who enjoy deep performance... 
    Performance

    Gimlet Labs

    San Francisco, CA
    2 days ago
  • $173.5k - $331.05k

     ...on a mission to build the modern, AI-powered video rendering pipeline behind...  ...looking for a senior, hands-on engineer to own and evolve the cross-platform GPU rendering platform at the heart of...  ...on Windows and Mac Own rendering performance, stability, and device... 
    Performance
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    3 days ago
  • $250k - $300k

     ...vertically integrated AI infrastructure...  ...be part of a high-performing team that believes...  ...the Role:As a Senior Staff/Principal Deployment Automation Engineer for the Compute Team...  ...-scale, multi-node GPU clusters. You will...  ...Knowledge of Linux kernel internals, specifically... 
    Performance
    Temporary work

    Crusoe

    San Francisco, CA
    1 day ago
  • $197.3k - $313.7k

     ...SalesforceSalesforce is the #1 AI CRM, where humans with...  ...OPPORTUNITIES*Slack is looking for a Staff Machine Learning Engineer with deep expertise in...  ...training pipelines on GPU infrastructure.Brainstorm with...  ..., assessment of job performance, discipline, termination, and... 
    Performance
    Full time

    Salesforce

    San Francisco, CA
    1 day ago
  • $350k

     ...Join a rapidly growing AI infrastructure...  ...across global cloud and GPU environments. The organization...  ...build reliable, high-performance platforms supporting...  ...opportunity is for a Staff Site Reliability Engineer to lead the...  ...expertise, including kernel tuning, CUDA lifecycle... 
    Performance
    Full time
    San Francisco, CA
    a month ago
  • $225k - $275k

     ...only vertically integrated AI infrastructure company built...  ..., and be part of a high-performing team that believes in each...  ...Cloud is seeking a Senior Staff Network Deployment Engineer to serve as the technical owner...  ...performance compute (HPC) and GPU-based AI infrastructure,... 
    Performance
    Temporary work
    Remote work

    Crusoe

    San Francisco, CA
    15 hours ago
  • Causal Labs is building a Large Physics foundation Model and GPU-driven compute environment to enable rapid research iteration at...  ...provisioning to observability, collaborating with researchers to optimize performance and placement. A strong systems background and experience with... 
    Performance

    Causal Labs

    San Francisco, CA
    1 day ago
  •  ...Francisco is hiring Members of Technical Staff to build systems that accelerate LLM...  ...end to end. You will work on high-performance kernels, inference engine internals, and production infrastructure...  ...collaborating with a fast-growing AI inference company. #J-18808-Ljbffr Simplify
    Performance

    Simplify

    San Francisco, CA
    2 days ago
  • Together AI in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms, architectures...  ...engines, including kernel backends and ATLAS-...  ...while profiling across GPU, networking, and memory...  ...pipelines, drive performance improvements, and mentor... 
    Performance

    Together

    San Francisco, CA
    3 days ago
  • $220k

    We build and run the inference engine behind every Perplexity query and deploy dozens...  ...management to support in API Gateway. GPU kernels migration to CuTe DSL. Port our in-house...  ...keep up with rapidly growing traffic. Performance optimisation. Profile and fix... 
    Performance

    Perplexity

    San Francisco, CA
    4 days ago
  •  ...Description Job Description Staff Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home...  ...trainable, deployable, observable, and performant under real-world constraints....  ...production deployment, including GPU optimization, memory efficiency,... 
    Performance
    Remote work
    Work from home

    Ginas Tech Jobs

    San Francisco, CA
    27 days ago
  •  ...Physics foundation Model and seeks an infrastructure engineer to design, deploy, and operate large distributed GPU clusters. You will extend schedulers, build self-...  ...will collaborate with researchers to optimize performance and placement, ensuring reliable, scalable... 
    Performance

    causal

    San Francisco, CA
    4 days ago
  •  ...The role: SoFi’s Staff AI Engineer is a hands-on AI engineering role in SoFi’s growing independent...  ...techniques to maximize model performance and minimize operational costs across the...  ...and managing the underlying Kubernetes/GPU orchestration for custom model deployments... 
    Performance
    Full time

    Sofi

    San Francisco, CA
    15 hours ago
  • $273k - $345k

     ...that. Atoms builds Physical AI- real-world robots for the...  ...scale. We are roboticists, engineers, operators, and builders. We believe...  ...and identifying systemic performance bottlenecks. What we're looking...  ...identify and eliminate CPU, GPU, and memory bandwidth bottlenecks... 
    Performance
    Full time
    Internship
    Work at office
    Flexible hours

    Atoms

    San Francisco, CA
    1 day ago
  • A tech-focused company in San Francisco seeks candidates with expertise in AI simulation development. The role emphasizes optimizing training efficiency, enhancing GPU performance, and ensuring low-latency inference. Applicants should be proficient in methodologies for... 
    Performance

    Embedding VC

    San Francisco, CA
    2 days ago
  • $193k - $234k

     ...only vertically integrated AI infrastructure company built...  ..., and be part of a high-performing team that believes in each...  ...high-energy, detail-oriented Staff Network Production Engineer to lead the physical and logical...  ...compute (HPC) and GPU-based AI infrastructure, you... 
    Performance
    Temporary work
    Remote work

    Crusoe

    San Francisco, CA
    2 days ago
  • Wafer is building AI-powered GPU optimization systems and is seeking engineers to join a small, highly collaborative team. You will work directly with the founders to implement the agent framework, profiling, and compiler tooling that power our GPU optimization platform... 

    Wafer

    San Francisco, CA
    1 day ago
  • $224k - $284k

     ...changing that. Atoms builds Physical AI — real-world robots for the...  ...work at scale. We are roboticists, engineers, operators, and builders. We believe...  ...team who will design and run the high-performance network that connects our GPU compute — building a fast, reliable... 
    Performance
    Full time
    Work at office
    Immediate start
    Flexible hours

    Atoms

    San Francisco, CA
    15 hours ago
  • $250k - $300k

     ...the only vertically integrated AI infrastructure company built...  ..., and be part of a high-performing team that believes in each other...  ...work directly with customer engineering teams to tailor deployments to...  ...vLLM and SGLang to the CUDA kernels underneath, profiling and running... 
    Performance
    Temporary work

    Crusoe

    San Francisco, CA
    15 hours ago
  • $141k - $249k

     ...Description Waabi, founded by AI visionary Raquel...  ...autonomy and algorithm engineers to scale safe self-...  ...and benchmark new CUDA kernels for inference. - Comprehensively...  ...and memory to pinpoint performance bottlenecks....  ...Skilled in profiling CPU and GPU code using tools such... 
    Performance
    Work at office
    Work from home
    Flexible hours

    Waabi

    San Francisco, CA
    2 days ago
  • $155k - $269k

     ...Description Waabi, founded by AI visionary Raquel Urtasun, is...  ...team of Research Scientists and Engineers building the content backbone...  ...designing, launching, and debugging GPU jobs in the cloud (AWS, GCP,...  ...awards and an annual performance bonus. Perks/Benefits: -... 
    Performance
    Full time
    Work at office
    Work from home
    Flexible hours

    Waabi

    San Francisco, CA
    15 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Engineer: GPU Kernels & AI Performance. Be the first to apply!