Staff Engineer: GPU Kernels & AI Performance
Gimlet Labs
A cutting-edge technology company in San Francisco is seeking a Member of Technical Staff focused on kernels and GPU performance. This role involves optimizing GPU and accelerator kernels for AI workloads by analyzing performance across various hardware. Ideal candidates have strong software engineering foundations and experience with performance-critical systems. Familiarity with tools like CUDA and performance profiling is preferred. This position offers a dynamic environment focused on real-world performance optimizations. #J-18808-Ljbffr Gimlet Labs
$250k - $300k
...vertically integrated AI infrastructure company... ...and be part of a high-performing team that believes in each... ...direction for Crusoe's Linux kernel team, owning the... ...mentoring and growing the engineers around you to deliver... ...technologies tailored for GPU-accelerated AI and HPC...PerformanceTemporary work$150k - $350k
...is seeking a Member of Technical Staff focused on optimizing GPU and accelerator kernels for AI workloads. This role involves analyzing and tuning performance across diverse execution platforms... ...a strong foundation in software engineering and experience with performance-critical...Performance- ...processing down to the lowest layers of the stack, optimize kernel performance, develop new request scheduling and parallelism strategies, and... ...and implement exotic parallelism schemes, write custom GPU kernels for regimes like cascade attention, and understand every...Performance
- B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role involves designing and operating high-performance systems for model inference, synthetic data generation... ...a passion for working in cutting-edge AI. Benefits include top-tier compensation...Performance
$220k - $280k
...the Role Together AI is building the best... ...We're looking for a Staff ML Engineer to drive the model serving... .... You'll profile GPU utilization, design batching... ...-in-class inference performance — architect and implement... ...analysis from GPU kernel behavior to framework-...PerformanceFull time- ...the next generation of AI-driven game experiences... ...Senior Machine Learning Engineer for On-Device & Mobile... ...export, quantization, and kernel-level tuning, to a... ...tuning across NPU, mobile GPU, and desktop/laptop GPU... ...bars.Do low-level performance work: write and tune WebGPU...PerformanceFull timeWork at officeRemote workWorldwide
- A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU kernels and using advanced techniques to boost computation efficiency. Ideal...Performance
- TypeSafe AI in San Francisco is seeking a GPU kernel engineer with deep CUDA expertise to accelerate our training and inference workloads. You will write, optimize, and maintain high-performance kernels close to the metal. Work alongside research and platform engineers...PerformanceWork at office
- ...for the world's most dynamic AI companies, like Cursor, Notion... ...and help build the platform engineers turn to to ship AI products.... ...THE ROLE We’re seeking a GPU Kernel Engineer to join our team at... ...your code directly impacts the performance of state-of-the-art machine...PerformanceFull timeFlexible hours
- ...heterogeneous neocloud for AI workloads. As AI systems... ...hardware that best fits its performance and efficiency needs. This... ...a Member of Technical Staff focused on kernels and GPU performance. In this role,... ...hardware. This role is ideal for engineers who enjoy deep performance...Performance
$173.5k - $331.05k
...on a mission to build the modern, AI-powered video rendering pipeline behind... ...looking for a senior, hands-on engineer to own and evolve the cross-platform GPU rendering platform at the heart of... ...on Windows and Mac Own rendering performance, stability, and device...PerformanceFull timeTemporary workLocal areaWorldwide$250k - $300k
...vertically integrated AI infrastructure... ...be part of a high-performing team that believes... ...the Role:As a Senior Staff/Principal Deployment Automation Engineer for the Compute Team... ...-scale, multi-node GPU clusters. You will... ...Knowledge of Linux kernel internals, specifically...PerformanceTemporary work$197.3k - $313.7k
...SalesforceSalesforce is the #1 AI CRM, where humans with... ...OPPORTUNITIES*Slack is looking for a Staff Machine Learning Engineer with deep expertise in... ...training pipelines on GPU infrastructure.Brainstorm with... ..., assessment of job performance, discipline, termination, and...PerformanceFull time$350k
...Join a rapidly growing AI infrastructure... ...across global cloud and GPU environments. The organization... ...build reliable, high-performance platforms supporting... ...opportunity is for a Staff Site Reliability Engineer to lead the... ...expertise, including kernel tuning, CUDA lifecycle...PerformanceFull time$225k - $275k
...only vertically integrated AI infrastructure company built... ..., and be part of a high-performing team that believes in each... ...Cloud is seeking a Senior Staff Network Deployment Engineer to serve as the technical owner... ...performance compute (HPC) and GPU-based AI infrastructure,...PerformanceTemporary workRemote work- Causal Labs is building a Large Physics foundation Model and GPU-driven compute environment to enable rapid research iteration at... ...provisioning to observability, collaborating with researchers to optimize performance and placement. A strong systems background and experience with...Performance
- ...Francisco is hiring Members of Technical Staff to build systems that accelerate LLM... ...end to end. You will work on high-performance kernels, inference engine internals, and production infrastructure... ...collaborating with a fast-growing AI inference company. #J-18808-Ljbffr SimplifyPerformance
- Together AI in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms, architectures... ...engines, including kernel backends and ATLAS-... ...while profiling across GPU, networking, and memory... ...pipelines, drive performance improvements, and mentor...Performance
$220k
We build and run the inference engine behind every Perplexity query and deploy dozens... ...management to support in API Gateway. GPU kernels migration to CuTe DSL. Port our in-house... ...keep up with rapidly growing traffic. Performance optimisation. Profile and fix...Performance- ...Description Job Description Staff Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home... ...trainable, deployable, observable, and performant under real-world constraints.... ...production deployment, including GPU optimization, memory efficiency,...PerformanceRemote workWork from home
- ...Physics foundation Model and seeks an infrastructure engineer to design, deploy, and operate large distributed GPU clusters. You will extend schedulers, build self-... ...will collaborate with researchers to optimize performance and placement, ensuring reliable, scalable...Performance
- ...The role: SoFi’s Staff AI Engineer is a hands-on AI engineering role in SoFi’s growing independent... ...techniques to maximize model performance and minimize operational costs across the... ...and managing the underlying Kubernetes/GPU orchestration for custom model deployments...PerformanceFull time
$273k - $345k
...that. Atoms builds Physical AI- real-world robots for the... ...scale. We are roboticists, engineers, operators, and builders. We believe... ...and identifying systemic performance bottlenecks. What we're looking... ...identify and eliminate CPU, GPU, and memory bandwidth bottlenecks...PerformanceFull timeInternshipWork at officeFlexible hours- A tech-focused company in San Francisco seeks candidates with expertise in AI simulation development. The role emphasizes optimizing training efficiency, enhancing GPU performance, and ensuring low-latency inference. Applicants should be proficient in methodologies for...Performance
$193k - $234k
...only vertically integrated AI infrastructure company built... ..., and be part of a high-performing team that believes in each... ...high-energy, detail-oriented Staff Network Production Engineer to lead the physical and logical... ...compute (HPC) and GPU-based AI infrastructure, you...PerformanceTemporary workRemote work- Wafer is building AI-powered GPU optimization systems and is seeking engineers to join a small, highly collaborative team. You will work directly with the founders to implement the agent framework, profiling, and compiler tooling that power our GPU optimization platform...
$224k - $284k
...changing that. Atoms builds Physical AI — real-world robots for the... ...work at scale. We are roboticists, engineers, operators, and builders. We believe... ...team who will design and run the high-performance network that connects our GPU compute — building a fast, reliable...PerformanceFull timeWork at officeImmediate startFlexible hours$250k - $300k
...the only vertically integrated AI infrastructure company built... ..., and be part of a high-performing team that believes in each other... ...work directly with customer engineering teams to tailor deployments to... ...vLLM and SGLang to the CUDA kernels underneath, profiling and running...PerformanceTemporary work$141k - $249k
...Description Waabi, founded by AI visionary Raquel... ...autonomy and algorithm engineers to scale safe self-... ...and benchmark new CUDA kernels for inference. - Comprehensively... ...and memory to pinpoint performance bottlenecks.... ...Skilled in profiling CPU and GPU code using tools such...PerformanceWork at officeWork from homeFlexible hours$155k - $269k
...Description Waabi, founded by AI visionary Raquel Urtasun, is... ...team of Research Scientists and Engineers building the content backbone... ...designing, launching, and debugging GPU jobs in the cloud (AWS, GCP,... ...awards and an annual performance bonus. Perks/Benefits: -...PerformanceFull timeWork at officeWork from homeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Engineer: GPU Kernels & AI Performance. Be the first to apply!
- software engineer staff San Francisco, CA
- assistant engineer San Francisco, CA
- engineering aide San Francisco, CA
- staff engineer San Francisco, CA
- staff security engineer San Francisco, CA
- assistant mechanical engineer San Francisco, CA
- assistant engineering manager San Francisco, CA
- senior staff systems engineer San Francisco, CA
- technology administrator San Francisco, CA
- project engineer assistant project manager San Francisco, CA



