Staff GenAI Kernel & Performance Engineer
Databricks Inc.
A leading data and AI company in San Francisco seeks a Staff Software Engineer to lead kernel-level performance engineering for GenAI workloads. The role involves designing and optimizing high-performance GPU kernels, mentoring engineers, and driving performance roadmaps for low-level compute paths. Ideal candidates should have advanced experience with GPU architectures and performance optimization techniques. This position offers a competitive salary and a chance to work with a talented team focused on pushing the frontier of inference performance. #J-18808-Ljbffr Databricks
- A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU kernels and using advanced techniques to boost computation efficiency. Ideal...Performance
- MakerMaker.AI in San Francisco is seeking a skilled Software Engineer to write and optimize GPU kernels. You will work on deep low-level tasks that directly impact the performance of machine learning models. The ideal candidate has over 4 years of experience with GPU kernels...Performance
$285k - $315k
SF Tensor is looking for a Founding GPU Kernel Engineer in San Francisco, specializing in GPU architecture and kernel optimization for machine... ...has deep expertise, proven capabilities in hand-optimizing performance-critical kernels, and strong programming skills in C++ and...PerformanceFull timeRelocation package- San Francisco Tensor Company is seeking a Founding GPU Kernel Engineer to enhance GPU performance for AI applications. You will optimize and write kernels while collaborating with compiler teams to improve efficiencies across architectures. The ideal candidate has deep...PerformanceWork at officeRelocation package
- Magic is hiring a Kernel Engineer in San Francisco to design, implement, and optimize high-performance kernels for long-context training and inference. You will tackle memory usage, data movement, and throughput challenges in real-time workloads. You’ll work across training...PerformanceVisa sponsorship
- Inception is seeking engineers and scientists to design, optimize, and maintain compute foundations for large‑scale language model training and inference. You will develop high‑performance ML kernels, enable efficient low‑precision arithmetic, and improve the distributed...Performance
- Mercor is seeking GPU kernel optimization experts to contribute to a project with a leading AI lab. This contract-based opportunity... ...practical GPU programming experience, and the ability to squeeze performance out of modern GPU architectures. You will analyze, optimize,...PerformanceContract workFreelance
$100k - $120k
...models. As training and inference workloads grow, we need kernel‑level innovations to reduce latency, memory usage, and... ...faster. Responsibilities Lead a team of kernel and system engineers focused on performance-critical code Design, implement, and optimize custom compute...Performance- ...ROLE You’ll write and optimize the GPU kernels and supporting systems software that makes... ...fast. This is deep, low-level work (performance counters, memory bandwidth, warp-level scheduling... ...models actually use. We hire kernel engineers because the gap between "this works" and...PerformanceShift work
- 1. Role Overview Mercor is seeking GPU kernel optimization experts to contribute to a project with a leading AI lab. This opportunity... ...GPU programming experience, and the ability to improve kernel performance using profiler-guided analysis. You’ll help evaluate, optimize...PerformanceContract workFreelance
$285k - $315k
About The Role We're looking for a Founding GPU Kernel Engineer who lives right at the boundary between hardware and software. Someone who... ...(matmuls, attention, normalization, etc.) to set the performance ceilings Profile at the microarchitectural level: look into...PerformanceFull timeWork at officeRelocation package$180k - $280k
...investors. Since mid-2024, we've been engineering the foundation for what comes after the... ...About the role We're looking for a GPU kernel engineer with deep, low-level CUDA expertise... ...: Write, optimize, and maintain high-performance GPU kernels (e.g., in CUDA / CuTe DSL)...PerformanceWork at officeVisa sponsorshipShift work$167.2k - $209k
...world. DigitalOcean is seeking a Senior Engineer 2 to play a key technical role in our AI... ...we can offer the industry-leading performance for our inference services. You will be... ...optimizations at the inference engine and GPU kernel layers, ensuring our infrastructure extracts...PerformanceLocal areaRemote workWorldwideFlexible hours$315k
...group of committed researchers, engineers, policy experts, and business... .... About the Role As a TPU Kernel Engineer, you'll be responsible... ...identifying and addressing performance issues across many different... ...policy: Currently, we expect all staff to be in one of our offices...PerformanceContract workFor contractorsFor subcontractorWork at officeRelocationVisa sponsorshipWork visaFlexible hours- ...in San Francisco seeks an experienced Software Engineer to help bring inference workloads to AWS... ...This deeply technical, cross‑stack role covers kernels, compilers, and model execution. You will develop high‑performance kernels, improve compiler support, and ensure...Performance
- ...efficiency serving platform. We seek an ML Engineer who operates at the intersection of... ...scale, and optimize end-to-end multimodal GenAI systems. You will design architectures that... ...to open-source projects while driving performance and reliability across delivering #J-18...Performance
$179.4k - $224.25k
...and private evaluations.About Data EngineOur Generative AI Data Engine powers the world’s most advanced LLMs and generative models... ...including job-related skills, experience, qualifications, interview performance, and relevant education or training. Scale employees in...PerformanceFull time$272k - $336k
...Waymo Systems Engineering Role Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver... ...hardware systems in groundbreaking new ways. We set the high performance standards that ensure our vehicles run smoothly and keep...PerformanceOdd jobFull timeRemote work$63k - $140k
...LevelAssociateJob Description & SummaryThe OpportunityAs a GenAI Python Systems Engineer - Experienced Associate, you will leverage advanced... ...and accepting feedback to enhance personal growth and team performance- Building commercial awareness by understanding how the business...PerformanceFull timeH1b- Fluidstack is seeking a hands-on hardware qualification engineer in Seattle to qualify next-generation compute, including GPUs, servers... ...fleet deployment. You will design and run benchmarks for performance, power, and thermals, and perform stress and fault-injection tests...Performance
- ...Larry Summers , and Jack Dorsey . Position: CUDA Engineering Expert Type: Contract Compensation: $500/hour... ...Remote Role Responsibilities Analyze and optimize GPU kernels for performance, efficiency, and hardware utilization. Use profiler metrics...PerformanceContract workSummer workRemote work
$190.9k - $232.8k
P-1285About This RoleAs a staff software engineer for GenAI Performance and Kernel, you will own the design, implementation, optimization, and correctness of the high-performance GPU kernels powering our GenAI inference stack. You will lead development of highly-tuned,...PerformanceLocal areaWorldwide$264.8k - $331k
Machine Learning Systems Research Engineer, Agent Post-training - Enterprise GenAI AI is becoming vitally important in every function of our society. At... ...state of the art post-training algorithms to reach the performance necessary for complex agents in enterprises around...PerformanceFull timeContract workFor contractorsFor subcontractorWork at office$140k - $230k
...service lives with minimal maintenance. As a Senior or Staff Electrical Hardware Reliability Engineer focused on power electronics, you will serve as the... ...test campaigns to validate long-term hardware performance. Author clear, technically rigorous test plans and...PerformanceLocal area$256k - $278k
...Bachelor's degree in Electrical Engineering, Computer Engineering,... ...years of experience in power or performance modeling or system performance... ...year of experience in linux kernel programming. Experience with... ...design team within DeepMind’s GenAI Systems, operating across Google...Performance$185k - $250.34k
...infrastructure consulting firm providing civil engineering and surveying services across California... ...scope, schedule, budget, and contractor performance Ensure compliance with environmental... ...Support proposal development, staff mentoring, and collaboration with regional...PerformanceContract workFor contractorsWork at officeLocal areaFlexible hours- ...of what's possible in video generation.We're seeking a GPU Performance Engineer to squeeze every last FLOP from our H100 infrastructure and... ...solutions that achieve 5-10x speedups. From writing custom CUDA kernels to eliminating cold start latency, you'll ensure our...Performance
$170k - $205k
...their AI strategies, and be part of a high-performing team that believes in each other, come... ..., hands-on Senior level performance engineer who thrives on deep technical challenges... ...candidate possesses a mastery of Linux kernel systems and low-level optimization, with...PerformanceFull timeTemporary work$155.6k - $306.8k
...Summary At Deloitte, Forward Deployed Engineers (FDE) don’t just build AI solutions,... ...prototype and deliver high-impact GenAI-enabled solutions. This requires a highly... ...limitation, individual and organizational performance. Deloitte is committed to providing...PerformanceLocal area$134.5k - $265.1k
...Senior Consultant - Forward Deployed Engineer (FDE) - DatabricksForward Deployed Engineers... ...and integration issues and optimize performance across Databricks and cloud-based enterprise... ...), generative artificial intelligence (GenAI), or large language model (LLM)-powered...PerformanceLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff GenAI Kernel & Performance Engineer. Be the first to apply!
- staff engineer San Francisco, CA
- assistant engineer San Francisco, CA
- research assistant engineering San Francisco, CA
- staff design engineer San Francisco, CA
- staff security engineer San Francisco, CA
- engineering aide San Francisco, CA
- senior staff engineer San Francisco, CA
- senior staff systems engineer San Francisco, CA
- assistant chief engineer San Francisco, CA
- assistant electrical engineer San Francisco, CA


