MTS - Kernel Engineer
Acceler8 Talent
Kernel Engineer — Compute and Accelerators Mountain View, CA About the Role An early-stage AI hardware company is seeking a Kernel Engineer to develop and optimize specialized compute kernels for a custom accelerator platform. This role sits at the critical boundary between machine learning workloads and silicon. Your work will directly influence how efficiently the hardware executes tensor operations, moves data, and uses the underlying memory hierarchy. You will work closely with architecture, compiler, simulation, and systems teams to define the kernel programming model, implement core operations, and build the profiling workflows used to evaluate hardware and software performance. What You'll Do Develop and optimize compute kernels for a custom AI accelerator. Implement tensor operations, data-movement patterns, and memory-hierarchy optimizations. Build and maintain profiling infrastructure to measure kernel performance against architectural targets. Define execution and data-shuffling patterns across general-purpose control cores, tensor-processing units, and specialized compute engines. Help shape the kernel programming model, including thread execution, register-passing conventions, synchronization, and memory-management strategies. Enable end-to-end kernel execution in simulation and pre-silicon environments. Collaborate with compiler engineers on intermediate representations, lowering strategies, and kernel integration. Use kernels as validation targets for compiler and architecture development. Create technical documentation, examples, and kernel-development guides for the broader engineering team. Investigate performance bottlenecks and recommend hardware or software improvements. What We're Looking For Strong C and C++ programming skills with experience writing production-quality systems or performance-critical code. Deep experience with CUDA or a comparable accelerator-programming model. Strong understanding of: Warp, wavefront, or thread-group execution Memory coalescing Shared or local memory Registers and caches Memory bandwidth and latency Ability to reason about computer architecture, including pipelines, execution units, memory hierarchies, and data-movement costs. Strong performance-profiling and optimization experience. Experience identifying bottlenecks, measuring throughput and latency, and iterating until performance targets are met. Practical understanding of tensor and numerical operations, including: GEMM Convolution Attention Reductions Scatter and gather Elementwise operations Python experience for scripting, tooling, automation, and integration work. Ability to collaborate effectively across architecture, compiler, and hardware teams. Preferred Experience Triton, CUTLASS, or similar kernel-development frameworks. MLIR, LLVM, or compiler infrastructure. RISC-V, x86, ARM64, or another instruction-set architecture. High-performance computing or scientific computing. Custom ASIC, GPU, NPU, or accelerator software. FPGA development or experience reading RTL. Verilog or SystemVerilog. Architectural simulators or instruction-set simulators. Hardware-software co-design. Kernel Engineer, Compute Kernels, Accelerator Kernels, GPU Kernels, CUDA, CUDA C++, C++, Python, Triton, CUTLASS, Tensor Operations, GEMM, Matrix Multiplication, Convolution, Attention, Reductions, Scatter/Gather, Elementwise Operations, Parallel Programming, SIMT, SIMD, Warp Execution, Wavefront Execution, Thread Blocks, Memory Coalescing, Shared Memory, Registers, Cache Optimization, Memory Hierarchy, Data Locality, Data Movement, Synchronization, Performance Profiling, Performance Optimization, Throughput, Latency, Roofline Analysis, Nsight Compute, Nsight Systems, Computer Architecture, Custom ASIC, AI Accelerator, NPU, GPU, MLIR, LLVM, Kernel DSL, Compiler Integration, Architectural Simulation, Instruction Set Simulator, RISC-V, ARM64, x86, HPC, Scientific Computing, Verilog, SystemVerilog, FPGA, Hardware-Software Co-Design, Member of Technical Staff, MTS, PMTS, Principal Member of Technical Staff #J-18808-Ljbffr Acceler8 Talent
$165k - $242k
...company (Nasdaq: CRWV) in March 2025. Learn more at . What You’ll Do CoreWeave is seeking a highly skilled and motivated Systems Kernel Engineer to join the HAVOCK Team, reporting to the Manager of Systems Engineering. In this role, you will be a key contributor to the...SuggestedPermanent employmentTemporary workCasual workWork at officeRemote workFlexible hours- ...AI hardware company is seeking a Compiler Engineer to develop the backend compiler... ...Collaborate with frontend, middle-end, runtime, kernel, and systems engineers. What We're Looking... ...Architecture, Member of Technical Staff, MTS, PMTS, Principal Member of Technical Staff...SuggestedRelocation
- CoreWeave is seeking a Senior Software Engineer for the Systems Engineering team to own kernel tracing and patching across Kubernetes and container runtimes. You will debug complex failures, trace root causes in the Linux kernel, and upstream fixes where appropriate. This...Suggested
- Google DeepMind in Mountain View, CA, USA is seeking a Staff Software Engineer for Performance and Kernel. You will develop low-level performance software for new accelerators, identify high-value algorithms, and collaborate across teams to deliver efficient, programmable...Suggested
$200k - $420k
Member of Technical Staff, Hardware, Kernel Engineer (Custom Silicon) At River, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, custom training...SuggestedLocal areaVisa sponsorshipWork visaRelocation package$184k - $287.5k
We are now looking for a Senior Formal Verification Engineer for GPU Kernels! Modern AI performance relies on highly optimized GPU kernels — performance-critical code where bugs can be hard to catch and expensive to miss. NVIDIA's Deep Learning Safety Team is hiring engineers...Full timeWork experience placement$260k - $320k
DensityAI is seeking an expert to develop compute kernels for a specialized AI accelerator in Mountain View, California. Your role will focus on writing performance-critical kernels while collaborating with architecture and compiler teams to enhance silicon design. Qualifications...- Tilde Research is seeking a Kernel Engineer in Palo Alto to design, implement, and optimize high-performance GPU kernels for training and inference workloads. You will collaborate with ML researchers and engineers to push hardware limits, improve throughput, and reduce...
- ...in California is looking for a Member of Technical Staff for Kernel/Compiler/Communication. This critical role requires strong expertise... ..., along with 5+ years of experience in performance engineering. The ideal candidate will design high-performance kernels and...
- ...pretraining science. We build foundational understanding of models to advance the frontier of intelligence. About The Role As a Kernel Engineer at Tilde, you'll design, implement, and optimize high-performance GPU kernels that are critical to scaling our training and...Full timeInternship
- RadixArk is seeking a deeply technical Member of Technical Staff to push the limits of performance at the kernel, compiler, and communication layers. You will optimize runtimes and libraries to unlock maximum efficiency on modern accelerators across large GPU clusters....
- CoreWeave is seeking a Senior Engineer for its Benchmarking & Performance team to own kernel-level optimization for LLM inference and end-to-end model serving, focusing on CUDA kernels and throughput/latency improvements. You will lead kernel design reviews, mentor engineers...
$260k - $320k
...permanent residents, asylees, or refugees) per 22 CFR 120.62 About the role You will write, evaluate, and profile specialized compute kernels that run on a custom AI accelerator. This is the critical interface between high-level ML workloads and silicon — your code...Permanent employmentH1bLive inWork visa- Acceler8 Talent is hiring a Kernel Engineer to develop and optimize compute kernels for a custom AI accelerator in Mountain View, CA. You will collaborate with architecture, compiler, simulation, and systems teams to define the kernel programming model and validate performance...
- River AI Inc. is hiring a kernel engineer to build the foundational compute engine for its custom silicon. You will design and implement robust kernel generators and emit optimized assembly for our greenfield architecture. You will collaborate with compiler engineers, silicon...
- Oracle in Santa Clara is seeking a systems engineer to enhance their technology solutions. This role involves patching systems and supporting... ...in C/C++ and Python, along with experience in Linux/UNIX kernel development. The position offers a competitive salary range and...
$184k - $287.5k
We're now looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to push every LLM inference operation to its performance ceiling? Our LLM Inference Performance Analysis and Optimization team builds the answer from the ground up. We develop...Full time$153k - $242k
A tech company specializing in AI is seeking a Senior Linux OS Automation Engineer to design and optimize bare-metal systems. The role involves developing and automating tooling for a variety of hardware platforms, collaborating with cross-functional teams, and maintaining...- ...scale, come make a difference at Fiserv.Job TitleStaff Systems Engineer - Device Engineering - In-person CommerceAbout your role:Clover... ...Senior Systems Engineer to serve as our resident Level 2 payment kernel expert at our Sunnyvale, CA campus, driving the engineering and...Full timeWork at officeMonday to Friday
$216k - $283k
...Overview:We are seeking a highly technical Director of Systems Engineering to lead the core electrical hardware, systems, and software... ...embedded stack and the OS-SW teams, driving OS-level configurations, kernel optimizations, and managing external contractors for board...For contractorsLocal area$120k - $400k
...programming language . Responsibilities include: Design and optimize kernels that interface directly with our hardware Work in partnership with our ML Research and Hardware Engineering teams Provide expertise and guidance on hardware architecture from a...Full timeWork experience placementLocal area$188k - $274k
...power.Minimum qualifications:Bachelor's degree in Electrical Engineering, Computer Engineering, Physics, a related field, or equivalent... ...equivalent practical experience.Experience with Android BSP, Kernel Drivers, Android Internals, Embedded Processors.Google is making...Work experience placement- ...Corporation in Santa Clara, CA is seeking outstanding AI systems engineers to develop groundbreaking inference technologies for the... ...accelerated stack. You will create libraries, code generators, and GPU kernel innovations for LLM workloads. Join a team that designs...
- NVIDIA is seeking outstanding AI systems engineers in Santa Clara to advance the inference software stack. You will build libraries, code generators, and GPU kernels for NVIDIA hardware, designing abstractions for LLM serving engines and JIT compilers to accelerate large...
- ...’d love to hear from you. What to Expect As a Software Engineer (Systems), you will be responsible for building out the core software... ...such as C, C++ or Rust ~ Strong understanding of linux: kernel tuning, scheduling, IPC, memory management and RTOS ~...Full time
- NVIDIA in Santa Clara, CA is seeking outstanding AI systems engineers to advance the inference software stack. You will design and optimize kernels, build new abstractions for LLM serving engines, and contribute to accelerators and runtimes that power large language models...
$184k - $356.5k
NVIDIA is hiring a Senior Engineer to join its GPU Software team in Santa Clara, California. This role focuses on designing and developing GPU kernel drivers and embedded software, impacting both datacenter and gaming markets. Candidates should have over 10 years of software...$120k - $275k
...primarily use the Rust programming language.What You'll Do HereDesign and optimize kernels that interface directly with our hardwareWork in partnership with our ML Research and Hardware Engineering teamsProvide expertise and guidance on hardware architecture from a...Full timeWork experience placementWork at officeLocal areaRemote workMonday to FridayFlexible hours3 days per week$212k - $259k
...highly motivated, experienced, and versatile Principal Verification Engineer to join our team. In this role, you will own the end-to-end... ....· Fundamental knowledge of Linux operating system internals, kernel, and device drivers.· Familiarity with advanced networking protocols...Full timeWorldwide- ...and Responsibility: AMD, Inc., is hiring MTS Software Development Eng. to Research, design... ..., applying principles and techniques of engineering, and mathematical analysis. Design,... ...experience in the following:Low-level GPU kernel optimization; and Performance profiling and...Remote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to MTS - Kernel Engineer. Be the first to apply!


