Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

MTS - Kernel Engineer

Acceler8 Talent

Kernel Engineer — Compute and Accelerators Mountain View, CA About the Role An early-stage AI hardware company is seeking a Kernel Engineer to develop and optimize specialized compute kernels for a custom accelerator platform. This role sits at the critical boundary between machine learning workloads and silicon. Your work will directly influence how efficiently the hardware executes tensor operations, moves data, and uses the underlying memory hierarchy. You will work closely with architecture, compiler, simulation, and systems teams to define the kernel programming model, implement core operations, and build the profiling workflows used to evaluate hardware and software performance. What You'll Do Develop and optimize compute kernels for a custom AI accelerator. Implement tensor operations, data-movement patterns, and memory-hierarchy optimizations. Build and maintain profiling infrastructure to measure kernel performance against architectural targets. Define execution and data-shuffling patterns across general-purpose control cores, tensor-processing units, and specialized compute engines. Help shape the kernel programming model, including thread execution, register-passing conventions, synchronization, and memory-management strategies. Enable end-to-end kernel execution in simulation and pre-silicon environments. Collaborate with compiler engineers on intermediate representations, lowering strategies, and kernel integration. Use kernels as validation targets for compiler and architecture development. Create technical documentation, examples, and kernel-development guides for the broader engineering team. Investigate performance bottlenecks and recommend hardware or software improvements. What We're Looking For Strong C and C++ programming skills with experience writing production-quality systems or performance-critical code. Deep experience with CUDA or a comparable accelerator-programming model. Strong understanding of: Warp, wavefront, or thread-group execution Memory coalescing Shared or local memory Registers and caches Memory bandwidth and latency Ability to reason about computer architecture, including pipelines, execution units, memory hierarchies, and data-movement costs. Strong performance-profiling and optimization experience. Experience identifying bottlenecks, measuring throughput and latency, and iterating until performance targets are met. Practical understanding of tensor and numerical operations, including: GEMM Convolution Attention Reductions Scatter and gather Elementwise operations Python experience for scripting, tooling, automation, and integration work. Ability to collaborate effectively across architecture, compiler, and hardware teams. Preferred Experience Triton, CUTLASS, or similar kernel-development frameworks. MLIR, LLVM, or compiler infrastructure. RISC-V, x86, ARM64, or another instruction-set architecture. High-performance computing or scientific computing. Custom ASIC, GPU, NPU, or accelerator software. FPGA development or experience reading RTL. Verilog or SystemVerilog. Architectural simulators or instruction-set simulators. Hardware-software co-design. Kernel Engineer, Compute Kernels, Accelerator Kernels, GPU Kernels, CUDA, CUDA C++, C++, Python, Triton, CUTLASS, Tensor Operations, GEMM, Matrix Multiplication, Convolution, Attention, Reductions, Scatter/Gather, Elementwise Operations, Parallel Programming, SIMT, SIMD, Warp Execution, Wavefront Execution, Thread Blocks, Memory Coalescing, Shared Memory, Registers, Cache Optimization, Memory Hierarchy, Data Locality, Data Movement, Synchronization, Performance Profiling, Performance Optimization, Throughput, Latency, Roofline Analysis, Nsight Compute, Nsight Systems, Computer Architecture, Custom ASIC, AI Accelerator, NPU, GPU, MLIR, LLVM, Kernel DSL, Compiler Integration, Architectural Simulation, Instruction Set Simulator, RISC-V, ARM64, x86, HPC, Scientific Computing, Verilog, SystemVerilog, FPGA, Hardware-Software Co-Design, Member of Technical Staff, MTS, PMTS, Principal Member of Technical Staff #J-18808-Ljbffr Acceler8 Talent

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the MTS - Kernel Engineer in Mountain View, CA vacancy
  • $165k - $242k

     ...company (Nasdaq: CRWV) in March 2025. Learn more at . What You’ll Do CoreWeave is seeking a highly skilled and motivated Systems Kernel Engineer to join the HAVOCK Team, reporting to the Manager of Systems Engineering. In this role, you will be a key contributor to the... 
    Suggested
    Permanent employment
    Temporary work
    Casual work
    Work at office
    Remote work
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    1 day ago
  •  ...AI hardware company is seeking a Compiler Engineer to develop the backend compiler...  ...Collaborate with frontend, middle-end, runtime, kernel, and systems engineers. What We're Looking...  ...Architecture, Member of Technical Staff, MTS, PMTS, Principal Member of Technical Staff... 
    Suggested
    Relocation

    Acceler8 Talent

    Mountain View, CA
    1 day ago
  • CoreWeave is seeking a Senior Software Engineer for the Systems Engineering team to own kernel tracing and patching across Kubernetes and container runtimes. You will debug complex failures, trace root causes in the Linux kernel, and upstream fixes where appropriate. This... 
    Suggested

    CoreWeave

    Sunnyvale, CA
    1 day ago
  • Google DeepMind in Mountain View, CA, USA is seeking a Staff Software Engineer for Performance and Kernel. You will develop low-level performance software for new accelerators, identify high-value algorithms, and collaborate across teams to deliver efficient, programmable... 
    Suggested

    Google Inc.

    Mountain View, CA
    3 days ago
  • $200k - $420k

    Member of Technical Staff, Hardware, Kernel Engineer (Custom Silicon) At River, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, custom training... 
    Suggested
    Local area
    Visa sponsorship
    Work visa
    Relocation package

    River AI Inc.

    Palo Alto, CA
    11 hours ago
  • $184k - $287.5k

    We are now looking for a Senior Formal Verification Engineer for GPU Kernels! Modern AI performance relies on highly optimized GPU kernels — performance-critical code where bugs can be hard to catch and expensive to miss. NVIDIA's Deep Learning Safety Team is hiring engineers... 
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    4 days ago
  • $260k - $320k

    DensityAI is seeking an expert to develop compute kernels for a specialized AI accelerator in Mountain View, California. Your role will focus on writing performance-critical kernels while collaborating with architecture and compiler teams to enhance silicon design. Qualifications... 

    DensityAI

    Mountain View, CA
    1 day ago
  • Tilde Research is seeking a Kernel Engineer in Palo Alto to design, implement, and optimize high-performance GPU kernels for training and inference workloads. You will collaborate with ML researchers and engineers to push hardware limits, improve throughput, and reduce... 

    Tilde Research

    Palo Alto, CA
    1 day ago
  •  ...in California is looking for a Member of Technical Staff for Kernel/Compiler/Communication. This critical role requires strong expertise...  ..., along with 5+ years of experience in performance engineering. The ideal candidate will design high-performance kernels and... 

    RadixArk

    Palo Alto, CA
    3 days ago
  •  ...pretraining science. We build foundational understanding of models to advance the frontier of intelligence. About The Role As a Kernel Engineer at Tilde, you'll design, implement, and optimize high-performance GPU kernels that are critical to scaling our training and... 
    Full time
    Internship

    Tilde Research

    Palo Alto, CA
    1 day ago
  • RadixArk is seeking a deeply technical Member of Technical Staff to push the limits of performance at the kernel, compiler, and communication layers. You will optimize runtimes and libraries to unlock maximum efficiency on modern accelerators across large GPU clusters.... 

    RadixArk

    Palo Alto, CA
    1 day ago
  • CoreWeave is seeking a Senior Engineer for its Benchmarking & Performance team to own kernel-level optimization for LLM inference and end-to-end model serving, focusing on CUDA kernels and throughput/latency improvements. You will lead kernel design reviews, mentor engineers... 

    Neura Market

    Sunnyvale, CA
    2 days ago
  • $260k - $320k

     ...permanent residents, asylees, or refugees) per 22 CFR 120.62 About the role You will write, evaluate, and profile specialized compute kernels that run on a custom AI accelerator. This is the critical interface between high-level ML workloads and silicon — your code... 
    Permanent employment
    H1b
    Live in
    Work visa

    DensityAI

    Mountain View, CA
    1 day ago
  • Acceler8 Talent is hiring a Kernel Engineer to develop and optimize compute kernels for a custom AI accelerator in Mountain View, CA. You will collaborate with architecture, compiler, simulation, and systems teams to define the kernel programming model and validate performance... 

    Acceler8 Talent

    Mountain View, CA
    1 day ago
  • River AI Inc. is hiring a kernel engineer to build the foundational compute engine for its custom silicon. You will design and implement robust kernel generators and emit optimized assembly for our greenfield architecture. You will collaborate with compiler engineers, silicon... 

    River AI Inc.

    Palo Alto, CA
    5 days ago
  • Oracle in Santa Clara is seeking a systems engineer to enhance their technology solutions. This role involves patching systems and supporting...  ...in C/C++ and Python, along with experience in Linux/UNIX kernel development. The position offers a competitive salary range and... 

    Oracle

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

    We're now looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to push every LLM inference operation to its performance ceiling? Our LLM Inference Performance Analysis and Optimization team builds the answer from the ground up. We develop... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $153k - $242k

    A tech company specializing in AI is seeking a Senior Linux OS Automation Engineer to design and optimize bare-metal systems. The role involves developing and automating tooling for a variety of hardware platforms, collaborating with cross-functional teams, and maintaining... 

    CoreWeave

    Sunnyvale, CA
    2 days ago
  •  ...scale, come make a difference at Fiserv.Job TitleStaff Systems Engineer - Device Engineering - In-person CommerceAbout your role:Clover...  ...Senior Systems Engineer to serve as our resident Level 2 payment kernel expert at our Sunnyvale, CA campus, driving the engineering and... 
    Full time
    Work at office
    Monday to Friday

    Fiserv

    Sunnyvale, CA
    3 days ago
  • $216k - $283k

     ...Overview:We are seeking a highly technical Director of Systems Engineering to lead the core electrical hardware, systems, and software...  ...embedded stack and the OS-SW teams, driving OS-level configurations, kernel optimizations, and managing external contractors for board... 
    For contractors
    Local area

    Ceribell

    Sunnyvale, CA
    2 days ago
  • $120k - $400k

     ...programming language . Responsibilities include: Design and optimize kernels that interface directly with our hardware Work in partnership with our ML Research and Hardware Engineering teams Provide expertise and guidance on hardware architecture from a... 
    Full time
    Work experience placement
    Local area

    Matx

    Mountain View, CA
    11 hours ago
  • $188k - $274k

     ...power.Minimum qualifications:Bachelor's degree in Electrical Engineering, Computer Engineering, Physics, a related field, or equivalent...  ...equivalent practical experience.Experience with Android BSP, Kernel Drivers, Android Internals, Embedded Processors.Google is making... 
    Work experience placement

    Google

    Mountain View, CA
    3 days ago
  •  ...Corporation in Santa Clara, CA is seeking outstanding AI systems engineers to develop groundbreaking inference technologies for the...  ...accelerated stack. You will create libraries, code generators, and GPU kernel innovations for LLM workloads. Join a team that designs... 

    NVIDIA Corporation

    Santa Clara, CA
    4 days ago
  • NVIDIA is seeking outstanding AI systems engineers in Santa Clara to advance the inference software stack. You will build libraries, code generators, and GPU kernels for NVIDIA hardware, designing abstractions for LLM serving engines and JIT compilers to accelerate large... 

    Segment (Twilio)

    Santa Clara, CA
    4 days ago
  •  ...’d love to hear from you. What to Expect As a Software Engineer (Systems), you will be responsible for building out the core software...  ...such as C, C++ or Rust ~ Strong understanding of linux: kernel tuning, scheduling, IPC, memory management and RTOS ~... 
    Full time

    Sunday

    Mountain View, CA
    11 hours ago
  • NVIDIA in Santa Clara, CA is seeking outstanding AI systems engineers to advance the inference software stack. You will design and optimize kernels, build new abstractions for LLM serving engines, and contribute to accelerators and runtimes that power large language models... 

    NVIDIA

    Santa Clara, CA
    2 days ago
  • $184k - $356.5k

    NVIDIA is hiring a Senior Engineer to join its GPU Software team in Santa Clara, California. This role focuses on designing and developing GPU kernel drivers and embedded software, impacting both datacenter and gaming markets. Candidates should have over 10 years of software... 

    NVIDIA

    Santa Clara, CA
    2 days ago
  • $120k - $275k

     ...primarily use the Rust programming language.What You'll Do HereDesign and optimize kernels that interface directly with our hardwareWork in partnership with our ML Research and Hardware Engineering teamsProvide expertise and guidance on hardware architecture from a... 
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work
    Monday to Friday
    Flexible hours
    3 days per week

    MatX

    Mountain View, CA
    2 days ago
  • $212k - $259k

     ...highly motivated, experienced, and versatile Principal Verification Engineer to join our team. In this role, you will own the end-to-end...  ....· Fundamental knowledge of Linux operating system internals, kernel, and device drivers.· Familiarity with advanced networking protocols... 
    Full time
    Worldwide

    Fortinet

    Sunnyvale, CA
    4 days ago
  •  ...and Responsibility: AMD, Inc., is hiring MTS Software Development Eng. to Research, design...  ..., applying principles and techniques of engineering, and mathematical analysis. Design,...  ...experience in the following:Low-level GPU kernel optimization; and Performance profiling and... 
    Remote work

    AMD

    Santa Clara, CA
    7 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to MTS - Kernel Engineer. Be the first to apply!