Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Kernel Engineer for AI Accelerator - Profiling

$260k - $320k

DensityAI

DensityAI is seeking an expert to develop compute kernels for a specialized AI accelerator in Mountain View, California. Your role will focus on writing performance-critical kernels while collaborating with architecture and compiler teams to enhance silicon design. Qualifications include strong proficiency in C/C++, CUDA, and performance optimization techniques. Compensation ranges from $260,000 to $320,000 USD, with additional benefits and equity grants offered. #J-18808-Ljbffr DensityAI

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Kernel Engineer for AI Accelerator - Profiling in Mountain View, CA vacancy
  • $142.8k - $274.8k

     ...Hardware, and Infrastructure Engineering (SCHIE) is the team...  ...is developing AI-native silicon and hyperscale...  ...a Principal AI Accelerator Tools Development Engineer...  ...to compiler-generated kernels, distributed communication...  ...including performance profiling, bottleneck analysis,... 
    Suggested
    Ongoing contract
    Work at office
    Local area
    Worldwide
    3 days per week

    Microsoft

    Mountain View, CA
    1 day ago
  • $260k - $320k

     ...permanent residents, asylees, or refugees) per 22 CFR 120.62 About the role You will write, evaluate, and profile specialized compute kernels that run on a custom AI accelerator. This is the critical interface between high-level ML workloads and silicon — your code directly... 
    Suggested
    Permanent employment
    H1b
    Live in
    Work visa

    DensityAI

    Mountain View, CA
    3 days ago
  • $147k - $210k

     ...Gemini models on hardware accelerators (TPUs and GPUs),...  ...dedicated compiler passes.Profile large-scale distributed...  ...performance, low-level custom kernels for critical model...  .... As a Software Engineer, you will be working with the cutting edge AI agents developed by our... 
    Suggested

    Google

    Mountain View, CA
    4 days ago
  • $182k - $242k

     ...The Essential Cloud for AI. Built for pioneers by...  ...technical expertise to accelerate breakthroughs and turn...  ...inference. Our stack is engineered for speed, scale, and cost...  ...team, focused on kernel authoring and optimization. You will write, profile, and tune the GPU kernels... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    3 days ago
  • CoreWeave is hiring a Senior Engineer for its Benchmarking & Performance team to write, profile, and optimize GPU kernels on the LLM inference path. You will improve latency and throughput and collaborate with product, orchestration, and hardware teams to achieve strict... 
    Suggested

    CoreWeave

    Sunnyvale, CA
    3 days ago
  • $200k - $420k

     ...Technical Staff, Hardware, Kernel Engineer (Custom Silicon) At River, our...  ...is to create personal AI owned and shaped by each individual...  ...based on kernel execution profiles. Performance Profiling & Validation...  ...Triton, CUTLASS, or custom accelerator assembly) and a proven track... 
    Local area
    Visa sponsorship
    Work visa
    Relocation package

    River AI Inc.

    Palo Alto, CA
    23 hours ago
  • $207k - $300k

    On-board emerging co-accelerators into Google's ML accelerator families...  ...daemons running on BMC/hosts and kernel drivers.Design and develop...  ...:Master’s degree or PhD in Engineering, Computer Science, or a related...  ...’s accelerator roadmap.The AI and Infrastructure team is redefining... 
    Worldwide

    Google

    Sunnyvale, CA
    3 days ago
  •  ...build great products that accelerate next-generation computing experiences—from AI and data centers, to...  ...high-performance GPU kernels, powering major AI...  ...architecture, and performance engineering. You are comfortable...  ...rigorous profiling and microbenchmarking... 

    AMD

    San Jose, CA
    23 hours ago
  • $174k - $252k

     ...software/hardware codesign, kernel optimization, performance benchmarking...  ...Learning (ML) hardware accelerators and ML inference software....  ...qualifications:PhD in Computer Engineering, Computer Science, or a...  ...compute platform for cutting-edge AI inference workloads, with an... 

    Google

    Mountain View, CA
    4 days ago
  • $174k - $252k

     ...ML compiler stack to bridge AI workloads and low-level Hardware...  ....Drive SW/HW codesign, kernel optimization, performance tuning...  ...machine learning (ML) hardware accelerators, ML compiler stacks (e.g.,...  ...qualifications:PhD degree in Computer Engineering, Computer Science, or a... 

    Google

    Mountain View, CA
    4 days ago
  • $120k - $275k

     ...designs hardware tailored for the world’s best AI models. Our hardware will make a given...  ...or subsystem level.MatX is seeking an AI Accelerator Compute Architect to help define the...  ...tradeoffs.Programming, code optimization, or kernel-level performance analysis experience.... 
    Daily paid
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work
    Monday to Friday
    Flexible hours

    MatX

    Mountain View, CA
    1 day ago
  • $150k - $230k

     ...researchers and veteran systems engineers who share a vision for...  ...computing. As AI workloads grow...  ...failures, and performance acceleration that dynamically routes...  ...them. Examples: Kernel subsystems, device drivers...  ...systems Runtimes, profilers, or performance... 

    Clockwork.io

    Palo Alto, CA
    18 days ago
  • NVIDIA in Santa Clara, CA is seeking outstanding AI systems engineers to advance the inference software stack. You will design and optimize kernels, build new abstractions for LLM serving engines, and contribute to accelerators and runtimes that power large language models... 

    NVIDIA

    Santa Clara, CA
    4 days ago
  • Senior AI Systems Performance Engineer Palo Alto, California, United States The era...  ...hidden value in their data, accelerate processes, reduce costs,...  ...SambaNova software stack. Profile and enhance model...  ...optimization Compiler, runtime, or kernel‑level optimization... 
    Full time
    Temporary work
    Local area
    Flexible hours

    SambaNova

    Palo Alto, CA
    3 days ago
  • $184k - $287.5k

     ...Senior Developer Technology Engineer for the Public Sector!...  ...techniques to GPU-accelerate leading applications in...  .../Climate Modeling, and AI in HPC. You will be performing...  ...for GPUs, including kernel optimization and strong...  .... Experience in profiling and optimizing applications... 
    Full time
    Work experience placement
    Remote work

    Nvidia

    Santa Clara, CA
    4 days ago
  • $200k - $350k

     ...talented Design Verification Engineers to help verify and deliver Velaura...  ...'s next generation Physical AI SoC. You will work closely...  ...This role focuses on our AI acceleration units. We are designing custom...  ...modeling, power analysis, profiling and workload characterization... 
    Full time
    Flexible hours

    Velaura AI

    Santa Clara, CA
    13 days ago
  • River AI Inc. is hiring a kernel engineer to build the foundational compute engine for its custom silicon. You will design and implement robust kernel generators and emit optimized assembly for our greenfield architecture. You will collaborate with compiler engineers, silicon... 

    River AI Inc.

    Palo Alto, CA
    2 days ago
  • Platform Recruitment is seeking a SIPI Characterization Engineer for an AI hardware client in Mountain View, CA. The role focuses on high-speed SerDes and PDN characterization across AI accelerator systems, from design through production. The candidate will build automated... 

    Platform Recruitment

    Mountain View, CA
    4 days ago
  • $184k - $287.5k

     ...skilled and motivated software engineers to join us and build AI inference systems that...  ...stacks, optimize GPU kernels and compilers, drive industry...  ...teams to push the frontier of accelerated computing for AI.What you’...  ...GPU hardware features; profile and optimize the inference... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $184k - $287.5k

     ...is possible on modern accelerator architectures.Join our...  ...performance for emerging AI workloads. You will be...  ...optimization, custom kernel development, and...  ...and runtime systems to profile and accelerate new paradigms...  ...Science, Computer Engineering, Electrical Engineering... 
    Full time

    Nvidia

    Santa Clara, CA
    23 hours ago
  • $90k - $180k

     ...Agentic WorkflowsBuild agentic AI services (planning, tool use,...  ...behind stable APIs and SDKs. Accelerated Compute & Data...  ...scalable workflow orchestration. Profile and optimize bottlenecks across...  ...contribute to design reviews and engineering best practices. Mentor peers... 
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    4 days ago
  • $200k - $360k

     ...professional to own MLIR dialect design and lowering passes for their AI accelerator. You'll work closely with chip-design and software teams....  ...knowledge of tensor compilation, and 5+ years of compiler engineering experience. The role includes a competitive compensation... 

    DensityAI

    Mountain View, CA
    3 days ago
  • A cutting-edge AI company in California is looking for a Member of Technical Staff for Kernel/Compiler/Communication. This critical role requires strong expertise in CUDA...  ...with 5+ years of experience in performance engineering. The ideal candidate will design high-... 

    RadixArk

    Palo Alto, CA
    23 hours ago
  • $120k - $275k

     ...designs hardware tailored for the world’s best AI models. Our hardware will make a given...  ...subsystem level. MatX is seeking an AI Accelerator Compute Architect to help define the...  ...tradeoffs. Programming, code optimization, or kernel-level performance analysis experience.... 
    Daily paid
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work
    Monday to Friday
    Flexible hours

    MatX Inc.

    Mountain View, CA
    2 days ago
  • $138k - $197k

     ...infrastructure and architecture for on-device AI/ML accelerators.Implement, model, analyze, and...  ...qualifications:Bachelor's degree in Electrical Engineering, Computer Engineering, Computer...  ...the use of simulation frameworks and profiling tools.Experience with CPU or ML... 
    Worldwide

    Google

    Mountain View, CA
    4 days ago
  • $207k - $300k

     ...platforms (Compute, Storage, hardware accelerators like TPUs and GPUs).Own the Linux kernel, networking stack, and...  ...SoCs.Partner directly with hardware engineering, accelerator (TPU/GPU) teams, and...  ...continue to push technology forward.The AI and Infrastructure team is... 
    Worldwide

    Google

    Sunnyvale, CA
    1 day ago
  •  ...build great products that accelerate next-generation computing experiences—from AI and data centers, to...  ...an influential software engineer who is passionate about...  ...the lowest-level GPU kernels to large-scale distributed...  ...experience using GPU profiling and performance analysis... 

    AMD

    Santa Clara, CA
    1 day ago
  • $120k - $280k

     ...potential of their data, from AI to multicloud....  ...experienced Systems Software Engineers across multiple NetApp...  ..., logs, tracing, and profiling ~Collaborate across...  ...AI-assisted tools to accelerate design, development,...  ..., OCI) ~Exposure to kernel subsystems, VFS, IO... 
    Part time
    Work at office
    Local area

    NetApp

    San Jose, CA
    19 hours ago
  • $218.8k - $335.3k

     ...General Motors, our Embodied AI teams are redefining what’s possible...  ...looking for a Staff Software Engineer to provide technical...  ...observability for on‑road incidents. Profile and optimize SDS components...  ...Hands‑on experience with GPU/accelerator‑based ML inference, model... 
    Full time
    Local area
    Remote work
    Work from home
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  • $90k - $125k

     ...security with the world’s most advanced AI-native platform. We work on large...  ...multiplier to proactively and continuously accelerate execution, build expertise, uncover...  ...macOS and Linux, in both user mode and kernel mode. Engineering software at that depth and that scale... 
    Full time
    Work experience placement
    Internship
    Work at office
    Local area
    Remote work
    Worldwide

    CrowdStrike

    Sunnyvale, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Kernel Engineer for AI Accelerator - Profiling. Be the first to apply!