Kernel Engineer for AI Accelerator - Profiling
$260k - $320kDensityAI
DensityAI is seeking an expert to develop compute kernels for a specialized AI accelerator in Mountain View, California. Your role will focus on writing performance-critical kernels while collaborating with architecture and compiler teams to enhance silicon design. Qualifications include strong proficiency in C/C++, CUDA, and performance optimization techniques. Compensation ranges from $260,000 to $320,000 USD, with additional benefits and equity grants offered. #J-18808-Ljbffr DensityAI
$142.8k - $274.8k
...Hardware, and Infrastructure Engineering (SCHIE) is the team... ...is developing AI-native silicon and hyperscale... ...a Principal AI Accelerator Tools Development Engineer... ...to compiler-generated kernels, distributed communication... ...including performance profiling, bottleneck analysis,...SuggestedOngoing contractWork at officeLocal areaWorldwide3 days per week$260k - $320k
...permanent residents, asylees, or refugees) per 22 CFR 120.62 About the role You will write, evaluate, and profile specialized compute kernels that run on a custom AI accelerator. This is the critical interface between high-level ML workloads and silicon — your code directly...SuggestedPermanent employmentH1bLive inWork visa$147k - $210k
...Gemini models on hardware accelerators (TPUs and GPUs),... ...dedicated compiler passes.Profile large-scale distributed... ...performance, low-level custom kernels for critical model... .... As a Software Engineer, you will be working with the cutting edge AI agents developed by our...Suggested$182k - $242k
...The Essential Cloud for AI. Built for pioneers by... ...technical expertise to accelerate breakthroughs and turn... ...inference. Our stack is engineered for speed, scale, and cost... ...team, focused on kernel authoring and optimization. You will write, profile, and tune the GPU kernels...SuggestedPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours- CoreWeave is hiring a Senior Engineer for its Benchmarking & Performance team to write, profile, and optimize GPU kernels on the LLM inference path. You will improve latency and throughput and collaborate with product, orchestration, and hardware teams to achieve strict...Suggested
$200k - $420k
...Technical Staff, Hardware, Kernel Engineer (Custom Silicon) At River, our... ...is to create personal AI owned and shaped by each individual... ...based on kernel execution profiles. Performance Profiling & Validation... ...Triton, CUTLASS, or custom accelerator assembly) and a proven track...Local areaVisa sponsorshipWork visaRelocation package$207k - $300k
On-board emerging co-accelerators into Google's ML accelerator families... ...daemons running on BMC/hosts and kernel drivers.Design and develop... ...:Master’s degree or PhD in Engineering, Computer Science, or a related... ...’s accelerator roadmap.The AI and Infrastructure team is redefining...Worldwide- ...build great products that accelerate next-generation computing experiences—from AI and data centers, to... ...high-performance GPU kernels, powering major AI... ...architecture, and performance engineering. You are comfortable... ...rigorous profiling and microbenchmarking...
$174k - $252k
...software/hardware codesign, kernel optimization, performance benchmarking... ...Learning (ML) hardware accelerators and ML inference software.... ...qualifications:PhD in Computer Engineering, Computer Science, or a... ...compute platform for cutting-edge AI inference workloads, with an...$174k - $252k
...ML compiler stack to bridge AI workloads and low-level Hardware... ....Drive SW/HW codesign, kernel optimization, performance tuning... ...machine learning (ML) hardware accelerators, ML compiler stacks (e.g.,... ...qualifications:PhD degree in Computer Engineering, Computer Science, or a...$120k - $275k
...designs hardware tailored for the world’s best AI models. Our hardware will make a given... ...or subsystem level.MatX is seeking an AI Accelerator Compute Architect to help define the... ...tradeoffs.Programming, code optimization, or kernel-level performance analysis experience....Daily paidFull timeWork experience placementWork at officeLocal areaRemote workMonday to FridayFlexible hours$150k - $230k
...researchers and veteran systems engineers who share a vision for... ...computing. As AI workloads grow... ...failures, and performance acceleration that dynamically routes... ...them. Examples: Kernel subsystems, device drivers... ...systems Runtimes, profilers, or performance...- NVIDIA in Santa Clara, CA is seeking outstanding AI systems engineers to advance the inference software stack. You will design and optimize kernels, build new abstractions for LLM serving engines, and contribute to accelerators and runtimes that power large language models...
- Senior AI Systems Performance Engineer Palo Alto, California, United States The era... ...hidden value in their data, accelerate processes, reduce costs,... ...SambaNova software stack. Profile and enhance model... ...optimization Compiler, runtime, or kernel‑level optimization...Full timeTemporary workLocal areaFlexible hours
$184k - $287.5k
...Senior Developer Technology Engineer for the Public Sector!... ...techniques to GPU-accelerate leading applications in... .../Climate Modeling, and AI in HPC. You will be performing... ...for GPUs, including kernel optimization and strong... .... Experience in profiling and optimizing applications...Full timeWork experience placementRemote work$200k - $350k
...talented Design Verification Engineers to help verify and deliver Velaura... ...'s next generation Physical AI SoC. You will work closely... ...This role focuses on our AI acceleration units. We are designing custom... ...modeling, power analysis, profiling and workload characterization...Full timeFlexible hours- River AI Inc. is hiring a kernel engineer to build the foundational compute engine for its custom silicon. You will design and implement robust kernel generators and emit optimized assembly for our greenfield architecture. You will collaborate with compiler engineers, silicon...
- Platform Recruitment is seeking a SIPI Characterization Engineer for an AI hardware client in Mountain View, CA. The role focuses on high-speed SerDes and PDN characterization across AI accelerator systems, from design through production. The candidate will build automated...
$184k - $287.5k
...skilled and motivated software engineers to join us and build AI inference systems that... ...stacks, optimize GPU kernels and compilers, drive industry... ...teams to push the frontier of accelerated computing for AI.What you’... ...GPU hardware features; profile and optimize the inference...Full time$184k - $287.5k
...is possible on modern accelerator architectures.Join our... ...performance for emerging AI workloads. You will be... ...optimization, custom kernel development, and... ...and runtime systems to profile and accelerate new paradigms... ...Science, Computer Engineering, Electrical Engineering...Full time$90k - $180k
...Agentic WorkflowsBuild agentic AI services (planning, tool use,... ...behind stable APIs and SDKs. Accelerated Compute & Data... ...scalable workflow orchestration. Profile and optimize bottlenecks across... ...contribute to design reviews and engineering best practices. Mentor peers...Full timeTemporary workPart time$200k - $360k
...professional to own MLIR dialect design and lowering passes for their AI accelerator. You'll work closely with chip-design and software teams.... ...knowledge of tensor compilation, and 5+ years of compiler engineering experience. The role includes a competitive compensation...- A cutting-edge AI company in California is looking for a Member of Technical Staff for Kernel/Compiler/Communication. This critical role requires strong expertise in CUDA... ...with 5+ years of experience in performance engineering. The ideal candidate will design high-...
$120k - $275k
...designs hardware tailored for the world’s best AI models. Our hardware will make a given... ...subsystem level. MatX is seeking an AI Accelerator Compute Architect to help define the... ...tradeoffs. Programming, code optimization, or kernel-level performance analysis experience....Daily paidFull timeWork experience placementWork at officeLocal areaRemote workMonday to FridayFlexible hours$138k - $197k
...infrastructure and architecture for on-device AI/ML accelerators.Implement, model, analyze, and... ...qualifications:Bachelor's degree in Electrical Engineering, Computer Engineering, Computer... ...the use of simulation frameworks and profiling tools.Experience with CPU or ML...Worldwide$207k - $300k
...platforms (Compute, Storage, hardware accelerators like TPUs and GPUs).Own the Linux kernel, networking stack, and... ...SoCs.Partner directly with hardware engineering, accelerator (TPU/GPU) teams, and... ...continue to push technology forward.The AI and Infrastructure team is...Worldwide- ...build great products that accelerate next-generation computing experiences—from AI and data centers, to... ...an influential software engineer who is passionate about... ...the lowest-level GPU kernels to large-scale distributed... ...experience using GPU profiling and performance analysis...
$120k - $280k
...potential of their data, from AI to multicloud.... ...experienced Systems Software Engineers across multiple NetApp... ..., logs, tracing, and profiling ~Collaborate across... ...AI-assisted tools to accelerate design, development,... ..., OCI) ~Exposure to kernel subsystems, VFS, IO...Part timeWork at officeLocal area$218.8k - $335.3k
...General Motors, our Embodied AI teams are redefining what’s possible... ...looking for a Staff Software Engineer to provide technical... ...observability for on‑road incidents. Profile and optimize SDS components... ...Hands‑on experience with GPU/accelerator‑based ML inference, model...Full timeLocal areaRemote workWork from homeFlexible hours$90k - $125k
...security with the world’s most advanced AI-native platform. We work on large... ...multiplier to proactively and continuously accelerate execution, build expertise, uncover... ...macOS and Linux, in both user mode and kernel mode. Engineering software at that depth and that scale...Full timeWork experience placementInternshipWork at officeLocal areaRemote workWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Kernel Engineer for AI Accelerator - Profiling. Be the first to apply!



