Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Member of Technical Staff, Hardware, Kernel Engineer (Custom Silicon)

$200k - $420k

Doist

At River, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, custom training infrastructure, next-generation UIs, and frontier deep learning research. Who we are We are scientists, engineers, and builders from the industry's top tech companies and AI labs. We bring a proven track record of scaling consumer systems for hundreds of millions of users and architecting the pre-training infrastructure behind today's frontier models. About the Role We are looking for exceptional performance and kernel generation engineers to build the foundational compute engine for our high-performance custom silicon. In this role, you will design and implement robust kernel generators that programmatically emit optimized low-level assembly code for our greenfield hardware architecture. You will bridge the gap between high-level compilation and raw hardware capability, pushing our custom architecture to its absolute theoretical limits for critical deep learning operations (including GEMMs, FlashAttention, and custom activations). You will collaborate closely up and down the stack with compiler engineers, silicon architects, and deep learning researchers to unlock maximum compute efficiency. What You’ll Do Kernel Generator Development : Design and build C++ code-generation frameworks and meta-programming toolchains that automatically emit optimized custom ISA assembly code. Low-Level Compute Optimization : Author and optimize core deep learning primitives (GEMM/MatMul, Attention mechanisms, Convolutions, and element-wise layers) directly targeted at our custom hardware. Microarchitectural Tuning : Hand-craft and automate instruction scheduling, register allocation, and software pipelining to maximize ALU utilization and hide execution latency on our silicon. Memory Hierarchy Management : Design sophisticated tiling, double-buffering, and data-movement strategies to optimize on-chip SRAM utilization and minimize memory bandwidth bottlenecks. HW/SW Co-Design : Partner with the RTL and architecture teams to evaluate hardware simulations, provide feedback on the ISA, and influence the design of future compute units based on kernel execution profiles. Performance Profiling & Validation : Benchmark generated assembly against hardware simulators and silicon, utilizing hardware performance counters to eliminate performance gaps and ensure mathematical correctness. Minimum Qualifications Bachelor’s degree in Computer Engineering, Computer Science, Electrical Engineering, or a related field, and 5+ years of practical industry experience in low-level performance programming. Deep understanding of hardware programming models (e.g., CUDA, Triton, CUTLASS, or custom accelerator assembly) and a proven track record of shipping highly optimized kernels. Advanced knowledge of Computer Architecture, including vector units, execution pipelines, register files, and complex memory hierarchies (caches, SRAM, HBM/DRAM). Proficiency in modern C++ for building robust, scalable meta-programming and code-generation frameworks. Strong mathematical foundation in linear algebra operations and deep learning primitives. A highly collaborative mindset to push boundaries and co-design effectively with hardware and compiler teams. Preferred Qualifications Deep familiarity with implementing microarchitectural optimizations for Tensor Cores, matrix multiply-accumulate units, or custom vector extensions. Experience utilizing advanced C++ template metaprogramming or code-generation techniques to automate the creation of heavily parameterized kernel variants. Advanced experience with low-level hardware profiling tools, execution tracing, and utilizing performance counters to identify cache misses, pipeline stalls, and ALU bubbles. Logistics Location: This role is based in Austin, Texas or Palo Alto, California. Compensation: Depending on background, skills, experience, and location, the expected annual salary range for this position is $200,000 - $420,000 USD. Visa Sponsorship: We sponsor visas. We can't guarantee success for every candidate or role, but if you're the right fit, we're committed to working through the visa process. Benefits: River AI offers generous health, dental, and vision benefits, unlimited PTO, and relocation support as needed. #J-18808-Ljbffr Doist

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Member of Technical Staff, Hardware, Kernel Engineer (Custom Silicon) in Palo Alto, CA vacancy
  • $200k - $420k

    Member of Technical Staff, Hardware, Compiler Engineer At River, our mission is to create personal AI owned...  ...for local inference, custom training infrastructure,...  ...high-performance custom silicon. You will create and...  ...will take ownership of kernel algorithms, intermediate... 
    Suggested
    Local area
    Visa sponsorship
    Work visa
    Relocation package
    Flexible hours

    River AI Inc.

    Palo Alto, CA
    2 days ago
  • Member of Technical Staff — Kernel / Compiler / Communication RadixArk is seeking a deeply technical engineer who pushes the limits of performance for frontier AI systems. You will work...  ...engineer who enjoys working close to hardware and solving performance problems that... 
    Suggested
    Flexible hours

    RadixArk

    Palo Alto, CA
    8 hours ago
  • Member of Technical Staff — Kernel / Compiler / Communication About the Role RadixArk is seeking a Member of Technical Staff — Kernel...  ...systems. This is a deeply technical role for engineers who enjoy working close to hardware and solving performance problems that most... 
    Suggested
    Flexible hours

    RadixArk

    Palo Alto, CA
    2 days ago
  • Member of Technical Staff — Developer Technology About the Role...  ...on modern GPU hardware. Our systems sit...  ...performance inference engine that serves...  ...hardware, build kernels, deliver day-0 model...  ...hardware enablement: custom CUDA/ROCm/Triton...  ...models on new silicon Training systems... 
    Suggested
    Flexible hours

    RadixArk

    Palo Alto, CA
    4 days ago
  • $120k - $200k

     ...With our headquarters in Silicon Valley—and teams in Paris,...  ...datasets, or full-cycle data engineering, Abaka AI provides the...  ...About the Role As a Member of Technical Staff, Platform, you'll build full...  ...feature in front of real customers and users. You'll work across... 
    Suggested
    Full time
    Flexible hours

    Abaka AI

    Mountain View, CA
    1 day ago
  •  ...tools for on-demand custom ASICs at scale. Our...  ...impossible with current hardware paradigms. Born out...  ...team blends AI with Silicon with a founding team...  .... What You’ll Do As Member of the Technical Staff - Software at...  ...This is a foundational engineering role where you’ll design... 

    Architect Labs

    Palo Alto, CA
    1 day ago
  •  ...tools for on-demand custom ASICs at scale. Our...  ...impossible with current hardware paradigms. Born out...  ...team blends AI with Silicon with a founding team...  ...ll Do As a Founding Member of the Technical Staff on the RTL Design...  ...accelerators, SIMD vector engines, DSPs, GPUs, on-chip... 

    Architect

    Palo Alto, CA
    1 day ago
  • RadixArk is seeking a Member of Technical Staff — Inference to push the limits...  ...the intersection of systems engineering, ML infrastructure, and...  ...enjoy working close to the hardware-software boundary and solving...  ...with CUDA, Triton, or custom kernel optimization Experience with... 
    Worldwide
    Flexible hours

    Dormont Manufacturing Co

    Palo Alto, CA
    4 days ago
  • We are looking for a Member of Technical Staff with strong Python skills and a...  ...architectures tailored to customer-driven use cases. Scale Distributed...  ...: Directly influence engineering strategy and roadmap...  ...optimization, CUDA or Triton kernels, or profiling AI workloads... 

    S27a

    Mountain View, CA
    2 days ago
  • $13 per hour

     ...work, but it will be worth it. Join us in building America’s mortgage rails . The Team We\'re looking for Full Stack Engineers to join our Customer Success team. This team exists for one reason: make sure the people using Pylon can close loans quickly, confidently, and... 
    Immediate start

    Pylon

    Palo Alto, CA
    2 days ago
  •  ...RadixArk is seeking a Member of Technical Staff: Accelerator Systems to...  .... Most performance engineering assumes a single vendor...  ...accelerators. That means porting kernels and runtimes onto unfamiliar hardware, designing the...  ...working directly with silicon and our partners,... 
    Flexible hours

    RadixArk

    Palo Alto, CA
    1 day ago
  •  ...tools for on-demand custom ASICs at scale. Our...  ...impossible with current hardware paradigms. Born out...  ...team blends AI with Silicon with a founding team...  ...ll Do As a Founding Member of the Technical Staff at Architect, you'll...  ...and production engineering for chip designs, implementing... 

    Architect Labs

    Palo Alto, CA
    1 day ago
  •  ...to design on-demand custom ASICs at scale. Our...  ...impossible with current hardware paradigms. Born out...  ...team blends AI with Silicon with a founding team...  ...ll Do As a Founding Member of the Technical Staff - Formal Methods at...  ...-tool maximalist. Engineering Rigor: Strong software... 

    Kindredventures

    Palo Alto, CA
    4 days ago
  •  ...tools for on-demand custom ASICs at scale. Our...  ...impossible with current hardware paradigms. Born out...  ...team blends AI with Silicon with a founding team...  ...ll Do As a Founding Member of the Technical Staff on the Design...  ...or PhD in Electrical Engineering, Computer Engineering... 
    Night shift

    Doist

    Palo Alto, CA
    1 day ago
  • Member of Technical Staff — Supercomputing About the Role RadixArk is...  ...at the intersection of engineering, deployment, reliability, and customer infrastructure. You will...  ...team has optimized kernels serving billions of tokens...  ...cloud providers, hardware companies, and frontier... 
    Flexible hours

    RadixArk

    Palo Alto, CA
    1 day ago
  • RadixArk is seeking a Member of Technical Staff — Training to build...  ..., and performance engineering. Your work will directly...  ...-LM, FSDP, or custom training stacks Experience...  ..., scalability, and hardware efficiency Improve...  ...GPUs and optimized kernels serving billions of... 
    Flexible hours

    RadixArk

    Palo Alto, CA
    1 day ago
  •  ...tools for on-demand custom ASICs at scale. Our...  ...impossible with current hardware paradigms. Born out...  ...team blends AI with Silicon with a founding team...  ...ll Do As a Founding Member of the Technical Staff (Applied AI) at...  ...translating deep hardware engineering expertise into... 

    Architect Labs

    Palo Alto, CA
    2 days ago
  •  ...founded by pioneers of AI silicon, systems, software, and infrastructure...  ...knit, comprised of senior engineering leaders—many of whom have...  ...for an exceptional Member of Technical Staff to help design, build, and...  ...components across hardware, systems, and/or software... 

    DensityAI

    Mountain View, CA
    4 days ago
  •  ...Role RadixArk is seeking a Member of Technical Staff — Training to build and...  ...systems, and performance engineering. Your work will directly impact...  ..., scalability, and hardware efficiency Improve reliability...  ...training. Our team has optimized kernels serving billions of tokens... 
    Flexible hours

    RadixArk

    Palo Alto, CA
    3 days ago
  • Member Of Technical Staff - Extreme-Scale Sparse Linear Algebra, Domain Decomposition...  ...infrastructure that modern hardware programs use to converge on...  ...just prototypes. Systems & Engineering Expectations CUDA first, HIP appreciated Kernel-level performance... 
    Full time
    Remote work

    Vinci4d

    Palo Alto, CA
    3 days ago
  • $180k

    Member of Technical Staff, Recommendation Systems About xAI xAI’s mission is to create AI systems that...  ..., highly motivated, and focused on engineering excellence. This organization is for...  ...candidates may be experienced in writing CUDA kernels Interview Process After submitting... 
    Temporary work
    Relocation

    Xai

    Palo Alto, CA
    8 hours ago
  • $180k - $250k

    Member of Technical Staff -- TPU Systems (JAX / XLA / PALLAS) About the Role RadixArk...  ...looking for a TPU Systems Engineer to build high-performance...  ...to their limits on TPU hardware, working on SGLang-JAX and...  ...Familiarity with Pallas for kernel development, or strong ability... 
    Full time
    Flexible hours

    RadixArk

    Palo Alto, CA
    4 days ago
  •  ...hiring a Performance Engineer in Palo Alto, CA — someone...  ...scheduling, batching, kernel behavior, distributed...  ...in production. Our customers care about real numbers...  ...cloud partners on deep technical evaluations Contribute...  ...across software, hardware, and infrastructure layers... 
    Flexible hours

    RadixArk

    Palo Alto, CA
    3 days ago
  • $180k

     ...small, highly motivated, and focused on engineering excellence. This organization is for individuals...  ...the full stack — from low‑level GPU kernel optimizations and Linux kernel internals...  ...isolation at cluster scale Build custom container orchestration, virtualization... 
    Temporary work

    xAI

    Palo Alto, CA
    1 day ago
  • About The Role RadixArk is seeking a Member of Technical Staff — Diffusion Model to advance the frontier...  ...deep research thinking with strong engineering execution—from designing novel...  ...and training. Our team has optimized kernels serving billions of tokens daily, designed... 
    Flexible hours

    RadixArk

    Palo Alto, CA
    8 hours ago
  • RadixArk is hiring a Member of Technical Staff — CI Engineer to own the infrastructure that keeps SGLang moving...  ...NVIDIA, AMD, Intel, and Ascend hardware pools, gating every commit to one of...  ...and training. Our team has optimized kernels serving billions of tokens daily, designed... 
    Flexible hours
    Night shift

    RadixArk

    Palo Alto, CA
    3 days ago
  • $148.5k - $223.9k

    Senior Member of Technical Staff - AI ResearchSkip to main content#Senior Member of Technical Staff...  ...CRM, where humans with agents drive customer success together. Here, ambition meets...  ...you are the future of Salesforce.*ware engineers, product managers and solution... 
    Work at office

    Salesforce, Inc.

    Palo Alto, CA
    8 hours ago
  • $180k

    Member of Technical Staff - Recommendation Systems SpaceXAI’s mission is to create AI systems that can...  ..., highly motivated, and focused on engineering excellence. This organization is for...  ...candidates may be experienced in writing CUDA kernels COMPENSATION AND BENEFITS: $180,000... 
    Temporary work

    SpaceXAI

    Palo Alto, CA
    1 day ago
  •  ...an AI assistant for hardware design. Our product helps engineers explore, analyze,...  ...Summary Hands-on technical role spanning validation...  ...training data, and customers. Vinci's physics...  ...and GPU kernel engineers building...  ...the right person — members of the team already... 
    Full time
    Remote work
    2 days per week
    3 days per week

    Vinci4d

    Palo Alto, CA
    2 days ago
  • $130k - $170k

     ...delivers freight daily for its customers across the southern...  ...looking for an experienced Hardware Quality Industrialization Engineer to join our team. Our...  ...person will be the founding member of the quality and will...  ...the base salary in our SF/Silicon Valley location, across several... 
    Temporary work
    Work at office
    Visa sponsorship

    Kodiak Robotics

    Mountain View, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Member of Technical Staff, Hardware, Kernel Engineer (Custom Silicon). Be the first to apply!