Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Founding Compiler Engineer - AI Models, CUDA & GPU

Slope

Luminal is hiring a Founding Compiler Engineer for an on-site role in downtown San Francisco. You will help shape the core compiler, implement CUDA kernels, and review model performance to accelerate AI pipelines. As a founding team member, you will contribute to production-grade deployment of AI models, with a focus on high-performance optimization and practical code paths for real-world workloads. #J-18808-Ljbffr Slope

Vacancy posted 20 hours ago
Similar jobs that could be interesting for youBased on the Founding Compiler Engineer - AI Models, CUDA & GPU in San Francisco, CA vacancy
  • $285k - $315k

     ...believe the future of AI and high-performance computing...  ...with researchers, engineers, and organizations who...  ...Role We're hiring a Founding GPU Compiler Engineer to build the...  .... That means taking models from PyTorch, JAX, and...  ...low-level optimization (CUDA, ROCm, or equivalent)... 
    Suggested
    Full time
    Work at office
    Relocation package

    San Francisco Tensor Company

    San Francisco, CA
    20 hours ago
  • $285k - $315k

    San Francisco Tensor Company is looking for a Founding GPU Compiler Engineer to be instrumental in building the core compilation infrastructure for AI workloads. This role involves designing compilation pipelines and optimizing for various GPU architectures. Join a passionate... 
    Suggested

    San Francisco Tensor Company

    San Francisco, CA
    20 hours ago
  • $285k - $315k

     ...believe the future of AI and high-...  ...with researchers, engineers, and organizations...  ...'re looking for a Founding GPU Kernel Engineer who...  ...that knowledge into compiler optimization passes that help every model we compile. What...  ...programming in C++ and CUDA (or ROCm/HIP)... 
    Suggested
    Full time
    Work at office
    Relocation package

    San Francisco Tensor Company

    San Francisco, CA
    20 hours ago
  •  ...are needed in generative modeling, reinforcement learning...  ...Role As a Software Engineer, you will help build AI systems that can...  ...Triton , a language and compiler for writing custom GPU kernels. The aim of Triton...  ...higher productivity than CUDA.  We frequently... 
    Suggested
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  •  ...patients worldwide.We’re a team of engineers, clinicians, and innovators...  ...PositionAs a Senior Systems GPU Engineer - AI & Robotics, you will be...  ....Responsibilities• GPU & Model Performance & Optimization:...  ...Expert in GPU Compute API - CUDA, OpenCL• Proficiency in multiple... 
    Suggested
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    San Francisco, CA
    2 days ago
  •  ...innovative company is seeking a talented software engineer to join their dynamic Inference team....  ...for large-scale multimodal models, focusing on high-performance delivery of...  ...product teams to push the boundaries of AI technology, ensuring reliable production... 

    Jobleads-US

    San Francisco, CA
    4 days ago
  • A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU...  ...candidates have 1-5 years of CUDA development experience and a strong... 

    Baseten

    San Francisco, CA
    3 days ago
  •  ...focused on optimizing AI models to accelerate and...  ...We are building an AI compiler that enhances model speeds...  ...on-site role for a Founding Compiler Engineer located in downtown...  ...will include writing CUDA kernels and conducting...  ...Strong proficiency in GPU model optimization... 
    Full time

    Luminal

    San Francisco, CA
    20 hours ago
  •  ...founders and senior engineers with deep expertise in...  ...We're looking for a Founding Engineer, ML Inference...  ...from generative media models. You'll work across...  ...optimizations using torch.compile, custom CUDA kernels, and...  ...Working knowledge of GPU hardware (NVIDIA) and... 
    Relocation
    Visa sponsorship
    Relocation package

    Reactor.am

    San Francisco, CA
    5 days ago
  • $100k - $120k

     ...generation robotic foundation models. As training and...  ...You will join Coda's founding team to architect and...  ...of kernel and system engineers focused on performance...  ...for CPU (AVX/ARM NEON), GPU (CUDA/ROCm), and hardware...  ...operations Experience with compiler design or code... 

    Coda Robotics

    San Francisco, CA
    1 day ago
  • San Francisco Tensor Company is seeking a Founding GPU Kernel Engineer to enhance GPU performance for AI applications. You will optimize and write kernels while collaborating with compiler teams to improve efficiencies across architectures. The ideal candidate has deep... 
    Work at office
    Relocation package

    San Francisco Tensor Company

    San Francisco, CA
    1 day ago
  • TypeSafe AI in San Francisco is seeking a GPU kernel engineer with deep CUDA expertise to accelerate our training and inference workloads. You will write, optimize, and maintain high-performance kernels close to the metal. Work alongside research and platform engineers... 
    Work at office

    TypeSafe AI

    San Francisco, CA
    20 hours ago
  •  ...for the world's most dynamic AI companies, like Cursor, Notion...  ...frontier of AI to bring cutting-edge models into production. We're growing...  ...and help build the platform engineers turn to to ship AI products....  ...engineers to lead our GPU Networking efforts, making RDMA... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    1 day ago
  • $342k

     ...demands of advanced AI workloads. The team...  ...integrated with AI models. In addition to...  ...About the RoleAs an Engineer on our hardware optimization...  ...with our kernel, compiler and machine...  ...designDeep understanding of GPU and/or other AI...  ...with CUDA, Triton or a related... 
    Work at office
    Local area
    Relocation package
    Flexible hours

    OpenAI

    San Francisco, CA
    3 days ago
  • $220k - $320k

     ..., diving deep into CUDA kernels, and turning...  ...language models for companies that...  ...need frontier-quality AI at a fraction of the...  ...ten‑person team of engineers who work in‑person...  ...10 years, and has founded and run their own software...  ...optimize CUDA kernels and GPU utilization across... 
    Work at office

    SOLANA FOUNDATION

    San Francisco, CA
    20 hours ago
  • $160k - $200k

     ...About Vast.ai Vast.ai runs one of the world's largest GPU marketplaces: 20,000+ GPUs...  ...tune, and serve AI models. We're profitable,...  ...role This is an engineering role that happens in...  ...GPU stack: driver/CUDA mismatches, OOM errors...  ...or projects that found a real audience... 
    Full time

    Vast.ai

    San Francisco, CA
    1 day ago
  • $130k - $200k

     ...us to apply for the Founding Engineer (Systems + ML) role at Partcl . Get AI-powered advice on this...  ...and optimize GPU‑accelerated engines (C++/CUDA) for timing analysis...  ...efficient file IO, training models, latency-sensitive...  ...EDA, chip design, compilers, physics-driven ML.... 
    Full time

    Partcl

    San Francisco, CA
    3 days ago
  •  ...serve OpenAI’s frontier models at massive scale. As part...  ...for a kernel-focused engineer to lead efforts in writing...  ...porting, and optimizing GPU kernels used in inference...  ...deep familiarity with CUDA or equivalent kernel programming...  ...OpenAI OpenAI is an AI research and deployment... 
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  •  ...that our most advanced models run efficiently, reliably...  ...the Role We’re hiring engineers to scale and optimize...  ...infrastructure across emerging GPU platforms. You’ll work...  ...GPU kernels using HIP, CUDA, or Triton, and care...  ...OpenAI OpenAI is an AI research and deployment... 
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  •  ...the world's most dynamic AI companies, like Cursor,...  ...AI to bring cutting-edge models into production. We're...  ...help build the platform engineers turn to to ship AI products...  ...We’re seeking a GPU Kernel Engineer to join...  ...and optimize code using CUDA, PTX assembly, and architecture... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    1 day ago
  •  ...enable any scientist to access AI-powered drug discovery....  ...feasible until now. New AI models are quickly eclipsing physics...  ...hiring three exceptional Founding Software Engineers to help us scale the...  ...EC2, S3, DynamoDB), Docker, CUDA, Conda, TensorFlow/PyTorch;... 
    Full time
    Relocation

    Tamarind Bio

    San Francisco, CA
    1 day ago
  •  ...Origin is building Physical AI for the built world -...  ...Construction Action Models which allows our modular...  ...models for edge: TensorRT compilation, latency profiling, memory...  ...like Isaac Sim. GPU profiling and optimization (TensorRT, ONNX, CUDA); you understand why 200... 

    Origin

    San Francisco, CA
    more than 2 months ago
  • $220k

    Perplexity is looking for an engineer to join their team in San Francisco. You will work on building and operating the inference engine, supporting new models, migrating GPU kernels, and developing a Rust-based serving runtime. The ideal candidate has 3+ years of experience... 

    Perplexity

    San Francisco, CA
    3 days ago
  •  ...and access our start-of-the-art AI models, allowing them to do things...  ...Role We are looking for an engineer who wants to take the world's...  ...utilize every FLOP and every GB of GPU RAM of our hardware. You...  ...that optimize them (e.g. NCCL, CUDA), as well as HPC technologies... 
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  •  ...the world's most dynamic AI companies, like Cursor,...  ...AI to bring cutting-edge models into production. We're...  ...help build the platform engineers turn to to ship AI products...  ...-LLM kernels, analyze CUDA kernel performance, implement...  ...patterns across multi-GPU setups Productionize... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    1 day ago
  •  ...is building next-generation speech models and inviting you to join its founding research team in San Francisco. You...  ...researchers who publish, present at top AI/ML conferences, and own long-term...  ...challenging research and hands-on engineering, apply now. #J-18808-Ljbffr Socket... 
    Relocation

    Socket.dev

    San Francisco, CA
    2 days ago
  • $166k - $225k

     ...the world's best data and AI infrastructure platform...  ...business. Databricks’ Model Serving product provides...  ...efficiency.As a Senior Engineer, you’ll play a critical...  ...inference across CPU and GPU workloads, influence architectural...  ...the globe and was founded by the original creators... 
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    2 days ago
  •  ...We are looking for a founding backend engineer with unusually strong engineering...  ...Hypercubic, we’re building AI systems that can understand,...  ...driven features end to end, from model interaction to user-facing...  ...especially mainframes or COBOL), compilers, program analysis techniques... 
    Full time
    H1b
    Visa sponsorship
    Work visa
    Relocation package

    Hypercubic

    San Francisco, CA
    1 day ago
  • $190k - $265k

     ...passionate about enabling data and AI teams to solve the world's...  ...to improve their business. Founded by engineers — and customer-obsessed — we...  ...everything from data apps, AI agents, model training, model serving, and...  ..., ML infrastructure, or GPU orchestrationFamiliarity with... 
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    2 days ago
  • $192k - $260k

     ...world's best data and AI infrastructure platform...  ...their business. Foundation Model Serving is the API...  ...necessary. We’re looking for engineers who have owned high...  ...low-latency inference on GPU workloads with frontier...  ...the globe and was founded by the original creators... 
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Founding Compiler Engineer - AI Models, CUDA & GPU. Be the first to apply!