GPU Kernel Engineer — Fast ML Training
MakerMaker.AI
MakerMaker.AI in San Francisco is seeking a skilled Software Engineer to write and optimize GPU kernels. You will work on deep low-level tasks that directly impact the performance of machine learning models. The ideal candidate has over 4 years of experience with GPU kernels, strong systems expertise, and a proven track record in kernel optimizations. This role requires on-site work in a collaborative environment. #J-18808-Ljbffr MakerMaker.AI
$167.2k - $209k
...bold, and are energized by the fast-paced environment of a true... ...DigitalOcean is seeking a Senior Engineer 2 to play a key technical... ...at the inference engine and GPU kernel layers, ensuring our infrastructure... ...for relevant conferences, training, and education. All employees...TrainingLocal areaRemote workWorldwideFlexible hours$180k - $280k
...in production. We're a small, fast-moving team from OpenAI,... .... Since mid-2024, we've been engineering the foundation for what comes... ...the role We're looking for a GPU kernel engineer with deep, low-level CUDA expertise to make our training and inference faster and more...TrainingWork at officeVisa sponsorshipShift work$285k - $315k
...We're looking for a Founding GPU Kernel Engineer who lives right at the boundary... ...-optimize GPU kernels for ML workloads (matmuls, attention... ...Experience with distributed training systems: collective ops like... ...wants to know why things are fast or slow on the hardware. You'...TrainingFull timeWork at officeRelocation package- A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU kernels and using advanced techniques to boost computation efficiency. Ideal...Suggested
$100k - $120k
...robotic foundation models. As training and inference workloads grow, we need kernel‑level innovations to... ...of kernel and system engineers focused on performance-critical... ...for CPU (AVX/ARM NEON), GPU (CUDA/ROCm), and... ...optimizations into distributed ML frameworks (e.g.,...Training- ...ll write and optimize the GPU kernels and supporting systems software that makes our training and inference workloads fast. This is deep, low-level work... ...use. We hire kernel engineers because the gap between "this... ...libraries, compilers, or ML frameworks Experience with...TrainingShift work
$285k - $315k
SF Tensor is looking for a Founding GPU Kernel Engineer in San Francisco, specializing in GPU architecture and kernel optimization for machine learning workloads. The ideal candidate has deep expertise, proven capabilities in hand-optimizing performance-critical kernels...Full timeRelocation package- San Francisco Tensor Company is seeking a Founding GPU Kernel Engineer to enhance GPU performance for AI applications. You will optimize and write kernels while collaborating with compiler teams to improve efficiencies across architectures. The ideal candidate has deep...Work at officeRelocation package
- Anthropic, a public benefit corporation headquartered in San Francisco, is seeking a TPU Kernel Engineer to identify and address performance issues across ML systems, including research, training, and inference. You will design and optimize kernels for the TPU and provide...Training
- Baseten is seeking an Engineering Manager to lead our GPU Kernel Engineering team, directing the low-level CUDA work that accelerates Baseten's inference stack... ...direction at the intersection of GPU architecture, ML systems, and production inference, while building processes...
- Inception is seeking engineers and scientists to design, optimize, and maintain compute foundations for large‑scale language model training and inference. You will develop high‑performance ML kernels, enable efficient low‑precision arithmetic, and improve the distributed...Training
$130k - $200k
...to apply for the Founding Engineer (Systems + ML) role at Partcl . Get AI-... ...Develop and optimize GPU‑accelerated engines (C++/... ...pipelines: high‑performance kernels, efficient file IO, training models, latency-sensitive... ...system design (N~>10M). Fast learner: able to absorb EDA...TrainingFull time- ...shipping excellence. We seek engineers with strong intrinsic... ...systems to optimize GPU performance at the... ...Design and optimize GPU kernels and tensor libraries Translate... ...with distributed training/inference frameworks (bonus... ...leaders Ambitious, fast-paced startup culture...TrainingFull timeWork at office
$179k - $218k
...seeking a Senior Staff Data Center Operations Engineer, GPU Hardware Architecture to be the... ...Predictive Operations & Telemetry: Leverage AI/ML methodologies to analyze fleet-wide... ...components before they impact customer training runs.Technical Sparing Architecture: Architect...TrainingTemporary work- ...is looking for a Senior Software Engineer to build scalable infrastructure for large‑scale training and fine-tuning of foundation... ...distributed training systems and optimize GPU utilization while collaborating... ...over 5 years of experience in ML infrastructure and a strong...Training
- ...possible in video generation.We're seeking a GPU Performance Engineer to squeeze every last FLOP from our H10... ...10x speedups. From writing custom CUDA kernels to eliminating cold start latency, you'... ..., and GPU utilizationCollaborate with ML engineers to optimize model...
$250k
...provider building a next-generation GPU platform designed for AI training, experimentation, and inference at scale... ...a Senior / Staff Site Reliability Engineer to support and scale large-scale HPC... ...working closely with platform, ML, and infrastructure teams to improve...TrainingFull timeRemote work$150k - $300k
...enables anyone to create, train, and deploy them. We... ...Solutions Architect for GPU Infrastructure, you'll... ...system performance from kernel parameters to CUDA... ...DeepSpeed, Megatron‑LM) ML framework optimization... ...collaborate with our world‑class engineering team while having...Training- Magic is hiring a Kernel Engineer in San Francisco to design, implement, and optimize high-performance kernels for long-context training and inference. You will tackle memory usage, data movement... ...Experience with AI accelerators and GPU kernel frameworks is essential; visa...TrainingVisa sponsorship
$285k - $315k
...portable. We are building a Kernel Optimizer that... ...partnering with researchers, engineers, and organizations who... ...'re hiring a Founding GPU Compiler Engineer to build... ...for large-scale AI pre-training. You'll own the entire... ...Work closely with ML researchers to understand...TrainingFull timeWork at officeRelocation package$315k
Performance Engineer, GPU Join to apply for the Performance Engineer, GPU... ...‑art techniques from custom kernel development to distributed system... ...improvements in production ML systems and will be excited... ...Production Systems: Large‑scale training infrastructure, fault...TrainingFull timeWork at officeVisa sponsorshipFlexible hours- ...patients worldwide.We’re a team of engineers, clinicians, and innovators... ...PositionAs a Senior Systems GPU Engineer - AI & Robotics, you... ...work alongside research, SW/ HW/ ML engineering, regulatory,... ...Virtualization: Development of Linux kernel internals, device drivers,...Local areaWorldwideFlexible hours
$110 per hour
...Dorsey . Position: MLOps Engineer (JAX, PyTorch, Pallas/... ...performance in MLOps , training infrastructure, and ML framework-level topics .... ...distributed systems reasoning, and kernel-level optimization across... ...or optimizing custom GPU kernels using Pallas (JAX...TrainingRemote jobContract workSummer workWeekday work$342k
...AI.About the RoleAs an Engineer on our hardware optimization... ...You will work with our kernel, compiler and machine... ...needs related to ML techniques, algorithms,... ...architectures towards efficient training and inference on our... ...understanding of GPU and/or other AI acceleratorsExperience...TrainingWork at officeLocal areaRelocation packageFlexible hours$206.3k - $388k
...We’re looking for a Principal ML Engineer to architect and scale the multimodal... ...ML building distributed, GPU-accelerated systems that turn billions of raw assets into training-ready data at scale. Your work will directly determine how fast and how well Adobe models can...TrainingFull timeTemporary workLocal areaWorldwide- ...us and help build the platform engineers turn to to ship AI products.... ...foundational engineers to lead our GPU Networking efforts, making... ...system behaviors. Optimize Kernels: You will work with communication... ...) ~ Exposure to a variety of ML startups, offering...Full timeFlexible hours
- Senior ML Systems Engineer, Frameworks & Tooling at Cohere Our mission is to... ...to serve humanity. We’re training and deploying frontier models... ...build and work hard and move fast to do what’s best for our customers... ...libraries, or custom kernels/fused ops. Experience with...TrainingFull timeWork at officeRemote workFlexible hours
- ...Francisco is seeking a Staff ML Systems Engineer to design and prototype... ...inference engines, including kernel backends and ATLAS-style systems... ..., while profiling across GPU, networking, and memory to improve... ...also co-design RL and post-training pipelines, drive performance...Training
$300 per month
...Role:At Crusoe, our Production Engineering team ensures the reliability... ...teams to optimize large-scale training and inference clustersAutomate... ...language models (LLMs) or AI/ML infrastructureSRE mindset and... ...skillsAbility to thrive in a fast-paced, mission-driven environmentBonus...TrainingTemporary work$132.1k - $258.4k
...our team thinks big and moves fast, and we're looking for smart,... ...applications? As a software quality engineer at WRITER, you'll play a... ...complex distributed systems or AI/ML applications.5+ years of... ...massage/chiropractor, personal training, etc.Learning and development...TrainingFull timeWork at officeLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to GPU Kernel Engineer — Fast ML Training. Be the first to apply!
- machine learning scientist San Francisco, CA
- machine learning part time San Francisco, CA
- machine learning intern San Francisco, CA
- machine learning remote San Francisco, CA
- machine learning researcher San Francisco, CA
- machine learning San Francisco, CA
- machine learning research scientist San Francisco, CA
- internship machine learning San Francisco, CA
- artificial intelligence - machine learning intern San Francisco, CA
- data engineer machine learning San Francisco, CA



