Founding Engineer - ML Performance
$250k - $395kuRun
Role Description
Performance is uRun's core differentiator. We're not chasing incremental gains — we're building infrastructure that runs 10–100x faster than the status quo. As our ML Performance Engineer, you will be the person who makes that true.
This is a founding technical hire. You will:
- Write custom CUDA kernels, pushing GPU utilization to its limits.
- Own inference latency end-to-end across the stack.
- Work directly with the founding team on the hardest performance problems in production AI infrastructure.
- Have your fingerprints on everything we ship.
What you'll actually be doing day-to-day:
- Write custom CUDA kernels that unlock performance headroom unavailable through off-the-shelf frameworks.
- Optimize model inference end-to-end, targeting sub-50ms latency across our inference platform.
- Drive 10x performance improvements across the stack: memory bandwidth, kernel fusion, operator scheduling, and beyond.
- Implement zero-copy distributed memory optimizations across multi-GPU and multi-node environments.
- Own GPU utilization and memory management, squeezing every available FLOP out of the hardware we run.
- Profile, benchmark, and instrument the full inference pipeline to find and eliminate bottlenecks systematically.
- Set the performance engineering bar for the team: define what fast looks like and build the tooling to measure it.
Qualifications
- Deep, hands-on CUDA expertise: you have written custom kernels in production, not just called into cuBLAS.
- Strong background in model inference and post-training optimization at scale.
- Fluency in GPU memory hierarchy, warp scheduling, kernel fusion, and hardware-aware algorithm design.
- Experience profiling and benchmarking complex inference pipelines: you know where the time goes and how to get it back.
- Able to operate at the frontier with minimal guidance — you identify the problem, design the approach, and ship the fix.
Requirements
- Public work in GPU optimization or inference efficiency — open source contributions, a published paper, or a side project that shows your depth (vLLM, Flash-Attention, TensorRT-LLM, PyTorch, or equivalent).
- Experience with hardware-aware optimization frameworks: CuTe, Triton, TileLang, or similar.
- Familiarity with distributed memory and communication primitives: NCCL, InfiniBand, NVLink, RoCE.
- Contributions to or deep familiarity with PyTorch Distributed, Ray core, or similar systems.
- Experience optimizing for video generation or other high-throughput, latency-sensitive generative workloads.
- Prior work at an inference-focused company or research lab pushing the boundary of what GPU hardware can do.
Benefits
- Competitive salary and meaningful equity in an early-stage AI infrastructure company.
- Health, dental, and vision — full coverage.
- 401(k) — company-supported retirement savings.
- FSA/HSA — flexible spending accounts for healthcare costs.
- Paid time off — we trust you to manage your time.
- Top-tier tooling — access to the best AI tools available: Claude, Codex, Kimi, and whatever else helps you move faster.
- MacBook Pro and AirPods — the hardware you need, on us.
- ...experiences. We're building our engineering hub in Argentina from the... ...standards, and the culture of a founding team, with real ownership... ...infrastructure (backend/core), or high-performance product interfaces at scale (... ...-cloud · probabilistic & ML systems. You don't need...PerformanceFull timeRemote work
- # Remote deep learning engineer jobsWe, at Turing, are looking for remote... ...Turing jobs a year back, I found the Turing test to be very... ...experience for me. Through the AI/ML-enabled tests, Turing reviews... ...aid in the development of high-performance machine learning models.If you...PerformanceFull timeRemote workWorldwide
$140.8k - $211.2k
...Qualcomm Technologies, Inc.Job Area:Engineering Group, Engineering Group >... ...background in AI and general ML techniquesProven hands-on... ...Generative AI workflows for accuracy, performance, and other key metricsML... ...Qualcomm's toll-free number found here. Upon request, Qualcomm...PerformanceWork experience placementWork from home- ...apply for the Developer Relations Engineer role at Canonical 2 days ago... ...members wherever they may be found - IRC, social media, product... ...geographical location, experience, and performance in shaping compensation... ...Software Engineer - Data, AI/ML & Analytics Software Engineer,...PerformanceFull timeLocal areaRemote workWorldwide
- ...Integration and Technology Development Engineer defines, develops, and... ...our company benefits can be found here:We are committed to sourcing... ..., attracting, and hiring high-performance innovators, while providing... ...structures. Familiarity with AI/ML techniques, data mining, or...PerformanceLocal area
- On The Stage (OTS) is at an inflection point. As our Founding GTM Engineer , you won’t just maintain systems — you’ll be the architect of how... ...Finance. Measurement & Continuous Improvement Own the performance of GTM automation and AI-powered workflows — identifying opportunities...PerformanceRemote jobShift work
$170k - $200k
...revolution. The Instawork Robotics ML Engineer will help build and scale the technology... ...and analytics to measure the quality and performance of our dataset and ML models.... ...job description. About Instawork Founded in 2015, Instawork is the nation’s leading...PerformanceHourly payFull timeInternshipLocal areaFlexible hoursShift work- ...integrated, and designed for scale. Our founding team has scaled category-defining... ...seeking a highly experienced Principal ML Engineer (Applied / Systems) to join our engineering... ...solutions, all with a strong focus on performance, maintainability, and impact. What...PerformanceFull time
$160k - $240k
...autonomy accessible to all. Founded in 2016, Nuro is building the... ...are looking for self-motivated engineers to build the next-generation... ...mission is to provide a high-performance, highly reliable foundation of... ...Robotics experience, ML inference optimization experience...PerformanceFull time$139.76k - $287.75k
...process here.The Production Engineering organization at Pinterest is... ...at scaleManage capacity and performance to help scale our infrastructure... ..., or reliability workflowsAI/ML infrastructure experience (LLM... ...for this position can be found here.US based applicants only...PerformanceWork at officeLocal areaRemote workRelocationRelocation package$193.93k - $291.15k
...all roads and all rides. Founded in 2016, Nuro is a physical AI... ...we are looking for a Software Engineer to join our Sensor Data and... ...about sensor data and autonomy performance Collaborate with... ...experience Deep understanding of ML fundamentals with hands-on experience...PerformanceFull timeImmediate startFlexible hours$150k - $250k
Job Title: Senior Analog Mixed-Signal Engineer - AI HardwareJob Location: Palo Alto, CA, Austin... ...compute building blocks including high-performance ADCs/DACs, data converter interfaces, and... ...of algorithms and circuits for ML workloads.Strong programming skills in Python...PerformanceRemote work$184k - $287.5k
...the world.We are seeking an AI Compiler Engineer with deep expertise in compiler technologies... ..., delivering measurable improvements in performance and efficiency, and advancing LLM-enabled... ....8+ years of software engineering and AI/ML experience, preferably in tools or...PerformanceFull timeRemote work$98k - $176k
...team members need and deserve. Our high-performing teams balance independence with collaboration... ...from the inside out.The Tech Inventory Engineering team:Develops user-first tooling that... ...to enable organization-wide, AI/ML-powered analyticsOur Tech Stack:TypeScript...PerformanceFull timeTemporary workWork experience placementFlexible hours$150k - $250k
...projects of this program since our founding in 2013. We've grown on this... ...providing the customer with Engineers who have done exceptional... ...Demonstrated experience with ML frameworks like scikit-learn,... ...correct errors or to improve performance Experience optimizing algorithms...PerformanceContract workWork at officeImmediate startRemote work$120k - $275k
...systems including hardware and software to train and run the largest ML workloads for AGI. MatX is seeking Emulation engineers to join our team as we create best-in-class silicon for high-performance and sustainable GenAI. Successful candidates for these roles will be...PerformanceFull timeWork experience placementLocal areaRemote workMonday to FridayFlexible hours$121.5k - $188.5k
...intersection of four disciplines: agent engineering, context engineering, evaluations and guardrails... ..., BigQuery, Redshift) or integrating ML models into production services.... ...ResourcesThis position may be eligible for performance-based incentives/bonuses. Benefits include...PerformanceFull time$116.96k - $160.82k
...signal processing, data analytics, and AI/ML-enabled algorithms are core... ...hands-on and experienced Senior Algorithms Engineer to lead the design, development, and optimization... ...algorithms, from concept through prototyping and performance evaluation.Apply signal processing...PerformancePermanent employmentFull timeContract workWork at officeRemote workDay shift$166k - $258k
...intersection of four disciplines: agent engineering, context engineering, evaluations and guardrails... ...solutions; drive accountability for performance, cost, accuracy, and security of feature... ..., BigQuery, Redshift) and integrating ML models into production services....PerformanceFull timeTemporary work- ...across Intelligence, Analytics, Engineering, Mission Support, and Communications disciplines. Founded in 2008, our mission is to... ...deploying advanced cloud-enabled AI/ML solutions that enhance... ...testing and evaluation of UI/UX performance against operational datasets....PerformanceRemote jobFor contractorsCasual work
$152k - $241.5k
NVIDIA is currently seeking a Senior Developer Technology Engineer for High-Performance Databases!Would you enjoy researching new algorithms and memory... ...and are becoming the bottleneck for Machine Learning (ML) and Deep Learning (DL) applications, as performance of the...PerformanceFull timeRemote work$154k - $231k
Company:Qualcomm Technologies, Inc.Job Area:Engineering Group, Engineering Group > DSP... ...IoT, and Automotive roadmap are our high-performance, multithreaded, low-power Hexagon cores.... ...com or call Qualcomm's toll-free number found here. Upon request, Qualcomm will provide...PerformanceWork experience placementWork from home$200k - $250k
...Software Engineer (AI Investor) – New York, Hybrid A seed‑stage Applied AI lab building a fully autonomous AI investor—a system designed... ...‑end What They’re Looking For: 4+ years of experience in high‑performance engineering environments Proven track record of shipping...PerformanceRemote work$216k - $270k
...world problems remains one of the hardest engineering challenges.As a Senior Frontier Agent... ...enterprise software.Unlike traditional ML roles that focus on a single model or product... ..., experience, qualifications, interview performance, and relevant education or training....PerformanceFull time$120k - $275k
...including hardware and software to train and run the largest ML workloads for AGI. MatX is seeking silicon micro-architects and design engineers to join our team as we create best-in-class silicon for high-performance and sustainable GenAI. Successful candidates for these...PerformanceFull timeWork experience placementLocal areaRemote workMonday to FridayFlexible hours$120k - $275k
...including hardware and software to train and run the largest ML workloads for AGI. MatX is seeking silicon micro-architects and design engineers to join our team as we create best-in-class silicon for high-performance and sustainable GenAI. Successful candidates for these...PerformanceFull timeWork experience placementLocal areaRemote workMonday to FridayFlexible hours$397.46k
...and a dynamic virtual economy. Our Economy ML team sits at the heart of this mission,... ...Avatar.We’re looking for a Distinguished Engineer/Technical Director to lead the strategy and... ...efforts to optimize ML system performance: low-latency serving, efficient GPU training...PerformanceFull timeWork experience placementH1bWork at officeLocal areaVisa sponsorshipMonday to Friday$300k
...enterprise partners, bridging research, ML infrastructure, and partner-facing product... ...ambiguous research questions, hands-on engineering, system architecture, and direct partner... ...to improve dataset structure and model performance. Develop LLM applications, including...PerformanceRemote jobFull time$95.38k - $160.85k
...Group has been named for ten consecutive years as a Top 50 performing P&C organization offering the stability of a large, profitable... ...about why you want to be here!PURPOSE OF JOBThe Quality Engineer for AI and ML Platforms establishes and evolves quality engineering practices...PerformanceFull timeWork experience placementWork at officeLocal areaRemote workFlexible hours- ...platforms Good judgment on choosing libraries, managing state, avoiding tech debt Experience debugging edge-case latency and performance in both frontend and backend Prior work with Cloudflare workers, Vercel, Next.js, or serverless-first architectures...PerformanceRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Founding Engineer - ML Performance. Be the first to apply!
- performance food group Remote
- system performance engineer Remote
- senior performance engineer Remote
- human performance consultant Remote
- performance improvement consultant Remote
- high performance computing engineer Remote
- senior performance tester Remote
- performance engineer Remote
- performance test architect Remote
- performance testing Remote




