Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Founding Engineer - ML Performance

$250k - $395k
Full-time

uRun

Role Description

Performance is uRun's core differentiator. We're not chasing incremental gains — we're building infrastructure that runs 10–100x faster than the status quo. As our ML Performance Engineer, you will be the person who makes that true.

This is a founding technical hire. You will:

  • Write custom CUDA kernels, pushing GPU utilization to its limits.
  • Own inference latency end-to-end across the stack.
  • Work directly with the founding team on the hardest performance problems in production AI infrastructure.
  • Have your fingerprints on everything we ship.

What you'll actually be doing day-to-day:

  • Write custom CUDA kernels that unlock performance headroom unavailable through off-the-shelf frameworks.
  • Optimize model inference end-to-end, targeting sub-50ms latency across our inference platform.
  • Drive 10x performance improvements across the stack: memory bandwidth, kernel fusion, operator scheduling, and beyond.
  • Implement zero-copy distributed memory optimizations across multi-GPU and multi-node environments.
  • Own GPU utilization and memory management, squeezing every available FLOP out of the hardware we run.
  • Profile, benchmark, and instrument the full inference pipeline to find and eliminate bottlenecks systematically.
  • Set the performance engineering bar for the team: define what fast looks like and build the tooling to measure it.

Qualifications

  • Deep, hands-on CUDA expertise: you have written custom kernels in production, not just called into cuBLAS.
  • Strong background in model inference and post-training optimization at scale.
  • Fluency in GPU memory hierarchy, warp scheduling, kernel fusion, and hardware-aware algorithm design.
  • Experience profiling and benchmarking complex inference pipelines: you know where the time goes and how to get it back.
  • Able to operate at the frontier with minimal guidance — you identify the problem, design the approach, and ship the fix.

Requirements

  • Public work in GPU optimization or inference efficiency — open source contributions, a published paper, or a side project that shows your depth (vLLM, Flash-Attention, TensorRT-LLM, PyTorch, or equivalent).
  • Experience with hardware-aware optimization frameworks: CuTe, Triton, TileLang, or similar.
  • Familiarity with distributed memory and communication primitives: NCCL, InfiniBand, NVLink, RoCE.
  • Contributions to or deep familiarity with PyTorch Distributed, Ray core, or similar systems.
  • Experience optimizing for video generation or other high-throughput, latency-sensitive generative workloads.
  • Prior work at an inference-focused company or research lab pushing the boundary of what GPU hardware can do.

Benefits

  • Competitive salary and meaningful equity in an early-stage AI infrastructure company.
  • Health, dental, and vision — full coverage.
  • 401(k) — company-supported retirement savings.
  • FSA/HSA — flexible spending accounts for healthcare costs.
  • Paid time off — we trust you to manage your time.
  • Top-tier tooling — access to the best AI tools available: Claude, Codex, Kimi, and whatever else helps you move faster.
  • MacBook Pro and AirPods — the hardware you need, on us.
Vacancy posted 7 days ago
Similar jobs that could be interesting for youBased on the Founding Engineer - ML Performance in Remote vacancy
  •  ...experiences. We're building our engineering hub in Argentina from the...  ...standards, and the culture of a founding team, with real ownership...  ...infrastructure (backend/core), or high-performance product interfaces at scale (...  ...-cloud · probabilistic & ML systems. You don't need... 
    Performance
    Full time
    Remote work

    Amperity

    Washington DC
    a month ago
  • # Remote deep learning engineer jobsWe, at Turing, are looking for remote...  ...Turing jobs a year back, I found the Turing test to be very...  ...experience for me. Through the AI/ML-enabled tests, Turing reviews...  ...aid in the development of high-performance machine learning models.If you... 
    Performance
    Full time
    Remote work
    Worldwide

    Turing Inc

    New York, NY
    22 hours ago
  • $140.8k - $211.2k

     ...Qualcomm Technologies, Inc.Job Area:Engineering Group, Engineering Group >...  ...background in AI and general ML techniquesProven hands-on...  ...Generative AI workflows for accuracy, performance, and other key metricsML...  ...Qualcomm's toll-free number found here. Upon request, Qualcomm... 
    Performance
    Work experience placement
    Work from home

    Qualcomm

    San Diego, CA
    2 days ago
  •  ...apply for the Developer Relations Engineer role at Canonical 2 days ago...  ...members wherever they may be found - IRC, social media, product...  ...geographical location, experience, and performance in shaping compensation...  ...Software Engineer - Data, AI/ML & Analytics Software Engineer,... 
    Performance
    Full time
    Local area
    Remote work
    Worldwide

    Canonical

    San Bernardino, CA
    4 days ago
  •  ...Integration and Technology Development Engineer defines, develops, and...  ...our company benefits can be found here:We are committed to sourcing...  ..., attracting, and hiring high-performance innovators, while providing...  ...structures. Familiarity with AI/ML techniques, data mining, or... 
    Performance
    Local area

    Onsemi

    Gresham, OR
    2 days ago
  • On The Stage (OTS) is at an inflection point. As our Founding GTM Engineer , you won’t just maintain systems — you’ll be the architect of how...  ...Finance. Measurement & Continuous Improvement Own the performance of GTM automation and AI-powered workflows — identifying opportunities... 
    Performance
    Remote job
    Shift work

    OnTheStage

    Austin, TX
    2 days ago
  • $170k - $200k

     ...revolution.  The Instawork Robotics ML Engineer will help build and scale the technology...  ...and analytics to measure the quality and performance of our dataset and ML models....  ...job description. About Instawork Founded in 2015,  Instawork is the nation’s leading... 
    Performance
    Hourly pay
    Full time
    Internship
    Local area
    Flexible hours
    Shift work

    Instawork

    Remote
    22 hours ago
  •  ...integrated, and designed for scale. Our founding team has scaled category-defining...  ...seeking a highly experienced Principal ML Engineer (Applied / Systems) to join our engineering...  ...solutions, all with a strong focus on performance, maintainability, and impact. What... 
    Performance
    Full time

    Soris

    Remote
    22 hours ago
  • $160k - $240k

     ...autonomy accessible to all. Founded in 2016, Nuro is building the...  ...are looking for self-motivated engineers to build the next-generation...  ...mission is to provide a high-performance, highly reliable foundation of...  ...Robotics experience, ML inference optimization experience... 
    Performance
    Full time

    Nuro

    Remote
    22 hours ago
  • $139.76k - $287.75k

     ...process here.The Production Engineering organization at Pinterest is...  ...at scaleManage capacity and performance to help scale our infrastructure...  ..., or reliability workflowsAI/ML infrastructure experience (LLM...  ...for this position can be found here.US based applicants only... 
    Performance
    Work at office
    Local area
    Remote work
    Relocation
    Relocation package

    Pinterest

    Chicago, IL
    4 days ago
  • $193.93k - $291.15k

     ...all roads and all rides. Founded in 2016, Nuro is a physical AI...  ...we are looking for a Software Engineer to join our Sensor Data and...  ...about sensor data and autonomy performance Collaborate with...  ...experience Deep understanding of ML fundamentals with hands-on experience... 
    Performance
    Full time
    Immediate start
    Flexible hours

    Nuro

    Remote
    22 hours ago
  • $150k - $250k

    Job Title: Senior Analog Mixed-Signal Engineer - AI HardwareJob Location: Palo Alto, CA, Austin...  ...compute building blocks including high-performance ADCs/DACs, data converter interfaces, and...  ...of algorithms and circuits for ML workloads.Strong programming skills in Python... 
    Performance
    Remote work

    CyberCoders

    Palo Alto, CA
    18 hours ago
  • $184k - $287.5k

     ...the world.We are seeking an AI Compiler Engineer with deep expertise in compiler technologies...  ..., delivering measurable improvements in performance and efficiency, and advancing LLM-enabled...  ....8+ years of software engineering and AI/ML experience, preferably in tools or... 
    Performance
    Full time
    Remote work

    Nvidia

    Redmond, WA
    2 days ago
  • $98k - $176k

     ...team members need and deserve. Our high-performing teams balance independence with collaboration...  ...from the inside out.The Tech Inventory Engineering team:Develops user-first tooling that...  ...to enable organization-wide, AI/ML-powered analyticsOur Tech Stack:TypeScript... 
    Performance
    Full time
    Temporary work
    Work experience placement
    Flexible hours

    Target

    Brooklyn Park, MN
    2 days ago
  • $150k - $250k

     ...projects of this program since our founding in 2013. We've grown on this...  ...providing the customer with Engineers who have done exceptional...  ...Demonstrated experience with ML frameworks like scikit-learn,...  ...correct errors or to improve performance Experience optimizing algorithms... 
    Performance
    Contract work
    Work at office
    Immediate start
    Remote work

    GRVTY

    Chantilly, Loudoun County, VA
    2 days ago
  • $120k - $275k

     ...systems including hardware and software to train and run the largest ML workloads for AGI. MatX is seeking Emulation engineers to join our team as we create best-in-class silicon for high-performance and sustainable GenAI. Successful candidates for these roles will be... 
    Performance
    Full time
    Work experience placement
    Local area
    Remote work
    Monday to Friday
    Flexible hours

    MatX

    Mountain View, CA
    18 hours ago
  • $121.5k - $188.5k

     ...intersection of four disciplines: agent engineering, context engineering, evaluations and guardrails...  ..., BigQuery, Redshift) or integrating ML models into production services....  ...ResourcesThis position may be eligible for performance-based incentives/bonuses. Benefits include... 
    Performance
    Full time

    Nordstrom

    Seattle, WA
    4 days ago
  • $116.96k - $160.82k

     ...signal processing, data analytics, and AI/ML-enabled algorithms are core...  ...hands-on and experienced Senior Algorithms Engineer to lead the design, development, and optimization...  ...algorithms, from concept through prototyping and performance evaluation.Apply signal processing... 
    Performance
    Permanent employment
    Full time
    Contract work
    Work at office
    Remote work
    Day shift

    Analog Devices

    Wilmington, MA
    3 days ago
  • $166k - $258k

     ...intersection of four disciplines: agent engineering, context engineering, evaluations and guardrails...  ...solutions; drive accountability for performance, cost, accuracy, and security of feature...  ..., BigQuery, Redshift) and integrating ML models into production services.... 
    Performance
    Full time
    Temporary work

    Nordstrom

    Seattle, WA
    1 day ago
  •  ...across Intelligence, Analytics, Engineering, Mission Support, and Communications disciplines. Founded in 2008, our mission is to...  ...deploying advanced cloud-enabled AI/ML solutions that enhance...  ...testing and evaluation of UI/UX performance against operational datasets.... 
    Performance
    Remote job
    For contractors
    Casual work

    Barbaricum

    Crane, IN
    24 days ago
  • $152k - $241.5k

    NVIDIA is currently seeking a Senior Developer Technology Engineer for High-Performance Databases!Would you enjoy researching new algorithms and memory...  ...and are becoming the bottleneck for Machine Learning (ML) and Deep Learning (DL) applications, as performance of the... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $154k - $231k

    Company:Qualcomm Technologies, Inc.Job Area:Engineering Group, Engineering Group > DSP...  ...IoT, and Automotive roadmap are our high-performance, multithreaded, low-power Hexagon cores....  ...com or call Qualcomm's toll-free number found here. Upon request, Qualcomm will provide... 
    Performance
    Work experience placement
    Work from home

    Qualcomm

    Austin, TX
    2 days ago
  • $200k - $250k

     ...Software Engineer (AI Investor) – New York, Hybrid A seed‑stage Applied AI lab building a fully autonomous AI investor—a system designed...  ...‑end What They’re Looking For: 4+ years of experience in high‑performance engineering environments Proven track record of shipping... 
    Performance
    Remote work

    Acceler8 Talent

    New York, NY
    3 days ago
  • $216k - $270k

     ...world problems remains one of the hardest engineering challenges.As a Senior Frontier Agent...  ...enterprise software.Unlike traditional ML roles that focus on a single model or product...  ..., experience, qualifications, interview performance, and relevant education or training.... 
    Performance
    Full time

    Scale AI

    New York, NY
    3 days ago
  • $120k - $275k

     ...including hardware and software to train and run the largest ML workloads for AGI. MatX is seeking silicon micro-architects and design engineers to join our team as we create best-in-class silicon for high-performance and sustainable GenAI. Successful candidates for these... 
    Performance
    Full time
    Work experience placement
    Local area
    Remote work
    Monday to Friday
    Flexible hours

    MatX

    Mountain View, CA
    18 hours ago
  • $120k - $275k

     ...including hardware and software to train and run the largest ML workloads for AGI. MatX is seeking silicon micro-architects and design engineers to join our team as we create best-in-class silicon for high-performance and sustainable GenAI. Successful candidates for these... 
    Performance
    Full time
    Work experience placement
    Local area
    Remote work
    Monday to Friday
    Flexible hours

    MatX

    Mountain View, CA
    3 days ago
  • $397.46k

     ...and a dynamic virtual economy. Our Economy ML team sits at the heart of this mission,...  ...Avatar.We’re looking for a Distinguished Engineer/Technical Director to lead the strategy and...  ...efforts to optimize ML system performance: low-latency serving, efficient GPU training... 
    Performance
    Full time
    Work experience placement
    H1b
    Work at office
    Local area
    Visa sponsorship
    Monday to Friday

    Roblox

    San Mateo, CA
    2 days ago
  • $300k

     ...enterprise partners, bridging research, ML infrastructure, and partner-facing product...  ...ambiguous research questions, hands-on engineering, system architecture, and direct partner...  ...to improve dataset structure and model performance. Develop LLM applications, including... 
    Performance
    Remote job
    Full time

    SaidGig

    Remote
    1 day ago
  • $95.38k - $160.85k

     ...Group has been named for ten consecutive years as a Top 50 performing P&C organization offering the stability of a large, profitable...  ...about why you want to be here!PURPOSE OF JOBThe Quality Engineer for AI and ML Platforms establishes and evolves quality engineering practices... 
    Performance
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work
    Flexible hours

    ICW Group

    San Diego, CA
    18 hours ago
  •  ...platforms Good judgment on choosing libraries, managing state, avoiding tech debt Experience debugging edge-case latency and performance in both frontend and backend Prior work with Cloudflare workers, Vercel, Next.js, or serverless-first architectures... 
    Performance
    Remote work

    Dench.com

    United States
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Founding Engineer - ML Performance. Be the first to apply!