Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Member of Technical Staff, ML Systems — Confidential AI Infrastructure Startup

Jobleads-US

GPU & ML Systems · Below the Application Layer

Member of Technical Staff, ML Systems
Make the model 10x faster.

About the Company

A deep-tech AI infrastructure startup rebuilding the training and inference stack for world models. Today's ML infrastructure was built for language models — this team is rebuilding it for video, image, and world-model workloads, co-designing across three layers at once: low-level GPU kernel optimization, distributed systems, and the algorithms and models themselves.

The company came out of stealth with public benchmarks already in hand: a leading open video-generation model running roughly 10x faster at half the cost on its stack, a 2K image-generation model running in about four seconds for three cents, and a real-time video model running faster than real time. The founding team — with prior experience across leading AI labs, hyperscalers, and infrastructure companies — works out of Menlo Park in person and has an API already in production. The company raised a $10M seed round and is approaching a Series A.

"You report to the CEO. He runs every screen himself and makes the hiring decision — there is no layer between the work and the person who decides. The work sits below the application layer: kernels, runtimes, and distributed engines for video and world models. Nothing here is agents or RAG."

The Opportunity

You’ll own speed and efficiency across the full ML systems stack — low-level kernels, distributed inference engines, and multi-node training and serving systems for image, video, and world-model workloads. You’ll work directly alongside a founding team that between them covers distributed systems, kernel optimization, cloud infrastructure, and research.

You’ll feel at home here if you’d rather make a video model ten times faster than train one.

What You’ll Do

  • Optimize GPU and system performance for training and inference across image, video, and world-model workloads
  • Profile and remove bottlenecks at the kernel, memory, system, and cluster level using Nsight and related tooling
  • Write low-level optimizations in CUDA and Triton on code paths that run in production
  • Build distributed inference and training engines for diffusion models across multiple GPUs and nodes
  • Own communication performance — NCCL, RDMA over InfiniBand or RoCE, and disaggregated serving
  • Build benchmarking and regression harnesses so performance gains don't slide back in production

Requirements

  • Worked on inference or training performance — GPU kernels, runtime, or distributed execution
  • Optimized diffusion, video, image, or other multimodal model workloads
  • Degree in Computer Science or a related quantitative field

Baseline

  • 1+ years of experience in deep learning inference or training systems, or distributed systems
  • Built inside or contributed to an inference engine or runtime — vLLM, SGLang, TensorRT-LLM, or equivalent

CUDA Triton PyTorch Nsight Systems / Compute NCCL RDMA (InfiniBand / RoCE)

Interview Process

1

30-minute conversation with the CEO on background, motivation, and a first read on GPU/distributed systems depth.

2

Domain Deep Dive

60-minute technical round with a member of the founding team on kernels, inference, or distributed execution.

3

System Design

60-minute systems design session with a member of the founding team.

4

Optional Onsite

If not already done in person, a chance to meet the full team on-site in Menlo Park.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Member of Technical Staff, ML Systems — Confidential AI Infrastructure Startup in Menlo Park, CA vacancy
  • $180k - $230k

     ...Job Description Job Description Member of Technical Staff, ML Systems Company: TensorScale Location: Menlo Park...  ...distributed systems, GPU kernel optimization, cloud infrastructure and research, with prior work at Fireworks AI, Meta, Google, Apple, Microsoft, Snowflake... 
    Suggested
    Full time
    H1b
    Work at office
    Visa sponsorship

    Transparent Search Group

    Menlo Park, CA
    12 days ago
  • $150k

     ...Member of Technical Staff Location: Palo Alto, CA Company Stage of Funding: Seed Stage AI Startup ($27M Raised) Office Type: Onsite (5 Days...  ...engineering team to build the infrastructure and product...  ...workflows, orchestration systems, and production AI runtime... 
    Suggested
    Work at office
    Visa sponsorship
    Relocation package

    Recruiting from Scratch

    Palo Alto, CA
    3 days ago
  •  .... Runta builds runtime infrastructure for long-running AI agents: efficient execution...  ...roadmap as an early technical leader, working directly with...  ...numbers Solid distributed systems fundamentals: consistency...  ...to have: background in AI/ML infrastructure, agent... 
    Suggested
    Local area

    Jobleads-US

    San Mateo, CA
    2 days ago
  • $180k

     ...Member of Technical Staff - Evaluation Infrastructure SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This... 
    Suggested
    Temporary work

    Jobleads-US

    Palo Alto, CA
    3 days ago
  • $230k - $350k

     ...Job Description Job Description Member of Technical Staff, Inference Systems Company: Photon Location:...  ...Photon is building a next-generation AI inference platform from the ground...  ...by Stanford alumni with deep AI infrastructure experience, including early work at... 
    Suggested
    Full time
    H1b
    Visa sponsorship

    Transparent Search Group

    Palo Alto, CA
    12 days ago
  • About the Role As a Member of Technical Staff [Research] at NeoCognition...  ...of LLM agents — systems that can reason, plan...  ...real world. We are an AI research lab focused...  ...familiarity with modern ML frameworks (e.g.,...  ...similar) and training infrastructure. Publications in top... 

    NeoCognition Inc.

    Palo Alto, CA
    1 day ago
  • $180k

     ...Description SpaceXAI's mission is to create AI systems that can accurately understand the...  ...-training, post- training/alignment, infrastructure/scaling, evaluation, tooling/demos, and...  ...building or optimizing large-scale distributed ML systems (training/inference... 
    Temporary work

    SpaceXAI

    Palo Alto, CA
    a month ago
  • $125k - $160k

     ...NETWORK ENGINEER, AI INFRASTRUCTURE (STARSHIELD) Starshield...  ..., designing secure systems to guarantee access to...  ...) Collaborate with ML training teams to translate...  ...’s degree in a technical or engineering discipline...  ...and maintained in a confidential file. As set forth... 
    Confidential
    Permanent employment
    Temporary work
    For contractors
    Internship
    Immediate start
    Weekend work

    Jobleads-US

    Palo Alto, CA
    3 days ago
  •  ...Poetiq Inc. is hiring a Member of the Technical Staff to advance RSI loops and high‑performance systems at scale. This role spans both core algorithms and the deployment infrastructure, with ownership of end-to-end reasoning chains and diagnostic pipelines. The position... 

    Jobleads-US

    Los Altos, CA
    2 days ago
  •  ...Machine Learning & AI for Mathematical Discovery and Reasoning...  ...formulate them into benchmarkable ML objectives, build reproducible...  ...researchers by providing technical guidance, rigorous code reviews...  ...combinatorics) and formal proof systems (Lean, Coq, Isabelle).... 

    Axiom Math

    Palo Alto, CA
    4 days ago
  •  ...most important frontier in AI research. We’re hiring a...  ...optimization, come apply to be a Member of the Technical Staff at Poetiq. This is a...  ...and the high-performance systems that let it run...  ...exact‑match RL tasks. Run infrastructure at scale — build and maintain... 
    Work at office
    Flexible hours

    Jobleads-US

    Los Altos, CA
    2 days ago
  •  ...At Coram AI, we’re reimagining video security for the...  ...Design scalable APIs and systems that interface with our ML pipelines and edge device fleet Collaborate on infrastructure and architecture decisions...  ...it works well Previous startup experience building systems... 
    Full time

    Coram AI

    Sunnyvale, CA
    1 day ago
  • About Architect Architect is an AI research and product lab for chip design. We build AI models and systems that can explore, design, optimize, and verify new hardware....  ...organizations. Publications (or submissions) in top ML venues (NeurIPS, ICLR, ICML) or EDA venues (DAC... 
    Internship

    Architect Labs

    Palo Alto, CA
    3 days ago
  •  ...redefine what proactive security can do in the AI era. Cyberattacks are becoming autonomous...  ...Coordinate at Swarm Scale: Build agentic systems with distributed execution, concurrency,...  ...customer trust. Publications in top ML venues (e.g., NeurIPS, ICML, ICLR, ACL).... 
    Full time
    Work at office
    Flexible hours

    Armadin

    Palo Alto, CA
    1 day ago
  • $137k - $169k

     ...directly with research and technical staff to define hiring...  ...and cultivate exceptional AI/ML research talent across industry...  ...candidate conversations and confidential searches A real instinct...  ...such as ML/AI, distributed systems, infrastructure, scientific computing, or... 
    Confidential
    Full time
    Immediate start
    Remote work

    Waymo

    Mountain View, CA
    more than 2 months ago
  • $180k

     ...Description Job Description SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity...  ...: Own backend engineering for scalable, low-latency voice infrastructure and model integrations. Collaborate directly with Grok... 
    Temporary work

    SpaceXAI

    Palo Alto, CA
    a month ago
  • $180k

     ...Description Job Description SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity...  ..., and low-latency at global scale. Architect robust infrastructure for real-time multi-modal interactions, including handling generation... 
    Temporary work
    Worldwide

    SpaceXAI

    Palo Alto, CA
    a month ago
  • $180k

     ...Description Job Description SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity...  ...knowledge with their teammates. ABOUT THE ROLE: The RL infrastructure team is looking for an engineer to help develop our RL training... 
    Temporary work

    SpaceXAI

    Palo Alto, CA
    a month ago
  •  ...Scientist to join our Central AI team located in...  ...the fundamental infrastructure, data pipeline, frameworks...  ...tasks such as designing system and model...  ...guidance to emerging ML engineers. Your role is...  ...information will be kept confidential according to EEO guidelines... 
    Confidential
    Work at office
    Local area

    Atlassian

    Mountain View, CA
    3 days ago
  •  ...Vice President, AI & Machine Learning Engineering...  ...high-quality, scalable systems in a short time frame....  ...strong background in AI and ML engineering, a deep...  ...and are capable of operating at a high level of technical excellence. Functions ~ Engineering Confidential
    Confidential

    Confidential

    Palo Alto, CA
    3 days ago
  •  ...At Coram AI, we’re reimagining video security for the modern world...  ...standards, and mentor team members. Skills and qualifications:...  ...Prior experience in fast-paced startup environments or developer-...  ...camera and sensor into a smart system that enhances safety, efficiency... 
    Full time

    Coram AI

    Sunnyvale, CA
    1 day ago
  • $180k - $210k

     ...redefining the future of legal work with AI-powered Augmented Intelligence, enabling Fortune...  ...to understand and address their AI/ML challenges, designing and deploying custom...  ...company technologies. Act as the primary technical contact, providing hands‑on support during... 

    Eudia

    Palo Alto, CA
    5 days ago
  •  ...entity.About the Central AI OrgOur organization is...  ...a robust Atlassian AI infrastructure for the future. Our...  ...and Conversational AI system that integrates seamlessly...  ....About the AI & ML Platform TeamOur team’...  ...information will be kept confidential according to EEO... 
    Confidential
    Work at office
    Local area

    Atlassian

    Mountain View, CA
    2 days ago
  • $180k

     ...Description Job Description SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity in...  ...-training checkpoints. BASIC QUALIFICATIONS: Expertise in ML and large model scaling, with familiarity across all kinds of scaling... 
    Temporary work

    SpaceXAI

    Palo Alto, CA
    17 days ago
  • $180k

     ...Description Job Description SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity...  ...Imagine team, you will build the critical safety systems and infrastructure that ensure Grok's multimodal generation capabilities (images... 
    Temporary work
    Worldwide

    SpaceXAI

    Palo Alto, CA
    7 days ago
  • $243.1k - $314.6k

     ...difference. Every member of Gilead’s...  ..., Applied AI Solutions, leads...  ...accountable for the technical and product...  ...services, infrastructure, and scaled production...  ...Development Systems. This role supplies...  ...protecting confidential data, intellectual...  ...applied AI, ML, or data engineering... 
    Confidential
    Full time
    For contractors
    Local area

    Gilead Sciences

    Foster, CA
    13 days ago
  • $200k - $350k

     ...human knowledge to advance the AI economy. Handshake AI works...  ...challenges, building the systems that turn expert human knowledge...  ...The Role We are hiring a Member of Technical Staff, Evals to help define how...  ...we're looking for PhD in ML/AI, computer science, data science... 
    Full time
    Work at office
    Flexible hours

    Handshake

    Mountain View, CA
    7 days ago
  • $200k - $300k

     ...on engineering role focused on building and improving autonomous AI agents that handle real executive work end to end, from...  ...science, machine learning, or analytics, with a focus on evaluation systems and quality measurement for production AI. ~ Demonstrated experience... 
    Permanent employment

    Jobleads-US

    Palo Alto, CA
    2 days ago
  • $180k

     ...Job Description Job Description SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for... 
    Temporary work

    SpaceXAI

    Palo Alto, CA
    a month ago
  •  ...deeply in Generative AI — pioneering...  ...Principal Machine Learning Systems Engineer (P60) to lead technical directions of GenAI...  ..., high-performance infrastructure.Working at...  ...ð Design and Build ML SystemsArchitect and...  ...information will be kept confidential according to EEO... 
    Confidential
    Work at office
    Local area

    Atlassian

    Mountain View, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Member of Technical Staff, ML Systems — Confidential AI Infrastructure Startup. Be the first to apply!