Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Founding Inference Engineer

General Compute

About us General Compute is the neocloud for alternative chips. Inference is fragmenting: purpose-built silicon from SambaNova, Cerebras, Positron, d-Matrix, and others already beats GPUs on decode, and we productionize that hardware — we buy the racks, find the data center space, and run it for our customers. Each piece of hardware runs the workload it's actually built for: prefill stays on GPUs, decode moves to the chip built for it, and today that means generating tokens 5–7× faster than existing GPU-based competitors. Our customers are frontier labs, fast-growing AI application companies, and asset-light clouds. We closed a $15M seed round in May 2026, and have since closed a $400M debt facility — $100M funded upfront by Upper90, with the balance available for drawdown — collateralized by our inference chips. About the Role Getting a model correct and fast on our silicon is only half the problem — the other half is serving it. You'll build and own the inference layer that sits between a bought-up model and a live customer request: request scheduling, batching, KV-cache management, autoscaling across our ASIC fleet, and the failure modes that only show up at real traffic and real scale. This is a founding role on a small team, which means the scope is wide and the ownership is real: there's no separate SRE org to hand reliability to and no platform team to hand infra to. You'll design the serving architecture, then be the person paged when it breaks. The bet is that a serving stack built specifically for our hardware — not adapted from a GPU-first framework — is a durable edge, and you're the person who proves that out in production. What You’ll Do: Own the inference serving stack end-to-end. Design and build the system that takes a bring-up-verified model and serves it in production: request routing, batching, scheduling, and autoscaling a single model's serving replicas. Push cost-per-token down. Continuously tune batching strategy, KV-cache handling, and hardware utilization to widen the throughput advantage over GPU-based serving. Build for reliability from day one. Put in place the monitoring, alerting, and failover that make a fast-moving inference stack trustworthy under real customer load — and be the one who responds when it isn't. Work at the boundary with the compiler and bring-up team. Define the interface between "a model is correct and compiled" and "a model is live and fast," and push issues back to the right side of that line. Shape the roadmap, not just the backlog. As a founding engineer, you'll help decide what we build next in serving — multi-tenant isolation, speculative decoding, new scheduling strategies — not just execute a spec someone else wrote. Set the technical bar for the team you're helping build. Early architecture and code-quality decisions you make here will shape how the serving team operates as it grows. What We Need From You: 5+ years building and operating production systems at the infrastructure layer, ideally including a high-throughput or low-latency serving system. Direct experience with LLM inference serving — request batching, KV-cache management, continuous batching, or similar — in a production environment, not just research code. Comfortable owning reliability: you've been on call for a system that mattered, and you design for failure rather than reacting to it after the fact. Strong systems fundamentals — concurrency, networking, scheduling — deep enough to reason about performance at the hardware level, not just the application level. Self-directed and comfortable with ambiguity. This is a founding role: there's no existing playbook to follow, and you'll help write it. Nice-to-Haves: Experience serving models on non-NVIDIA accelerators (TPU, Trainium/Inferentia, Tenstorrent, Groq, Cerebras, or similar). Familiarity with serving frameworks such as vLLM, TGI, TensorRT-LLM, or SGLang, and an opinion on where they fall short. Experience running infrastructure at a small company or in a founding/early-engineer capacity before. Exposure to capacity planning or fleet management for specialized hardware. #J-18808-Ljbffr General Compute

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Founding Inference Engineer in San Francisco, CA vacancy
  • $200k - $400k

    Founding Engineer Location - San Francisco, CA (On-site) - *On-site role in San Francisco, CA.* Compensation - $200,000 -- $400,000 Base...  ...AI systems preferred Experience with model post-training or inference systems preferred Experience building low-latency production... 
    Suggested
    Internship
    H1b
    Visa sponsorship

    Recruiting From Scratch

    San Francisco, CA
    2 days ago
  • Hivemind is building a social AI that talks to many people at once, with one-on-one conversations. We are seeking founding members of technical staff to co-own the foundation from day zero and shape the team as it grows. You will define technical standards, build core... 
    Suggested

    Hivemind

    San Francisco, CA
    3 days ago
  • $150k - $220k

     ...Compensation & Benefits $150,000–$220,000 base salary + meaningful founding-team equity. The Stack The stack is TypeScript hosted on...  ...to eight people, founded by Girum Tihtina (ex-Meta payments engineering), Maurice Landers III (CTO, Harvey Mudd CS, design-engineer),... 
    Suggested
    Full time

    Doctours

    San Francisco, CA
    9 days ago
  •  ...from across IT and security. Several design partners today and a founding team of three. Full-Time, In-person, San Francisco, CA....  ...About the Role: This is a backend- and systems-heavy founding engineering role where "full-stack" means owning the product and system end... 
    Suggested
    Full time
    H1b
    Visa sponsorship

    Thomas Talent Network

    San Francisco, CA
    a month ago
  •  ...Job Description Job Description SF Bay Area, California | Fully On-site We are seeking an Founding Computer Vision Engineer – Multimodal AI & VLMs to join an early-stage technology company building AI systems for real-world industrial environments. This role focuses... 
    Suggested
    Work at office
    Visa sponsorship

    MaxIT Consulting - Max Corporate Group

    San Francisco, CA
    8 days ago
  • $160k - $252.5k

     ...circulation and driving the energy transition. Founded in 2017, we're delivering low-cost and...  ...the past. Power Electronics Controls Engineer As a Power Electronics Controls...  ...or employment information, and inferences drawn from your PI. We collect your PI for... 
    Full time
    Shift work

    Redwood Materials

    San Francisco, CA
    2 days ago
  • $187.5k - $247.5k

     ...circulation and driving the energy transition. Founded in 2017, we're delivering low-cost and...  ...have. Staff Mechanical Design Engineer, EPC Redwood Materials is hiring...  ...or employment information, and inferences drawn from your PI. We collect your PI for... 
    Full time
    Work experience placement

    Redwood Materials

    San Francisco, CA
    1 day ago
  • Trajectory, based in San Francisco, is seeking a Founding Generalist Engineer to craft the product experience for continual learning. The role involves owning the SDK and API surfaces while working directly with founders and industry-leading partners. Ideal candidates... 

    Trajectory

    San Francisco, CA
    4 days ago
  •  ...About Loop Loop AI is a San Francisco–based tech company founded in 2022. We provide a Delivery Intelligence Platform for data-driven...  ...first restaurants. About The Role We are looking for a Founding Engineer with 4+ years of backend engineering experience to join our fast... 
    Temporary work

    Loop AI - Delivery Intelligence Platform

    San Francisco, CA
    18 hours ago
  • $150k - $250k

     ...cockpit to run the business. We just raised an oversubscribed 2,5mio angel round in 7 days and are moving at rapid speed. Founding Engineer You'd be the first hire. This is a once-in-a-generation shot to define how service companies get built. You'll own the... 
    Visa sponsorship

    Orpex

    San Francisco, CA
    18 hours ago
  • $250k

     ...Our client is looking for an experienced Founding Engineer to build a production-grade Voice AI legal intake platform. This is the first engineering hire . The immediate challenge isn't simply building another AI prototype. It's taking an existing prototype and... 
    Work at office
    Immediate start
    Relocation

    JeffreyM Consulting

    San Francisco, CA
    4 days ago
  •  ...180 LPA. Min Experience: 1 years. Location: San Francisco Bay Area. JobType: full-time. We're looking for a highly motivated Founding Engineer to join us at the ground floor and help build our product from zero to one. This is a hands-on role for someone who thrives in... 
    Full time
    Weekday work

    Weekday AI (YC W21)

    San Francisco, CA
    2 days ago
  •  ...About The Role We’re hiring a Founding Engineer to help architect and scale the core platform at Proaction. This is a hands‑on, high‑ownership role for someone who operates at a staff level but is excited to build in a fast‑moving, early‑stage environment. About The Role... 

    Pro Action

    San Francisco, CA
    2 days ago
  •  ...Francisco. The Opportunity Silmaril brings research and production engineering into the same loop. You will choose important technical...  ...carry successful ideas into production. Your work could reduce inference latency, develop faster and more capable security models, create... 
    Shift work

    Silmaril

    San Francisco, CA
    18 hours ago
  •  ...We’re looking for a founding engineer to build the system that runs industrial commerce. Avent is turning the global supply chain into a machine-to-machine system where products are quoted, sourced, and fulfilled at software speed. This is not about building features... 
    Relocation

    Avent (YC S25)

    San Francisco, CA
    3 days ago
  • $147.5k - $235k

     ...Redwood Materials Redwood Materials was founded in 2017 by JB Straubel, co-founder and...  ...using proprietary energy architecture engineered entirely in-house. The company is profitable...  ...or employment information, and inferences drawn from your PI. We collect your PI for... 
    Full time

    Redwood Materials

    San Francisco, CA
    1 day ago
  • $150k - $250k

     ...Agentmail Engineer Opportunity Location: San Francisco, CA Type: Full-time, on-site We are seeing a surge in market demand as...  ...for your email, it's email for your AI. We're looking for a founding engineer with strong backend and infra instincts to help build... 
    Full time

    Agentmail (yc s25)

    San Francisco, CA
    18 hours ago
  • $160k - $180k

     ...About the job Founding Engineer Location: San Francisco, CA - In Person Compensation: $160,000-$180,000 base + meaningful early-stage equity Employment: Full-Time Experience: Early-career to ~4 years Visa Sponsorship: Not available at this time... 
    Full time

    Rheaction

    San Francisco, CA
    18 hours ago
  • $175k - $250k

     ...compensation Excellent benefits (healthcare, vision, dental) Job Details We're looking for an engineer to help build and maintain a high-performance inference library designed to support modern AI models across a variety of compute architectures. The role is... 
    Local area

    Jobot

    San Francisco, CA
    3 days ago
  • $160k - $210k

     ...About the job Founding Engineer Founding Engineer Global Placement Firm is conducting a confidential search for Founding Engineers to join a small, high-growth healthcare technology company in San Francisco focused on transforming how speech, occupational,... 
    Full time
    Private practice
    H1b
    Work at office
    Relocation
    Visa sponsorship

    Global Placement Firm

    San Francisco, CA
    4 days ago
  •  ...users constantly, and are just getting started. We're looking for engineers who want to own product from day one and help shape the future...  ...parts of the codebase, and ship product fast. This is true founding-level ownership. You'll shape both the product and the company... 
    Relocation

    Socket

    San Francisco, CA
    3 days ago
  •  ...Founding Engineer (Full Stack) Location: San Francisco Bay Area Type: Full-Time Compensation: Competitive salary + meaningful equity (founding tier) Backed by 8VC, we're building a world-class team to tackle one of the industry's most critical infrastructure problems... 
    Full time

    Fabrion

    San Francisco, CA
    3 days ago
  •  ...connections will form the foundation of the agentic internet – and Caylex will be the infrastructure powering it. The Role As a Founding Engineer, you’ll play a central role in shaping both our product and engineering culture. You will: Work full-stack across... 
    Full time
    Work at office

    Caylex

    San Francisco, CA
    2 days ago
  • $150k - $250k

     ...AgentMail is building the identity and communication infrastructure for AI agents, starting with email. We are looking for a Founding Engineer with backend and infra focus to architect systems for agents as the end users. Responsibilities: Design APIs that are... 
    Work experience placement
    Internship

    AgentMail

    San Francisco, CA
    1 day ago
  • $180k - $280k

     ...About the Role This is a founding engineer role at an early-stage AI startup building intelligent agent platforms for top-tier insurance carriers. You will own large swaths of the product end-to-end, working at the intersection of LLMs, unstructured data, and complex... 

    Wintermeyer Ventures

    San Francisco, CA
    2 days ago
  • $180k - $230k

     ...(CTO) and Chirag (CEO) are the team today. You'd be the second engineer: real-time voice, generative UI that redraws itself based on who...  ...without being asked. Comp Base: $180-230K Equity: Founding-caliber Location: In-person, San Francisco Process 4... 
    Live in
    Work at office

    HOBBES Inc

    San Francisco, CA
    18 hours ago
  • $180k

     ...Culture Talent. We're excited to help connect talented professionals with this exceptional team! The Role We're hiring a Founding Engineer to join an early-stage company building AI-powered software and data infrastructure for healthcare organizations operating in... 
    Work experience placement

    People Culture Talent

    San Francisco, CA
    1 day ago
  •  ...Join The Founding Engineering Team At Adam We're building the founding engineering team at Adam. At Adam, we're tackling a frontier problem: building a new way to interface with CAD via AI. This demands creativity, deep technical ability, and novel thinking.... 

    adam.ai

    San Francisco, CA
    3 days ago
  • $350k

     ...conversations consistent and personalized across all channels. Reverse engineering legacy healthcare systems that don't have public APIs to...  ...who desperately need them. We're looking for exceptional founding engineers to help us realize this ambition. If you're an... 
    Contract work
    Night shift

    Alleviate Health Inc

    San Francisco, CA
    4 days ago
  •  ...Senior Systems Engineer San Francisco, California Onsite or Remote At Evidently, we...  ...Engineer who owns reliability the way a founding engineer owns their product: with a sense...  ...posture, database performance, AI inference infrastructure: you can cover these areas... 
    Remote work
    Work from home

    Evidently

    San Francisco, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Founding Inference Engineer. Be the first to apply!