Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Edge Inference Developer Tooling Founder

$250k

Forum Ventures

Edge Ai Founders OpportunityEdge AI is a production requirement across automotive, robotics, and industrial verticals, but the infrastructure underneath it doesn't exist. Every team deploying models on edge devices rebuilds memory management, platform abstraction, and monitoring logic from scratch. The hardware is fragmenting across dozens of chipsets and runtimes. No single vendor's tools cover the stack. The developer experience is years behind what cloud teams take for granted.The opportunity: Build the software and tooling layer that makes edge hardware usable — frameworks that abstract across platforms, memory managers that optimize dynamically, observability stacks that surface what's happening inside a deployed model running on a sensor or a vehicle. Start with the highest-cost-of-failure verticals (autonomous vehicles, industrial robotics), then expand into the full edge AI developer stack. A few examples:Edge Observability and Developer ToolsTeams deploying models at the edge have no visibility into what those models are doing in the field. Inference latency, memory pressure, thermal headroom: none of it is legible. When a model degrades, debugging is manual and slow.Memory Management and Testing for Edge DeploymentsEdge devices operate in conditions data center memory was never designed for: temperature extremes, vibration, extended duty cycles. Memory allocation across a fleet is static and manual. Performance degradation goes undetected until failure.Hardware-Agnostic FrameworksA model that runs on one edge chipset requires substantial rework to run on another. There is no universal SDK, no abstraction layer that makes inference logic portable. Vendor lock-in is the default, not a choice.Memory Architectures and Processing-in-MemoryFor vision processing, sensor fusion, and real-time control, the cost of moving data between memory and compute is measured in power, latency, and heat. Data center DRAM was not designed for these tradeoffs. Application-specific memory systems for edge workloads don't yet exist at scale.We're looking for Founders to leverage their domain expertise, and care deeply about the developer experience.You've shipped inference at the edge and hit the wall on memory, latency, or platform lock-in. You've debugged a model that ran fine in the cloud and failed on the device. You know the gap isn't the model. It's the infrastructure underneath it. Are convinced AI can deliver the outcomes better than humans to build full-stack, AI Native services in industries where outcomes depend on expertise, regulation, and trust.While we're actively building out these ideas, Forum Ventures is always open to hearing from founders with bold, original ideas. If you're working on an early-stage B2B SaaS company, even at the idea stage, pitch us still.Your Partners in Co-Creation:Forum's AI Studio brings together ambitious people, brilliant ideas, and capital to build the best B2B SaaS businesses in the world, from 0 to 1. In addition to capital and an idea, we provide founders with access to investor networks, fundraising support, and the resources needed to build transformational companies. We've launched 17 companies since 2023, and we're launching 7 more in 2026. We're designed to help Founders move faster, develop better insights, and build companies that have a higher success rate than startups built in any other way.Forum derisks and accelerates the Founder path by:A $250K USD investment, and an end-to-end fundraising playbook and network to raise your seed and Series AA validation partner with proven playbooks, networks, and the systems to validate, unlock your ICP, ignite your sales pipeline, and beyondThe support of our experienced operating team - from business design, product development, growth, recruitment, legal and moreExecuting a winning GTM strategy to get to $50-100K ARR, and the right MVP to ensure early tractionA community full of mentors, peers, and leaders$100K worth of business perksAbout you:Early-stage operator experience, and expertise in the problem space you're building in.The ability to attract, hire and lead world-class teamsA demonstrated passion for technology's ability to change the worldA desire to be a venture-backed founderCanada or US-based, and the ability to collaborate in North American time zones.Please note, this is not an employment opportunity with Forum Ventures. If you are currently fundraising your pre-seed or Seed round, please pitch our investment team instead here.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Edge Inference Developer Tooling Founder in San Francisco, CA vacancy
  •  ...BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies...  ...research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing... 
    Suggested
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    13 hours ago
  •  ...BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies...  ...research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing... 
    Suggested
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    13 hours ago
  •  ...BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies...  ...research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing... 
    Suggested
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    13 hours ago
  • $300k

     ...systems. About the Role The Cloud Inference team scales and optimizes Claude to serve the massive audiences of developers and enterprise companies across AWS, GCP, Azure...  ...regressions  Design interfaces and tooling abstractions across CSPs that enable cost-effective... 
    Suggested
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    1 day ago
  •  ...heterogeneous compute targets, including GPUs, FPGAs, ASICs, and edge SoCs. We support automotive, aerospace, defense, and...  ...software engineer to help us build a new generation of transpilation tools enabled by AI and modern verification techniques — bridging the... 
    Suggested
    Full time
    Remote work
    Relocation package
    Flexible hours

    Code Metal

    San Francisco, CA
    13 hours ago
  • $170k - $205k

     ...Senior Software Engineer to build a platform for network tooling that supports edge, backbone, and data center operations and deployment tasks...  ...You'll Be Working On:Network State Management: Design and develop the systems that manage network state for Crusoe Cloud, including... 
    Temporary work

    Crusoe

    San Francisco, CA
    1 day ago
  •  ...About the Team OpenAI’s Inference team powers the deployment of our most advanced models...  ...engineers focused on delivering a world-class developer experience while pushing the boundaries...  .... Have familiarity with inference tooling like vLLM, TensorRT-LLM, or custom model... 
    Full time

    OpenAI

    San Francisco, CA
    13 hours ago
  •  ...fine-tuned extraction models where legacy OCR and other parsing tools consistently fail. We are a small, fast-growing team of...  ...About the Role Specialize in low-latency, high-throughput inference for OCR and multimodal models. Own profiling, batching, and autoscaling... 
    Full time
    Work at office
    Visa sponsorship
    Relocation package

    Pulse

    San Francisco, CA
    13 hours ago
  •  ...serve OpenAI’s frontier models at massive scale. As part of the inference team, you’ll be responsible for unlocking every last FLOP from...  ...Contribute to and extend internal GPU libraries and runtime tools. Work closely with hardware-specific profiling tools (e.g.,... 
    Full time

    OpenAI

    San Francisco, CA
    13 hours ago
  •  ...operations management. Our co-founders Xerxes and Philip are...  ...power real-time perception and inference across our edge-cloud platform. This role...  ..., ONNX, CoreML). Developing continuous training and evaluation...  ...of edge nodes. Building tools and infrastructure for... 
    Full time

    S.e. Specter

    San Francisco, CA
    13 hours ago
  • $140k - $225k

     ...ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark...  ...As the Senior Software Engineer, Tooling and Development Infrastructure, you will...  ...to influence the development of cutting-edge automation frameworks, foster a culture of... 
    Full time
    Temporary work
    Local area
    Flexible hours

    HP IQ

    San Francisco, CA
    13 hours ago
  •  ...VPNs, model routers, and oodles of other tools, developers solve every networking problem with one...  ...universal gateway for API delivery, AI inference, device fleets, and site-to-site...  ...team builds the software that sits at the edge of every ngrok connection. The Agent is... 
    Remote job
    Permanent employment
    Full time
    Work at office
    Local area
    Immediate start
    Home office
    Flexible hours

    Ngrokinc

    San Francisco, CA
    2 days ago
  •  ...About the Team OpenAI’s Inference team ensures that our most advanced models run efficiently, reliably, and at scale. We build and optimize...  ...the systems that power our production APIs, internal research tools, and experimental model deployments. As model architectures and... 
    Full time

    OpenAI

    San Francisco, CA
    13 hours ago
  • $170k - $216k

     ...technical challenges to build services and tools for a broad range of customers Software...  ...You will: Build and evolve ML inference infrastructure for simulations. Be responsible...  ...experience ~ Experience in developing and maintaining distributed systems.... 
    Full time
    Remote work

    Waymo

    San Francisco, CA
    13 hours ago
  •  ...About the Team We’re hiring a Developer Productivity engineer to support OpenAI’s Inference Runtime teams. These teams own the systems responsible for serving models...  ...inference systems reliability. You’ll work on the tooling and operational foundations that support model... 
    Full time

    OpenAI

    San Francisco, CA
    13 hours ago
  •  ...industrial enzymes, and create cutting edge molecules that weren’t...  ...eclipsing physics-based tools in computational drug discovery...  ...collaborate closely with the founders to design, build, and scale our...  ...problems ranging from scaling ML inference on AWS for hundreds of GPUs to... 
    Full time
    Relocation

    Tamarind Bio

    San Francisco, CA
    13 hours ago
  •  ...We empower consumers, enterprises, and developers alike to access state-of-the-art AI models...  .... We focus on high-performance model inference and accelerating research through efficient...  ...scalable inference pipelines. Build tooling and observability to detect bottlenecks,... 
    Full time

    OpenAI

    San Francisco, CA
    13 hours ago
  •  ...our daily lives. Our team brings together founders from Oculus and Ubiquity6, alongside...  ..., with the ability to quickly learn new tools and services, and with a strong intuition...  ...tooling, and data management. Improve our developer productivity and our application... 
    Full time
    Contract work
    Flexible hours

    Sesame

    San Francisco, CA
    13 hours ago
  •  ...Our team brings together founders from Oculus and...  ...complex machine learning inference, scalable agentic workflows...  ...systems design to cutting-edge applied AI. At the...  ...opportunities to improve developer efficiency within your...  ...You might prototype a tool or workflow improvement... 
    Full time
    Contract work

    SESAME

    San Francisco, CA
    1 day ago
  •  ...with an early-stage, cutting-edge AI startup in San Francisco to...  ...You will work closely with the founders to turn ideas into production-...  ...modelsBuild AI agents that can use tools, complete workflows, and...  ...scalability, security, and AI inference costsWork directly with founders... 
    Live in
    Work at office
    Local area
    Remote work
    Relocation

    Joseph Michaels International

    San Francisco, CA
    9 hours ago
  •  ...training and deploying frontier models for developers and enterprises who are building AI...  ...they influence latency and throughput of inference. ~ Strong understanding or working experience...  ...Work closely with a team on the cutting edge of AI research  Weekly lunch stipend,... 
    Full time
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Cohere

    San Francisco, CA
    13 hours ago
  • A cutting-edge technology startup in San Francisco is seeking a first engineer to collaborate with founders in building innovative AI products. This role requires strong full-stack development skills, particularly with React, TypeScript, and Node.js. Candidates should have... 

    TraceRoot.AI (YC S25)

    San Francisco, CA
    3 days ago
  •  ...About the Team Our team analyzes inference stack performance across the application, model, and fleet layers to identify bottlenecks...  ...build cost-to-serve estimates from microbenchmarks and create tools that help cross-functional teams reason about latency, capacity... 
    Full time

    OpenAI

    San Francisco, CA
    13 hours ago
  • $320k

     ...beneficial AI systems. About the Role Our mandate is to make inference deployment boring and unattended. Anthropic serves Claude to...  ...sizes Extend deployment observability — dashboards and tooling that answer "what code is running in production," "where is my... 
    Full time
    Work at office
    Visa sponsorship
    Flexible hours
    Shift work

    Anthropic

    San Francisco, CA
    13 hours ago
  •  ...OpenAI safely brings cutting-edge technology to the world. We have...  ...also manages large-scale inference and platform infrastructure that...  ...offerings. Build tools and surfaces that allow teams...  ...Bring significant experience in developing (and redeveloping) production... 
    Full time

    OpenAI

    San Francisco, CA
    13 hours ago
  • $200k - $250k

     ...career, join us.Attributes We ValueWe hire successful builders with founder-like energy who want real impact, accelerated learning, and...  ...the teamYou’d be joining the team building Airwallex’s AI tooling platform. A product that helps employees turn ideas into working... 
    Temporary work
    Local area
    Worldwide

    Airwallex

    San Francisco, CA
    3 days ago
  • $165k - $195k

     ...Roboflow, we’re building the tools, community, and...  ...models. Today, over 1M developers, including those from...  ...is made up of former founders and thrive in the level...  ...our enterprise Jetson inference pipeline, so a solid foundation...  ...products like our edge inference server). A... 
    Full time
    Second job
    Remote work
    Work from home
    Relocation package
    Flexible hours

    Roboflow

    San Francisco, CA
    3 days ago
  • $175k - $220k

     ...Document Processing (IDP), developing the necessary tooling and infrastructure to...  ...interfaces that use cutting‑edge computer use models in real...  ...at a FAANG company Serial Founder/Co‑founder If you are a passionate...  ...Staffing and Recruiting Inferred from the description for... 
    Full time

    TogetherWeTech

    San Francisco, CA
    2 days ago
  •  ...Persona Edge Networking EngineerPersona is the configurable identity platform built for businesses in a digital-first world. Verifying...  ..., and you'll shape what it looks like — from SDLC, usage of AI tooling, static analysis, chaos engineering, and usage of simulation. We... 
    Full time
    For contractors
    Internship

    Persona

    San Francisco, CA
    4 days ago
  •  ...products. We empower consumers, enterprise and developers alike to use and access our start-of-the...  ...focus on performant and efficient model inference, as well as accelerating research...  ...production. Introduce new techniques, tools, and architecture that improve the... 
    Full time

    OpenAI

    San Francisco, CA
    13 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Edge Inference Developer Tooling Founder. Be the first to apply!