Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Member of Technical Staff: Agent Runtime

ego AI (YC W24)

Member of Technical Staff: Agent Runtime (Ego) Location: San Francisco, CA (hybrid) Reports to: Founding team About The Role Ego AI is a YC-backed applied AI research lab building the behavioral infrastructure for AI companions and agents. We work at the intersection of real-time conversational instincts, memory, and persistent identity; the layer that makes AI feel genuinely alive. We're a small, fast-moving team and we're defining a new category of human-AI interaction. What you'll do Own the agent harness. Design and maintain our core agentic loop (plan → tool call → observe → repeat → finalize) as a small, legible state machine. Errors are first-class: malformed tool calls, hallucinated tool names, and throwing tools get fed back to the model. They never crash the loop. Make execution durable. Build checkpointed, resumable sessions backed by a database: step logs with pending/committed status, idempotency keys on side-effecting tools, and a clear account of the at-least-once vs exactly-once boundary. A crash between the LLM call and the tool call, or mid-compaction, must never corrupt a session or double-fire a POST. Solve long-horizon context. Own our compaction strategy: pinned goals, running summaries, sliding windows of recent turns, and retrieval over older state. A 30+-step task should finish aimed at the original goal, within budget. Design for extensibility. Ship a tool registry (name + JSON schema + handler) that teammates and partners can extend without touching the loop. Validate schemas before handlers run. Keep model adapters swappable in one place. Ship deployable systems. Deliver one-command local bring-up, persistence behind an interface (Postgres SQLite Cloudflare storage), clean APIs, and secrets hygiene. Build per-user durable agent instances that wake on triggers (cron, webhooks, inbound events). Cloudflare Durable Objects experience is a real plus. Support the team. Unblock product engineers building on the harness, pair on extensions, and review agent-adjacent designs. Projects this hire will own Productionize the agent runtime ("Ronin core"). Take our harness from working prototype to the shared runtime every Ego product sits on: durable step log, budget enforcement, compaction, tool registry, and observability/tracing. Per-user durable personal agents. Each user gets a long-lived agent instance with its own memory, triggers, and budget. It wakes on webhooks or cron, survives restarts, and stays isolated per tenant. Likely on Cloudflare Durable Objects or an equivalent single-writer model. Evaluation and reliability harness. Build repeatable crash tests (kill -9 mid-task and mid-compaction), race tests on concurrent session writes, cheap long-horizon stubs that exercise compaction without burning tokens, and per-step tracing. What we're looking for Must have 3+ years building backend or infrastructure systems, with real production ownership of stateful services or relevant projects that they have worked on. Hands-on experience building LLM agent systems: tool-calling loops, defensive handling of model output, and step/token/cost budgets. Deep grasp of durable execution: checkpointing, idempotency, outbox/step-log patterns, and the ability to reason precisely about failure windows (what if it dies after the side effect but before it is recorded?). A deliberate context-management strategy for long-horizon tasks, and the ability to defend its trade-offs rather than only describe it. Strong API and data-model design; persistence behind clean seams; comfort being judged on docker compose up working cold from a README. Clear written communication. Architecture docs a reviewer understands before reading the code. Honest scoping (what you cut and why). Nice to have Cloudflare Workers / Durable Objects / D1 in production. Multi-model experience (DeepSeek, Claude, open-weight models) and eval tooling. Long-term memory systems for agents (personalization layers, retrieval over older state, hierarchical summaries). Concurrency safety on shared session state; streaming output; structured per-step observability. Use AI tools freely. We do. You will walk us through the design and defend every trade-off, so own each decision in the repo. #J-18808-Ljbffr ego AI (YC W24)

Vacancy posted 11 hours ago
Similar jobs that could be interesting for youBased on the Member of Technical Staff: Agent Runtime in San Francisco, CA vacancy
  • Member of Technical Staff - Agents at Prime Intellect - San Francisco Building the Future of Open Source + Decentralized AI At Prime Intellect, we are on a mission to accelerate open and decentralized AI progress by enabling anyone to contribute compute, code or capital... 
    Suggested
    Remote work
    Flexible hours

    Victrays

    San Francisco, CA
    1 day ago
  • $256k - $276k

     ...The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity As a Member of Technical Staff and AI Agent Development Lead, you will lead the design, development, and deployment of next-generation AI agents that interact... 
    Suggested
    Work at office
    Flexible hours
    3 days per week

    Jobzhr

    San Francisco, CA
    4 days ago
  • Member of Technical Staff - Agent Engineer About Phylo Phylo is an applied research lab building agentic intelligence to accelerate discovery for...  ...discipline. Familiarity with LLM APIs, tool calling, agent runtimes, or workflow orchestration. Strong quantitative... 
    Suggested
    Work at office

    Phylo, Inc.

    South San Francisco, CA
    1 day ago
  • $208k - $312k

    About Vercel: Vercel is the agentic infrastructure company. We free people and agents to ship what’s next. For more than a decade, Vercel has shaped how the web is built. As the team behind Next.js, v0, and AI SDK, we create products that help builders move from idea to... 
    Suggested
    Full time
    Work from home
    Worldwide
    Flexible hours

    vercel.com

    San Francisco, CA
    2 days ago
  •  ...About This Role We’re hiring an LLM Engineer to build intelligent agent systems that interact with robotic platforms. This role blends...  ...agent frameworks, tooling, and development workflows. High technical bar, with strong judgment around code quality, system design, and... 
    Suggested

    dimensional

    San Francisco, CA
    4 days ago
  • $350k

     ...penchant for building your own tools, this role will be a good fit. Some example areas you might work on: Build and innovate on the agent harness: agent loop architecture, tool integrations, prompt scaffolding, execution environments, and capability primitives Design... 

    Mirendil

    San Francisco, CA
    4 days ago
  • $100k - $300k

     ...Cogent is an Applied AI Lab building the next generation of AI agents for cybersecurity. AI has fundamentally changed how attacks happen...  ...who thrive in high-impact environments, love solving hard technical problems, and want to see the tangible results of their work in... 

    Cogent

    San Francisco, CA
    3 days ago
  •  ...week in our SF Mission district office. Your Role As a Member of Technical Staff, you will be responsible for building Eventual's core...  ...inference optimization (batching, GPU utilization, TensorRT/ONNX Runtime, streaming data loaders) to minimize latency and maximize... 
    Work at office
    Immediate start
    Flexible hours
    Night shift

    Eventual

    San Francisco, CA
    13 hours ago
  •  ...feature store on top that makes the data agent-readable. An action layer that runs...  ...precedents to copy from. About the Role Members of Technical Staff (MTS) are the senior engineers who...  ...fast as portco 5. Workflow and action runtime. The execution layer that runs operational... 

    BEACON SOFTWARE COMPANY

    San Francisco, CA
    3 days ago
  •  ...to gigawatt-class AI datacenters. Gimlet Labs is seeking a Member of Technical Staff focused on ML systems and inference. In this role, you will...  .... You will work at the intersection of model architecture, runtime behavior, and system performance to ensure inference is fast... 

    Gimlet Labs

    San Francisco, CA
    2 days ago
  • $200k - $300.09k

     ...of the world's largest open source AI agent project. Its mission is grounded in a...  ...OpenClaw Foundation is seeking exceptional Members of Technical Staff (MTS) to serve as full‑time...  ...ecosystem Areas of Ownership AI Agent Runtime Systems Distributed Systems & Infrastructure... 
    Full time

    The OpenClaw Foundation

    San Francisco, CA
    4 days ago
  • $300 per month

     ...About us Edison Scientific builds and deploys AI scientist agents to accelerate science and the development of new medicines. We are...  ...Engineering, Mathematics, Physics, Data Science, or a related technical field. Proficiency in Python and/or TypeScript, with comfort... 
    Full time
    Work at office

    Edison Scientific Inc.

    San Francisco, CA
    21 hours ago
  •  ...market. Inference is just one piece of an effective background agent. Let's design and build the rest of the system, that turns billions...  ...the CTO, who will ask about your experience, and share as much technical detail about Sail as you want to hear. Come in to Sail's SF... 
    Work at office
    Immediate start

    Sail Research

    San Francisco, CA
    2 days ago
  •  ...become an expert; nothing less will do in our immensely competitive market. Inference is just one piece of an effective background agent. Let's design and build the rest of the system, that turns billions of tokens into the best possible answers. What you’ll do Design... 

    SAIL

    San Francisco, CA
    3 days ago
  • $150k - $300k

     ...systems into our RL training stack. Core Technical Responsibilities LLM Serving Multi‑...  .... GPU & Networking: Architecture, CUDA runtime, NCCL, InfiniBand; GPU‑aware bin‑packing...  ...in open development and encourage team members to contribute to the broader AI community... 
    Work at office
    Remote work
    Visa sponsorship
    Relocation package
    Flexible hours
    Shift work

    Prime Intellect

    San Francisco, CA
    4 days ago
  • $150k - $350k

     ...gigawatt‑class AI datacenters. Mission Gimlet Labs is seeking a Member of Technical Staff focused on kernels and GPU performance. In this role, you...  ..., and instruction scheduling Work with compilers and runtimes to ensure kernels integrate cleanly and perform well in... 

    Gimlet Labs, Inc.

    San Francisco, CA
    2 days ago
  • Job Title Member of Technical Staff, Research Salary Not Disclosed + Equity Company Description Primitive...  ...of software development to create agents that model real user decision-making....  ...translating findings into model, memory, and runtime improvements. Build complex... 

    Jack & Jill

    San Francisco, CA
    2 days ago
  • $150k - $350k

    Mission Gimlet Labs is seeking a Member of Technical Staff focused on distributed systems. In this role, you will build the core platform that...  ...observability and debugging at scale Work closely with compilers, runtimes, and hardware to ensure end‑to‑end system correctness and... 

    Gimlet Labs, Inc.

    San Francisco, CA
    2 days ago
  •  ...hardware architectures. You will work across compiler infrastructure, runtime systems, scheduling, memory movement, kernel orchestration, and...  ...every part of how we build and run this company. As an early member of the team, you will have significant ownership over your work... 

    The Consensus

    San Francisco, CA
    2 days ago
  •  ...to gigawatt-class AI datacenters. Gimlet Labs is seeking a Member of Technical Staff focused on distributed systems. In this role, you will...  ...observability and debugging at scale Work closely with compilers, runtimes, and hardware to ensure end-to-end system correctness and... 

    Gimlet Labs

    San Francisco, CA
    2 days ago
  •  ...gigawatt-class AI datacenters. Gimlet Labs is seeking a Member of Technical Staff focused on compilers. In this role, you will work on the core...  ...through multiple IRs, and target a range of execution runtimes and accelerators. This is a role for engineers who enjoy... 

    Gimlet Labs

    San Francisco, CA
    2 days ago
  • Perplexity is seeking energetic engineers to join our highly driven Agents engineering team. The Agents team consists of backend, full-stack, and AI/ML engineers who collaborate to build harnesses and AI systems powering delightful agentic experiences. These experiences... 
    Flexible hours

    The Consensus

    San Francisco, CA
    3 days ago
  • $150k - $250k

     ...servicing with the industry’s most advanced AI credit-servicing agents. We are backed by Long Journey Ventures (Arielle Zuckerberg,...  ...including Ryan Hoover (Founder, Product Hunt), Charlie Songhurst (Board Member, Meta), and Michael Jones (Former Chair, Huntington Bank... 
    Full time
    Work experience placement
    Internship
    Worldwide

    Krew

    San Francisco, CA
    more than 2 months ago
  • $250k - $300k

     ...systems spanning storage, networking, virtualisation, and container runtimes You will be working at the lower levels of the Linux stack,...  ...Skills / Must Have: A track record of impressive technical work you can speak to in depth; the years matter less than the... 
    Full time
    Remote work
    San Francisco, CA
    10 days ago
  •  ...people actually work, translating that expertise directly back into agents. Mercor is creating a new category of work where expertise...  ...in agent engineering and evaluation, including how agent runtimes and harnesses produce a trajectory and where it fails. Experience... 
    Work at office
    Relocation package

    Mercor

    San Francisco, CA
    1 day ago
  • $120k - $170k

     ...with teams across the United States to help them hire. Member of Technical Staff (Founding Engineer) Location: San Francisco, CA (North Beach...  ...the commerce layer for AI. The platform enables AI agents such as ChatGPT, Claude, Gemini, and other frontier models... 
    Work at office
    Remote work
    Visa sponsorship
    Relocation package

    Recruiting from Scratch

    San Francisco, CA
    3 days ago
  •  ...Bal and Varun Krishnan, who met at Harvard University while studying Computer Science. Mustafa is a core contributor to ONNX Runtime and DeepSpeed with deep expertise in distributed systems and large-scale model training infrastructure Varun is an INFORMS Wagner... 

    Pear VC

    San Francisco, CA
    2 days ago
  • $240k - $300k

     ...Member of Technical Staff Salary: $240,000-$300,000 + Competitive Equity Location: San Francisco, CA (5 days onsite) One of the most...  ..., and transaction monitoring for banks and fintechs. Their agents work directly inside browsers and internal systems,... 

    Xpertalent

    San Francisco, CA
    21 hours ago
  •  ...Member of Technical Staff @ Lotus AI Who we are Lotus AI is a groundbreaking primary care app that integrates your medical records, AI,...  ...issues, and what's breaking. What you'll do AI Agents and Product Intelligence Build and iterate on AI agent... 

    Lotus Health AI, Inc

    San Francisco, CA
    4 days ago
  • $120k - $300k

     ...(Seed) · Industry: FinTech The Role A backend-leaning Member of Technical Staff building the end-to-end systems that power the client's production...  ...or at an early-stage startup Why Join Frontier AI agents in production: build browser-based agents for the "... 
    Full time
    H1b
    Visa sponsorship

    David Joseph & Company

    San Francisco, CA
    21 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Member of Technical Staff: Agent Runtime. Be the first to apply!