Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Member of Technical Staff - Infrastructure

Gimlet Labs

About Us Gimlet is building the next generation of AI infrastructure: large-scale AI datacenters and the orchestration platform that coordinates them. The future of AI will require vastly more compute than exists today. But as AI workloads become more complex and new hardware architectures emerge, simply deploying more GPUs isn't enough. The challenge is making increasingly diverse compute work together. Gimlet's platform intelligently partitions and routes workloads across heterogeneous hardware, enabling step-function improvements in performance and efficiency. Customers deploy through production-grade APIs without needing to think about hardware selection, placement, or optimization. We work with foundation labs, hyperscalers, and AI-native companies to power production workloads at massive scale and help define the infrastructure layer for the future of AI. About this Role We are looking for an Infrastructure platform Engineer to design, build, and operate the cluster infrastructure behind Gimlet’s heterogeneous inference cloud. Unlike traditional cloud platforms built around a single hardware ecosystem, Gimlet's infrastructure spans multiple accelerator vendors and architectures. Infrastructure engineers play a key role in bringing new hardware platforms online, building the operational abstractions that make heterogeneous infrastructure manageable at scale, and ensuring new silicon can serve production workloads reliably from day one. This role is highly hands‑on. You will work across bare metal, Linux, Kubernetes or cluster schedulers, high‑speed networking, observability, provisioning, and incident response. You will partner closely with distributed systems, runtime, compiler, and hardware teams to ensure Gimlet’s infrastructure can support demanding AI workloads at production scale. What you will work on Design, deploy, and operate large‑scale CPU, GPU, and accelerator clusters powering production AI inference. Build automation for provisioning, configuration, upgrades, validation, and lifecycle management. Design and scale provisioning systems for heterogeneous bare‑metal infrastructure across multiple datacenters and hardware vendors. Operate cluster scheduling, resource allocation, isolation, quotas, and utilization systems. Debug complex production issues across Linux, networking, storage, drivers, firmware, and orchestration layers. Build and operate high‑performance networking infrastructure, including RDMA‑enabled environments and accelerator interconnects. Build observability for cluster health, capacity, performance, failures, and workload behavior. Improve reliability, availability, and recovery across multi‑node production systems. Work with distributed systems and runtime teams to support low‑latency, high‑throughput inference workloads. Evaluate and integrate new hardware platforms, accelerators, networking technologies, and datacenter designs. Create runbooks, operational standards, and incident response practices as the fleet scales. You may be a good fit if Experience in infrastructure, cluster engineering, platform engineering, SRE, HPC, or distributed systems. Deep Linux systems experience, including debugging performance, networking, storage, processes, and kernel‑level issues. Experience operating Kubernetes, Slurm, Nomad, or similar orchestration and scheduling systems. Strong automation skills using tools such as Terraform, Ansible, Helm, Python, Go, or equivalent. Experience with GPU or accelerator infrastructure, including drivers, firmware, CUDA/ROCm stacks, or hardware validation. Familiarity with high‑performance networking such as InfiniBand, RoCE, high‑speed Ethernet, or datacenter fabrics. Strong operational judgment: you know how to build systems that are observable, recoverable, and boring in production. Comfort working in a fast‑moving startup environment with high ownership and ambiguity. Strong candidates may also have Experience building or operating AI inference, training, HPC, or neocloud infrastructure. Experience with bare‑metal provisioning, PXE/iPXE, image pipelines, BIOS/firmware management, or rack bring‑up. Experience with multi‑tenant cluster isolation, quota systems, fair scheduling, or usage accounting. Experience debugging distributed workload performance across compute, memory, network, and storage bottlenecks. Experience building observability platforms using technologies such as Prometheus, OpenTelemetry, Grafana, or similar tooling. Familiarity with heterogeneous hardware environments across NVIDIA, AMD, Intel, ARM, or emerging accelerators. #J-18808-Ljbffr

Vacancy posted 20 hours ago
Similar jobs that could be interesting for youBased on the Member of Technical Staff - Infrastructure in San Francisco, CA vacancy
  •  ...Infrastructure Platform Engineer We are looking for an Infrastructure platform Engineer to design, build, and operate the cluster infrastructure behind Gimlet's heterogeneous inference cloud. Unlike traditional cloud platforms built around a single hardware ecosystem... 
    Suggested

    Gimlet Labs

    San Francisco, CA
    2 days ago
  •  ...agent for enterprise computer automation. Our developer platform writes, tests, and maintains automation code on fully-managed infrastructure – cutting dev time by 90%. We're starting with healthcare, where legacy systems make reliable automation a genuinely hard problem... 
    Suggested
    Immediate start
    Remote work

    CloudCruise Inc.

    San Francisco, CA
    1 day ago
  •  ...requirements, and very few precedents to copy from. About the Role Members of Technical Staff (MTS) are the senior engineers who build the platform...  ...is systems engineering at its core. Multi‑tenant data infrastructure across very different portcos. Event‑driven pipelines... 
    Suggested

    BEACON SOFTWARE COMPANY

    San Francisco, CA
    20 hours ago
  •  ...Member Of Technical Staff We're looking for a member of technical staff to build and deploy production-grade AI systems. In this role, you...  ...Familiarity with data pipelines, APIs, and cloud infrastructure (AWS, GCP) Experience working with machine learning models... 
    Suggested

    Eragon

    San Francisco, CA
    1 day ago
  • $140k - $200k

     ...Member of Technical Staff Harper is an AI-native commercial insurance company in San Francisco. We're not bolting AI onto insurance - we...  ...building software, full-stack by instinct (frontend, backend, infrastructure). ~ You've shipped AI to production - real systems... 
    Suggested
    Work at office
    Relocation

    Harper Group

    San Francisco, CA
    4 days ago
  •  ...Member of Technical Staff @ Lotus AI Who we are Lotus AI is a groundbreaking primary care app that integrates your medical records, AI, and...  ...Experience with system refactors, schema migrations, and data infrastructure simplification Familiarity with PostgreSQL (including... 

    Lotus Health AI

    San Francisco, CA
    2 hours ago
  •  ...Member of Technical Staff – Full Stack / AI Systems Company : AdsGency AI Relocation : San Francisco City Required Authorization : Applicants...  ...’ll be one of the core builders of AdsGency’s AI‑driven infrastructure — designing, architecting, and scaling systems that orchestrate... 
    Full time
    Work experience placement
    Relocation
    Visa sponsorship

    AdsGency AI

    San Francisco, CA
    20 hours ago
  •  ...environments for Fortune 500 companies. About the Role As a Member of Technical Staff, you will be part of the team responsible for the work...  ...entire stack. The MTS works vigorously on the underlying infrastructure, core features, agent configurations, and user experience... 
    Work experience placement
    H1b
    Work at office
    Visa sponsorship

    Ersilia

    San Francisco, CA
    4 days ago
  • $200k

     ...Join to apply for the Member of Technical Staff role at Listen Labs . TL;DR: We are seeing strong market demand and an aggressive 6...  ...product and must make decisions across the LLM pipeline, infrastructure, backend, and UX. You have a high bar for quality: In... 
    Flexible hours

    Listen Labs

    San Francisco, CA
    1 day ago
  •  ...Catalog is building the commerce layer for AI - the missing infrastructure that lets agents not just search the web, but understand,...  ...reshape how people discover and buy online. Role As a Member of Technical Staff, you will ship core systems, set engineering culture, and... 
    Work at office

    Getcatalog

    San Francisco, CA
    1 day ago
  •  ...Member of Technical Staff @ Lotus AI Lotus AI is a groundbreaking primary care app that integrates your medical records, AI, and real doctors...  ...with system refactors, schema migrations, and data infrastructure simplification Familiarity with PostgreSQL (including... 

    Lotus Health

    San Francisco, CA
    3 days ago
  •  ...We are looking for a Member of Technical Staff with strong Python skills and a passion for building scalable platforms for AI and ML workloads...  ..., and play a critical role in shaping Activeloop's AI infrastructure. What You Will Be Doing Architect and Develop:... 

    Renaissance Academy

    San Francisco, CA
    20 hours ago
  •  ...Pixeltable Inc. Member of Technical Staff San Francisco, CA·Full time Apply for Member of Technical Staff As a founding member of the engineering...  ...lies in empowering teams to focus on innovation, not on infrastructure. We aim to simplify the AI development stack, allowing... 
    Full time
    Part time
    Work at office
    Work from home
    Flexible hours
    2 days per week

    Pixeltable, Inc.

    San Francisco, CA
    1 day ago
  • $250k

     ...leaves their servers. The team is small, technical, and moving fast, with strong early...  ...· Industry: AI Tools. The Role Member of Technical Staff who can handle everything from modeling...  ...data pipelines, APIs, and cloud infrastructure (AWS, GCP) Experience working with machine... 
    Full time

    David Joseph & Company

    San Francisco, CA
    1 day ago
  • $227.5k - $401k

     ...motivated individuals who tackle unique technical challenges at scale and solve them...  ...financial technology sector. As a Member of Technical Staff, you will operate with a high degree...  ...Experience in AI‑enabled fintech or infrastructure companies. Familiarity with classical... 
    Work at office
    Immediate start
    Relocation
    Flexible hours

    Adyen

    San Francisco, CA
    20 hours ago
  • $200k

     ...role. You can refer people through our form. We’re hiring a member of technical staff to work closely with the founding team. You'll shape both...  ...Backend engineering: Design and build scalable infrastructure and systems for AI agents, including containerization, and... 
    Immediate start

    Fulcrum

    San Francisco, CA
    1 day ago
  • $150k - $300k

     ...pioneering biologists, Phylo is building the next generation of AI systems for the life sciences. About The Role We’re looking for an Infrastructure Engineer to build, operate, and scale the deployment foundations for Phylo’s enterprise AI platform. This role is focused on... 
    Work at office

    Phylo

    South San Francisco, CA
    3 days ago
  •  ...As a Member of Technical Staff (MTS), you'll build production-grade systems that power continuous optimization loops for AI agents—from evaluation pipelines and data/trace infrastructure to APIs that deploy improved policies. This role is a blend of MLE + backend engineering... 

    VizopsAI

    San Francisco, CA
    20 hours ago
  •  ...longer whether quantum and AI converge; it is who builds the infrastructure to make that convergence reliable, scalable, and useful....  ...ours at the frontier of science. Role Overview As a Member of Technical Staff you will shape Conductor's core offerings: AI software that... 

    Conductor Quantum

    San Francisco, CA
    1 day ago
  •  ...uses Shapes every single day, and everyone talks to users. Member of Technical Staff is the title we use for engineers who own hard problems...  ...scale High-performance Python backends at scale Realtime infrastructure (WebRTC, WebSockets, streaming) for chat, voice, or video... 

    Shapes

    San Francisco, CA
    1 day ago
  • Member of Technical Staff Shared Context is building adaptive personal AI that understands the texture of real life: our relationships, routines...  ..., IDEO. What you’ll do Build core agent, product, and infrastructure systems. Create simulation environments for users,... 

    Shared Context Lab

    San Francisco, CA
    20 hours ago
  •  ...structures that shaped the open web and modern computing infrastructure, the Foundation serves as a neutral, multi-stakeholder convening...  ...The Role The OpenClaw Foundation is seeking exceptional Members of Technical Staff (MTS) to serve as full-time maintainers, builders, and... 
    Full time

    OpenClaw Foundation

    San Francisco, CA
    4 days ago
  • $150k - $250k

     ...people take ownership, grow together, and share both the challenges and the wins. What you'll do Build the supercomputing infrastructure that runs our agents. Our agents tackle long-horizon, high-performance workloads, and you'll design the cloud compute,... 
    Work at office
    Remote work
    Flexible hours

    Asari AI

    San Francisco, CA
    11 days ago
  •  ...built brag to your friends about your hyper-optimized AI coding workflows tinker and build software for the love of the game feel equally strong obligations to both 1) choose good and 2) to win think that this role should be renamed "member of tomo staff"... 
    Immediate start

    Tomo

    San Francisco, CA
    20 hours ago
  •  ...kernels, but will expand into every corner of ML systems and AI infrastructure. We're a small team (4 people) backed by Fifty Years, Y...  ...etc.) What We Look For You're a strong fit if you: Have deep technical intuition and can learn new domains quickly Are comfortable working... 
    Remote work

    WAFER INC

    San Francisco, CA
    1 day ago
  • $140k - $200k

     ...real users You’re full-stack by instinct (frontend, backend, infrastructure) You ship fast and iterate—meaningful features in days that others...  ...mission and pace If in SF: Super Day on-site If outside SF: Technical phone screen, then on-site To Apply If you want to build... 
    Work at office
    Relocation

    Harper (YC W25)

    San Francisco, CA
    1 day ago
  •  ...interacting with third‑party systems. But today's financial infrastructure was built for humans, not autonomous systems. Companies and developers...  ..., and beyond. You'll operate as a true owner, making real technical tradeoffs and shaping product direction alongside the... 
    Shift work

    Sapiom

    San Francisco, CA
    20 hours ago
  • $150k - $300k

     ...powered agents. You'll design and own zero‑to‑one systems across LLM pipelines, document parsing, browser automation, and backend infrastructure. You’ll shape the product and the company. What you’ll do Build and deploy AI agents that parse documents, extract entities,... 
    Work at office

    Modus

    San Francisco, CA
    1 day ago
  • $125k - $200k

     ...our core AI agent system from the ground up Making critical technical decisions that will shape our product's future Building...  ...monitoring and safety systems for AI agents Architect cloud-native infrastructure ready for enterprise scale About You: Voracious... 
    Full time
    Temporary work
    Currently hiring
    Immediate start
    Flexible hours

    burnt

    San Francisco, CA
    3 days ago
  • $150k - $250k

     ...our own DC design to automate the generation of an end to end site design and schedule. We're building the live model of a AI infrastructure project’s entire delivery schedule, durable workflows handing off across every team, to replace static Gantt charts with a plan... 
    Local area

    AI Chopping Block, Inc.

    San Francisco, CA
    20 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Member of Technical Staff - Infrastructure. Be the first to apply!