Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Software Engineer - Infrastructure Storage

Lambda Inc.

Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. *Note: This position requires presence in our San Francisco/San Jose/Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday. In the world of distributed AI, raw GPU and CPU horsepower is just a part of the story. High-performance networking and storage are the critical components that enable and unite these systems, making groundbreaking AI training and inference possible. The Lambda Infrastructure Engineering organization forges the foundation of high-performance AI clusters by welding together the latest in AI storage, networking, GPU and CPU hardware. Our expertise lies at the intersection of: High-Performance Distributed Storage Solutions and Protocols: We engineer the protocols and systems that serve massive datasets at the speeds demanded by modern clustered GPUs. Dynamic Networking: We design advanced networks that provide multi-tenant security and intelligent routing without compromising performance, using the latest in AI networking hardware. Compute Virtualization: We enable cutting-edge virtualization and clustering that allows AI researchers and engineers to focus on AI workloads, not AI infrastructure, unleashing the full compute bandwidth of clustered GPUs. About the Role We are seeking a seasoned Staff Storage Software Engineer with deep experience designing and deploying storage protocol solutions at scale across object, block, and file paradigms. This is a unique opportunity to work at the intersection of large-scale distributed systems and the rapidly evolving field of artificial intelligence infrastructure. This is an opportunity to have a significant impact on the future of AI. You will be building the foundational infrastructure that powers some of the most advanced AI research and products in the world. What You’ll Do Technical Leadership: Set technical direction for storage software architecture across the Infrastructure Engineering organization, influencing decisions that span petabyte-scale deployments. Author and review design documents for new storage systems, protocols, and integrations; raise the technical bar across the team. Mentor and develop senior engineers, providing guidance on systems design, debugging complex distributed systems issues, and navigating technical tradeoffs. Serve as a technical anchor for cross-functional initiatives involving storage, networking, compute, and control plane teams. Represent the storage software team in architectural reviews, roadmap planning, and customer-facing technical discussions where needed. Execution: Design, develop, and maintain high-performance storage systems software with a focus on performance, scalability, reliability, and operational simplicity. Implement and optimize storage protocol APIs across file (NFS, SMB, Lustre), block (NVMe-oF, iSCSI, Fibre Channel), and object (S3) access patterns. Develop distributed systems for managing and orchestrating storage resources across multiple solutions and redundant arrays. Collaborate with hardware and system architects to integrate software with storage solutions including NVMe, GPU-direct storage, and DPU-accelerated data paths. Troubleshoot and resolve complex issues in production data center environments, including performance regressions, protocol mismatches, and hardware failures. Contribute across the full software development lifecycle — from requirements gathering and system design through deployment, monitoring, and long‑term maintenance. Build and maintain tooling for storage benchmarking, performance profiling, and capacity planning. Collaboration: Work closely with storage software and networking teams to execute cross-functional infrastructure initiatives and new data center deployments, including integration of storage protocols across a variety of on-prem solutions. Partner with the control plane and Kubernetes teams to meet customer and product requirements for usability, reliability, and telemetry. Work with the observability team to define, build, and track SLOs/SLIs for storage systems. Coordinate with Networking, Compute, and Storage Engineering teams to deploy high-performance distributed storage solutions that serve AI/ML workloads. Partner with the Fleet Engineering team to ensure seamless deployment, monitoring, and ongoing maintenance of distributed storage infrastructure. Innovate: Stay current with the latest research and developments in AI and HPC storage technologies, and bring relevant advances into Lambda's infrastructure. Work with the Lambda product team to identify emerging trends in AI inference and training that will shape next-generation storage requirements. Evaluate and prototype new storage solutions, protocols, and hardware integrations — from open-source distributed filesystems to vendor-specific accelerated storage products. Optimize storage protocol solutions for AI workloads, including checkpoint I/O for training, high-throughput dataset serving, and latency-sensitive inference pipelines. You Have Experience 10+ years of experience in storage systems engineering, with at least 5 years in a technical lead or Staff+ IC role. Proven track record designing and operating storage infrastructure at scale (multi-petabyte environments preferred) in production data center or cloud settings. Experience leading technical projects end-to-end, from architecture through delivery with cross-functional stakeholders. Background working in high-performance computing, AI/ML infrastructure, or large-scale cloud storage environments. Systems-Level Programming Strong proficiency in one or more low-level systems programming languages: C, C++, Rust, or Go. Demonstrated ability to write high-performance, concurrent, production-grade systems code and conduct thorough code reviews. Experience with kernel-level storage drivers, user-space I/O frameworks, or storage daemon development is a strong plus. Familiarity with DPDK and SPDK and their role in building high-performance, kernel-bypass storage and networking data paths. Storage Protocol & API Expertise Deep hands-on experience with two or more storage protocols: object (S3 or similar), block (iSCSI, Fibre Channel, NVMe-oF), or file (NFS, SMB, Lustre, DAOS). Experience implementing or maintaining storage protocol servers or clients in production, not just consuming them. Familiarity with storage API performance characteristics such as latency, throughput, IOPS and the ability to diagnose and resolve bottlenecks at the protocol level. Storage Performance Optimization Experience profiling and tuning storage systems for throughput, latency, and IOPS under real production workloads. Familiarity with tools such as fio, blktrace, perf, eBPF/bpftrace, or equivalent for storage performance analysis. Understanding of I/O scheduling, caching layers, write amplification, and related performance tradeoffs. Modern Storage Technologies Familiarity with NVMe, NVMe-oF, and RDMA (RoCE or InfiniBand) and their impact on storage system architecture. Working knowledge of DPUs (e.g., NVIDIA BlueField) and their role in offloading storage and networking data paths. Experience with GPU-direct storage or similar zero-copy data paths is a plus. Physical Infrastructure & Operational Acumen Comfort working in a physical data center environment — understanding rack-scale infrastructure, storage array hardware, cabling, and failure domains. Experience building and operating storage systems with strong reliability expectations: designing for failure, building runbooks, and driving incident response. Familiarity with storage observability tooling — metrics pipelines (Prometheus, Grafana), log aggregation, and tracing in distributed storage environments. Nice to Have Experience with NVIDIA BlueField DPUs or SuperNICs for accelerated storage data paths, including GPUDirect Storage implementation. Deep production experience with enterprise or HPC storage platforms: Vast Data, Weka, NetApp, or IBM Spectrum Scale. Experience deploying and operating Ceph at scale (100PB+) in an HPC or AI infrastructure environment. Familiarity with emerging storage technologies such as CXL memory pooling, computational storage, or ZNS (Zoned Namespace) SSDs. Experience contributing to or maintaining open-source storage projects (e.g., Ceph, DAOS, Lustre, MinIO). Salary Range Information The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description. Benefits We offer generous cash & equity compensation Health, dental, and vision coverage for you and your dependents Wellness and commuter stipends for select roles 401k Plan with 2% company match (USA employees) Flexible paid time off plan that we all actually use A Final Note You do not need to match all of the listed expectations to apply for this position. We are committed to building a team with a variety of backgrounds, experiences, and skills. Equal Opportunity Employer Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law. #J-18808-Ljbffr Lambda Inc.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Staff Software Engineer - Infrastructure Storage in San Francisco, CA vacancy
  • $300 per month

     ...and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each...  ...in each other, come build with us at Crusoe. Senior Staff Software Engineer, Storage San Francisco, Sunnyvale, or Bellevue (Onsite) About This... 
    Suggested

    Crusoe Energy Systems LLC

    San Francisco, CA
    1 day ago
  • $215k - $265k

    Data Direct Networks is seeking a Senior Staff Software Engineer to enhance their S3 compliant high-performance file system. The role demands expertise in C/C++, with over 12 years of software development experience. Candidates should possess strong communication skills... 
    Suggested
    Remote job

    DataDirect Networks Inc

    San Francisco, CA
    2 days ago
  •  ...London and Amsterdam. The Online Storage team is growing! We build...  ...Data Models used by all of engineering. The goal is to evolve Plaid...  ...query performance and infrastructure cost. You will wield terraform...  ...confidently Qualifications Strong software engineering experience with... 
    Suggested
    Work experience placement
    Local area

    Plaid

    San Francisco, CA
    4 days ago
  • $200k - $400k

     ...and grow as a team. About the Team The Infrastructure team builds and operates the...  ...foundational cloud stack—networking, compute, storage, security, and infrastructure‑as‑code—to...  ...Role We’re hiring a Senior Infrastructure Engineer to design, build, and operate... 
    Suggested
    Full time
    Work at office
    Local area

    Decagon AI, Inc.

    San Francisco, CA
    1 day ago
  • $200k - $230k

     ...looking for an experienced engineer with deep expertise in distributed...  ...shape the future of Gusto’s storage layer. You’ll manage complex...  ...the Team The Datastores Infrastructure Engineering team designs,...  ...re looking for 12+ years of software engineering experience building... 
    Suggested
    Work at office
    Local area
    Remote work
    2 days per week
    3 days per week

    Prudence Holdings Inc

    San Francisco, CA
    1 day ago
  • Reific is seeking a Member of Technical Staff for a full-time role in San Francisco. You'll be responsible for building the Reific interface and backend, transforming operational data into forecasts and decision records. The role entails creating product flows, designing... 
    Full time

    Reific

    San Francisco, CA
    17 hours ago
  • $207k - $345k

    Senior Staff Software Engineer, Infrastructure About this Position Rippling gives businesses one place to run HR, IT, and Finance. It brings together all of the workforce systems that are normally scattered across a company, like payroll, expenses, benefits, and computers... 
    Work at office
    Local area
    3 days per week

    Rippling

    San Francisco, CA
    4 days ago
  •  ...general intelligence benefits all of humanity. The Identity Infrastructure Engineering team sits at the core of this effort, designing and...  ...innovative AI research. About the Role We’re looking for a Staff+ Software Engineer to help build and evolve the identity... 
    Work at office
    Relocation package

    Aimling

    San Francisco, CA
    17 hours ago
  • $197.3k - $313.7k

    ## Staff Software Engineer, Electron & Browser Infrastructure - Slack DesktopApplyremote type: Office Tech-Flexiblelocations: Georgia - Atlanta: Washington - Seattle Metro - Remote: Washington - Seattle: California - Remote: California - San Franciscotime type: Full timeposted... 
    Work at office
    Remote work

    Slack Enterprise

    San Francisco, CA
    4 days ago
  • A leading AI infrastructure company in San Francisco seeks a Staff Software Engineer to lead software-defined networking initiatives. You will oversee the development of cutting-edge networking solutions and drive technology adoption within the engineering team. Candidates... 

    Crusoe Energy Systems LLC

    San Francisco, CA
    3 days ago
  • $300 per month

     ...energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each...  ...Systems is seeking a highly skilled and motivated Senior Staff Software Engineer - SoftwareDefinedNetworking to lead the development and... 
    Temporary work

    Crusoe Energy Systems LLC

    San Francisco, CA
    4 days ago
  • $185k - $224k

     ...and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each...  ...Role: Crusoe Cloud seeks a highly skilled and experienced Staff Software Engineer to lead the development and execution of our cutting‑edge... 
    Temporary work

    Crusoe Energy Systems LLC

    San Francisco, CA
    3 days ago
  • $207k - $300k

    Staff Software Engineer, Firestore, Google Cloud Google San Francisco, CA, USA Required Qualifications...  ...ideation. Experience with systems infrastructure and distributed systems. Preferred...  ...system design, networking, data storage, security, artificial intelligence, natural... 
    Full time

    Google Inc.

    San Francisco, CA
    1 day ago
  • $300 per month

     ...intelligence. We’re crafting the engine that powers a world...  ..., transformative cloud infrastructure. About This Role: The Crusoe Cloud Software Development team is...  ...and experienced Senior Staff Software Engineer...  ...accelerating AI compute, storage, and networking resources... 
    Full time
    Temporary work

    Crusoe Energy Systems LLC

    San Francisco, CA
    3 days ago
  • $320k - $405k

     ...growing group of committed researchers, engineers, policy experts, and business leaders...  .... About the role the company's infrastructure footprint — datacenters, networking, compute...  ...that need, and we’re looking for a Staff Software Engineer to join the build and own its... 
    Visa sponsorship
    Flexible hours

    United States Digital Space LLC

    San Francisco, CA
    1 day ago
  • $189k - $330.75k

     ...from @ Rippling.com addresses. About the role Rippling’s Infrastructure organization builds the technical backbone that powers...  ...significant surge in daily events and system complexity. As a Staff Software Engineer on the Cloud Infrastructure team, you'll be in a high-... 
    Work at office
    3 days per week

    Rippling

    San Francisco, CA
    4 days ago
  • $180k - $200k

     ...follow us on LinkedIn. AI Engineering @ Ironclad Ironclad is...  ...reliable, highly scalable software and services designed for a...  ...Deliver and Optimize AI Infrastructure: Work with platform teams to...  ...0,000 Base Salary Range - Staff: $210,000 - $235,000 The base... 
    Contract work

    Ironclad

    San Francisco, CA
    4 days ago
  •  ...computing. About the Role We're seeking a Platform Engineer to design and build the control plane that powers Hyperbolic...  ...planes for cloud platforms, developer platforms, or infrastructure services Expert-level software engineering skills in Go (Golang) with a track record... 
    Worldwide

    Hyperbolic Labs

    San Francisco, CA
    17 hours ago
  • Golunar in San Francisco is looking for a software engineer to develop innovative healthcare solutions. You will design modern cloud architectures and access control systems, ensuring 24/7 hospital operations. The ideal candidate has 10+ years of experience, proficiency... 
    Flexible hours
    3 days per week

    Golunar

    San Francisco, CA
    17 hours ago
  •  ...privacy, and financial audits. We build software for the people who enable trust...  ...‑critical work. About the Role As a Staff Platform Engineer at Fieldguide, you will design and build...  ...reducing complexity for product, AI, and infrastructure engineers. Your work will directly... 
    Remote work
    Flexible hours

    Fieldguide.ai

    San Francisco, CA
    2 days ago
  • $200k - $271.5k

    Job Summary The Staff Software Engineer, Monetization Platform serves as a technical leader with a primary focus on the systems that turn...  ...CD practices At least one major cloud platform and modern infrastructure tooling A strong product and systems mindset: you can model... 
    Contract work
    Flexible hours

    Touring Capital LLC

    San Francisco, CA
    4 days ago
  • Job Description Slack is looking for a Staff Software Engineer to join the Data Infrastructure team within the broader Data Engineering organization. The mission is to build secure, reliable, performant, scalable, and cost‑efficient infrastructure that powers Slack’s data... 

    100 Salesforce, Inc.

    San Francisco, CA
    1 day ago
  •  ...Especially when it’s hard. Courage: Embrace the challenge. Together: Build bridges and lift up your colleagues. Job Summary As a Staff Software Engineer, you will play a key role in the entire engineering lifecycle from design, documentation, build, test and maintain our... 
    Temporary work
    Work at office
    Remote work
    Flexible hours

    SmithRx

    San Francisco, CA
    2 days ago
  • Description Slack is looking for a Staff Software Engineer to join the Data Infrastructure team within the broader Data Engineering organization. The mission...  ...services that underpin data ingestion, transformation, storage, compute and orchestration at a massive scale. As a... 
    Permanent employment

    B Capital

    San Francisco, CA
    2 days ago
  • $189k - $330.75k

    Staff Software Engineer - Spend Platform About this position Rippling gives businesses one place to run HR, IT, and Finance. It brings together all of the workforce systems that are normally scattered across a company, like payroll, expenses, benefits, and computers. For... 

    Rippling

    San Francisco, CA
    3 days ago
  • A leading technology company based in San Francisco is seeking a Staff Software Engineer for its Cloud Infrastructure team. This role involves designing resilient cloud infrastructure to support significant system load increases. The ideal candidate has over 8 years of... 

    Rippling

    San Francisco, CA
    4 days ago
  • $320k - $405k

     ...group of committed researchers, engineers, policy experts, and...  ...Platform. We build and operate the infrastructure that turns product usage...  ...mission. We’re looking for a software engineer to join Billing....  ...policy: Currently, we expect all staff to be in one of our offices... 
    Contract work
    Visa sponsorship
    Shift work

    United States Digital Space LLC

    San Francisco, CA
    1 day ago
  •  ...that redefines work itself. The Role As a founding Applied AI Engineer at Valence, you will help define and build the future of AI-powered...  ...Looking For Technical foundation: 8+ years of experience in software engineering, AI/ML, data-intensive systems, AI/ML development (... 
    Full time
    Freelance
    Remote work

    Valence

    San Francisco, CA
    1 day ago
  • $200k - $271.5k

    # Staff Software Engineer, Monetization PlatformHybrid - San FranciscoApply**Our Mission & Values:** At Drata, we help companies earn and...  ...practices + At least one major cloud platform and modern infrastructure tooling* A strong product and systems mindset: you can model... 
    Contract work
    Work at office
    Immediate start
    Worldwide
    Monday to Friday
    Flexible hours

    Careers at Drata

    San Francisco, CA
    4 days ago
  •  ...world. The Role: We are looking for an experienced Senior Staff Software Engineer to join our Builder Tools engineering organization with a...  ..., and elevate product reliability through testing infrastructure innovations and practices. You will get the chance to lead... 
    Remote work

    Israelvcforum

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Software Engineer - Infrastructure Storage. Be the first to apply!