Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer, Infrastructure & Performance

$180k - $250k
Full-time

Dedalus Labs, Inc.

About the role

Dedalus Labs builds persistent computers for AI agents. Our flagship product, Dedalus Machines, gives agents an isolated environment where they can run software, keep files and state, and work over time.

We're looking for an infrastructure engineer who's bothered by a slow build, an unexplained latency spike, or a workflow that makes engineers do the same work twice. You want to understand where the time went. Then you want to fix it, measure the improvement, and make sure everyone who uses that system benefits from it.

At a small startup, those details affect how quickly the whole company can move. A build that wastes ten minutes wastes them repeatedly. A confusing API slows every product built on it. A deployment that needs someone to supervise it interrupts work elsewhere. You'll seek out these problems across our systems and codebase, including the ones people have learned to tolerate. Your work can bring product releases and customer commitments forward by weeks, and over time, months.

You'll work directly with our CTO and engineers, with substantial freedom to investigate, experiment, and choose an approach. We want someone who takes pride in the craft of software engineering and cares about the systems they'll leave for the next person. That means useful abstractions, clear failure behavior, tests that establish the right guarantees, and documentation that explains the decisions.

What you'll work on

Make the development loop faster. Profile builds, tests, CI queues, artifact transfers, and local workflows. Find repeated work, unnecessary dependencies, poor scheduling, and cache misses. Follow a bottleneck into a compiler, linker, filesystem, network, or upstream dependency when the evidence takes you there. Establish a baseline, make the change, and verify the result on a representative workload.

You'll use and improve our open-source tooling: Bessemer (BSMR), our build system, and Hollywood, which generates GitHub Actions from typed TypeScript definitions and runs actions locally. This includes working on the tools themselves when their implementation limits what we can do. Experience with Buck2, Bazel, Blaze, remote execution, or build cache design is particularly relevant.

Build platforms other engineers can build on. Design APIs, controllers, CLIs, and services that support multiple internal products and consumers. Think through resource lifecycles, concurrency, cancellation, permissions, and failure reporting. Work with the engineers using those interfaces, understand their constraints, and make common operations straightforward without hiding important system behavior.

Make delivery reliable and understandable. Improve CI/CD, GitOps, environment provisioning, staging, production rollout, and recovery. Connect the source change, build inputs, artifact, and running software so engineers can trace a release and diagnose a failure. Instrument the path with useful metrics, logs, and traces. Use incidents and recurring manual work to identify what needs a software fix.

Work with hardware we control. We have an in-office homelab where we test and debug on real machines and networks we control. Depending on your strengths, you can help assemble, provision, and maintain machines, configure networking, and build test infrastructure so engineers can reproduce failures and measure changes under conditions we understand.

You'd be a good fit if

  • A problem keeps your attention when the first few explanations turn out to be wrong. You read source, build a smaller reproduction, ask a better question, or collect another trace. You ask for help when it will move the investigation forward and keep responsibility for the outcome.

  • You notice slowness and repeated effort even outside your immediate assignment. You investigate proactively and can explain why fixing it matters to the people using the system.

  • You enjoy the last 20% of performance work, including the part that takes 80% of the effort. You can also judge when that effort is worth spending and when a deadline calls for a smaller, complete improvement.

  • You're detail oriented and a little perfectionist. You care about API names, error messages, correctness, performance, and the next engineer's ability to understand your work. You can make a decision, finish it, and ship.

  • You respect the craft and stay open to better ways of practicing it. You use modern AI coding tools to accelerate exploration, implementation, testing, and review. You understand the resulting code and take responsibility for its behavior.

  • You work well with ambiguity. You can turn an incomplete goal into a concrete problem, agree on the constraints, and choose a useful next step. You enjoy the freedom to tinker and can turn an experiment into something the team depends on.

Experience that helps

You should have strong software engineering skills and experience building or operating infrastructure that other people rely on. We're particularly interested in depth in Go, Rust, C++, or another systems language, practical Linux debugging, and thoughtful API design. You'll also work with TypeScript in our tooling.

Build systems, CI/CD, GitOps, Kubernetes, observability, and infrastructure as code are relevant areas of experience. Familiarity with Terraform is useful. Tell us where you've gone deep and what you learned from operating the system after you built it.

Hands-on hardware and networking experience is a plus. You may have assembled and maintained machines, installed storage or network cards, provisioned Linux, or configured switches, routing, and VLANs. We're interested in people who understand how the parts fit together and can trace a problem across application code, the operating system, storage, and the network. Deeper experience with Ethernet, SFP+/QSFP optics and direct-attach cables, RDMA, or RoCE is useful additional depth.

This role may be a poor fit if

  • You need a complete specification and a fixed twelve-month roadmap before you can make progress.

  • You prefer an assignment limited to one tool or layer and find it frustrating to follow a problem across software, infrastructure, and hardware.

  • You routinely accept slow or manual workflows because they've always worked that way.

  • You keep polishing after the agreed deadline or declare a performance improvement before measuring it.

Location and compensation

This is a full-time, in-person role in San Francisco. The annual base salary range is $180,000-$250,000 USD, depending on relevant experience and role scope. Relocation assistance and visa sponsorship are available.

Show us your work

Tell us about a system you improved because something about it bothered you. What did you notice? How did you find the cause? What changed, how did you measure it, and what did you choose to leave alone?

A code sample, technical write-up, open-source contribution, or homelab project is welcome. A description of a private production project works too. Please keep confidential details out of your application. We'd like to understand your judgment, your persistence, and the care you put into the result.

Explore our open-source repositories, particularly Hollywood and Bessemer (BSMR), and star them if you'd like to follow their development. We especially welcome substantive contributions: a useful bug fix, a measured performance improvement, a thoughtful API change, or tests that establish an important guarantee. Start by discussing the problem and proposed scope with the maintainers.

Working through an issue, implementation, and code review gives us a direct sense of how you investigate, explain your decisions, and respond to feedback. It also gives you a chance to experience how we work and whether you'd enjoy building with us. Contributing is an optional way to show your work, and we're equally interested in the work you've already done.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Software Engineer, Infrastructure & Performance in San Francisco, CA vacancy
  • $148.5k - $223.9k

     ...control and data planes, transforming the software stack to adopt cloud-native primitives,...  ...distributed systems software engineers passionate about new challenges and capable...  ...benefits, training, assessment of job performance, discipline, termination, and everything... 
    Performance
    Full time

    Salesforce

    San Francisco, CA
    1 day ago
  • $165k - $185k

     ...together. About The Role We’re looking for a Senior Software Engineer, Infrastructure to own the platform that VSCO product and data teams...  ...) and is eligible for discretionary bonuses based on performance. The benefits available for this position include flexible... 
    Performance
    Temporary work
    Work at office
    Local area
    Worldwide
    Flexible hours

    VSCO

    San Francisco, CA
    2 days ago
  • $300k

     ...opportunity?  Join a stealth-mode hyperscale infrastructure startup building a 300MW+ AI compute...  ...under development. The Principal Software Engineer will take ownership of the software...  ...connects thousands of GPUs, high-performance networking fabrics, storage systems,... 
    Performance
    Full time
    Remote work
    Flexible hours
    San Francisco, CA
    more than 2 months ago
  • $208k - $260k

     ...institutions Work together with engineers, scientists, operators, and...  ...Human data is the core infrastructure to AI advancement. Frontier...  ...Handshake is hiring a Senior Software Engineer to join Backend...  ...architecture, APIs, data systems, performance, reliability, and developer... 
    Performance
    Full time
    Work at office
    Remote work
    Flexible hours

    Handshake

    San Francisco, CA
    12 days ago
  •  ...ll help design, deploy and operate the infrastructure that makes this possible. This is a...  ...of infrastructure, platform engineering and security. You’ll define how Modern...  ...engineering standards uplift Optimize performance, reliability, and observability across... 
    Performance
    Full time
    Local area
    Immediate start
    Remote work

    Modern Treasury

    San Francisco, CA
    3 days ago
  • $180k - $250k

     ...generation of AI products. We build the infrastructure, tools, and model access that teams...  ...: a unified platform where high-performance inference, orchestration, and...  ...build on. You are a hands-on engineer who builds the software and processes that keep a large fleet... 
    Performance
    Temporary work
    Local area
    Remote work
    Relocation package

    The Consensus

    San Francisco, CA
    4 days ago
  •  ...About the Team We’re hiring Software Engineers to join our broader Infrastructure organization, which supports multiple high-impact teams. Depending on...  ...complex engineering problems at scale, ensuring their performance, scalability and reliability Team Focus Areas... 
    Performance

    openai

    San Francisco, CA
    23 hours ago
  • $190k - $221k

    Senior Software Engineer, Infrastructure TRM Labs is a blockchain intelligence company committed to fighting crime and creating a safer world. By...  ...appropriate tradeoffs between simplicity, readability, and performance. Provides mentorship to junior engineers, and enhances... 
    Performance
    Remote work

    TRM Labs

    San Francisco, CA
    5 days ago
  •  ...address real-world challenges. The Infrastructure Engineering team is crucial to the overall...  ...Spearhead critical architectural decisions, perform deep-dive code reviews, and evaluate...  ...Standards: Establish and enforce elite software engineering and DevOps standards,... 
    Performance
    Shift work

    Hayden AI Technologies, Inc.

    San Francisco, CA
    23 hours ago
  • $184k - $259.44k

     ...Scale AI is seeking a highly skilled and motivated Software Engineer, Frontier AI Infrastructure to join our dynamic Public Sector Engineering team....  ...related skills, experience, qualifications, interview performance, and relevant education or training. Scale employees... 
    Performance
    Full time
    Work at office
    3 days per week
    Early shift

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  • Serval Software Engineer, Infrastructure Serval is building an AI platform to automate complex IT workflows for modern enterprises. As a Software...  ...self-hosted instances of Serval. Ensure high availability, performance, and reliability of production systems through... 
    Performance

    Serval

    San Francisco, CA
    5 days ago
  •  ...the data and intelligence infrastructure that helps businesses work...  ...Employers lists. About Middesk Engineering: "Velocity" is the rate at...  ...that provide insight into performance, reliability, and usage....  ...+ years of experience as a software engineer, with genuine interest... 
    Performance
    Work at office
    Local area
    2 days per week

    Middesk

    San Francisco, CA
    6 days ago
  •  ...Think of this as a product engineer role - but the product is everything...  ...real actions in the world, infrastructure is the user experience. When...  ...) for reliability, performance, and cost Lead infrastructure...  ...'re looking for 4+ years of software engineering experience, with... 
    Performance
    Local area

    Blockit AI, Inc

    San Francisco, CA
    6 days ago
  •  ...Doppel is building the future of social engineering defense. Our AI-native platform uses...  ...brands detect and dismantle attacker infrastructure while strengthening employee resilience...  ...tracing), improving our ability to debug performance issues and maintain reliability at... 
    Performance
    Work at office
    Flexible hours

    Doppel

    San Francisco, CA
    6 days ago
  • Software Engineer Voxel's perception system is the technical core of everything we ship. Our...  ...strong software engineer to own the ML Infrastructure that powers how Voxel trains and...  ...for ML workloads. Strong Python. Write performant code that scales well in production environments... 
    Performance
    Work at office
    Flexible hours

    Voxel

    San Francisco, CA
    4 days ago
  • $168k - $213k

    Senior Infrastructure Engineer Bretton runs AI-native operations for the financial back office. OCC, Fed and FDIC regulated banks use Bretton...  ...cloud environments to meet the strictest regulatory and performance requirements, including SOC 2 compliance. Your work will be... 
    Performance
    Flexible hours

    Bretton

    San Francisco, CA
    5 days ago
  • Platform Engineer HUD is building infrastructure to create RL training data and evals for frontier AI agents...  ...who can own the reliability, scale, performance, and developer experience of HUD's...  ...code and apply strong software engineering judgment across product... 
    Performance
    Full time
    Work at office
    Remote work
    Relocation
    Visa sponsorship

    Hud (yc W25)

    San Francisco, CA
    4 days ago
  •  ...will own and architect core infrastructure systems that power our...  ...the opportunity to shape the engineering organization and lead major...  ...researchers. Overview As a Senior Software Engineer - Infrastructure...  ...of system scalability, performance, and reliability Experience... 
    Performance
    Local area

    AfterQuery

    San Francisco, CA
    4 days ago
  •  ...computing and make it accessible to software developers of all skill levels....  ...is looking for a Software Engineer to join the Platform and Infrastructure team. Anyscale aims to provide the...  ...the data plane, which ensures high-performance execution of distributed workloads... 
    Performance
    Flexible hours

    Anyscale, Inc

    San Francisco, CA
    4 days ago
  • $160k - $250k

     ...on placing the best product managers, software, and hardware talent at innovative...  ...States to help them hire. Software Engineer, Infrastructure Location: San Francisco, CA Company...  ...stability, security, scalability, and performance. Improve the reliability and efficiency... 
    Performance
    Full time
    For contractors
    H1b
    Work at office
    Remote work
    Visa sponsorship
    3 days per week

    Recruiting from Scratch

    San Francisco, CA
    2 days ago
  • $115k - $158k

     ...assembling a diverse, world-class team—engineers, designers, researchers, and...  ...About The Role As the Software Engineer, Tooling and Development Infrastructure, you will play a critical role...  ...insights into system behavior and performance. Experience with consumer... 
    Performance
    Full time
    Temporary work
    Local area
    Flexible hours

    HP IQ

    San Francisco, CA
    2 days ago
  • $137.1k - $299.3k

     ...Platform and builds the shared infrastructure that helps DoorDash, Wolt,...  ...model serving and inference engines, fine-tuning and training...  ...systems and pushing the cost/performance frontier of GPU inference and...  ...years of industry experience in software engineering ~ Deep... 
    Performance
    Hourly pay
    Full time
    Work at office
    Local area
    Remote work
    Flexible hours

    DoorDash USA

    San Francisco, CA
    3 days ago
  •  ...with Airflow; and support ML feature engineering tooling such as Chronon. Our mission is...  ...analytics. We’re not just scaling infrastructure – we’re redefining how people interact...  ...and storage systems, designed for high performance and scalability. You’ll help design, build... 
    Performance
    Work at office
    Relocation package

    OpenAI

    San Francisco, CA
    5 days ago
  • $162k - $216k

     ...and ship category-defining software that enables Ryder and its 5...  ...and creating impact for the engine of the American economy, you...  ...Role: Software Engineer - Infrastructure Department: Data Platform...  ...familiarity with SQL: Query Performance Analysis, Optimization, and... 
    Performance
    Full time
    Work at office
    Immediate start
    Remote work
    Monday to Friday

    Baton (A Ryder Technology Lab)

    San Francisco, CA
    28 days ago
  •  ...Job Description Job Description Senior Software Engineer – Backend Systems & Infrastructure Location: San Francisco - on-site Compensation: Competitive...  ..., and infrastructure teams Improve system performance, observability, reliability, and scalability Contribute... 
    Performance

    MRINetwork Jobs

    San Francisco, CA
    a month ago
  • $130k - $230k

     ...About You and The Role  Zipline is looking for software engineers to build the tools and infrastructure that power our systems validation and flight test organizations...  ...software and hardware is deployed globally to perform critical, real-world deliveries. What You'll Do... 
    Performance
    Local area

    Zipline

    South San Francisco, CA
    28 days ago
  •  ...role: As a member of the infrastructure team, you’ll play a key...  ...operability of our systems Build performance-critical, user-facing...  ...growth Partner with other engineering teams to ensure we build...  ...9+ years of experience in software engineering, and you’ve worked... 
    Performance
    Work at office
    Remote work
    3 days per week

    Orb

    San Francisco, CA
    2 days ago
  •  ...in the physical world. Critical infrastructure is constrained by labor shortages, hazardous...  ...Role At Watney, ML Infrastructure engineers turn data collected from a live fleet...  ...identifying and troubleshooting GPU performance bottlenecks in large-scale training environments... 
    Performance

    Watney

    San Francisco, CA
    a month ago
  •  ..., Ditto's peer-to-peer sync engine ensures devices stay connected...  ...stack and building high-performance solutions in next-generation...  ...perspectives, and talents. AS A SOFTWARE ENGINEER – NETWORKING, YOU...  ...systems, and building low-level infrastructure that operates at scale and... 
    Performance
    Full time
    Remote work

    Ditto

    San Francisco, CA
    12 days ago
  • $117k - $138k

     ...As the only vertically integrated AI infrastructure company built from the ground up, we...  ...AI strategies, and be part of a high-performing team that believes in each other, come...  ...Network team at Crusoe is looking for a Software Engineer I to join our team building and... 
    Performance
    Temporary work
    Internship

    Crusoe

    San Francisco, CA
    14 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer, Infrastructure & Performance. Be the first to apply!