Distributed Systems Engineer
Fluidstack
Fluidstack
We exist to make humanity more free. For most of human history, you farmed or you starved. Technology gave people more time for the things they wanted to do, instead of things they had to do. Powerful AI will be the biggest lever for human choice we've ever built - but only if models are aligned with what humanity actually wants. There are groups building AI who don't share these goals. Whoever deploys frontier compute infrastructure fastest will decide whether AI expands human freedom or shrinks it.
We're singularly focused on delivering 10 to 100s of GWs of compute faster than anyone else, rethinking every layer of the stack. We acquire power, design and build data centers, and operate them - with teams spanning hardware and software. Speed and scale are our key differentiators. Come be a part of building civilization-scale infrastructure for AI.
We hire people who care deeply about this problem space. If that is you, please apply!
How We Operate
Be a barrel. Full autonomy. Own things end to end, take on scope without being asked, no permission required to operate outside your core role.
Insane urgency. We drive everything forward as fast as possible.
Reason from first principles. Challenge every assumption. Zero analogy thinking, no egos, the best idea wins.
Love of the game. The frontier of AI is the most interesting problem of our time. We put in long hours at high intensity to push the frontier forward.
Build something that actually matters. If you're going to spend your time, spend it on something that matters to the world.
The Production Engineering Team
Examples of key exciting problems the team is working on
Make tens of thousands of GPUs legible in real time: build the observability platform that turns raw telemetry into signal, from site-level health down to individual device and link. At this scale, you cannot operate what you cannot see.
Build the control plane every team at Fluidstack depends on: replace one-off tooling with a stable, versioned API surface that covers unified machine management, actual state inspection, and distributed command execution. One interface for the whole company, not a hundred scripts.
Make the system's view of itself always match reality: integrate fleet state as a machine-readable source of truth across provisioning, operations, and customer-facing platforms, so every new site and GPU generation lands cleanly from day zero.
Role Scope
Own the observability platform. Build and operate the data pipelines, decoration and correlation engine, and healthcheck framework that make the fleet legible — from site down to device and link. No other team should need to scrape production directly to answer a question.
Define and build the API surface for infrastructure. Design the contracts between production infrastructure and every tool that touches it. All other teams at Fluidstack use your tooling to manage and operate our hyperscale fleet.
Build the production control plane. Unified machine management, actual state inspection, distributed command execution — and the Kubernetes-based infrastructure that underpins it all.
Own fleet state as source of truth. SLOs, site lifecycle state, and integration with internal infrastructure management and customer-facing operations platforms. What the system says about itself should match reality, and you're accountable when it doesn't.
Land new hardware into the platform cleanly. ZTP, DHCP, DNS, artifacts — every new XPU generation and site integration goes through IaaS before production.
What We're Looking For
The below is a starting point. We always make space for exceptional people, so if you don't fit this role exactly, tell us where you would.
You treat toil as a bug. If something requires a human to do it twice, you build the thing that makes it not require a human.
You design APIs that age well. You've felt the pain of a leaky abstraction at scale and you don't repeat it.
You move toward ambiguity, not away from it. You walk into the fog, build the map, and explain it to everyone else.
You learn at a steep slope. You reach real competence in an unfamiliar domain fast. We value this over existing expertise.
You carry a pager without flinching. You run the incident, write the postmortem, fix the systemic cause, and move on.
You're fluent with AI tooling. LLM APIs, MCP servers, and agentic frameworks, and you drive Claude Code, Cursor, or similar every day.
You've shipped production services that other teams depend on at scale, and you're comfortable in any language using AI coding tools.
Bonus: Distributed systems and data pipeline engineering. Time-series observability stacks (Prometheus, Thanos, VictoriaMetrics). API design and versioning at scale. Workflow and orchestration engines (Temporal, Cadence). BMC/Redfish or hardware telemetry. Go, Python, and Postgres.
Salary & Benefits
Competitive total compensation package (salary + equity).
Retirement or pension plan, in line with local norms.
Health, dental, and vision insurance.
Generous PTO policy, in line with local norms.
We are committed to pay equity and transparency.
Fluidstack is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans' status, or any other characteristic protected by law. Fluidstack will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
You will receive a confirmation email once your application has successfully been accepted. If there is an error with your submission and you did not receive a confirmation email, please email View email address on click.appcast.io with your resume/CV, the role you've applied for, and the date you submitted your application-- someone from our recruiting team will be in touch.
- ...that matters to the world.The Production Engineering TeamExamples of key exciting problems... ...management, actual state inspection, and distributed command execution. One interface for... ...company, not a hundred scripts.Make the system's view of itself always match reality:...SuggestedLocal area
- ...fast, self-hosted CI infrastructure and a scalable worker platform in NYC. We’re seeking engineers with 5+ years in software development and deep experience with distributed systems and performance tuning. You’ll work closely on low-level systems in C, C++, Go, Rust, or...Suggested
$388k
...engaged audiences.Our TeamThe Ads Platform Engineering team is the heartbeat of our... ...advertising technology. We don't just design systems; we ship production code that handles massive... ...You’re comfortable navigating complex distributed systems, identifying bottlenecks, and...SuggestedHourly payFull timeImmediate startFlexible hours$153k - $376k
...heart of everything we build. As a Software Engineer on our Infrastructure team, you’ll help design, build, and operate the systems that power our real-time collaborative... ...scaling fast, and we’re looking for experienced distributed systems engineers across a variety of...SuggestedFull timeRemote workWork from homeWorldwide$192k - $240k
Distributed Systems engineers at Datadog design, implement and run in production the foundational platforms powering our applications. Your data pipelines will ingest, store, analyze and query in real-time billions of events per second from companies all over the globe....SuggestedWork at office$200k - $300k
...you can achieve.ProfileYou’re able to talk both high-level distributed systems design trade-offs and low-level OS-level detailsYou’re able... ...of expertise: mathematics and computer science, physics and engineering, media and tech. We’re a community of self-starters who are...Work at officeLocal areaImmediate start- ...understanding. We are looking for a Senior Software Engineer with strong modern C++ expertise to help design and evolve the distributed engine powering a global blockchain ledger.... ...layer, working on high-impact distributed systems that secure consensus, transaction...Full timeImmediate startRelocation
$227.2k - $324.5k
About the Role:As a Staff Software Engineer on the ML Infrastructure team, you will collaborate... ...low-latency ML model serving systems that support Deep Learning, LLM, and Search... ...scalable, high throughput, and low latency distributed systems using ScalaBuild reusable...Full timeTemporary workLocal areaFlexible hours$227k - $303k
...more at About the role We are looking for a Principal Engineer to provide technical leadership across Security Products.... ...architecture, guide execution across multiple teams, and solve complex distributed-systems problems in security-critical infrastructure. The...Permanent employmentFull timeTemporary workCasual workWork at officeRemote workFlexible hours$180k - $320k
...Software Engineer, Distributed Systems (Core)Title of Role: Software Engineer, Distributed Systems (Core)Location: New York, remoteCompany Stage of Funding: Series C — Software DevelopmentOffice Type: RemoteSalary: $180K–$320KCompany DescriptionWe're representing a dynamic...Remote work$150k - $300k
Hudson River Trading (HRT) is looking for Systems Engineers to join our growing Research & Development team. This team builds and maintains exceptionally large and growing distributed compute clusters, multi petabyte-scale storage layers, operating systems, automation software...Full timeWork at officeLocal areaImmediate startRemote workWorldwide$135.4k - $181.6k
Lead Content Distribution Network EngineerDisney Entertainment and ESPN Product & TechnologyTechnology is at the heart of Disney’s past... ...Entertainment and ESPN Product & Technology is a global organization of engineers, product developers, designers, technologists, data...Remote work$150k - $300k
Hudson River Trading (HRT) is looking for GPU Systems Engineers to help scale and evolve our exceptionally sophisticated HPC/AI research environment... ...-scale storage and massive CPU and GPU clusters in globally distributed data centers. As such, this is a high-impact role with broad...Work at officeLocal areaImmediate start- Systems Engineer, Servers/Cloud/DataLocation: Must be able to work onsite Tue, Wed, & Thursday out of one of these office locations: New... ...orchestration, and high-performance computing (HPC) solutions across distributed systems, collaborate with development and operations teams,...Work at office
$59.15k - $106.93k
...you grow your career is good business. At Leidos is seeking Distribution Engineers & Designers in the Eastern, PA region who are passionate... ...for OH, UG, URD, and Make Ready, using customer GIS and WMS systems such as EFD, AUD, Smallworld, ArcGIS, Infor, EAM, STORMS, and...Work at officeLocal areaImmediate startRemote workFlexible hours$285k - $325k
...IT Systems Engineer, Mobile Client Platform EngineerBoston, MA; Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New... ...the infrastructure underneath them. We treat the fleet as a distributed platform, not a collection of devices, and every piece of configuration...Work at officeRemote workVisa sponsorshipFlexible hours$88k - $120k
...Our Systems Engineering team owns the infrastructure that keeps a ~2,500-person global firm running around the clock: cloud (primarily AWS... ...~ Clear communicator who collaborates well across a distributed team. ~ Based on a 9-5 ET schedule, but realistic...Full timeLocal areaFlexible hoursShift work- ...Distributed SpectrumDS creates systems that power the next generation of radio spectrum intelligence. We collect radio data from all over the world... ...specialistsBS or higher or equivalent experience Electrical Engineering or Physics or similarNice to HaveExperience working...Temporary workWork at office
$130k - $250k
At Goldman Sachs, our Engineers don't just make things - we make things possible. Change the... ...build massively scalable software and systems, architect low latency infrastructure solutions... ...configurations and rules within a distributed, high-availability environment.Basic...Full timeTemporary workPart timeImmediate start$200k
...validate innovative technology solutions for the best business outcomes and then deploys them at scale through its global warehousing, distribution and integration capabilities. With over 10,000 employees and more than 55 locations around the world, WWT's culture, built on a...Base plus commissionFull time$130k - $225k
Senior Systems Engineer - Windows/SQL Services Location New York Business Area Engineering and CTO Ref # 10052871... ...ll help engineer, operate, and improve Bloomberg’s globally distributed Windows hosting and SQL Server database services. You will bring...Temporary workFor contractorsWork experience placementWorldwide$150k - $250k
...Senior Systems Engineer Tower Research Capital is a leading quantitative trading firm founded in 1998. Tower has built its business on... ...Lead on-prem to cloud and cloud-to-cloud migrations of HPC and distributed workloads Collaborate with research/engineering teams to...Casual workWork at office- ...Description VAST Data is looking for a Senior Systems Engineer to join our growing team! This is a great opportunity to be part... ...leadership and subject matter expertise on storage products, distributed storage architectures, file systems, and competitive storage...Traineeship
- ...Engineering ManagerWe're looking for an Engineering Manager to lead a team of highly experienced... ...as your team tackles hard problems in distributed computing, large-scale data handling,... ...building high-performance distributed systems at scaleStrong background in cloud...
$180k - $240k
Senior Systems EngineerWe are looking for a Senior Systems Developer to lead the design,... ...development, and operation of large-scale distributed systems powering high-performance... ...production and ongoing operationMentor and guide engineers, raising the technical bar across the...Flexible hours$200k - $300k
...of the world's leading trading firms on an interesting GPU Systems Engineer position. This is a hands-on infrastructure role working at... ...infrastructure, GPUDirect RDMA, CUDA, C/C++, Python, networking and distributed systems, with the opportunity to get involved from hardware...$160k - $220k
...dedicated cloud for CI workloads. Role Snapshot YOE: 3 - 10 years of experience as a Senior/Staff level Systems Engineer who has developed and scaled distributed systems Compensation: $160K - $220K About The Company They enable companies to run GitHub Actions significantly...$200k - $300k
...the world’s best systematic trading and engineering talent. We empower portfolio managers... ...responsible for the compute, storage, operating systems, and automation behind that work at... ...Design, deploy, and scale distributed GPU clusters, from hardware selection and...Work experience placementCasual workWork at office$99.2k - $143.9k
...through integrity. Skills and Competencies 5+ years in an engineering role, with a focus on AV systems and infrastructure Proven experience in AV design... ...Visual and Voice Engineering team is a globally distributed group responsible for overseeing both the Voice and Audio...Full time$99.2k - $143.9k
...integrity. Skills and Competencies 5+ years in an engineering role, with a focus on AV systems and infrastructure Proven experience in AV design engineering... ...Visual and Voice Engineering team is a globally distributed group responsible for overseeing both the Voice and...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Distributed Systems Engineer. Be the first to apply!
- system support engineer New York, NY
- healthcare systems engineer New York, NY
- system engineer contract New York, NY
- broadcast systems engineer New York, NY
- senior linux systems engineer New York, NY
- systems engineer New York, NY
- computer system validation engineer New York, NY
- distributed systems engineer New York, NY
- space systems engineer New York, NY
- visual systems engineer New York, NY


