Platform Engineer
Replay
About Replay
At its core, Replay was founded to help founders. We started by supporting startups through shutting down, but we have since expanded into unlocking a new revenue stream for all types of businesses.
In 2025, we had a unique insight: the data every company generates each day through collaboration, communication, and building is some of the most valuable training data in the world. Public and synthetic data can only get frontier models so far, so the next generation of model progress depends on real, proprietary data grounded in how actual businesses operate. We are a primary source of it, partnering directly with the frontier AI labs building what comes next.
Why Join Replay Now
We have scaled from $0 to a multi-eight-figure run rate in a matter of months
We have raised from top-tier investors, including Floodgate, Afore, Ludlow, and Hustle Fund
We are small enough that you will carry outsized responsibility and grow as quickly as the company does
You will partner with and build for some of the fastest and most important companies in the world
You will help build a massive, category-defining business from the ground floor
The Role
Replay operates customer-facing SaaS products, connector and ingestion services, asynchronous workers, high-volume data pipelines, model-backed systems, review tools, and customer-delivery paths. These workloads have different shapes, but they need a coherent foundation for infrastructure, delivery, observability, recovery, access, and cost.
You will build and operate the shared platform that lets our product, data, and AI teams ship reliable, secure, observable, and cost-aware systems without manual infrastructure work or operational risk growing linearly. You will write software and infrastructure, improve real engineering workflows, lead through incidents, and create paved roads teams can use without waiting on you.
This is not a deployment-operator or internal-IT role. Product, data, and ML teams remain responsible for the systems they build. You will give them the runtime, delivery, visibility, recovery, and operating patterns to own those systems well. You will partner closely with our Security Lead, but you will not be expected to run the entire security or compliance program.
Problems You Might Own
Make several workload shapes feel like one coherent platform
Create a small set of supported patterns for customer-facing services, connectors, scheduled jobs, data-processing pipelines, model-backed workloads, and evaluation runs. Define the contracts for environments, compute, state, networking, delivery, secrets, telemetry, failure handling, and recovery without forcing every workload into an inappropriate stack.
Turn delivery and operations into product-quality experiences
Make it straightforward for an engineer to create an environment, ship a safe change, understand a failed deploy or job, get the right access, recover a system, and know who owns the result. Build useful self-service and escape hatches while making unsupported paths and exceptions explicit.
Make reliability visible from customer request to completed workload
Connect service, queue, job, pipeline, and model telemetry to the outcome that matters. Establish practical objectives, alerts, incident mechanics, replay and recovery paths, and reviews that remove recurring failure classes instead of only documenting them.
Make infrastructure cost and control evidence part of normal operation
Expose cost and capacity in workload-relevant units, then improve them without hiding reliability, security, quality, or developer time. Work with Security to implement least privilege, secrets, logging, backup, deployment, and audit controls whose evidence comes from the systems that actually enforce them.
What You'll Do
Establish Replay's current platform, workload, reliability, ownership, toil, recovery, cost, and technical-control baseline
Build reusable infrastructure-as-code modules, runtime templates, deployment workflows, environment contracts, and operational tooling
Create supported paths for customer-facing services, asynchronous and batch jobs, data pipelines, and model-backed workloads
Improve deploy safety, workload visibility, backup and recovery, incident response, replay, rollback, and durable remediation
Work with engineering teams to define useful service and pipeline objectives, ownership, escalation, and recovery paths
Build self-service for common infrastructure, environment, access, deploy, debugging, and recovery work without becoming a central approval queue
Make cloud and vendor cost understandable by service and workload and improve efficiency within explicit reliability and security bounds
Partner with Security on cloud identity, secrets, isolation, audit logging, vulnerability response, incident readiness, and automated control evidence
Support employees and contractors through bounded access, safe environments, release controls, documentation, and timely removal of authority
Use AI tools deeply for platform engineering and operations while verifying generated code, plans, queries, state changes, and incident conclusions
What Success Looks Like
Replay's environments, runtimes, deploy paths, service and pipeline owners, reliability risks, recovery gaps, manual work, and infrastructure costs are visible and prioritized
One consequential failure or toil class is materially reduced in your first 90 days, and another team can use the resulting paved road without case-by-case help
Product, data, and AI teams can ship and understand their systems faster while retaining clear operating ownership
Priority services and pipelines have useful objectives, actionable telemetry, tested recovery paths, and incident learning that removes recurring failures
Common platform work becomes self-service while exceptions remain explicit, owned, monitored, and time-bounded
Cloud cost and capacity are understandable in workload-relevant units and improve without hidden reliability, security, and developer-time regressions
Security and customer-trust evidence becomes easier to produce because it reflects current technical controls
You Might Thrive Here If
You have personally owned production cloud infrastructure and delivery or reliability systems across multiple services, including an asynchronous, batch-data, or model-backed workload
You are a strong software engineer who is comfortable changing application, platform, and infrastructure code and operating the result in production
You can reason from user impact through dependencies, state, telemetry, incident response, recovery, and durable remediation
You have built paved roads other engineers adopted because they made real work easier, not because a platform team required them
You understand both long-running services and high-volume or scheduled workloads and know where their reliability models should differ
You can make pragmatic tradeoffs among delivery speed, least privilege, isolation, recovery, developer experience, and unit cost
You are effective in an early-stage environment where the first step is often to establish ownership and a trustworthy baseline
You can lead calmly through ambiguous incidents, communicate clearly, and leave the system and operating model stronger afterward
You use modern AI engineering tools fluently and verify generated infrastructure, queries, code, and operational conclusions before they affect production
This Role May Not Be for You If
You want a deployment or cloud-administration role where product teams hand systems to you to operate permanently
You prefer designing a platform in isolation to learning how engineers, services, pipelines, and customer deliveries actually work
You measure platform success by migration, ticket, dashboard, or uptime counts without connecting them to adoption, reliability, recovery, and user impact
You want to standardize every workload on one stack regardless of its state, scale, failure, or recovery requirements
You do not want AI tools to be part of your daily engineering and operational workflow
Bonus
Experience as an early platform or SRE hire at a fast-growing company
Experience with AWS, Terraform, container runtimes, workflow orchestration, and observability systems
Experience with high-volume data processing, model serving, evaluation jobs, GPU workloads, or machine-learning platforms
Experience improving developer environments, preview systems, CI/CD, progressive delivery, or internal developer platforms
Experience with replayable pipelines, backup and restore, disaster recovery, capacity planning, or cloud-cost allocation
Experience implementing technical controls and automated evidence for SOC 2 or enterprise customer requirements
$106.9k - $147k
Become a part of our caring communityHumana is looking for a Senior DevOps Platform Engineer to design, implement, and support secure, compliant cloud and DevOps platforms across Azure and GCP. You will ensure enterprise-scale software delivery through modern CI/CD pipelines...SuggestedFull timeTemporary workApprenticeshipWork at officeRemote workHome office- Seeking a Senior Platform DevOps Engineer with Site Reliability Engineering (SRE) experience, the full-time remote position will design and maintain cloud infrastructure, automate deployment processes, and implement observability solutions to enhance system reliability...SuggestedFull timeRemote work
$176k - $191k
...the menu at Wonder. Except compromise. Wonder is the mealtime platform built to feed every craving in one order. With Wonder, you can... ...is a new team chartered to own developer experience for Wonder Engineering — 400+ engineers across multiple converging organizations —...SuggestedFull timeTemporary workWork at officeFlexible hours3 days per week- ...digital partner that combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions that deliver... ...Practical action. Endless possibilities.We’re hiring AWS Cloud Platform Engineers to build and own the platform’s foundation. That means...Suggested
$131.75k - $170.5k
...model. While the internal title for this position is Senior Linux Engineer, this role has been posted externally under a different title... ...and skill sets.Role Overview We are seeking a Senior Platform Engineer to join our Systems Platform Engineer team. In this role...SuggestedFull timeWork at officeImmediate start$176k - $191k
...practical experience. We need 5+ years of professional software engineering experience, including recent hands-on work with LLM-based... ...system design capability and a track record of owning tools or platforms end to end, from design through adoption. We need comfort...Full timeWork at office3 days per week- ...Software Engineer (Platform Engineer III Build & CI Infrastructure) Location : New York, NY (Hybrid) Duration : 6 month Interview: Virtual round Note: LinkedIn with location Local only with recent local project Ideal Candidate Profile...Local area
- ...Senior Veritas eDiscovery Platform (eDP) EngineerEmployment Type: Full-Time, Executive-LevelCGS is seeking a dedicated Senior Veritas eDiscovery Platform (eDP) Engineer to join a fast-paced and hard-working team to assist with any legal accounts. As a Veritas eDiscovery...Full timeFor contractorsRemote work
$87.5k - $145k
...physical infrastructure and delivery layer for a new commercial data platform serving institutional clients. This means cloud platform... ...technical consent enforcement Collaboration: Work with the Data Engineer (who builds pipelines that run on your platform), the Data...Local areaFlexible hours- ...Platform Engineer Employment: Contract-to-Hire (C2H) Location: Plano/Frisco, TX preferred (Draper, UT | Columbus, OH | Chadds Ford, PA | Wilmington, DE | NYC/Manhattan.) Candidates near other client locations may also be considered Role Overview We are looking...Contract work
$190k - $210k
...Reports to: Lead Platform Engineer Location: Remote in the US Compensation: $190,000 - $210,000 About ReflexAI At ReflexAI, we are building the operating system for conversation performance. We believe that people will always be at the center of high-stakes...Contract workLocal areaRemote work$91k - $118k
...Your Team Responsibilities As a Kubernetes/Linux Engineer and subject matter expert for MSCI’s Infrastructure Engineering group,... ...challenges · Keep pace with emerging tools, techniques, and cloud platforms Your skills and experience that will help you excel · BSc...Full timeWork at officeLocal areaFlexible hours- ...Senior Platform Engineer Nomic is the domain-specific AI platform for the Architecture, Engineering, and Construction (AEC) industry. We help enterprise teams extract structured knowledge from decades of drawings, specs, and project files — combining embedding models...Remote workFlexible hours
- ...IoT Platform Engineer Experience: 10+ years Certification: CCIE Lab (Required) Primary Skill: Routing & Switching for IoT Key Responsibilities: Design and develop lab frameworks to connect IoT endpoints with campus and branch networks, while certifying hardware...
$180k - $230k
...broken, it's depressing if you don't laugh with us. About the role: We are looking for an experienced engineer to join as the third member of our Platform team. This unit operates across the stack, developing the resources and tooling that allow our engineers to...Work at officeRemote workFlexible hours- ...Platform Engineer Platform Engineers at Fractal are responsible for developing, maintaining, and optimizing the infrastructure that powers enterprise analytics platforms. In this role, you will manage dbt and Airflow instances, handle system upgrades, resolve infrastructure...
- ...Full TimeWorking Type HybridJob Reference 0000021462Salary Type AnnuallyIndustry Law PracticeSelling Points Lead innovative AI platform engineering projects at a forward-thinking organization. Collaborate on cutting-edge Azure AI/ML solutions and cloud infrastructure....
$90k - $105k
...,000 per year Requirements: Strong hands-on software development experience, with additional experience in SRE, DevOps, Platform Engineering, or Infrastructure Engineering. Strong Terraform skills, including writing, deploying, maintaining, and troubleshooting infrastructure...Hourly payFull timeContract workRemote work$147k - $170k
...- Design, build, and operate a secure and reliable internal platform for running business applications and machine learning workloads... ...technical guidance, reviews, and mentorship to peers and junior engineers. - Own your professional development, continuously learning...Temporary workWork experience placementShift work$193.4k - $290k
...operate. By combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise, we’re reshaping how critical... ...started. Role Overview We’re looking for experienced frontend engineers to help build and evolve Harvey’s frontend platform- the...$176k - $179.5k
...firm leaders, CST leaders, practice leaders and cutting-edge engineers. Your team is led by some of our firm’s most senior leaders and... ...customization and additional features required to deploy this platform for Defense and other Government clients. Your expertise in Python...ApprenticeshipEasy work- ...digital partner that combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions that deliver... ...together a dedicated delivery pod to build and run an internal agent platform on AWS Bedrock AgentCore for a global life sciences client....
$136k - $253k
...transaction data into structured, searchable market intelligence. Its platform helps legal and finance professionals analyze deal terms,... ...with greater speed and confidence. As a Senior Data Platform Engineer, you will design, build, and maintain the infrastructure that ingests...Full timeWork at officeLocal areaFlexible hours- ...Director, Platform Engineering The Director, Platform Engineering will own the strategy, roadmap, and execution of InvestCloud's shared platform capabilities across on-prem and AWS. This role provides both technical and people leadership across environments, infrastructure...Shift work
- ...Sequence Holdings is seeking a Product Engineer in New York, NY to own the Atlas platform and scale the AI-driven transformation across portfolio companies. You will build agent infrastructure, workflow orchestration, and data tooling, working closely with Forward Deployed...
- ...Adonis in New York is seeking a Senior Infrastructure Engineer to own and scale our cloud platform, developer tooling, and data infrastructure. This hands-on IC role sets reliability, security, and performance standards across healthcare products serving enterprise physician...
- ...Filevine is a Legal AI company delivering LOIS-powered operating intelligence. We are seeking a Senior CI/CD Platform Engineer to own end-to-end pipelines across .NET, Python, and Svelte, transforming delivery speed and reliability. You will build internal tooling, establish...Work at officeRemote work
- ...digital partner that combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions that deliver... ...Practical action. Endless possibilities. We’re hiring AWS Cloud Platform Engineers to build and own the platform’s foundation. That...
- ...Senior Database Platform Engineer Our Database Engineering team is transforming how Balyasny Asset Management delivers database services through a standardized platform built on declarative, code-defined provisioning, automated lifecycle management, and self-healing...
- ...We hold a bachelors or masters degree in computer science, engineering, or a related discipline, or have equivalent practical experience... ...databases, highly available distributed non-relational storage platforms, scalable distributed data and caching platforms, and...Full timeApprenticeshipWork at officeWork from homeWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Platform Engineer. Be the first to apply!
- client platform engineer New York, NY
- platform engineering manager New York, NY
- data platform engineer New York, NY
- senior platform engineer New York, NY
- platform engineer New York, NY
- platform developer New York, NY
- power platform New York, NY
- platform product manager New York, NY
- digital platform specialist New York, NY
- director of digital platform New York, NY




