Platform Engineer
Sunsets HQ Corp.
About Replay At its core, Replay was founded to help founders. We started by supporting startups through shutting down, but we have since expanded into unlocking a new revenue stream for all types of businesses. In 2025, we had a unique insight: the data every company generates each day through collaboration, communication, and building is some of the most valuable training data in the world. Public and synthetic data can only get frontier models so far, so the next generation of model progress depends on real, proprietary data grounded in how actual businesses operate. We are a primary source of it, partnering directly with the frontier AI labs building what comes next.
Why Join Replay Now
Problems You Might Own
Make several workload shapes feel like one coherent platform Create a small set of supported patterns for customer-facing services, connectors, scheduled jobs, data-processing pipelines, model-backed workloads, and evaluation runs. Define the contracts for environments, compute, state, networking, delivery, secrets, telemetry, failure handling, and recovery without forcing every workload into an inappropriate stack.
Turn delivery and operations into product-quality experiences Make it straightforward for an engineer to create an environment, ship a safe change, understand a failed deploy or job, get the right access, recover a system, and know who owns the result. Build useful self-service and escape hatches while making unsupported paths and exceptions explicit.
Make reliability visible from customer request to completed workload Connect service, queue, job, pipeline, and model telemetry to the outcome that matters. Establish practical objectives, alerts, incident mechanics, replay and recovery paths, and reviews that remove recurring failure classes instead of only documenting them.
Make infrastructure cost and control evidence part of normal operation Expose cost and capacity in workload-relevant units, then improve them without hiding reliability, security, quality, or developer time. Work with Security to implement least privilege, secrets, logging, backup, deployment, and audit controls whose evidence comes from the systems that actually enforce them.
What You'll Do
Why Join Replay Now
- We have scaled from $0 to a multi-eight-figure run rate in a matter of months
- We have raised from top-tier investors, including Floodgate, Afore, Ludlow, and Hustle Fund
- We are small enough that you will carry outsized responsibility and grow as quickly as the company does
- You will partner with and build for some of the fastest and most important companies in the world
- You will help build a massive, category-defining business from the ground floor
Problems You Might Own
Make several workload shapes feel like one coherent platform Create a small set of supported patterns for customer-facing services, connectors, scheduled jobs, data-processing pipelines, model-backed workloads, and evaluation runs. Define the contracts for environments, compute, state, networking, delivery, secrets, telemetry, failure handling, and recovery without forcing every workload into an inappropriate stack.
Turn delivery and operations into product-quality experiences Make it straightforward for an engineer to create an environment, ship a safe change, understand a failed deploy or job, get the right access, recover a system, and know who owns the result. Build useful self-service and escape hatches while making unsupported paths and exceptions explicit.
Make reliability visible from customer request to completed workload Connect service, queue, job, pipeline, and model telemetry to the outcome that matters. Establish practical objectives, alerts, incident mechanics, replay and recovery paths, and reviews that remove recurring failure classes instead of only documenting them.
Make infrastructure cost and control evidence part of normal operation Expose cost and capacity in workload-relevant units, then improve them without hiding reliability, security, quality, or developer time. Work with Security to implement least privilege, secrets, logging, backup, deployment, and audit controls whose evidence comes from the systems that actually enforce them.
What You'll Do
- Establish Replay's current platform, workload, reliability, ownership, toil, recovery, cost, and technical-control baseline
- Build reusable infrastructure-as-code modules, runtime templates, deployment workflows, environment contracts, and operational tooling
- Create supported paths for customer-facing services, asynchronous and batch jobs, data pipelines, and model-backed workloads
- Improve deploy safety, workload visibility, backup and recovery, incident response, replay, rollback, and durable remediation
- Work with engineering teams to define useful service and pipeline objectives, ownership, escalation, and recovery paths
- Build self-service for common infrastructure, environment, access, deploy, debugging, and recovery work without becoming a central approval queue
- Make cloud and vendor cost understandable by service and workload and improve efficiency within explicit reliability and security bounds
- Partner with Security on cloud identity, secrets, isolation, audit logging, vulnerability response, incident readiness, and automated control evidence
- Support employees and contractors through bounded access, safe environments, release controls, documentation, and timely removal of authority
- Use AI tools deeply for platform engineering and operations while verifying generated code, plans, queries, state changes, and incident conclusions
- Replay's environments, runtimes, deploy paths, service and pipeline owners, reliability risks, recovery gaps, manual work, and infrastructure costs are visible and prioritized
- One consequential failure or toil class is materially reduced in your first 90 days, and another team can use the resulting paved road without case-by-case help
- Product, data, and AI teams can ship and understand their systems faster while retaining clear operating ownership
- Priority services and pipelines have useful objectives, actionable telemetry, tested recovery paths, and incident learning that removes recurring failures
- Common platform work becomes self-service while exceptions remain explicit, owned, monitored, and time-bounded
- Cloud cost and capacity are understandable in workload-relevant units and improve without hidden reliability, security, or developer-time regressions
- Security and customer-trust evidence becomes easier to produce because it reflects current technical controls
- You have personally owned production cloud infrastructure and delivery or reliability systems across multiple services, including an asynchronous, batch-data, or model-backed workload
- You are a strong software engineer who is comfortable changing application, platform, and infrastructure code and operating the result in production
- You can reason from user impact through dependencies, state, telemetry, incident response, recovery, and durable remediation
- You have built paved roads other engineers adopted because they made real work easier, not because a platform team required them
- You understand both long-running services and high-volume or scheduled workloads and know where their reliability models should differ
- You can make pragmatic tradeoffs among delivery speed, least privilege, isolation, recovery, developer experience, and unit cost
- You are effective in an early-stage environment where the first step is often to establish ownership and a trustworthy baseline
- You can lead calmly through ambiguous incidents, communicate clearly, and leave the system and operating model stronger afterward
- You use modern AI engineering tools fluently and verify generated infrastructure, queries, code, and operational conclusions before they affect production
- You want a deployment or cloud-administration role where product teams hand systems to you to operate permanently
- You prefer designing a platform in isolation to learning how engineers, services, pipelines, and customer deliveries actually work
- You measure platform success by migration, ticket, dashboard, or uptime counts without connecting them to adoption, reliability, recovery, and user impact
- You want to standardize every workload on one stack regardless of its state, scale, failure, or recovery requirements
- You do not want AI tools to be part of your daily engineering and operational workflow
- Experience as an early platform or SRE hire at a fast-growing company
- Experience with AWS, Terraform, container runtimes, workflow orchestration, and observability systems
- Experience with high-volume data processing, model serving, evaluation jobs, GPU workloads, or machine-learning platforms
- Experience improving developer environments, preview systems, CI/CD, progressive delivery, or internal developer platforms
- Experience with replayable pipelines, backup and restore, disaster recovery, capacity planning, or cloud-cost allocation
- Experience implementing technical controls and automated evidence for SOC 2 or enterprise customer requirements
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Platform Engineer in New York, NY vacancy
$176k - $191k
...the menu at Wonder. Except compromise. Wonder is the mealtime platform built to feed every craving in one order. With Wonder, you can... ...is a new team chartered to own developer experience for Wonder Engineering — 400+ engineers across multiple converging organizations —...SuggestedFull timeTemporary workWork at officeFlexible hours3 days per week- ...Experience ProfessionalsContact: Ashley RezinJob ID: REQ8583Platform Engineering owns the end-to-end experience of how engineers, investment... ...systems that underpin our live trading environments, to the platforms that keep tens of thousands of builds, applications and tasks...Suggested
$165k - $242k
...Cloud for AI. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to... ...Learn more at .What You’ll Do:We are seeking a Senior Platform Engineer to join our Kubernetes Infrastructure team. This role involves...SuggestedPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$150k - $200k
...investor access to every asset, in every market, through a unified platform built for speed, transparency and scale.We give our clients the... ...tomorrow.For more information, visit .The RoleAs a Platform Engineer, your customers are our engineers.You'll build the internal...SuggestedWork at officeLocal areaImmediate start- ...digital partner that combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions that deliver... ...Practical action. Endless possibilities.We’re hiring AgentCore Platform Engineers to build the agent-facing layer of the platform. That...SuggestedContract workTemporary work
$120k - $170k
...Rockstar Games is on the lookout for a passionate and talented Engineer to help enable safe, scalable, and effective adoption of cloud... ...sits within IT and acts as the technical authority for cloud platforms, bridging emerging vendor capabilities with real-world production...Full timeWork at office$141k
...Prepared and Carbyne under Axon. Together, we are creating the only platform that combines modern 911 infrastructure with an AI intelligence... ....Position OverviewWe are looking for an experienced platform engineer to join Prepared's platform engineering team under Axon 911. As...Work experience placementWork at office- ...Experience ProfessionalsContact: Ashley RezinJob ID: REQ8582Platform Engineering owns the end-to-end experience of how engineers, investment... ...systems that underpin our live trading environments, to the platforms that keep tens of thousands of builds, applications and tasks...
$83.3k - $149.88k
We’re hiring a Google Cloud Platform Engineer to design, deploy, and operate scalable, secure cloud infrastructure and platform services. You’ll partner with product and engineering teams to build automated, observable, and cost-efficient solutions on GCP—moving workloads...Work at officeLocal area$153k - $207k
...life-changing mission to develop education for our half a billion (and growing!) learners around the world.About the role...As a Platform Engineer on the Compute team, you’ll improve how engineers build and run services, helping ensure our compute platform is reliable at...Work experience placementInternship$175k - $350k
Role Overview Citadel Securities is seeking an exceptional Senior Platform Infrastructure Engineer to join one of our Platform Engineering teams in Miami or New York. Our team is responsible for building and maintaining the foundational platform that orchestrates application...Worldwide$131.75k - $170.5k
...model. While the internal title for this position is Senior Linux Engineer, this role has been posted externally under a different title... ...and skill sets.Role Overview We are seeking a Senior Platform Engineer to join our Systems Platform Engineer team. In this role...Full timeWork at officeImmediate start$113.9k - $189.9k
Summary of Business Unit/Function:The One Policy Engine is a unified policy platform for multi-cloud environments that keeps Infrastructure‑as‑Code (IaC) compliant from code to cloud. We apply policy enforcement early and consistently across the SDLC so security and compliance...Full timePart timeInternship$175k - $250k
Senior Platform EngineerAbout MillenniumMillennium is a global, diversified alternative investment firm, founded in 1989. Defined by... ...TeamMillennium’s Infrastructure organization is dedicated to designing, engineering, supporting, and managing a robust server estate, systems...- ...ChicagoDepartment: TechnologyExperience Level: Experience ProfessionalsContact: Ashley RezinJob ID: REQ8338We are seeking a Trading Platform Engineer to join our Systematic Technology team, focused on the reliability, observability, and day-to-day support of a high-...
- We are looking for a Databricks Platform Manager/Engineer responsible for designing, administering, securing, and optimizing the enterprise Databricks platform on AWS. The ideal candidate will own the platform lifecycle, ensuring scalability, security, governance, cost...
$84k - $126k
...deliver purposeful work and meaningful impact every day. Learn more about what makes us different and how you can thrive as a Platform Engineer at MMA. Marsh McLennan Agency (MMA) provides business insurance, employee health & benefits, retirement, and private client...Minimum wageLocal areaRemote workNight shift- Seeking a Senior Platform DevOps Engineer with Site Reliability Engineering (SRE) experience, the full-time remote position will design and maintain cloud infrastructure, automate deployment processes, and implement observability solutions to enhance system reliability...Full timeRemote work
- Who we areAbout StripeStripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises... ...while protecting user data.What you’ll doAs a software engineer on Secure Devices, you will work at the intersection of software...
$100.2k - $167k
SummaryWe are looking for a Senior Associate Software & Platform Engineer to join our FX Engineering group as an individual contributor. This role is suited for an engineer who has strong foundational experience in Java application development, Spring Boot, build systems...Full timePart timeInternship$189k - $236k
...requirements vary by role and will be assessed during the interview process.About the Role:We’re hiring seasoned engineers to join our teams that work on core platform capabilities, improving our existing systems for extensibility and scalability, and building the future of...Full timeWork at officeLocal areaRemote work2 days per week3 days per week$131k - $164k
Help shape the technology that enables a global organisation to do its best work. As Senior Manager, Platform Engineering, you’ll lead the team responsible for Diligent’s Atlassian and Microsoft platforms while setting the architectural direction for the wider internal...Work at officeLocal areaWorldwideVisa sponsorshipFlexible hours$187k - $240k
As a Platform Security Engineer you will partner with different stakeholders across the organization to secure our infrastructure and application components of the Datadog platform. As part of the Platform Security organization we secure the building blocks of Datadog’...Work at office$160k - $185k
...Senior Azure DevOps & Kubernetes Platform Engineer Company Overview: Business Integration Partners (BIP) is Europe’s fastest growing digital consulting company and are on track to reach the Top 20 by 2030, with an expanding global footprint in the US (New York, Charlotte...Temporary workRemote workWorldwide- BNY is seeking a Senior Vice President to lead the Wealth Services Platform Cloud Nerve Center. This leader will establish the production environment, engineering patterns, and operational capabilities needed to migrate WSP applications to the cloud securely, efficiently...WorldwideFlexible hours
$168k - $200k
...grow our team to meet the needs of more companies, teams, and innovators in this way. The Role: As a Senior Software Engineer on the Broker Platform team, you will play a key role in delivering high-quality experiences to our Brokers, working with modern, low-latency...Work experience placementWork at officeLocal area2 days per week3 days per week$129.7k - $216.1k
SummaryWe are looking for a hands-on Lead Software and Platform Engineer to drive the design, development, delivery, and operational reliability of large-scale, low-latency microservices-based trading platforms within our FX Engineering group.This role combines strong Java...Full timePart timeInternship- ...Experience ProfessionalsContact: Ashley RezinJob ID: REQ8278About UsOur Database Engineering team is transforming how Balyasny Asset Management delivers database services through a standardized platform built on declarative, code-defined provisioning, automated lifecycle...
$210k - $270k
...Zocdoc’s marketplace topower access to care wherever patients search, from provider websites and insurance directories to search engines, AI platforms, and more.Healthcare still lacks something every other major consumer industry takes for granted: a seamless way to go from...Remote workFlexible hours£45k - £65k per year
...i2 Group , a Harris Computer company , is seeking a Platform Software Engineer on a full-time, permanent, remote basis with a monthly requirement of attending the i2 Group offices in Cambridge. The Role The Platform team solves one of the harder problems in...Permanent employmentFull timeImmediate startRemote workShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Platform Engineer. Be the first to apply!
Related searches
- client platform engineer New York, NY
- data platform engineer New York, NY
- senior platform engineer New York, NY
- platform engineering manager New York, NY
- platform developer New York, NY
- platform engineer New York, NY
- power platform New York, NY
- platform manager New York, NY
- digital platform specialist New York, NY
- platform product manager New York, NY



