Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Platform Engineer

Sunsets HQ Corp.

About Replay

At its core, Replay was founded to help founders. We started by supporting startups through shutting down, but we have since expanded into unlocking a new revenue stream for all types of businesses.

In 2025, we had a unique insight: the data every company generates each day through collaboration, communication, and building is some of the most valuable training data in the world. Public and synthetic data can only get frontier models so far, so the next generation of model progress depends on real, proprietary data grounded in how actual businesses operate. We are a primary source of it, partnering directly with the frontier AI labs building what comes next.
Why Join Replay Now
  • We have scaled from $0 to a multi-eight-figure run rate in a matter of months
  • We have raised from top-tier investors, including Floodgate, Afore, Ludlow, and Hustle Fund
  • We are small enough that you will carry outsized responsibility and grow as quickly as the company does
  • You will partner with and build for some of the fastest and most important companies in the world
  • You will help build a massive, category-defining business from the ground floor
The Role

Replay operates customer-facing SaaS products, connector and ingestion services, asynchronous workers, high-volume data pipelines, model-backed systems, review tools, and customer-delivery paths. These workloads have different shapes, but they need a coherent foundation for infrastructure, delivery, observability, recovery, access, and cost.

You will build and operate the shared platform that lets our product, data, and AI teams ship reliable, secure, observable, and cost-aware systems without manual infrastructure work or operational risk growing linearly. You will write software and infrastructure, improve real engineering workflows, lead through incidents, and create paved roads teams can use without waiting on you.

This is not a deployment-operator or internal-IT role. Product, data, and ML teams remain responsible for the systems they build. You will give them the runtime, delivery, visibility, recovery, and operating patterns to own those systems well. You will partner closely with our Security Lead, but you will not be expected to run the entire security or compliance program.
Problems You Might Own
Make several workload shapes feel like one coherent platform

Create a small set of supported patterns for customer-facing services, connectors, scheduled jobs, data-processing pipelines, model-backed workloads, and evaluation runs. Define the contracts for environments, compute, state, networking, delivery, secrets, telemetry, failure handling, and recovery without forcing every workload into an inappropriate stack.
Turn delivery and operations into product-quality experiences

Make it straightforward for an engineer to create an environment, ship a safe change, understand a failed deploy or job, get the right access, recover a system, and know who owns the result. Build useful self-service and escape hatches while making unsupported paths and exceptions explicit.
Make reliability visible from customer request to completed workload

Connect service, queue, job, pipeline, and model telemetry to the outcome that matters. Establish practical objectives, alerts, incident mechanics, replay and recovery paths, and reviews that remove recurring failure classes instead of only documenting them.
Make infrastructure cost and control evidence part of normal operation

Expose cost and capacity in workload-relevant units, then improve them without hiding reliability, security, quality, or developer time. Work with Security to implement least privilege, secrets, logging, backup, deployment, and audit controls whose evidence comes from the systems that actually enforce them.
What You'll Do
  • Establish Replay's current platform, workload, reliability, ownership, toil, recovery, cost, and technical-control baseline
  • Build reusable infrastructure-as-code modules, runtime templates, deployment workflows, environment contracts, and operational tooling
  • Create supported paths for customer-facing services, asynchronous and batch jobs, data pipelines, and model-backed workloads
  • Improve deploy safety, workload visibility, backup and recovery, incident response, replay, rollback, and durable remediation
  • Work with engineering teams to define useful service and pipeline objectives, ownership, escalation, and recovery paths
  • Build self-service for common infrastructure, environment, access, deploy, debugging, and recovery work without becoming a central approval queue
  • Make cloud and vendor cost understandable by service and workload and improve efficiency within explicit reliability and security bounds
  • Partner with Security on cloud identity, secrets, isolation, audit logging, vulnerability response, incident readiness, and automated control evidence
  • Support employees and contractors through bounded access, safe environments, release controls, documentation, and timely removal of authority
  • Use AI tools deeply for platform engineering and operations while verifying generated code, plans, queries, state changes, and incident conclusions
What Success Looks Like
  • Replay's environments, runtimes, deploy paths, service and pipeline owners, reliability risks, recovery gaps, manual work, and infrastructure costs are visible and prioritized
  • One consequential failure or toil class is materially reduced in your first 90 days, and another team can use the resulting paved road without case-by-case help
  • Product, data, and AI teams can ship and understand their systems faster while retaining clear operating ownership
  • Priority services and pipelines have useful objectives, actionable telemetry, tested recovery paths, and incident learning that removes recurring failures
  • Common platform work becomes self-service while exceptions remain explicit, owned, monitored, and time-bounded
  • Cloud cost and capacity are understandable in workload-relevant units and improve without hidden reliability, security, or developer-time regressions
  • Security and customer-trust evidence becomes easier to produce because it reflects current technical controls
You Might Thrive Here If
  • You have personally owned production cloud infrastructure and delivery or reliability systems across multiple services, including an asynchronous, batch-data, or model-backed workload
  • You are a strong software engineer who is comfortable changing application, platform, and infrastructure code and operating the result in production
  • You can reason from user impact through dependencies, state, telemetry, incident response, recovery, and durable remediation
  • You have built paved roads other engineers adopted because they made real work easier, not because a platform team required them
  • You understand both long-running services and high-volume or scheduled workloads and know where their reliability models should differ
  • You can make pragmatic tradeoffs among delivery speed, least privilege, isolation, recovery, developer experience, and unit cost
  • You are effective in an early-stage environment where the first step is often to establish ownership and a trustworthy baseline
  • You can lead calmly through ambiguous incidents, communicate clearly, and leave the system and operating model stronger afterward
  • You use modern AI engineering tools fluently and verify generated infrastructure, queries, code, and operational conclusions before they affect production
This Role May Not Be for You If
  • You want a deployment or cloud-administration role where product teams hand systems to you to operate permanently
  • You prefer designing a platform in isolation to learning how engineers, services, pipelines, and customer deliveries actually work
  • You measure platform success by migration, ticket, dashboard, or uptime counts without connecting them to adoption, reliability, recovery, and user impact
  • You want to standardize every workload on one stack regardless of its state, scale, failure, or recovery requirements
  • You do not want AI tools to be part of your daily engineering and operational workflow
Bonus
  • Experience as an early platform or SRE hire at a fast-growing company
  • Experience with AWS, Terraform, container runtimes, workflow orchestration, and observability systems
  • Experience with high-volume data processing, model serving, evaluation jobs, GPU workloads, or machine-learning platforms
  • Experience improving developer environments, preview systems, CI/CD, progressive delivery, or internal developer platforms
  • Experience with replayable pipelines, backup and restore, disaster recovery, capacity planning, or cloud-cost allocation
  • Experience implementing technical controls and automated evidence for SOC 2 or enterprise customer requirements
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Platform Engineer in New York, NY vacancy
  • $176k - $191k

     ...the menu at Wonder. Except compromise. Wonder is the mealtime platform built to feed every craving in one order. With Wonder, you can...  ...is a new team chartered to own developer experience for Wonder Engineering — 400+ engineers across multiple converging organizations —... 
    Suggested
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    New York, NY
    4 days ago
  •  ...Experience ProfessionalsContact: Ashley RezinJob ID: REQ8583Platform Engineering owns the end-to-end experience of how engineers, investment...  ...systems that underpin our live trading environments, to the platforms that keep tens of thousands of builds, applications and tasks... 
    Suggested

    Balyasny Asset Management

    New York, NY
    2 days ago
  • $165k - $242k

     ...Cloud for AI. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to...  ...Learn more at .What You’ll Do:We are seeking a Senior Platform Engineer to join our Kubernetes Infrastructure team. This role involves... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    New York, NY
    4 days ago
  • $150k - $200k

     ...investor access to every asset, in every market, through a unified platform built for speed, transparency and scale.We give our clients the...  ...tomorrow.For more information, visit .The RoleAs a Platform Engineer, your customers are our engineers.You'll build the internal... 
    Suggested
    Work at office
    Local area
    Immediate start

    Clear Street

    New York, NY
    1 day ago
  •  ...digital partner that combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions that deliver...  ...Practical action. Endless possibilities.We’re hiring AgentCore Platform Engineers to build the agent-facing layer of the platform. That... 
    Suggested
    Contract work
    Temporary work

    Appnovation

    New York, NY
    4 days ago
  • $120k - $170k

     ...Rockstar Games is on the lookout for a passionate and talented Engineer to help enable safe, scalable, and effective adoption of cloud...  ...sits within IT and acts as the technical authority for cloud platforms, bridging emerging vendor capabilities with real-world production... 
    Full time
    Work at office

    Rockstar Games

    New York, NY
    3 days ago
  • $141k

     ...Prepared and Carbyne under Axon. Together, we are creating the only platform that combines modern 911 infrastructure with an AI intelligence...  ....Position OverviewWe are looking for an experienced platform engineer to join Prepared's platform engineering team under Axon 911. As... 
    Work experience placement
    Work at office

    Axon

    New York, NY
    4 days ago
  •  ...Experience ProfessionalsContact: Ashley RezinJob ID: REQ8582Platform Engineering owns the end-to-end experience of how engineers, investment...  ...systems that underpin our live trading environments, to the platforms that keep tens of thousands of builds, applications and tasks... 

    Balyasny Asset Management

    New York, NY
    2 days ago
  • $83.3k - $149.88k

    We’re hiring a Google Cloud Platform Engineer to design, deploy, and operate scalable, secure cloud infrastructure and platform services. You’ll partner with product and engineering teams to build automated, observable, and cost-efficient solutions on GCP—moving workloads... 
    Work at office
    Local area

    Perficient

    New York, NY
    4 days ago
  • $153k - $207k

     ...life-changing mission to develop education for our half a billion (and growing!) learners around the world.About the role...As a Platform Engineer on the Compute team, you’ll improve how engineers build and run services, helping ensure our compute platform is reliable at... 
    Work experience placement
    Internship

    Duolingo

    New York, NY
    2 days ago
  • $175k - $350k

    Role Overview Citadel Securities is seeking an exceptional Senior Platform Infrastructure Engineer to join one of our Platform Engineering teams in Miami or New York. Our team is responsible for building and maintaining the foundational platform that orchestrates application... 
    Worldwide

    Citadel Securities

    New York, NY
    1 day ago
  • $131.75k - $170.5k

     ...model. While the internal title for this position is Senior Linux Engineer, this role has been posted externally under a different title...  ...and skill sets.Role Overview We are seeking a Senior Platform Engineer to join our Systems Platform Engineer team. In this role... 
    Full time
    Work at office
    Immediate start

    Cboe Exchange

    New York, NY
    4 days ago
  • $113.9k - $189.9k

    Summary of Business Unit/Function:The One Policy Engine is a unified policy platform for multi-cloud environments that keeps Infrastructure‑as‑Code (IaC) compliant from code to cloud. We apply policy enforcement early and consistently across the SDLC so security and compliance... 
    Full time
    Part time
    Internship

    London Stock Exchange Group

    New York, NY
    3 days ago
  • $175k - $250k

    Senior Platform EngineerAbout MillenniumMillennium is a global, diversified alternative investment firm, founded in 1989. Defined by...  ...TeamMillennium’s Infrastructure organization is dedicated to designing, engineering, supporting, and managing a robust server estate, systems... 

    Millennium Management

    New York, NY
    3 days ago
  •  ...ChicagoDepartment: TechnologyExperience Level: Experience ProfessionalsContact: Ashley RezinJob ID: REQ8338We are seeking a Trading Platform Engineer to join our Systematic Technology team, focused on the reliability, observability, and day-to-day support of a high-... 

    Balyasny Asset Management

    New York, NY
    4 days ago
  • We are looking for a Databricks Platform Manager/Engineer responsible for designing, administering, securing, and optimizing the enterprise Databricks platform on AWS. The ideal candidate will own the platform lifecycle, ensuring scalability, security, governance, cost... 

    EXL Service

    New York, NY
    2 days ago
  • $84k - $126k

     ...deliver purposeful work and meaningful impact every day. Learn more about what makes us different and how you can thrive as a Platform Engineer at MMA. Marsh McLennan Agency (MMA) provides business insurance, employee health & benefits, retirement, and private client... 
    Minimum wage
    Local area
    Remote work
    Night shift

    Marsh McLennan

    New York, NY
    7 hours ago
  • Seeking a Senior Platform DevOps Engineer with Site Reliability Engineering (SRE) experience, the full-time remote position will design and maintain cloud infrastructure, automate deployment processes, and implement observability solutions to enhance system reliability... 
    Full time
    Remote work

    Virtual Vocations Inc

    New York, NY
    2 days ago
  • Who we areAbout StripeStripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises...  ...while protecting user data.What you’ll doAs a software engineer on Secure Devices, you will work at the intersection of software... 

    Stripe

    New York, NY
    3 days ago
  • $100.2k - $167k

    SummaryWe are looking for a Senior Associate Software & Platform Engineer to join our FX Engineering group as an individual contributor. This role is suited for an engineer who has strong foundational experience in Java application development, Spring Boot, build systems... 
    Full time
    Part time
    Internship

    London Stock Exchange Group

    New York, NY
    2 days ago
  • $189k - $236k

     ...requirements vary by role and will be assessed during the interview process.About the Role:We’re hiring seasoned engineers to join our teams that work on core platform capabilities, improving our existing systems for extensibility and scalability, and building the future of... 
    Full time
    Work at office
    Local area
    Remote work
    2 days per week
    3 days per week

    Gusto

    New York, NY
    5 days ago
  • $131k - $164k

    Help shape the technology that enables a global organisation to do its best work. As Senior Manager, Platform Engineering, you’ll lead the team responsible for Diligent’s Atlassian and Microsoft platforms while setting the architectural direction for the wider internal... 
    Work at office
    Local area
    Worldwide
    Visa sponsorship
    Flexible hours

    Diligent

    New York, NY
    5 days ago
  • $187k - $240k

    As a Platform Security Engineer you will partner with different stakeholders across the organization to secure our infrastructure and application components of the Datadog platform. As part of the Platform Security organization we secure the building blocks of Datadog’... 
    Work at office

    Datadog

    New York, NY
    5 days ago
  • $160k - $185k

     ...Senior Azure DevOps & Kubernetes Platform Engineer Company Overview: Business Integration Partners (BIP) is Europe’s fastest growing digital consulting company and are on track to reach the Top 20 by 2030, with an expanding global footprint in the US (New York, Charlotte... 
    Temporary work
    Remote work
    Worldwide

    BIP US

    New York, NY
    18 hours ago
  • BNY is seeking a Senior Vice President to lead the Wealth Services Platform Cloud Nerve Center. This leader will establish the production environment, engineering patterns, and operational capabilities needed to migrate WSP applications to the cloud securely, efficiently... 
    Worldwide
    Flexible hours

    The Bank of New York Mellon

    New York, NY
    3 days ago
  • $168k - $200k

     ...grow our team to meet the needs of more companies, teams, and innovators in this way. The Role:   As a Senior Software Engineer on the Broker Platform team, you will play a key role in delivering high-quality experiences to our Brokers, working with modern, low-latency... 
    Work experience placement
    Work at office
    Local area
    2 days per week
    3 days per week

    Forge Global

    New York, NY
    5 days ago
  • $129.7k - $216.1k

    SummaryWe are looking for a hands-on Lead Software and Platform Engineer to drive the design, development, delivery, and operational reliability of large-scale, low-latency microservices-based trading platforms within our FX Engineering group.This role combines strong Java... 
    Full time
    Part time
    Internship

    London Stock Exchange Group

    New York, NY
    2 days ago
  •  ...Experience ProfessionalsContact: Ashley RezinJob ID: REQ8278About UsOur Database Engineering team is transforming how Balyasny Asset Management delivers database services through a standardized platform built on declarative, code-defined provisioning, automated lifecycle... 

    Balyasny Asset Management

    New York, NY
    1 day ago
  • $210k - $270k

     ...Zocdoc’s marketplace topower access to care wherever patients search, from provider websites and insurance directories to search engines, AI platforms, and more.Healthcare still lacks something every other major consumer industry takes for granted: a seamless way to go from... 
    Remote work
    Flexible hours

    ZocDoc

    New York, NY
    3 days ago
  • £45k - £65k per year

     ...i2 Group , a Harris Computer company , is seeking a Platform Software Engineer on a full-time, permanent, remote basis with a monthly requirement of attending the i2 Group offices in Cambridge. The Role The Platform team solves one of the harder problems in... 
    Permanent employment
    Full time
    Immediate start
    Remote work
    Shift work

    Remote Worker LTD.

    New York, NY
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Platform Engineer. Be the first to apply!