Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Platform Engineer

Jobleads-US

About Replay

At its core, Replay was founded to help founders. We started by supporting startups through shutting down, but we have since expanded into unlocking a new revenue stream for all types of businesses.

In 2025, we had a unique insight: the data every company generates each day through collaboration, communication, and building is some of the most valuable training data in the world. Public and synthetic data can only get frontier models so far, so the next generation of model progress depends on real, proprietary data grounded in how actual businesses operate. We are a primary source of it, partnering directly with the frontier AI labs building what comes next.

Why Join Replay Now

  • We have scaled from $0 to a multi-eight-figure run rate in a matter of months

  • We have raised from top-tier investors, including Floodgate, Afore, Ludlow, and Hustle Fund

  • We are small enough that you will carry outsized responsibility and grow as quickly as the company does

  • You will partner with and build for some of the fastest and most important companies in the world

  • You will help build a massive, category-defining business from the ground floor

The Role

Replay operates customer-facing SaaS products, connector and ingestion services, asynchronous workers, high-volume data pipelines, model-backed systems, review tools, and customer-delivery paths. These workloads have different shapes, but they need a coherent foundation for infrastructure, delivery, observability, recovery, access, and cost.

You will build and operate the shared platform that lets our product, data, and AI teams ship reliable, secure, observable, and cost-aware systems without manual infrastructure work or operational risk growing linearly. You will write software and infrastructure, improve real engineering workflows, lead through incidents, and create paved roads teams can use without waiting on you.

This is not a deployment-operator or internal-IT role. Product, data, and ML teams remain responsible for the systems they build. You will give them the runtime, delivery, visibility, recovery, and operating patterns to own those systems well. You will partner closely with our Security Lead, but you will not be expected to run the entire security or compliance program.

Problems You Might Own

Make several workload shapes feel like one coherent platform

Create a small set of supported patterns for customer-facing services, connectors, scheduled jobs, data-processing pipelines, model-backed workloads, and evaluation runs. Define the contracts for environments, compute, state, networking, delivery, secrets, telemetry, failure handling, and recovery without forcing every workload into an inappropriate stack.

Turn delivery and operations into product-quality experiences

Make it straightforward for an engineer to create an environment, ship a safe change, understand a failed deploy or job, get the right access, recover a system, and know who owns the result. Build useful self-service and escape hatches while making unsupported paths and exceptions explicit.

Make reliability visible from customer request to completed workload

Connect service, queue, job, pipeline, and model telemetry to the outcome that matters. Establish practical objectives, alerts, incident mechanics, replay and recovery paths, and reviews that remove recurring failure classes instead of only documenting them.

Make infrastructure cost and control evidence part of normal operation

Expose cost and capacity in workload-relevant units, then improve them without hiding reliability, security, quality, or developer time. Work with Security to implement least privilege, secrets, logging, backup, deployment, and audit controls whose evidence comes from the systems that actually enforce them.

What You'll Do

  • Establish Replay's current platform, workload, reliability, ownership, toil, recovery, cost, and technical-control baseline

  • Build reusable infrastructure-as-code modules, runtime templates, deployment workflows, environment contracts, and operational tooling

  • Create supported paths for customer-facing services, asynchronous and batch jobs, data pipelines, and model-backed workloads

  • Improve deploy safety, workload visibility, backup and recovery, incident response, replay, rollback, and durable remediation

  • Work with engineering teams to define useful service and pipeline objectives, ownership, escalation, and recovery paths

  • Build self-service for common infrastructure, environment, access, deploy, debugging, and recovery work without becoming a central approval queue

  • Make cloud and vendor cost understandable by service and workload and improve efficiency within explicit reliability and security bounds

  • Partner with Security on cloud identity, secrets, isolation, audit logging, vulnerability response, incident readiness, and automated control evidence

  • Support employees and contractors through bounded access, safe environments, release controls, documentation, and timely removal of authority

  • Use AI tools deeply for platform engineering and operations while verifying generated code, plans, queries, state changes, and incident conclusions

What Success Looks Like

  • Replay's environments, runtimes, deploy paths, service and pipeline owners, reliability risks, recovery gaps, manual work, and infrastructure costs are visible and prioritized

  • One consequential failure or toil class is materially reduced in your first 90 days, and another team can use the resulting paved road without case-by-case help

  • Product, data, and AI teams can ship and understand their systems faster while retaining clear operating ownership

  • Priority services and pipelines have useful objectives, actionable telemetry, tested recovery paths, and incident learning that removes recurring failures

  • Common platform work becomes self-service while exceptions remain explicit, owned, monitored, and time-bounded

  • Cloud cost and capacity are understandable in workload-relevant units and improve without hidden reliability, security, and developer-time regressions

  • Security and customer-trust evidence becomes easier to produce because it reflects current technical controls

You Might Thrive Here If

  • You have personally owned production cloud infrastructure and delivery or reliability systems across multiple services, including an asynchronous, batch-data, or model-backed workload

  • You are a strong software engineer who is comfortable changing application, platform, and infrastructure code and operating the result in production

  • You can reason from user impact through dependencies, state, telemetry, incident response, recovery, and durable remediation

  • You have built paved roads other engineers adopted because they made real work easier, not because a platform team required them

  • You understand both long-running services and high-volume or scheduled workloads and know where their reliability models should differ

  • You can make pragmatic tradeoffs among delivery speed, least privilege, isolation, recovery, developer experience, and unit cost

  • You are effective in an early-stage environment where the first step is often to establish ownership and a trustworthy baseline

  • You can lead calmly through ambiguous incidents, communicate clearly, and leave the system and operating model stronger afterward

  • You use modern AI engineering tools fluently and verify generated infrastructure, queries, code, and operational conclusions before they affect production

This Role May Not Be for You If

  • You want a deployment or cloud-administration role where product teams hand systems to you to operate permanently

  • You prefer designing a platform in isolation to learning how engineers, services, pipelines, and customer deliveries actually work

  • You measure platform success by migration, ticket, dashboard, or uptime counts without connecting them to adoption, reliability, recovery, and user impact

  • You want to standardize every workload on one stack regardless of its state, scale, failure, or recovery requirements

  • You do not want AI tools to be part of your daily engineering and operational workflow

Bonus

  • Experience as an early platform or SRE hire at a fast-growing company

  • Experience with AWS, Terraform, container runtimes, workflow orchestration, and observability systems

  • Experience with high-volume data processing, model serving, evaluation jobs, GPU workloads, or machine-learning platforms

  • Experience improving developer environments, preview systems, CI/CD, progressive delivery, or internal developer platforms

  • Experience with replayable pipelines, backup and restore, disaster recovery, capacity planning, or cloud-cost allocation

  • Experience implementing technical controls and automated evidence for SOC 2 or enterprise customer requirements

#J-18808-Ljbffr Jobleads-US
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Platform Engineer in New York, NY vacancy
  • $176k - $191k

     ...the menu at Wonder. Except compromise. Wonder is the mealtime platform built to feed every craving in one order. With Wonder, you can...  ...is a new team chartered to own developer experience for Wonder Engineering — 400+ engineers across multiple converging organizations —... 
    Suggested
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    New York, NY
    4 days ago
  • $131.75k - $170.5k

     ...model. While the internal title for this position is Senior Linux Engineer, this role has been posted externally under a different title...  ...and skill sets.Role Overview We are seeking a Senior Platform Engineer to join our Systems Platform Engineer team. In this role... 
    Suggested
    Full time
    Work at office
    Immediate start

    Cboe Exchange

    New York, NY
    4 days ago
  •  ...digital partner that combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions that deliver...  ...Practical action. Endless possibilities.We’re hiring AgentCore Platform Engineers to build the agent-facing layer of the platform. That... 
    Suggested
    Contract work
    Temporary work

    Appnovation

    New York, NY
    4 days ago
  • Seeking a Senior Platform DevOps Engineer with Site Reliability Engineering (SRE) experience, the full-time remote position will design and maintain cloud infrastructure, automate deployment processes, and implement observability solutions to enhance system reliability... 
    Suggested
    Full time
    Remote work

    Virtual Vocations Inc

    New York, NY
    1 day ago
  • $189k - $236k

     ...requirements vary by role and will be assessed during the interview process.About the Role:We’re hiring seasoned engineers to join our teams that work on core platform capabilities, improving our existing systems for extensibility and scalability, and building the future of... 
    Suggested
    Full time
    Work at office
    Local area
    Remote work
    2 days per week
    3 days per week

    Gusto

    New York, NY
    4 days ago
  • $160k - $185k

     ...Senior Azure DevOps & Kubernetes Platform Engineer Company Overview: Business Integration Partners (BIP) is Europe’s fastest growing digital consulting company and are on track to reach the Top 20 by 2030, with an expanding global footprint in the US (New York, Charlotte... 
    Temporary work
    Remote work
    Worldwide

    BIP US

    New York, NY
    8 hours ago
  • £45k - £65k per year

     ...i2 Group , a Harris Computer company , is seeking a Platform Software Engineer on a full-time, permanent, remote basis with a monthly requirement of attending the i2 Group offices in Cambridge. The Role The Platform team solves one of the harder problems in... 
    Permanent employment
    Full time
    Immediate start
    Remote work
    Shift work

    Remote Worker LTD.

    New York, NY
    3 days ago
  • $180k - $230k

     ...broken, it's depressing if you don't laugh with us. About the role: We are looking for an experienced engineer to join as the third member of our Platform team. This unit operates across the stack, developing the resources and tooling that allow our engineers to... 
    Work at office
    Remote work
    Flexible hours

    Camber

    New York, NY
    2 days ago
  •  ...We’re looking for a Platform Engineer to define and lead the technical direction of our platform as we scale 10x+ in enterprise customers This is a hands-on, high-impact role responsible for architecture, reliability, and developer experience across our Kubernetes... 

    Rogo

    New York, NY
    3 days ago
  •  ...recently raised our $1.5B Series F, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Cloud Platform Engineer, you’ll envision and build robust systems and... 
    Flexible hours

    The Consensus

    New York, NY
    1 day ago
  •  ...Full TimeWorking Type HybridJob Reference 0000021462Salary Type AnnuallyIndustry Law PracticeSelling Points Lead innovative AI platform engineering projects at a forward-thinking organization. Collaborate on cutting-edge Azure AI/ML solutions and cloud infrastructure.... 

    Green Key Resources

    New York, NY
    16 hours ago
  •  ...at scale by combining deep industry expertise, proven software platforms, and innovative AI-driven solutions. A global market leader in...  ...day.We are seeking an experienced Senior ServiceNow Platform Engineer to manage, enhance, and optimize the ServiceNow platform across... 
    Minimum wage
    Full time

    DXC Technology

    New York, NY
    1 day ago
  •  ...Job Description Job Description Genesis10 is currently seeking a Site Reliability / Platform Engineer - Remote position with a Leading Asset Management Firm. This role is open to Remote US based resources, however candidates that are able to work hybrid in either New... 
    Hourly pay
    Permanent employment
    Contract work
    Temporary work
    Remote work

    Genesis10

    New York, NY
    4 days ago
  • $50k

     ...backed AI software company building a next-generation enterprise platform that leverages artificial intelligence, automation, and cloud-...  ...operations. Position Summary The Principal Platform Engineer is the organization's most senior individual contributor and technical... 

    Affinity Inc

    New York, NY
    3 days ago
  • $200k - $230k

     ...The Director, Platform Engineering will own the strategy, roadmap, and execution of InvestCloud's shared platform capabilities across on-prem and AWS. This role provides both technical and people leadership across environments, infrastructure, CI/CD, and platform services... 
    Flexible hours
    Shift work

    InvestCloudOct2024

    New York, NY
    2 days ago
  • $176k - $179.5k

     ...firm leaders, CST leaders, practice leaders and cutting-edge engineers. Your team is led by some of our firm’s most senior leaders and...  ...customization and additional features required to deploy this platform for Defense and other Government clients. Your expertise in Python... 
    Apprenticeship
    Easy work

    McKinsey & Company

    New York, NY
    16 hours ago
  •  ...digital partner that combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions that deliver...  ...together a dedicated delivery pod to build and run an internal agent platform on AWS Bedrock AgentCore for a global life sciences client.... 

    Appnovation

    New York, NY
    4 days ago
  • $136k - $253k

     ...transaction data into structured, searchable market intelligence. Its platform helps legal and finance professionals analyze deal terms,...  ...with greater speed and confidence. As a Senior Data Platform Engineer, you will design, build, and maintain the infrastructure that ingests... 
    Full time
    Work at office
    Local area
    Flexible hours

    Thomson Reuters

    New York, NY
    2 days ago
  • $130k - $185k

     ...shape your future with confidence.Within EY‑Parthenon, the Growth Platforms team focuses on building durable technology and AI capabilities that power long‑term enterprise growth. The Software Engineering Director for AI Tooling plays a pivotal role in translating... 
    Work experience placement
    Summer holiday
    Work at office
    Flexible hours

    EY (Ernst & Young)

    New York, NY
    3 days ago
  • $155k - $215k

     ...to develop a firmwide Artificial Intelligence (AI) Development Platform that aligns with the firm's Technology principles and drives...  ...adoption of AI across our businesses.This role is for a platform engineering specialist who will help build a firmwide AI Development... 
    Temporary work

    Morgan Stanley

    New York, NY
    2 days ago
  • $107.5k - $204.5k

     ...Enterprise Services Data (ES-Data) group is seeking a Responsible AI Engineer to support RTX’s focus on Responsible AI policy implementation...  .... The role will focus on configuration of AI governance platforms, automation and integration with enterprise systems, reporting... 
    Temporary work
    Work experience placement
    Work at office
    Remote work
    Work from home
    Flexible hours

    Raytheon

    New York, NY
    16 hours ago
  • $130.2k - $195.3k

     ...that matter – both for our audiences and our employees – and aim to leave a positive mark on culture. We are seeking a Sr. ML Platform Engineer who is excited to deploy MLOps products that shape business strategy, optimize content, inform marketing investment decisions,... 

    Paramount Pictures

    New York, NY
    1 day ago
  •  ...digital partner that combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions that deliver...  ...Practical action. Endless possibilities. We’re hiring AWS Cloud Platform Engineers to build and own the platform’s foundation. That... 

    Jobleads-US

    New York, NY
    3 days ago
  • $161k - $193k

     ...Columbia, United States-based Platform Engineer role (TS/SCI with Polygraph) seeks expertise in Linux and Ansible to build and maintain internal developer platforms. The role emphasizes automation, CI/CD, observability, and strong collaboration with engineering teams.... 

    Jobleads-US

    New York, NY
    2 days ago
  •  ...Sequence Holdings is seeking a Product Engineer in New York, NY to own the Atlas platform and scale the AI-driven transformation across portfolio companies. You will build agent infrastructure, workflow orchestration, and data tooling, working closely with Forward Deployed... 

    Jobleads-US

    New York, NY
    4 days ago
  •  ...Astronomer is redefining how companies run Apache Airflow at scale. We seek a Senior Software Engineer to join Platform Engineering, shaping production systems, CI/CD pipelines, and reliability practices for Astro, Observe, and IDE products. Lead testing, deployment... 

    Jobleads-US

    New York, NY
    2 days ago
  •  ...Filevine is a Legal AI company delivering LOIS-powered operating intelligence. We are seeking a Senior CI/CD Platform Engineer to own end-to-end pipelines across .NET, Python, and Svelte, transforming delivery speed and reliability. You will build internal tooling, establish... 
    Work at office
    Remote work

    Jobleads-US

    New York, NY
    3 days ago
  •  ...Replay is building a shared platform to empower product, data, and AI teams to ship reliable, secure, and cost-aware systems without...  ...diverse workloads. You will collaborate with teams across engineering and security, drive end-to-end reliability, and help reduce toil... 

    Jobleads-US

    New York, NY
    2 days ago
  •  ...Adonis in New York is seeking a Senior Infrastructure Engineer to own and scale our cloud platform, developer tooling, and data infrastructure. This hands-on IC role sets reliability, security, and performance standards across healthcare products serving enterprise physician... 

    Jobleads-US

    New York, NY
    3 days ago
  •  ...Scale, building reliable AI systems for high-stakes decisions, seeks a Senior Software Engineer to help evolve the Orchestration Platform. You will design primitives, services, and developer tools to enable Scale teams to author, operate, observe, and safely scale distributed... 

    Jobleads-US

    New York, NY
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Platform Engineer. Be the first to apply!