Platform Delivery & Reliability Engineer (Remote) - 29337
$119.57k - $235kHII's Mission Technologies division
Enlighten, honored as a Top Workplace from USA Today, is a leader in big data solution development and deployment, with expertise in cloud-based services, software and systems engineering, cyber capabilities, and data science. Enlighten provides continued innovation and proactivity in meeting our customers’ greatest challenges.
We recognize that the most effective environment for your projects doesn’t always look the same. Our hybrid work approach ensures that you can make lasting relationships with your team and collaborate in-person to get the job done—while having the flexibility to work from home when needed to achieve focused results.
Why Enlighten?
At Enlighten, our team’s unwavering work ethic, top talent and celebration of innovative ideas have helped us thrive. We know that our employees are essential to our company’s success, so we seek to take care of you as much as you take care of us. Here are a few highlights of our benefits package:
• 100% paid employee premium for healthcare, vision and dental plans.
• 10% 401k benefit.
• Generous PTO + 10 paid holidays.
• Education/training allowances.
Anticipated Salary Range: $119,574.00 - $235,000.00. The salary range for this role is intended as a good faith estimate based on the role's location, expectations, and responsibilities. When extending an offer, Enlighten takes a variety of factors into consideration which include, but are not limited to, the role's function, internal equity and a candidate's education or training, work experience, certifications and key skills. Occasionally positions/roles may include additional non-recurrent compensation and will be addressed by the recruiter during the interview process.
Job Description
Enlighten is looking for a Senior Platform Delivery & Reliability Engineer to own the rollout of our data lakehouse platform across a large, multi-site government enterprise, currently ~50 production Kubernetes clusters and growing. This role is a rare hybrid of platform engineer, SRE, and delivery lead. You can deploy the platform, debug anything you encounter in the field, feed what you learn back to the engineering teams, and help fix underlying issues in the code base. Just as importantly, you can step back from any individual issue and fix the system that produced it by building the processes, tooling, and communication channels inside Enlighten that make every deployment faster and less painful than the one before it.
You will be a full member of the Infrastructure team, working daily with our Ingest, Query, Application and Testing teams as well as government customers and site personnel. Success in this role looks like: rollouts across the enterprise happen predictably and efficiently, issues found in the field are cataloged, communicated, and resolved quickly, and the friction that slows deployments steadily disappears.
#LI-DS1 #Senior Level
Essential Job Responsibilities
- Plan, coordinate, and execute deployments and upgrades of the data lakehouse platform across 50 production Kubernetes clusters in customer environments.
- Debug and troubleshoot critical issues anywhere in the stack (infrastructure, Kubernetes, platform services, data services, and applications) and drive them to root cause.
- Contribute patches, configuration changes, and automation improvements directly back to the platform.
- Catalog and triage issues discovered in the field, communicate them clearly to the Infrastructure, Ingest, Query, and Application teams, and maintain a living knowledge base of failure modes, fixes, and runbooks.
- Enable and improve site reliability for fielded environments: monitoring, alerting, incident response, and continuous reliability improvement.
- Identify friction and dysfunction in how deployments happen, such as unclear handoffs, communication gaps, and repeated manual work; then design, implement, and institutionalize the processes that eliminate them (release checklists, readiness reviews, escalation paths, cross-team communication cadences).
- Continuously improve deployment tooling and automation so rollouts become faster, safer, and more repeatable.
- Coordinate with a large set of stakeholders (the engineering teams, government programs, security, and site personnel) and keep them informed.
- Mentor engineers, both junior and senior, on debugging, deployment, and operational excellence.
- Other duties as assigned.
Minimum Qualifications
- Clearance Requirement: Must obtain and maintain a U.S. Government Security Clearance, but not required on day one; U.S. Citizenship required.
- 9 years relevant experience with Bachelors in related field; 7 years relevant experience with Masters in related field; or High School Diploma or equivalent and 13 years relevant experience.
- Deep, hands-on experience deploying, operating, and debugging production Kubernetes clusters and their ecosystem (networking, storage, service meshes, observability, volume management).
- Proven record of delivering complex distributed systems into production across many environments or sites: not just building platforms, but landing them with customers.
- Elite troubleshooting and analytical skills across the full stack: Linux systems, hosts, networks, security, containers, and application services.
- Experience with infrastructure as code (e.g., Terraform) and modern cloud environments (e.g., AWS, Azure, GCP).
- Experience with CI/CD pipelines (e.g., GitLab CI) and proficiency in scripting or programming (e.g., Go, Python, Bash).
- Working knowledge of SRE practices: monitoring and alerting, incident management, blameless postmortems, and runbook development.
- Demonstrated experience creating or improving engineering and delivery processes that other teams actually adopted; you can point to a workflow that exists because you built it.
- Excellent verbal and written communication skills; able to translate deep technical issues for engineers, leadership, and customers, and comfortable coordinating a large number of people across organizational boundaries.
- Work Location: *Remote or Hybrid. This role is fully remote unless you are located near one of our offices in Columbia, MD; San Antonio, TX; Boise, ID; Greenville, SC; or Augusta, GA, where a hybrid schedule applies. Note: Work models are subject to change based on business needs.
Preferred Requirements
- Experience deploying or operating large-scale data platforms and lakehouse technologies (e.g., Spark, Trino/Presto, Kafka, NiFi, object storage, Iceberg/Delta/Hudi)
- Experience delivering into DoD, IC, or other federal environments, including STIG-hardened, disconnected, or air-gapped deployments and familiarity with the RMF/ATO process
- Experience with Kubernetes Operators/Controllers development
- Prior release management, delivery lead, field engineering, or deployment engineering experience on a multi-team program
- Understanding of agile software development methodologies and use of standard software development tool suites (e.g., YouTrack, GitLab, Nexus)
- DoD 8140 / 8570 compliance certifications may be required in this position as directed by the customer
We have many more additional great benefits/perks that you can find on our website at
- About the RoleAs Senior Reliability Engineer, you will ensure the resilience and availability of Kohl’s systems and applications, collaborate... ...troubleshooting and performance tuningWorking experience with one cloud platform (GCP, AWS, or Azure)Working experience with monitoring...Remote workPlatformWork experience placement
$110k - $150k
...Phoenix, AZ that is seeking a few Senior Reliability Engineers. We are partnering directly with the... ...key hires. This position is Hybrid Remote.Key Responsibilities:* Drive equipment... ...environments* Experience with Electro-Mechanical Platforms or systems* Proven success developing...Remote workPlatformWork at office- ...everything possible.The Senior Reliability Engineer is responsible for... ...therapeutics from discovery to delivery.What you will do:Analyze equipment... ...equipment, access platforms and components 5–6 feet off... ...the benefits of flexible, remote working arrangements for eligible...Remote workPlatformHourly payFull timeLocal areaWork from homeFlexible hours
$71.2k - $142.8k
...and we’re looking for top-tier Systems Reliability Engineers who share our passion for customer... ...out of their investment in the Nutanix platform.What You Will BringExcellent written and... ...hybrid capacity, blending the benefits of remote work with the advantages of in-person...Remote workPlatformWork at officeWorldwideRelocation package3 days per weekWeekday work$170k - $250k
...completed 10B lifetime deliveries. We’re focused on how... ...motivated Senior/Staff Test Engineer to join our team. This... ...of our unmanned platforms at the system and component... ...scripts to automate reliability tests. Experience... ...Located in NYC or Remote Jobs Associated With Office...Remote workPlatformHourly payWork at officeLocal areaFlexible hours$139.4k - $205k
...has completed 10B lifetime deliveries. We’re focused on how to... ...a highly motivated Senior Reliability & Test Engineer to join our team. This individual... ...of our unmanned platforms at the system and component... ...for Jobs Located in NYC or Remote Jobs Associated With Office...Remote workPlatformHourly payWork at officeLocal areaFlexible hours$100k - $120k
...brighter way forward. JLL is seeking a Reliability Engineer to join our team! In JLL Work... ...Services Reliability & Asset Management platform for instruction in deploying client’s... ...and internal considerations.Location:Remote -Andover, MA, Boston, MA, Chicago, ILIf...Remote workPlatformFull timeWork at officeLocal areaFlexible hours$230k - $250k
GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability... ...based in Arlington, VA, as a hybrid/remote position.ResponsibilitiesResponsibilities... ...toolsFamiliarity with monitoring platforms and incident management practicesExperience...Remote workPlatform$100k - $150k
...Systems Reliability Engineer – Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud... ...infrastructure and operations problems, and continually pushing the platform toward higher reliability with lower operational toil....Remote workPlatformFull timeH1bLocal areaImmediate startVisa sponsorship- ...Building cloud-native reliability and automation solutions, the full-time Staff Reliability Engineer will design and operate engineering platforms for software validation and release, while working flexibly or remotely to enhance developer productivity and deployment...Remote workPlatformFull time
$96k - $140k
...right data to the people who need it, our platforms empower our partners to develop... ...missing children, and more.The RoleProduct Reliability Engineers (PREs) are responsible for the health,... ...there are a few roles that allow for “Remote” work on an exceptional basis. If you...Remote workPlatformPermanent employmentFull timeFixed term contractWork experience placementWork at officeWork from homeRelocation packageShift work- ...seeking a skilled and proactive Site Reliability Engineer to join our team, ensuring the stability... ...to the government. This is a fully remote position for candidates in the continental... ...a particular focus on Google Cloud Platform.ResponsibilitiesGoogle Cloud Integration...Remote workPlatform
- ...SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8 to 15 yearsSkillsKubernetes... ...skilled Senior Site Reliability Engineer (SRE) to join our dynamic team.... ....Proficient in AWS or GCP cloud platforms.Hands on experience with...Remote workPlatform
$104.9k - $174.7k
...:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and... ...schedule. If not, this role is fully remote. We do not restrict applicants based on... ...environmentsExperience operating monitoring and uptime platforms such as Grafana, Pingdom, and...Remote workPlatformFull timeWork at officeLocal areaWork from home- ...Joining a remote-first team, the full-time SRE Platform Engineer will design, build, and evolve platform infrastructure while enhancing developer experience... ...production infrastructure Strong understanding of reliability, security, and observability in systems...Remote workPlatformFull time
$210k - $230k
...currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and... ...Arlington, VA and is a hybrid remote/onsite position.ResponsibilitiesKey... ...infrastructure and application delivery• Build self-service tools and platforms to enable development...Remote workPlatformCurrently hiring$125k - $130k
...industry-leading unified DataOps platform powered by Apache Airflow®. Astro accelerates building reliable data products that unlock... ...Astronomer Customer Reliability Engineering (CRE) team is responsible for... ...production. Participate remotely within a fully distributed...Remote workPlatformFull timeWeekend work$67.2k - $100.8k
...Workstyle:Blue Bell, PA: Primarily remote; candidates should be within... ...teams to keep the ADT platform running and customers... ...partners (IT, Security, DevOps, Engineering) to improve operational health... ...SRE best practicesSupport the reliability, availability, scalability,...Remote workPlatformTemporary workH1bWork at office- ...Apex Systems is seeking a Site Reliability Analyst for a fully remote contract position. The role focuses on... ...excellence of a large-scale analytics platform. Initially, you will support system... ...then transition into performance engineering to troubleshoot complex issues. Key...Remote workPlatformContract work
- ...United States is seeking a Senior Site Reliability Engineer to strengthen reliability, observability, and operations for the Qira platform. You will work across device, edge, and... ...deployments and robust incident response. Open to remote US with Chicago as the preferred...Remote workPlatform
- ...days in the office/two days remote).Job Summary:With a "document... ..., scalability, and reliability of systems and applications.... ...CloudFormation.Mentor junior engineers and provide technical guidance... ...experience.Experience with cloud platforms (AWS or Azure).Experience with...Remote workPlatformWork at office
$15k
...lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our... ...Your work will provide a world-class HPC platform for researchers to focus on cutting-edge... ...: $205K - $235KLocationBerkeley, CA; Remote, United StatesEmployment TypeFull...Remote workPlatformWork at officeLocal area$80k - $133k
...Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and... ...obligated to pay a placement fee.SummaryLocation: US - TX, San Antonio; US - VA, McLean; US - Remote (Any location)Type: Full timeRemote workPlatformPermanent employmentFull timeContract workFlexible hours- ...apply now.We are currently seeking a Site Reliability Engineer to join our team in Westlake, Texas (... ..., Infrastructure-as-Code (IaC), cloud platforms, CI/CD pipelines, and operational... ...roles. The starting pay range for this remote role is 85,000-140,000 This range reflects...Remote workPlatformFull timeTemporary workWork at officeFlexible hours
- ...A leading livestream shopping platform is seeking a Senior Software Engineer for the Logistics Platform team. This role focuses on improving logistical data... ...debugging. The position offers flexibility for remote work and benefits including health insurance and generous...Remote workPlatform
- ...Passionate about designing and automating cloud platforms, the full-time Senior Azure Site Reliability Engineer will focus on improving reliability through Infrastructure as Code, automation, and observability while managing Azure cloud infrastructure and collaborating...Remote workPlatformFull time
- ...shaping the future of software delivery. From generative AI and cloud-native platforms to advanced release engineering practices, our teams are... ...- 2 days onsite, 3 days remote per weekSponsorship Notice:... ...accelerate development and improve reliability. Your work will directly...Remote workPlatformH1bWork at officeVisa sponsorshipFlexible hours2 days per week
$148.9k - $260.6k
...are seeking an exceptional Staff Site Reliability Engineer to lead critical infrastructure initiatives... ...the next generation identity security platform for the multi-cloud era - will you... ...and trust. Work personas (flexible, remote, or required in office) are categories...Remote workPlatformWork at officeFlexible hours$146k - $194k
...lead provider of specialized engineering and products for... ...requirements.ABOUT THE JOBAs a Site Reliability Engineer, your primary mission... ...automated solutions that increase platform resilience. We are looking... ...Patch/update management and remote management tooling.DevOps...Remote workPlatformFull timeWork experience placementImmediate start$125k - $185k
...right data to the people who need it, our platforms empower our partners to develop... ...’re looking for Forward Deployed Site Reliability Engineers who can help us build, operate, and maintain... ...there are a few roles that allow for “Remote” work on an exceptional basis. If you...Remote workPlatformFull timeWork experience placementWork at officeWork from homeRelocation package
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Platform Delivery & Reliability Engineer (Remote) - 29337. Be the first to apply!
- data platform engineer Remote
- client platform engineer Remote
- platform engineering manager Remote
- platform developer Remote
- senior platform engineer Remote
- platform engineer Remote
- database reliability engineer Remote
- senior reliability engineer Remote
- reliability maintenance engineering technician Remote
- network reliability engineer Remote



