Senior Platform Operations Engineer, Infrastucture
Viome Life Sciences
Job Description
Job Description
At Viome, we are driven by a singular mission: to help people live a healthy, disease-free life. This mission guides our actions, fuels our passion, and shapes the impact we aim to have in the world. Our core values - Be Bold, Be Collaborative, Be Frugal, and Grow Continuously - underpin our approach to achieving this goal. If you are motivated by the idea of working in an environment that prioritizes bold innovation, teamwork, efficient resource use, and continuous learning, all towards promoting health and preventing disease, we warmly invite you to apply. Join us in our journey to transform lives and create a healthier future for all.
We are looking for a Senior Platform Operations Engineer, Infrastructure to take ownership of Viome’s Azure-based service platform with a clear mandate: make it dramatically simpler through robust Unix-based systems engineering.
This is a consolidation role for a seasoned systems administrator. The work involves migrating a suite of abstracted services onto a deliberately stable target platform: Linux VMs, systemd-supervised services, Apache, HAProxy, and Nginx routing. You must possess strong Unix knowledge to operate and eventually transform the current stack safely. We are looking for a traditional operations background where success is measured by the removal of unnecessary layers and a genuine bias toward simplicity and host-level stability.
ResponsibilitiesSimplification & Migration (core mandate)
Plan and execute zero-downtime migrations of services from AKS/containers to VM-based hosting: systemd unit authoring and service supervision (restart policy, resource limits, sandboxing), Apache, HAProxy and/or nginx as reverse proxy and TLS terminator, certificate automation.
Build and own the deployment scripts for the target platform: build → test → archive → ship → symlink-flip → health-check → rollback, scripted in bash.
Preserve two non-negotiable invariants while simplifying everything else: immutable, commit-traceable artifacts and scripted rollback.
Inventory and retire stale infrastructure: dormant deployments, unused DNS records, orphaned firewall rules, unpinned image tags.
Maintain host hygiene for consolidated services: OS patching discipline, runtime vendoring, log rotation, centralized log aggregation.
Familiarity with UptimeKuma, Nagios or similar.
Operate and migrate data stores: PostgreSQL and/or MySQL.
Network and Infrastructure Operations
Operate and troubleshoot workloads during the transition, focusing on the networking layer, load balancing, and core platform services.
Administer the hub network: Azure Firewall rules, VPN gateways, VNet peering, public and private DNS zones and reason about a packet’s full path from public IP to service.
Support the existing release process and network and infrastructure operations until each service is migrated to the new VM-based standard.
Keep the observability stack healthy (OpenTelemetry, ELK, Grafana, uptime and cost monitoring) and carry its essentials forward to the simplified platform.
Coordinate cross-cloud dependencies with AWS: DNS/edge routing, queue consumers, and egress IP allowlists.
L2+ operations support; manage runbooks for external L0, L1 support.
External Integrations & Security
Own the external integrations most at risk during migration: e-commerce and subscription platforms, messaging/notification providers, and clinical/health-data partners — webhook delivery, signature verification, idempotency, and retry semantics.
Raise the security baseline as you consolidate: secrets management, webhook authentication, least-privilege network access, and data-retention hygiene. Findings from an internal review are ready for you to remediate; the instinct to spot and close this class of issue — and not create more — is part of the job.
Required
7+ years operating production Unix/Linux systems, with deep knowledge of systemd, process supervision; Apache, HAProxy
Strong shell plus one scripting language (bash/Python/PHP) with a track record of building deploy and rollback tooling, not just using it.
Demonstrated reverse-engineering ability: taking ownership of an undocumented production service and recovering its real dependencies and failure modes.
Message-queue literacy: Azure Service Bus, SQS, RabbitMQ, or equivalent — ordering, lock/ack semantics, dead-letter handling.
PostgreSQL operations: replicas/clones, connection proxying, network-restricted access, production diagnostics.
Working proficiency with Kubernetes and cloud networking — enough to operate private AKS clusters, an Istio-style ingress layer, and Azure firewall/DNS during the transition. We will onboard you on our specifics; deep specialization is not required.
Migration experience: consolidating or re-platforming production services with zero-downtime cutover and tested rollback.
A demonstrable record of reducing system surface area — services consolidated, infrastructure retired.
Clear written communication for runbooks, migration plans, and cross-functional coordination.
Strongly Preferred
Azure networking at landing-zone depth: hub-spoke VNets, Azure Firewall, Private Link/Private DNS, VPN gateways.
Administration of external-dns, cert-manager, and policy engines such as Kyverno.
ELK and OpenTelemetry pipeline operations.
Webhook-heavy integration experience (e-commerce platforms such as Shopify, subscription billing, marketing/notification platforms, or healthcare data exchanges).
Experience in a regulated or health-data environment: PHI handling, audit trails, least-privilege network design.
Terraform or equivalent IaC for cloud network and compute resources.
Enough AWS to manage cross-cloud seams (Route53, SQS, egress allowlists).
AI minded; proficient with Claude or similar.
First 30 days: Trace and fix a production issue end-to-end (DNS → firewall → ingress → service → database) with guidance; produce a true-dependency inventory for one candidate service.
First 90 days: First service migrated off AKS to the VM/systemd platform with a passing rollback drill; firewall rules, DNS zones, and cluster add-ons documented and reproducible; stale-resource inventory complete and retirement underway.
First 6 months: Migration cadence established with multiple services consolidated; release and rollback drills routine on both platforms; measurable reduction in infrastructure footprint and spend; security-baseline remediations closed.
Become part of a team that values your health, growth, and contributions to the health industry.
We believe that transparency in pay is essential to promoting fairness and accountability in our workplace. Therefore, we will provide candidates with the salary range for this position during the interview process. We also welcome and encourage candidates to discuss their compensation expectations with us during this process.
We are an equal opportunity employer and do not discriminate based on race, color, religion, sex, national origin, age, disability, or any other legally protected status. We value diversity, equity, and inclusion, and are committed to creating a workplace that reflects these values.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
$166k - $208k
...Anduril’s family of systems is powered by Lattice OS, an AI-powered operating system that turns thousands of data streams into a realtime, 3... ...and success of our customer accounts. Working across product, engineering, sales, and logistics teams, our Mission Operations Engineers...SeniorFull timeContract workWork experience placementImmediate start$160k - $200k
...in every domain. Umbra’s ecosystem operates through three business units: Remote... ...), and Mission Solutions (the platforms). Together, our teams develop capabilities... ...are looking to add a talented Senior Cyber Threat Operations Engineer to become a key player in our vibrant...SeniorPermanent employmentFull timeWork at officeLocal areaRemote workWorldwide$107.9k - $195.05k
Leidos is excited to present an opportunity for a TS/SCI‑cleared Platform Operations Engineer to join a high‑impact team driving the design, development, and deployment of a modern technology stack supporting the DOMEX Technology Platform (DTP). This role directly supports...SuggestedFull timeRemote workFlexible hours- Computer Technologies Consultants (CTC) is looking for a Senior Cybersecurity Operations Engineer to work onsite in Washington D.C. The role requires a minimum of six continuous years of experience and Public Trust clearance. You will conduct security assessments, develop...Senior
- GEICO in Bethesda, MD is seeking a Sr Staff Engineer to monitor security gaps, drive timely resolutions, and create visuals on the current state of security across CSIRT, GRC, and partner teams. You will organize security documentation, lead discussions on best practices...Senior
- DAn Solutions, Inc. is seeking an experienced MS Exchange Operations Engineer to administer and optimize enterprise-grade Exchange services within a secure DoD environment. The role requires strong collaboration with IT, cybersecurity, and AD teams to ensure secure messaging...Senior
- ...Job Description Job Title: Senior Security Operations Engineer Location: Washington, DC Note: This is an onsite position Place at NIGC... ...supporting cyber operations and maintaining operational security platforms across on-premises, hybrid, and cloud infrastructures....Senior
$103k - $116k
...Strawn LLP in Washington, DC is seeking an experienced IT Service Management professional. The role involves managing the Freshservice platform, designing workflows, and leading junior staff. Candidates should have a background in IT and service management methodologies....Senior- Stream Realty Partners seeks a Lead Building Engineer to oversee maintenance and repair activities in Washington, D.C. The role demands at least 5 years of experience in commercial building operations and proficiency in mechanical and electrical systems. Candidates should...Senior
- Job Posting TitleSenior Software Developer/Operations Engineer (CIS O&M)Job DescriptionJob Overview:Customer: Internal Revenue Services (IRS) Location: Remote, with occasional onsite support for meetings, deployments, production activities, or knowledge transfer as required...SeniorContract workFor contractorsFor subcontractorRemote workWorldwideRelocation package
- ...Senior Security Engineer (Security Operations) Sword Health is shifting healthcare from human-first to AI-first through its AI Care platform, making world-class healthcare available anytime, anywhere, while significantly reducing costs for payers, self-insured employers...SeniorFull timeRemote workFlexible hoursShift work
- ...for an innovative organization and the opportunity to learn and grow professionally? We can help! We are seeking a Senior Cybersecurity Operations Engineer to provide on-demand Cybersecurity and IT services to support the National Indian Gaming Commission (NIGC) mission...SeniorFull timePart time
$88.2k - $190.9k
Operations Engineer, Senior TS Clearance REQUIRED Position Description CGI Federal has an exciting opportunity for an Operations Engineer within our Intel sector advancing the national security mission through cutting edge technology. You must have a passion for...SeniorLocal area$140k - $165k
Job DescriptionEverforth ECS is seeking a Senior Databricks Platform Engineer to work in our Arlington, VA office (Hybrid).We are seeking a highly... ...building scalable infrastructure, governance frameworks, and operational excellence—with secondary responsibilities in data...SeniorWork at office$148.5k - $223.9k
...for an experienced and hands-on software engineer to join our team to build and scale the... ...of infrastructure management service and operation solutions. Your charter will be to... ..., frameworks, workflows, and validation platforms, applying AI tools, that help Salesforce...SeniorFull time$176k - $276k
...Infrastructure (GNI) organization. We deploy, integrate, and operate the Kubernetes-based platform and shared services used to provision, monitor, and... ...across environments.We are looking for a hands-on senior engineer to own the lifecycle and automation of the Kubernetes...SeniorFull timeRemote workWeekend work- ...A leading security solutions company is seeking a Senior Software Engineer to revolutionize offensive security through AI. The role involves developing innovative AI capabilities for vulnerability discovery and enhancing security workflows. Candidates should have over...SeniorRemote workFlexible hours1 day per week
- ...Peregrine is seeking a Senior Software Engineer to join our Platform team in Washington, DC. You will own moderately complex infrastructure problems, ship initiatives to improve performance, reliability, and security, and help make the platform developer-friendly. This...Senior
- ...Job Description Job Description **CONTINGENT UPON CONTRACT AWARD**Overview: Job Title: Security Operations Engineer – Senior Location : Washington, DC (Due to the nature of the work and contract requirements, U.S. Citizenship is required. ) Description:...SeniorContract work
$131k - $164k
...Help shape the technology that enables a global organisation to do its best work. As Senior Manager, Platform Engineering, you’ll lead the team responsible for Diligent’s Atlassian and Microsoft platforms while setting the architectural direction for the wider internal...SeniorFull timeWork at officeLocal areaWorldwideVisa sponsorshipFlexible hours$120k - $260k
...Great Careers. As a Sr Staff Engineer, you will: Monitor and track... ...Organize, store and manage operational best practices documentation... ...security solutions to protect our platforms including endpoint, cloud,... ...communicating and presentating to senior and junior staff with the...SeniorHourly payWork experience placementLocal area$120k - $165k
...The Position The Sr. Infrastructure Operations Engineer will work as part of a dynamic team to... ...support across all areas of the Windows platform, virtualization, storage, SQL Server, and... ...in Azure. This role serves as the senior technical escalation point for the Infrastructure...SeniorFull timeWork experience placementWork at office- EnDepth seeks a Senior Kubernetes Platform Systems Engineer to support U.S. Government programs in Laurel, MD. You will administer, automate, secure, monitor, and maintain Kubernetes platform infrastructure across development, staging, and production environments for mission...Senior
- ...USC - Interview - Video + Inperson Position: Sr. Platform Engineer (VMware & AWS GovCloud) Location: DC (Onsite) Requirement: Seeking a Senior Platform Engineer with 7+ years of experience in VMware, AWS GovCloud, and Kubernetes to support...Senior
- ...connecting and securing critical operations across the globe, keeping our... ...reliable, and scalable cloud platforms that enable mission teams to... ...an experienced Platform Engineer Sr Principal to lead the architecture... ...IMPACT We are seeking a senior Platform Engineering...SeniorContract work
- ZoomInfo is looking for a Senior Software Engineer to join their API & MCP Platform team in Bethesda, Maryland. This role involves designing and maintaining RESTful APIs using Java and NestJS, ensuring high performance and scalability. Ideal candidates have at least 5 years...Senior
- ...moderately complex components within platform services or SDKs; leads team-... ...AI team, you will build and operate cloud services that leverage... ...of experienced, hands-on engineers with the expertise and... ...critical applications. As a senior developer in Generative AI team...SeniorWorldwideFlexible hours
- EnDepth Solutions, LLC in Laurel, MD is seeking a Senior Kubernetes Platform Systems Engineer to administer, automate, secure, monitor, and maintain Kubernetes... ..., CI/CD pipelines, and security‑driven platform operations for government workloads. #J-18808-Ljbffr EnDepth...Senior
- ...Senior Veritas eDiscovery Platform (eDP) Engineer Employment Type: Full-Time, Executive-Level Department: Legal CGS is seeking a dedicated Senior Veritas eDiscovery Platform (eDP) Engineer to join a fast-paced and hard-working team to assist with any legal accounts...SeniorFull timeFor contractorsRemote workFlexible hours
$100.8k - $170k
SiriusXM is looking for a Senior Software Engineer to enhance our platform's tooling and infrastructure in Washington, DC. In this role, you will develop reusable platform components and streamline cloud management, directly collaborating with internal teams. The ideal...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Platform Operations Engineer, Infrastucture. Be the first to apply!
- senior platform engineer Washington DC
- platform developer Washington DC
- platform engineer Washington DC
- client platform engineer Washington DC
- platform engineering manager Washington DC
- data platform engineer Washington DC
- production network engineer Washington DC
- network operations center engineer Washington DC
- production operations engineer Washington DC
- application operations engineer Washington DC


