Senior Deployment and Site Reliability Engineer
$113.2k - $188.8kOpenkyber
Job Description Summary Grid Automation offers its customers a complete range of innovative products, systems and services covering design and manufacture, as well as commissioning and long-term maintenance of Substation Protection and Automation solutions. This includes, amongst others, Protection relays, control systems, cyber security solutions, device management, asset performance, and many more. Grid Automation Software operates at the boundaries of Grid technologies, embedded software and advanced analytics solutions. If you're passionate about Software and want to work on improving Grid resilience through Software, this is the job for you. We are looking for a Deployment and Site Reliability Engineer to join the GridBeats team. The GridBeats Software Portfolio aggregates all Grid Automation applications which monitor, optimize and manage grid connected devices. It is a very complex and exciting portfolio which is setting the basis for grid connected devices management for the future. The Deployment and Site Reliability Engineer takes GridBeats software into customer operation and keeps it running there. The role deploys, commissions, and upgrades our software at customer sites and remotely, and is the technical escalation point when something critical breaks in a live installation. It is the engineer customers and Regional Operations rely on when a deployment has to succeed inside a maintenance window, and when a production issue has to be understood and resolved rather than merely logged. The portfolio runs on Kubernetes, so deep Kubernetes skill is the central technical requirement for this position. This is a customer-facing, hands-on role, and it is not a product engineering role. The software is built by the product teams; this role gets it installed, proven, and supported in real substation and utility environments, and feeds what it learns in the field back into the products. Travel to customer sites, work inside planned maintenance windows, and availability for critical issues outside normal hours are an integral part of the job. Job Description Roles and Responsibilities:
- Deployment & Commissioning: Lead end-to-end installation, configuration, and commissioning of GridBeats software across diverse architectures (cloud, on-premise, hybrid, and air-gapped) while ensuring seamless technical handover to customers.
- System Upgrades & Recovery: Execute software upgrades, patching, and data migrations; maintain rigorous backup/restore protocols and rehearsed disaster recovery plans to ensure system integrity.
- Critical Issue Resolution: Serve as the technical escalation point for high-severity incidents, conducting deep-dive analysis of logs and network behavior to restore service and drive permanent resolutions with product teams.
- Staging & Validation: Prepare and manage staging and pre-production environments that mirror customer architectures to validate release packages, configurations, and upgrade paths prior to field deployment.
- Reliability & Monitoring: Implement proactive monitoring, logging, and alerting systems; analyze system health and resource consumption to recommend preventive maintenance and capacity planning.
- Automation & Efficiency: Utilize infrastructure-as-code and deployment automation to minimize manual effort, advocating for platform improvements that enhance field deployability and repeatability.
- Operational Documentation: Author and maintain critical technical documentation, including runbooks, installation guides, and troubleshooting procedures to support consistent service delivery.
- Knowledge Transfer: Provide training and technical support to Regional Operations, service teams, and customer personnel to ensure operational excellence across the installed base.
- Cross-Functional Advocacy: Champion the "field perspective" within product development, contributing insights on deployability and supportability to influence product design and future release quality.
Required Must Have Qualifications
- Education: Bachelor's or Master's degree in Software Engineering, Computer Science, or a related technical discipline.
- Systems Deployment: 5-8 years of experience deploying, commissioning, and supporting complex software systems in production or customer environments across cloud (Azure/AWS) and on-premise architectures.
- Kubernetes & Containerization: 5-8 years of deep expertise in full-stack Kubernetes operations (Helm, RBAC, Networking) and container/Docker proficiency (building, inspecting, and troubleshooting runtime behavior).
- Incident Management: 5-8 years of experience managing critical production issues end-to-end, including triage, service restoration, root cause analysis, and corrective action under customer-facing pressure.
- Technical Infrastructure: 5-8 years of experience with relational databases (e.g., PostgreSQL), networking/security configurations (TLS, firewalls, load balancing, VPNs), and scripting automation (Python, Bash).
Desired Qualifications
- Operational Technology (OT): Experience supporting software in industrial, substation, or utility environments, including familiarity with IEC 62443 standards and hardened or air-gapped systems.
- Travel & Availability: Ability to travel to customer sites (approx. 30% of time) and flexibility to support maintenance windows or critical issues outside normal business hours.
- Communication: Exceptional ability to communicate technical findings clearly to both technical and non-technical stakeholders, maintaining customer confidence during high-pressure incidents.
- Environment Management: Proven experience maintaining staging/pre-production environments integrated with release pipelines and artifact repositories (e.g., Artifactory).
- Advanced Orchestration: Familiarity with GitOps-based delivery (Argo CD or Flux CD) and service mesh technologies such as Istio.
- Service Management: Experience with structured incident/problem management practices (e.g., ITIL) and service management tools like Jira or ServiceNow.
- High Availability: Experience with disaster recovery architectures, database clustering, and connection pooling (e.g., pgpool).
- Continuous Improvement: A proactive mindset focused on turning field experiences into permanent improvements for products, deployment procedures, and automation.
- Global Collaboration: Ability to interface effectively within international, multi-cultural, and matrixed organizations to deliver cohesive technical solutions.
GE Vernova offers a great work environment, professional development, challenging careers, and competitive compensation. GE Vernova is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, national or ethnic origin, sex, sexual orientation, gender identity or expression, age, disability, protected veteran status or other characteristics protected by law. GE Vernova will only employ those who are legally authorized to work in the United States for this opening. Any offer of employment is conditioned upon the successful completion of a drug screen (as applicable). Relocation Assistance Provided: No #LI-Remote - This is a remote position For candidates applying to a U.S. based position, the pay range for this position is between $113,200.00 and $188,800.00. The Company pays a geographic differential of 110%, 120% or 130% of salary in certain areas. The specific pay offered may be influenced by a variety of factors, including the candidate's experience, education, and skill set. Bonus eligibility: discretionary annual bonus. This posting is expected to remain open for at least seven days after it was posted on September 25, 2026. Available benefits include medical, dental, vision, and prescription drug coverage; access to Health Coach from GE Vernova, a 24/7 nurse-based resource; and access to the Employee Assistance Program, providing 24/7 confidential assessment, counseling and referral services. Retirement benefits include the GE Vernova Retirement Savings Plan, a tax-advantaged 401(k) savings opportunity with company matching contributions and company retirement contributions, as well as access to Fidelity resources and financial planning consultants. Other benefits include tuition assistance, adoption assistance, paid parental leave, disability benefits, life insurance, 12 paid holidays, and permissive time off. GE Vernova Inc. or its affiliates (collectively or individually, "OpenKyber") sponsor certain employee benefit plans or programs OpenKyber reserves the right to terminate, amend, suspend, replace, or modify its benefit plans and programs at any time and for any reason, in its sole discretion. No individual has a vested right to any benefit under a OpenKyber welfare benefit plan or program. This document does not create a contract of employment with any individual.
For applications and inquiries, contact:View email address on us.fitly.work
- ...Site Reliability Engineer We're looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability,... ...proactive capacity planning Own and evolve CI/CD pipelines, deployment automation, rollback mechanisms, and config management...Senior
- ...seeking an experienced AWS solution design engineer/architect to join our infrastructure... ...and confidently them into production.As Senior SRE, you will be responsible for providing... ...best practices of architecture, review deployment architecture and ensure that costs are managed...Senior
- ...Lead Engineer, Site Reliability Engineering Team As a lead engineer with Retail, Site Reliability Engineering team, you will be at the forefront... ...lifecycle of services from inception and design through deployment, operation, and refinement. Support capacity planning,...Senior
- ...Senior Site Reliability Engineer Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient,... ...reduce human intervention Improve CI/CD pipelines and deployment safety (canary, rollback, blue-green) Support Infrastructure...SeniorWorldwide
$168k - $200k
...healthcare. What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the... ...services. Drive CI/CD for Data & ML : Automate deployments of data pipelines, ML workflows, and infra components using...SeniorRemote work$136.2k - $214.01k
...Exceptional in execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of... ...infrastructure powering the Proofpoint services, including deployment, maintenance, troubleshooting, performance tuning, and...SeniorFull timeFlexible hours- ...As the Senior Site Reliability Engineer, you will serve as a trusted technical resource responsible for deploying, validating, and operationalizing AI, HPC, Kubernetes, and enterprise infrastructure environments. This role transforms newly installed hardware into production...SeniorWork at officeImmediate startWorldwide
$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The... ...analyzing recurring failures, building automation, supporting deployments, and contributing to capacity planning, disaster recovery,...SeniorTemporary workImmediate startFlexible hoursShift work$178.13k - $205.4k
...Bachelor’s degree or foreign degree equivalent in Computer Engineering, Computer Science, Engineering, or related field plus five (5... ...containerization with Docker and Kubernetes. 5+ years experience in deployment automation and monitoring solutions. 5+ years experience in...SeniorWork at officeRemote workFlexible hours$123.4k - $222.53k
...journey at T-Mobile? Our team is searching for our next Sr. Site Reliability Engineer to strengthen the reliability and resilience of the... ...Automate processes to accelerate software development and deployment while minimizing manual interventions Conduct root cause...SeniorFull timeTemporary workPart timeWork experience placementLocal areaFlexible hours- ...customers in more than 35 countries worldwide.Site Reliability EngineerOnsite: Atlanta, GAJob... ...we're looking for a Site Reliability Engineer II to help build, support, and scale the... ...infrastructure, application, and deployment issues in partnership with engineering...Full timeWorldwideFlexible hours
- ...We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production systems hosted on AWS. The role focuses on production... .../Unix environments. Support CI/CD pipelines and AWS deployment processes. Apply reliability patterns and continuously...Contract work
$169.3k - $304.7k
...maintaining fast, efficient, scalable, and reliable routing software and infrastructure... ...global platform. As a Principal Site Reliability Engineer - Network, you will be responsible... ...teams on standards and processes for deployments and updating documentation accordingly...Work experience placementWork at office$100k - $120k
...OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and... ...architecture enables AI-driven bundling, automation, and real-time deployment.Solutions from Origami Risk and Dais Technology are backed...Full timeTemporary workWork experience placementFlexible hours- ...review the following job description:The Site Reliability Engineering Lead role focuses on enhancing the... ...and on-premises environments. This senior technical leader drives improvements... ...software development lifecycle, testing, deployment, and security practices.Preferred...Permanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
- ...such as public cloud, data science, AI, engineering innovation, and IoT. Our customers... ...profitable, and growing. We are hiring a Site Reliability Engineer Our goal is to perfect... ...: Globally remote role The role We deploy and run OpenStack, Kubernetes, storage...Work at officeLocal areaRemote workWork from homeWorldwide
- ...Intercontinental Exchange (NYSE:ICE), we engineer technology, exchanges and clearing... ...to join our team. We are seeking a Site Reliability Engineer to bring 3+ years of hands-on... ...Product and Engineering teams to plan and deploy product releases with operational rigor...
- ...cloud-native systems. As a Staff Platform Engineer, you will play a critical role in... ...technical leadership role. You will own reliability for major platform domains, design scalable... ...teams will integrate with and leverage to deploy and operate their applications Architect...Senior
- ...integration, production and automate deployment for production • Develop integrations... ...monitoring and alerting metrics so the support engineers can proactively and timely validate,... ...components. • 1+ Years in Site Reliability Engineering organization preferred •...Work experience placement
$75.7k - $136.3k
...and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages... ...CD. You will design, develop, and operate infrastructure deployment for the Akamai Cloud. As a Site Reliability Engineer,...Work experience placementWork at office$130k - $145k
...Back Site Reliability Engineer Cloud/Infrastructure Atlanta , GA Sep 2, 2026 Site Reliability Engineer Atlanta, GA / Hybrid Blu Omega... ...modules, manage Terraform state, and troubleshoot deployment failures Support the reliability, performance, scalability...Temporary work- ...Application/Infrastructure Performance, and availability. • Automation Experience with Build/deployment, Software Configuration/Continuous Integration/Continuous Delivery/Release Engineering related tasks in JavaEE/C++ Environments. • Experience in automating manual...Immediate start
$60 - $68 per hour
...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with long-term... ...Continuous Integration (CI), Continuous Delivery and Continuous Deployment (CD) pipeline; Ability to design, build, implement and...Contract workLocal areaImmediate start- ...Site Reliability Engineering (SRE) Architect Location: Atlanta, GA Duration: 12Months+ Extension... ..., and blueprints for service design, deployment, monitoring, and operational... ...Leadership & Consultation: Act as a senior technical advisor and subject matter...Hourly payPermanent employmentContract workLocal areaEarly shift
$145k - $160k
...We are seeking a specialized Observability & Infrastructure Engineer to lead the monitoring, telemetry, and platform-as-code initiatives... ...Responsibilities Disaster Recovery Observability: Design, deploy, and validate monitoring and telemetry strategies supporting...Temporary workRemote workFlexible hours- ...effectively. The Lead Systems Engineer is a senior individual contributor responsible for the reliability, scalability, and... ...direction for how our platform runs, deploys, and scales. This is a highly... ...DevOps, platform engineering, site reliability engineering, or a...Remote workFlexible hours
- ...the way our development teams build and deploy applications. Our mission is to create an... ...of our Kubernetes Development team, the Senior Developer will work hand-in-hand with... ...This individual will collaborate across engineering, development, and operations teams to ensure...SeniorWork experience placement
- ...our talented Team.Job Title: Senior Software EngineerLocation(s):... ...are seeking a Senior Software Engineer to join our agile development... ..., develop, and deploy robust, scalable software solutions... ...ensure software quality and reliability.Support functional testing efforts...SeniorWork at officeShift work3 days per week
$120k - $175k
...together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering... ...in languages such as python, ruby, or go. ~ Deploying and supporting applications in Kubernetes at scale....SeniorFull timeRemote workWork visaFlexible hours- ...Role: The Mission Systems Engineering team develops the Mission Management... ...and ground systems. As a Senior/Principal Mission Software... ...involving performance, reliability, latency, fault tolerance, and... ..., integration, testing, and deployment Preferred Skills and...SeniorWeekly payPermanent employmentWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Deployment and Site Reliability Engineer. Be the first to apply!
- site reliability engineer Atlanta, GA
- site reliability engineer remote Atlanta, GA
- site reliability engineer sre Atlanta, GA
- senior living director Atlanta, GA
- senior php developer remote Atlanta, GA
- senior manager customer operations Atlanta, GA
- senior support engineer Atlanta, GA
- senior product manager mobile Atlanta, GA
- senior java developer Atlanta, GA
- senior software engineer ruby on rails Atlanta, GA




