Senior Deployment and Site Reliability Engineer
Openkyber
Job Description Summary Grid Automation offers its customers a complete range of innovative products, systems and services covering design and manufacture, as well as commissioning and long-term maintenance of Substation Protection and Automation solutions. This includes, amongst others, Protection relays, control systems, cyber security solutions, device management, asset performance, and many more. Grid Automation Software operates at the boundaries of Grid technologies, embedded software and advanced analytics solutions. If you're passionate about Software and want to work on improving Grid resilience through Software, this is the job for you.
We are looking for a Deployment and Site Reliability Engineer to join the GridBeats team. The GridBeats Software Portfolio aggregates all Grid Automation applications which monitor, optimize and manage grid connected devices. It is a very complex and exciting portfolio which is setting the basis for grid connected devices management for the future. The Deployment and Site Reliability Engineer takes GridBeats software into customer operation and keeps it running there. The role deploys, commissions, and upgrades our software at customer sites and remotely, and is the technical escalation point when something critical breaks in a live installation. It is the engineer customers and Regional Operations rely on when a deployment has to succeed inside a maintenance window, and when a production issue has to be understood and resolved rather than merely logged. The portfolio runs on Kubernetes, so deep Kubernetes skill is the central technical requirement for this position. This is a customer-facing, hands-on role, and it is not a product engineering role. The software is built by the product teams; this role gets it installed, proven, and supported in real substation and utility environments, and feeds what it learns in the field back into the products.
Travel to customer sites, work inside planned maintenance windows, and availability for critical issues outside normal hours are an integral part of the job.
Job Description Roles and Responsibilities- Deployment & Commissioning: Lead end-to-end installation, configuration, and commissioning of GridBeats software across diverse architectures (cloud, on-premise, hybrid, and air-gapped) while ensuring seamless technical handover to customers.
- System Upgrades & Recovery: Execute software upgrades, patching, and data migrations; maintain rigorous backup/restore protocols and rehearsed disaster recovery plans to ensure system integrity.
- Critical Issue Resolution: Serve as the technical escalation point for high-severity incidents, conducting deep-dive analysis of logs and network behavior to restore service and drive permanent resolutions with product teams.
- Staging & Validation: Prepare and manage staging and pre-production environments that mirror customer architectures to validate release packages, configurations, and upgrade paths prior to field deployment.
- Reliability & Monitoring: Implement proactive monitoring, logging, and alerting systems; analyze system health and resource consumption to recommend preventive maintenance and capacity planning.
- Automation & Efficiency: Utilize infrastructure-as-code and deployment automation to minimize manual effort, advocating for platform improvements that enhance field deployability and repeatability.
- Operational Documentation: Author and maintain critical technical documentation, including runbooks, installation guides, and troubleshooting procedures to support consistent service delivery.
- Knowledge Transfer: Provide training and technical support to Regional Operations, service teams, and customer personnel to ensure operational excellence across the installed base.
- Cross-Functional Advocacy: Champion the "field perspective" within product development, contributing insights on deployability and supportability to influence product design and future release quality.
The successful candidate will sit onsite at OpenKyber Electrification Lab 511N John Rodes Blvd Melbourne Fl, 32934 No relocation assistance is provided for this role.
Required Must Have Qualifications
- Education: Bachelor's or Master's degree in Software Engineering, Computer Science, or a related technical discipline.
- Systems Deployment: 5-8 years of experience deploying, commissioning, and supporting complex software systems in production or customer environments across cloud (Azure/AWS) and on-premise architectures.
- Kubernetes & Containerization: 5-8 years of deep expertise in full-stack Kubernetes operations (Helm, RBAC, Networking) and container/Docker proficiency (building, inspecting, and troubleshooting runtime behavior).
- Incident Management: 5-8 years of experience managing critical production issues end-to-end, including triage, service restoration, root cause analysis, and corrective action under customer-facing pressure.
- Technical Infrastructure: 5-8 years of experience with relational databases (e.g., PostgreSQL), networking/security configurations (TLS, firewalls, load balancing, VPNs), and scripting automation (Python, Bash).
Desired Qualifications
- Operational Technology (OT): Experience supporting software in industrial, substation, or utility environments, including familiarity with IEC 62443 standards and hardened or air-gapped systems.
- Travel & Availability: Ability to travel to customer sites (approx. 30% of time) and flexibility to support maintenance windows or critical issues outside normal business hours.
- Communication: Exceptional ability to communicate technical findings clearly to both technical and non-technical stakeholders, maintaining customer confidence during high-pressure incidents.
- Environment Management: Proven experience maintaining staging/pre-production environments integrated with release pipelines and artifact repositories (e.g., Artifactory).
- Advanced Orchestration: Familiarity with GitOps-based delivery (Argo CD or Flux CD) and service mesh technologies such as Istio.
- Service Management: Experience with structured incident/problem management practices (e.g., ITIL) and service management tools like Jira or ServiceNow.
- High Availability: Experience with disaster recovery architectures, database clustering, and connection pooling (e.g., pgpool).
- Continuous Improvement: A proactive mindset focused on turning field experiences into permanent improvements for products, deployment procedures, and automation.
- Global Collaboration: Ability to interface effectively within international, multi-cultural, and matrixed organizations to deliver cohesive technical solutions.
Additional Information OpenKyber offers a great work environment, professional development, challenging careers, and competitive compensation. OpenKyber is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, national or ethnic origin, sex, sexual orientation, gender identity or expression, age, disability, protected veteran status or other characteristics protected by law. OpenKyber will only employ those who are legally authorized to work in the United States for this opening. Any offer of employment is conditioned upon the successful completion of a drug screen (as applicable). Relocation Assistance Provided: No
For applications and inquiries, contact:View email address on us.fitly.work
- ...seeking an experienced AWS solution design engineer/architect to join our infrastructure... ...and confidently them into production.As Senior SRE, you will be responsible for providing... ...best practices of architecture, review deployment architecture and ensure that costs are managed...Senior
- ...Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high... ...reduce human intervention Improve CI/CD pipelines and deployment safety (canary, rollback, blue-green) Support...SeniorWorldwide
- ...Site Reliability Engineer We're looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability,... ...proactive capacity planning Own and evolve CI/CD pipelines, deployment automation, rollback mechanisms, and config management...Senior
$136.2k - $214.01k
...Exceptional in execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of... ...infrastructure powering the Proofpoint services, including deployment, maintenance, troubleshooting, performance tuning, and...SeniorFull timeFlexible hours- ...Lead Engineer, Site Reliability Engineering Team As a lead engineer with Retail, Site Reliability Engineering team, you will be at the forefront... ...lifecycle of services from inception and design through deployment, operation, and refinement. Support capacity planning,...Senior
$168k - $200k
...healthcare. What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the... ...services. Drive CI/CD for Data & ML : Automate deployments of data pipelines, ML workflows, and infra components using...Senior$109.5k
...highly motivated, diligent, and skillful Site Reliability Engineer to join the Cyber Security... ...function, is responsible for designing, deploying, maintaining, and optimizing the tool... ...be remote anywhere in the U.S. The Senior Site Reliability Engineer will be responsible...SeniorTemporary workLocal areaRemote work- ...AbbVie Information Security seeks a Senior Site Reliability Engineer for the Cyber Security Engineering team to design, deploy, and maintain security infrastructure supporting cybersecurity operations. This remote U.S. position focuses on ensuring production system reliability...SeniorRemote work
$146.4k - $263.6k
...diverse multi-national team of engineering talents? Join our highly skilled Site Reliability team Our team designs, develops... ...be responsible for: As a Senior II Site Reliability Engineer -... ...support development, testing, and deployment workflows. Defining SLOs for...SeniorWork experience placementWork at office- ...rate for C2C/1099/W2. Job Description: Job Title : Sr. Site Reliability Engineer Location : Atlanta, GA - Hybrid Duration : 6+ Months... ...Extensive/Strong AWS experience: experience in designing, deploying managing scalable/reliable cloud-based infrastructure...SeniorContract workLocal areaImmediate start
$178.13k - $205.4k
...~Bachelor’s degree or foreign degree equivalent in Computer Engineering, Computer Science, Engineering, or related field plus five (5... ...Kubernetes. ~ Five (5) years (60 months) of experience in deployment automation and monitoring solutions. ~ Five (5) years (60 months...SeniorWork at officeRemote workFlexible hours$123.4k - $222.53k
...Responsibilities Enhance system reliability and resilience by identifying issues and implementing... ...to accelerate software development and deployment while minimizing manual interventions... ...of study include Computer Science, Engineering or related field (Required) ~4-7 years...SeniorFull timeTemporary workPart timeWork experience placementLocal areaFlexible hours- ...Intercontinental Exchange (NYSE:ICE), we engineer technology, exchanges and clearing... ...people to join our team.We are seeking a Site Reliability Engineer to bring 3+ years of hands-on... ...Product and Engineering teams to plan and deploy product releases with operational rigor...
$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and... ...architecture enables AI-driven bundling, automation, and real-time deployment. Solutions from Origami Risk and Dais Technology are...Full timeTemporary workWork experience placementFlexible hours- ...We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production systems hosted on AWS. The role focuses on production... .../Unix environments. Support CI/CD pipelines and AWS deployment processes. Apply reliability patterns and continuously...Contract work
- ...customers in more than 35 countries worldwide.Site Reliability EngineerOnsite: Atlanta, GAJob... ...we're looking for a Site Reliability Engineer II to help build, support, and scale the... ...infrastructure, application, and deployment issues in partnership with engineering...Full timeWorldwideFlexible hours
- ...review the following job description:The Site Reliability Engineer role focuses on enhancing the... ...cloud and on-premises environments. This senior technical leader drives improvements... ...software development lifecycle, testing, deployment, and security practices.Preferred...Permanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
- ...AI Solutions Engineer Lead Senior Location: This role requires associates to be in-office... ...planning, design, implementation, testing, deployment, and production support). The ideal... ...able to articulate the scaling and reliability limits. Owns the change request...SeniorWork at office3 days per week
$169.3k - $304.7k
...maintaining fast, efficient, scalable, and reliable routing software and infrastructure... ...global platform. As a Principal Site Reliability Engineer - Network, you will be responsible... ...teams on standards and processes for deployments and updating documentation accordingly...Work experience placementWork at office$130k - $150k
...We are seeking a highly skilled Site Reliability Engineer (SRE) to join our team and help build... ...Contribute to CI/CD pipelines and deployment automation Create and maintain documentation... ...lifecycle and CI/CD principles Seniority level ~ Seniority level Mid-...Full timeRemote work- ...Cloud gives AI teams the infrastructure they need to build, deploy, and scale on one unified cloud. One cloud for compute... ...agents. Role Overview We are seeking a skilled Site Reliability Engineer to join the GMI Global Infrastructure team. This role is hands...
- ...Job Title :- Site Reliability Engineer (SRE) Employment Type :- W2 Duration :- Long Term Visa Type :- All Visa applicable which are ready... ...AEM applications while ensuring seamless integrations and deployments. Key Responsibilities: AEM Reliability & Performance...
- ...Join to apply for the Site Reliability Engineer role at Motion Recruitment Join to apply for... ...Experience developing for data streaming, deploying/monitoring high availability critical... .... Posted By: VMS Sourcing Seniority level ~ Seniority level Mid-Senior...Contract workWorldwide
- ...cloud-native systems. As a Staff Platform Engineer, you will play a critical role in... ...technical leadership role. You will own reliability for major platform domains, design scalable... ...teams will integrate with and leverage to deploy and operate their applications Architect...Senior
- ...OpenShift - Site Reliability Engineer Atlanta , GA / Onsite Qualifications: This position is 60 % SRE and 40% SDE.... ...automate delivery to system integration, production and automate deployment for production • Develop integrations between the...Work experience placement
- ...re Looking For We’re looking for a proactive, hands‑on Site Reliability Engineer who thrives in building and scaling cloud infrastructure in... ...modern CI/CD pipelines with GitLab to enable fast, safe deployments Implementing and evolving monitoring, alerting, and observability...Work experience placementFlexible hours
- ...Application/Infrastructure Performance, and availability. • Automation Experience with Build/deployment, Software Configuration/Continuous Integration/Continuous Delivery/Release Engineering related tasks in JavaEE/C++ Environments. • Experience in automating manual...Immediate start
$71.6k - $119.4k
...Our services provide applications with reliability, security, and better customer... ...applications with provisioning needs, deployment support, and security improvements. You... ...troubleshoot issues, and work closely with senior engineers to learn and apply best practices. You...Full timeTemporary workInternshipLocal areaWork from home$95k - $171k
...Technology Group. We design, implement, deploy and operate AI platforms that enable... ...infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for:...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$60 - $68 per hour
...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with long-term... ...Continuous Integration (CI), Continuous Delivery and Continuous Deployment (CD) pipeline; Ability to design, build, implement and...Contract workLocal areaImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Deployment and Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre Atlanta, GA
- site reliability engineer Atlanta, GA
- site reliability engineer remote Atlanta, GA
- senior service associate Atlanta, GA
- senior safety specialist Atlanta, GA
- senior vice president of business development Atlanta, GA
- senior service designer Atlanta, GA
- senior sales recruiter Atlanta, GA
- senior mulesoft developer Atlanta, GA
- senior media manager Atlanta, GA





