Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Deployment and Site Reliability Engineer

Openkyber

Job Description Summary Grid Automation offers its customers a complete range of innovative products, systems and services covering design and manufacture, as well as commissioning and long-term maintenance of Substation Protection and Automation solutions. This includes, amongst others, Protection relays, control systems, cyber security solutions, device management, asset performance, and many more. Grid Automation Software operates at the boundaries of Grid technologies, embedded software and advanced analytics solutions. If you're passionate about Software and want to work on improving Grid resilience through Software, this is the job for you.

We are looking for a Deployment and Site Reliability Engineer to join the GridBeats team. The GridBeats Software Portfolio aggregates all Grid Automation applications which monitor, optimize and manage grid connected devices. It is a very complex and exciting portfolio which is setting the basis for grid connected devices management for the future. The Deployment and Site Reliability Engineer takes GridBeats software into customer operation and keeps it running there. The role deploys, commissions, and upgrades our software at customer sites and remotely, and is the technical escalation point when something critical breaks in a live installation. It is the engineer customers and Regional Operations rely on when a deployment has to succeed inside a maintenance window, and when a production issue has to be understood and resolved rather than merely logged. The portfolio runs on Kubernetes, so deep Kubernetes skill is the central technical requirement for this position. This is a customer-facing, hands-on role, and it is not a product engineering role. The software is built by the product teams; this role gets it installed, proven, and supported in real substation and utility environments, and feeds what it learns in the field back into the products.

Travel to customer sites, work inside planned maintenance windows, and availability for critical issues outside normal hours are an integral part of the job.

Job Description Roles and Responsibilities
  • Deployment & Commissioning: Lead end-to-end installation, configuration, and commissioning of GridBeats software across diverse architectures (cloud, on-premise, hybrid, and air-gapped) while ensuring seamless technical handover to customers.
  • System Upgrades & Recovery: Execute software upgrades, patching, and data migrations; maintain rigorous backup/restore protocols and rehearsed disaster recovery plans to ensure system integrity.
  • Critical Issue Resolution: Serve as the technical escalation point for high-severity incidents, conducting deep-dive analysis of logs and network behavior to restore service and drive permanent resolutions with product teams.
  • Staging & Validation: Prepare and manage staging and pre-production environments that mirror customer architectures to validate release packages, configurations, and upgrade paths prior to field deployment.
  • Reliability & Monitoring: Implement proactive monitoring, logging, and alerting systems; analyze system health and resource consumption to recommend preventive maintenance and capacity planning.
  • Automation & Efficiency: Utilize infrastructure-as-code and deployment automation to minimize manual effort, advocating for platform improvements that enhance field deployability and repeatability.
  • Operational Documentation: Author and maintain critical technical documentation, including runbooks, installation guides, and troubleshooting procedures to support consistent service delivery.
  • Knowledge Transfer: Provide training and technical support to Regional Operations, service teams, and customer personnel to ensure operational excellence across the installed base.
  • Cross-Functional Advocacy: Champion the "field perspective" within product development, contributing insights on deployability and supportability to influence product design and future release quality.

The successful candidate will sit onsite at OpenKyber Electrification Lab 511N John Rodes Blvd Melbourne Fl, 32934 No relocation assistance is provided for this role.

Required Must Have Qualifications

  • Education: Bachelor's or Master's degree in Software Engineering, Computer Science, or a related technical discipline.
  • Systems Deployment: 5-8 years of experience deploying, commissioning, and supporting complex software systems in production or customer environments across cloud (Azure/AWS) and on-premise architectures.
  • Kubernetes & Containerization: 5-8 years of deep expertise in full-stack Kubernetes operations (Helm, RBAC, Networking) and container/Docker proficiency (building, inspecting, and troubleshooting runtime behavior).
  • Incident Management: 5-8 years of experience managing critical production issues end-to-end, including triage, service restoration, root cause analysis, and corrective action under customer-facing pressure.
  • Technical Infrastructure: 5-8 years of experience with relational databases (e.g., PostgreSQL), networking/security configurations (TLS, firewalls, load balancing, VPNs), and scripting automation (Python, Bash).

Desired Qualifications

  • Operational Technology (OT): Experience supporting software in industrial, substation, or utility environments, including familiarity with IEC 62443 standards and hardened or air-gapped systems.
  • Travel & Availability: Ability to travel to customer sites (approx. 30% of time) and flexibility to support maintenance windows or critical issues outside normal business hours.
  • Communication: Exceptional ability to communicate technical findings clearly to both technical and non-technical stakeholders, maintaining customer confidence during high-pressure incidents.
  • Environment Management: Proven experience maintaining staging/pre-production environments integrated with release pipelines and artifact repositories (e.g., Artifactory).
  • Advanced Orchestration: Familiarity with GitOps-based delivery (Argo CD or Flux CD) and service mesh technologies such as Istio.
  • Service Management: Experience with structured incident/problem management practices (e.g., ITIL) and service management tools like Jira or ServiceNow.
  • High Availability: Experience with disaster recovery architectures, database clustering, and connection pooling (e.g., pgpool).
  • Continuous Improvement: A proactive mindset focused on turning field experiences into permanent improvements for products, deployment procedures, and automation.
  • Global Collaboration: Ability to interface effectively within international, multi-cultural, and matrixed organizations to deliver cohesive technical solutions.

Additional Information OpenKyber offers a great work environment, professional development, challenging careers, and competitive compensation. OpenKyber is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, national or ethnic origin, sex, sexual orientation, gender identity or expression, age, disability, protected veteran status or other characteristics protected by law. OpenKyber will only employ those who are legally authorized to work in the United States for this opening. Any offer of employment is conditioned upon the successful completion of a drug screen (as applicable). Relocation Assistance Provided: No

For applications and inquiries, contact:View email address on us.fitly.work

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior Deployment and Site Reliability Engineer in Atlanta, GA vacancy
  •  ...seeking an experienced AWS solution design engineer/architect to join our infrastructure...  ...and confidently them into production.As Senior SRE, you will be responsible for providing...  ...best practices of architecture, review deployment architecture and ensure that costs are managed... 
    Senior

    Black Knight Financial Services

    Atlanta, GA
    6 hours ago
  •  ...Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high...  ...reduce human intervention Improve CI/CD pipelines and deployment safety (canary, rollback, blue-green) Support... 
    Senior
    Worldwide

    Inspire Brands Inc

    Atlanta, GA
    1 day ago
  •  ...Site Reliability Engineer We're looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability,...  ...proactive capacity planning Own and evolve CI/CD pipelines, deployment automation, rollback mechanisms, and config management... 
    Senior

    Alembic Technologies

    Atlanta, GA
    3 days ago
  • $136.2k - $214.01k

     ...Exceptional in execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of...  ...infrastructure powering the Proofpoint services, including deployment, maintenance, troubleshooting, performance tuning, and... 
    Senior
    Full time
    Flexible hours

    Proofpoint

    Atlanta, GA
    1 day ago
  •  ...Lead Engineer, Site Reliability Engineering Team As a lead engineer with Retail, Site Reliability Engineering team, you will be at the forefront...  ...lifecycle of services from inception and design through deployment, operation, and refinement. Support capacity planning,... 
    Senior

    Next Level Business Services, Inc.

    Atlanta, GA
    14 hours ago
  • $168k - $200k

     ...healthcare. What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the...  ...services. Drive CI/CD for Data & ML : Automate deployments of data pipelines, ML workflows, and infra components using... 
    Senior

    Datavant

    Atlanta, GA
    2 days ago
  • $109.5k

     ...highly motivated, diligent, and skillful Site Reliability Engineer to join the Cyber Security...  ...function, is responsible for designing, deploying, maintaining, and optimizing the tool...  ...be remote anywhere in the U.S. The Senior Site Reliability Engineer will be responsible... 
    Senior
    Temporary work
    Local area
    Remote work

    AbbVie

    Atlanta, GA
    1 day ago
  •  ...AbbVie Information Security seeks a Senior Site Reliability Engineer for the Cyber Security Engineering team to design, deploy, and maintain security infrastructure supporting cybersecurity operations. This remote U.S. position focuses on ensuring production system reliability... 
    Senior
    Remote work

    AbbVie

    Atlanta, GA
    2 days ago
  • $146.4k - $263.6k

     ...diverse multi-national team of engineering talents? Join our highly skilled Site Reliability team Our team designs, develops...  ...be responsible for: As a Senior II Site Reliability Engineer -...  ...support development, testing, and deployment workflows. Defining SLOs for... 
    Senior
    Work experience placement
    Work at office

    Jobleads-US

    Atlanta, GA
    2 days ago
  •  ...rate for C2C/1099/W2. Job Description: Job Title : Sr. Site Reliability Engineer Location : Atlanta, GA - Hybrid Duration : 6+ Months...  ...Extensive/Strong AWS experience: experience in designing, deploying managing scalable/reliable cloud-based infrastructure... 
    Senior
    Contract work
    Local area
    Immediate start

    Navtech

    Atlanta, GA
    2 days ago
  • $178.13k - $205.4k

     ...~​Bachelor’s degree or foreign degree equivalent in Computer Engineering, Computer Science, Engineering, or related field plus five (5...  ...Kubernetes. ~ Five (5) years (60 months) of experience in deployment automation and monitoring solutions. ~ Five (5) years (60 months... 
    Senior
    Work at office
    Remote work
    Flexible hours

    Workday

    Atlanta, GA
    3 days ago
  • $123.4k - $222.53k

     ...Responsibilities Enhance system reliability and resilience by identifying issues and implementing...  ...to accelerate software development and deployment while minimizing manual interventions...  ...of study include Computer Science, Engineering or related field (Required) ~4-7 years... 
    Senior
    Full time
    Temporary work
    Part time
    Work experience placement
    Local area
    Flexible hours

    T-Mobile

    Atlanta, GA
    1 day ago
  •  ...Intercontinental Exchange (NYSE:ICE), we engineer technology, exchanges and clearing...  ...people to join our team.We are seeking a Site Reliability Engineer to bring 3+ years of hands-on...  ...Product and Engineering teams to plan and deploy product releases with operational rigor... 

    Black Knight Financial Services

    Atlanta, GA
    4 days ago
  • $100k - $120k

    OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and...  ...architecture enables AI-driven bundling, automation, and real-time deployment. Solutions from Origami Risk and Dais Technology are... 
    Full time
    Temporary work
    Work experience placement
    Flexible hours

    Origami Risk

    Atlanta, GA
    3 days ago
  •  ...We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production systems hosted on AWS. The role focuses on production...  .../Unix environments. Support CI/CD pipelines and AWS deployment processes. Apply reliability patterns and continuously... 
    Contract work

    2T Consulting

    Atlanta, GA
    a month ago
  •  ...customers in more than 35 countries worldwide.Site Reliability EngineerOnsite: Atlanta, GAJob...  ...we're looking for a Site Reliability Engineer II to help build, support, and scale the...  ...infrastructure, application, and deployment issues in partnership with engineering... 
    Full time
    Worldwide
    Flexible hours

    NCR

    Atlanta, GA
    4 days ago
  •  ...review the following job description:The Site Reliability Engineer role focuses on enhancing the...  ...cloud and on-premises environments. This senior technical leader drives improvements...  ...software development lifecycle, testing, deployment, and security practices.Preferred... 
    Permanent employment
    Full time
    Part time
    H1b
    Work at office
    Local area
    Immediate start
    Work visa
    Monday to Friday
    Shift work
    Day shift

    Truist

    Atlanta, GA
    4 days ago
  •  ...AI Solutions Engineer Lead Senior Location: This role requires associates to be in-office...  ...planning, design, implementation, testing, deployment, and production support). The ideal...  ...able to articulate the scaling and reliability limits. Owns the change request... 
    Senior
    Work at office
    3 days per week

    Elevance Health

    Atlanta, GA
    4 days ago
  • $169.3k - $304.7k

     ...maintaining fast, efficient, scalable, and reliable routing software and infrastructure...  ...global platform. As a Principal Site Reliability Engineer - Network, you will be responsible...  ...teams on standards and processes for deployments and updating documentation accordingly... 
    Work experience placement
    Work at office

    Akamai

    Atlanta, GA
    3 days ago
  • $130k - $150k

     ...We are seeking a highly skilled Site Reliability Engineer (SRE) to join our team and help build...  ...Contribute to CI/CD pipelines and deployment automation Create and maintain documentation...  ...lifecycle and CI/CD principles Seniority level ~ Seniority level Mid-... 
    Full time
    Remote work

    Prestige Staffing

    Atlanta, GA
    3 days ago
  •  ...Cloud gives AI teams the infrastructure they need to build, deploy, and scale on one unified cloud. One cloud for compute...  ...agents. Role Overview We are seeking a skilled Site Reliability Engineer to join the GMI Global Infrastructure team. This role is hands... 

    GMI Cloud

    Atlanta, GA
    20 hours ago
  •  ...Job Title :- Site Reliability Engineer (SRE) Employment Type :- W2 Duration :- Long Term Visa Type :- All Visa applicable which are ready...  ...AEM applications while ensuring seamless integrations and deployments. Key Responsibilities: AEM Reliability & Performance... 

    Highbrow

    Atlanta, GA
    14 hours ago
  •  ...Join to apply for the Site Reliability Engineer role at Motion Recruitment Join to apply for...  ...Experience developing for data streaming, deploying/monitoring high availability critical...  .... Posted By: VMS Sourcing Seniority level ~ Seniority level Mid-Senior... 
    Contract work
    Worldwide

    Motion Recruitment

    Atlanta, GA
    2 days ago
  •  ...cloud-native systems. As a Staff Platform Engineer, you will play a critical role in...  ...technical leadership role. You will own reliability for major platform domains, design scalable...  ...teams will integrate with and leverage to deploy and operate their applications   Architect... 
    Senior

    Saviynt

    Atlanta, GA
    3 days ago
  •  ...OpenShift - Site Reliability Engineer Atlanta , GA / Onsite Qualifications: This position is 60 % SRE and 40% SDE....  ...automate delivery to system integration, production and automate deployment for production • Develop integrations between the... 
    Work experience placement

    Fisec Global

    Atlanta, GA
    14 hours ago
  •  ...re Looking For We’re looking for a proactive, hands‑on Site Reliability Engineer who thrives in building and scaling cloud infrastructure in...  ...modern CI/CD pipelines with GitLab to enable fast, safe deployments Implementing and evolving monitoring, alerting, and observability... 
    Work experience placement
    Flexible hours

    Rainforest

    Atlanta, GA
    3 days ago
  •  ...Application/Infrastructure Performance, and availability. • Automation Experience with Build/deployment, Software Configuration/Continuous Integration/Continuous Delivery/Release Engineering related tasks in JavaEE/C++ Environments. • Experience in automating manual... 
    Immediate start

    Navtech

    Atlanta, GA
    2 days ago
  • $71.6k - $119.4k

     ...Our services provide applications with reliability, security, and better customer...  ...applications with provisioning needs, deployment support, and security improvements. You...  ...troubleshoot issues, and work closely with senior engineers to learn and apply best practices. You... 
    Full time
    Temporary work
    Internship
    Local area
    Work from home

    RELX

    Atlanta, GA
    2 days ago
  • $95k - $171k

     ...Technology Group. We design, implement, deploy and operate AI platforms that enable...  ...infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for:... 
    Permanent employment
    Work experience placement
    Work at office
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Akamai

    Atlanta, GA
    1 day ago
  • $60 - $68 per hour

     ...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with long-term...  ...Continuous Integration (CI), Continuous Delivery and Continuous Deployment (CD) pipeline; Ability to design, build, implement and... 
    Contract work
    Local area
    Immediate start

    Pyramid Corporation

    Atlanta, GA
    14 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Deployment and Site Reliability Engineer. Be the first to apply!