Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

DevOps Manager for Large Scale SaaS Software Application

ENSYTE Energy Software Int'l, Inc.

Job Description

Job Description

About the Role

We are seeking an experienced DevOps Manager to lead the infrastructure, automation, and release engineering efforts behind our SaaS software solution for the energy industry. This role blends hands-on technical leadership with people management, overseeing a team responsible for the uptime, scalability, and security of the production systems our customers depend on 24/7. You'll be central to how we ship features quickly and safely while maintaining the reliability, performance, and data isolation our SaaS customers expect. The ideal candidate combines deep technical expertise in cloud infrastructure and CI/CD practices with strong leadership skills and a track record of running production systems at a large SaaS scale.

 

Key Responsibilities

  • Lead, mentor, and grow a team of DevOps/Site developers, setting clear goals and providing regular feedback and career development support.
  • Own uptime and performance of our production SaaS platform against defined SLAs/SLOs, ensuring a reliable experience for all customers.
  • Design, implement, and maintain scalable CI/CD pipelines that support frequent, zero-downtime deployments to a live, multi-tenant production environment.
  • Own the architecture and health of our multi-tenant cloud Azure infrastructure, ensuring high availability, tenant data isolation, security, and cost efficiency as the customer base grows.
  • Drive infrastructure-as-code practices to support consistent, repeatable environment provisioning across dev, staging, and production.
  • Establish and maintain monitoring, alerting, and on-call incident response processes to minimize customer-facing downtime and meet up time commitments.
  • Partner with Engineering, Product, Customer Success, and Security teams to align infrastructure strategy with product roadmap and customer growth.
  • Manage container orchestration platforms supporting horizontal scaling as tenant load and usage patterns change.
  • Design and maintain multi-region/multi-AZ architecture and disaster recovery plans to meet SaaS customer uptime and data durability expectations.
  • Support scalable multi-tenancy patterns, including tenant provisioning/deprovisioning automation and per-tenant resource isolation.
  • Define and track key operational metrics (uptime, deployment frequency, MTTR, error budgets, etc.) and drive continuous improvement.
  • Manage cloud vendor relationships and cost optimization initiatives as infrastructure scales with customer growth.
  • Ensure compliance with security best practices and SaaS-relevant regulatory/compliance frameworks (SOC 2).
  • Support customer-facing status pages, uptime reporting, and communication processes during incidents.
  • Lead incident postmortems and champion a blameless culture of continuous learning.

 

Required Qualifications

  • 5+ years of experience in DevOps, Site Reliability Engineering, or Infrastructure Engineering roles, ideally supporting a production SaaS product.
  • 2+ years of experience in a people management or team lead capacity.
  • Experience operating and scaling multi-tenant cloud architecture, including tenant isolation and per-tenant scaling considerations.
  • Strong hands-on experience with at least one major cloud provider (AWS, Azure, or GCP).
  • Proven expertise with containerization and orchestration tools (Docker, Kubernetes).
  • Experience with infrastructure-as-code tools (Terraform, Ansible, CloudFormation, or similar).
  • Solid understanding of CI/CD tools and practices (Jenkins, GitLab CI, GitHub Actions, CircleCI, etc.).
  • Experience with monitoring and observability tools (Prometheus, Grafana, Datadog, New Relic, etc.).
  • Scripting/programming proficiency.
  • Strong understanding of networking, security, and system administration fundamentals.
  • Excellent communication skills and experience working cross-functionally with engineering, product, and security teams.

 

Preferred Qualifications

  • Experience in a high-growth SaaS startup or scaling subscription-based engineering organization.
  • Relevant certifications (AWS Certified DevOps Engineer, CKA, etc.).
  • Experience with service mesh, GitOps workflows, or serverless architectures.
  • Background in incident management frameworks and on-call rotation design.
  • Experience building / maintaining public-facing status pages &  customer uptime reporting.
  • Familiarity with SaaS-specific compliance frameworks (SOC 2 Type II) and audit processes.

 

What We Offer

  • Competitive salary package
  • Comprehensive health, dental, and vision coverage
  • Flexible PTO policy
Company Description

Our company is a leading provider of Energy Trading Risk Management (ETRM) solutions for natural gas companies. We have over 40 years of experience delivering energy software solutions to the energy industry, but we operate as a young, dynamic company in a phase of significant growth. Our culture is fun, exciting, challenging, and everyone on our team works together towards a common goal.

Company Description

Our company is a leading provider of Energy Trading Risk Management (ETRM) solutions for natural gas companies. We have over 40 years of experience delivering energy software solutions to the energy industry, but we operate as a young, dynamic company in a phase of significant growth. Our culture is fun, exciting, challenging, and everyone on our team works together towards a common goal.

Vacancy posted more than 2 months ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to DevOps Manager for Large Scale SaaS Software Application. Be the first to apply!