Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer I

Full-time

PagerDuty

Responsibilities

  • Support and improve foundational networking, compute, Kubernetes, and ingress or traffic-management infrastructure.
  • Harden existing systems and support the rollout of new infrastructure capabilities to improve platform reliability and scalability.
  • Monitor system health through metrics, logs, and alerts.
  • Participate in 24/7 on-call rotations to detect, respond to, and resolve incidents.
  • Participate in team planning, standups, retrospectives, and progress or risk communication.

Requirements

  • Require 0 to 1+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering roles.
  • Require hands-on experience operating Linux-based systems in production.
  • Require working knowledge of networking fundamentals including load balancing, DNS, TLS, and ingress traffic flow.
  • Require experience with container orchestration such as EKS or Kubernetes.
  • Require experience with cloud-native infrastructure such as AWS, GCP, or Azure, including networking and compute concepts.
  • Require proficiency in at least one programming language such as Python, Ruby, or Go.
  • Require experience with Infrastructure as Code such as Terraform or CloudFormation.
  • Prefer experience with AWS cloud networking concepts including VPCs, subnets, routing, security groups, and load balancers.
  • Prefer experience operating or contributing to production Kubernetes platforms, including cluster upgrades, networking, or ingress configuration.
  • Prefer experience with monitoring, observability, and logging platforms such as Datadog, New Relic, SumoLogic, Splunk, Prometheus, or Grafana.
  • Prefer familiarity with service meshes, ingress controllers, or API gateways such as Envoy, Istio, or NGINX.

Benefits

  • Hybrid work model based in PagerDuty’s Atlanta office, with location eligibility restrictions.
  • Competitive salary and comprehensive benefits package.
  • Flexible work arrangements, company equity, and an Employee Stock Purchase Program, subject to eligibility.
  • Retirement or pension plan, paid vacation, paid holidays, and sick leave.
  • Dutonian Wellness Days and HibernationDuty companywide paid days off.
  • Paid parental leave and 20 hours of paid volunteer time off per year.
  • Company-wide hack weeks and mental wellness programs.
Vacancy posted more than 2 months ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer I. Be the first to apply!