Site Reliability Engineer I
Full-time
PagerDuty
Responsibilities
- Support and improve foundational networking, compute, Kubernetes, and ingress or traffic-management infrastructure.
- Harden existing systems and support the rollout of new infrastructure capabilities to improve platform reliability and scalability.
- Monitor system health through metrics, logs, and alerts.
- Participate in 24/7 on-call rotations to detect, respond to, and resolve incidents.
- Participate in team planning, standups, retrospectives, and progress or risk communication.
Requirements
- Require 0 to 1+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering roles.
- Require hands-on experience operating Linux-based systems in production.
- Require working knowledge of networking fundamentals including load balancing, DNS, TLS, and ingress traffic flow.
- Require experience with container orchestration such as EKS or Kubernetes.
- Require experience with cloud-native infrastructure such as AWS, GCP, or Azure, including networking and compute concepts.
- Require proficiency in at least one programming language such as Python, Ruby, or Go.
- Require experience with Infrastructure as Code such as Terraform or CloudFormation.
- Prefer experience with AWS cloud networking concepts including VPCs, subnets, routing, security groups, and load balancers.
- Prefer experience operating or contributing to production Kubernetes platforms, including cluster upgrades, networking, or ingress configuration.
- Prefer experience with monitoring, observability, and logging platforms such as Datadog, New Relic, SumoLogic, Splunk, Prometheus, or Grafana.
- Prefer familiarity with service meshes, ingress controllers, or API gateways such as Envoy, Istio, or NGINX.
Benefits
- Hybrid work model based in PagerDuty’s Atlanta office, with location eligibility restrictions.
- Competitive salary and comprehensive benefits package.
- Flexible work arrangements, company equity, and an Employee Stock Purchase Program, subject to eligibility.
- Retirement or pension plan, paid vacation, paid holidays, and sick leave.
- Dutonian Wellness Days and HibernationDuty companywide paid days off.
- Paid parental leave and 20 hours of paid volunteer time off per year.
- Company-wide hack weeks and mental wellness programs.
Vacancy posted more than 2 months ago
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer I. Be the first to apply!
Related searches
- site reliability engineer Atlanta, GA
- site reliability engineer sre Atlanta, GA
- site reliability engineer remote Atlanta, GA
- website coordinator Atlanta, GA
- on-site clinical research associate (traveling/remote) Atlanta, GA
- site safety Atlanta, GA
- junior website developer Atlanta, GA
- construction site safety Atlanta, GA
- IT site lead Atlanta, GA
- website content developer Atlanta, GA
