Site Reliability Engineer II
$113k - $171.6kPagerDuty Inc.
PagerDuty, Inc. (NYSE: PD) is the global leader in AI-first digital operations. By automatically detecting, diagnosing, and remediating issues, the PagerDuty Platform orchestrates AI agents and automated workflows with context from over 750 integrations. Trusted by approximately two-thirds of the Fortune 100 and nearly half of the Fortune 500, PagerDuty is the industry standard for organizations scaling resilient, autonomous operations. Notable customers include Chipotle, Cloudflare, Docusign, Fox, Nvidia, Salesforce, Spotify, Zoom and more. We are growing rapidly and hiring top talent with leading AI skills across engineering, sales, product, marketing, and beyond as we build the leading digital operations platform.
As a Site Reliability Engineer II on the Core Infrastructure team in our Atlanta office,
you'll help build and operate the foundational infrastructure that powers PagerDuty's
real-time digital operations platform. Our systems support millions of events and alerts
daily, enabling customers to detect, respond to, and resolve incidents quickly and
reliably.
You'll work at the intersection of platform evolution and operational excellence, building
and evolving foundational network, compute, and ingress infrastructure while scaling
and hardening existing systems. Your work will directly impact the reliability, scalability,
and security of the services our customers rely on to keep their businesses running as
PagerDuty continues to grow across products, regions, and customer use cases.
Key Responsibilities
● Support and improve foundational infrastructure, including networking, compute
platforms, Kubernetes clusters, and ingress/traffic management systems.
● Contribute to the reliability and scalability of PagerDuty's core platform by
hardening existing systems and supporting the rollout of new infrastructure
capabilities.
● Participate in agile rituals (standups, planning, retros) and communicate
progress/risks early
● You stay current on technical trends to suggest innovative tools and approaches
to interesting problems
● Monitor system health using metrics, logs, and alerts, and participate in 24/7
on-call rotations to help detect, respond to, and resolve incidents.
Basic Qualifications
● 3+ years of experience in Site Reliability Engineering, DevOps, or Platform
Engineering roles
● Hands-on experience operating Linux-based systems in production
environments
● Working knowledge of networking fundamentals, such as load balancing, DNS,
TLS, and ingress traffic flow
● Experience with container orchestration (e.g., EKS, Kubernetes)
● Experience working on cloud-native infrastructure (e.g., AWS, GCP, Azure),
including networking and compute concepts
● Proficiency in at least one programming language (e.g., Python, Ruby, Go, etc.)
● Experience with Infrastructure as Code (e.g., Terraform, CloudFormation)
Preferred Qualifications
● Experience with AWS cloud networking concepts such as VPCs, subnets,
routing, security groups, and load balancers
● Experience operating or contributing to production Kubernetes platforms (e.g.,
EKS), including cluster upgrades, networking, or ingress configuration
● Experience with monitoring, observability, and logging platforms (e.g., DataDog,
New Relic, SumoLogic, Splunk, Prometheus, Grafana)
● Familiarity with service meshes, ingress controllers, or API gateways (e.g.,
Envoy, Istio, NGINX)
Salary Range: $113,000 to $171,600
Hesitant to apply?
We encourage you to submit your resume even if you don't meet every requirement. We value potential and consider each candidate's full professional story. Whether you're exploring a career change or taking your next step, we look forward to reviewing your application. If this just isn’t the right role or time - sign up for job alerts ( !
Where we work
PagerDuty operates a hybrid work model with offices ( in 8 major cities: Atlanta, Lisbon, London, San Francisco, Santiago, Sydney, Tokyo, and Toronto. While we offer flexibility within our established locations, we cannot employ candidates residing in:
Location restrictions:
Australia: Northern Territory, Queensland, South Australia, Tasmania, Western Australia
Canada: Alberta, Manitoba, Newfoundland, Northwest Territories, Nunavut, PEI, Quebec, Saskatchewan, Yukon
United States: Alaska, Hawaii, Iowa, Louisiana, Mississippi, Nebraska, New Mexico, Oklahoma, Rhode Island, South Dakota, West Virginia, Wyoming
Candidates must reside in an eligible location, which vary by role.
How we work
Our values ( guide how we support customers, collaborate with colleagues, develop products, and foster a culture of belonging. They define not just our actions, but what it means to be Dutonian.
People Leaders at PagerDuty are responsible for creating high performance environments that drive accountability. PagerDuty has four key dimensions that define our Leadership Impact: Lead Self, Lead the Team, Lead the Business, and Lead the Future. Each dimension has three associated competencies to give leaders a shared language for guiding their development, career, promotion, and succession planning discussions. Our Manager Expectations serve as a practical guide for managers to understand their responsibilities, prioritize their efforts, and drive engagement and performance.
What we offer
As a global organization, our total rewards approach is competitive with industry standards and aligned with local laws and regulations. Learn more, including country-specific offerings, on our benefits site ( .
Your package may include:
Competitive salary
Comprehensive benefits package
Flexible work arrangements
Company equity*
ESPP (Employee Stock Purchase Program)*
Retirement or pension plan*
Generous paid vacation time
Paid holidays and sick leave
Dutonian Wellness Days & HibernationDuty - companywide paid days off in addition to PTO
Paid parental leave: 22 weeks for pregnant parent, 12 weeks for non-pregnant parent (some countries have longer leave standards and we comply with local laws)*
Paid volunteer time off: 20 hours per year
Company-wide hack weeks
Mental wellness programs
*Eligibility may vary by role, region, and tenure
About PagerDuty
PagerDuty, Inc. (NYSE:PD) is a global leader in digital operations management. The PagerDuty Operations Cloud is an AI-powered platform that empowers business resilience and drives operational efficiency for enterprises. With a generative AI assistant at its core, PagerDuty empowers teams to detect and resolve issues in real time, orchestrate complex workflows, and drive continuous improvement across their digital operations. Trusted by nearly half of both the Fortune 500 and the Forbes AI 50, as well as approximately two-thirds of the Fortune 100, PagerDuty is essential for delivering always-on digital experiences to modern businesses
PagerDuty is Great Place to Work-certified™, a Fortune Best Workplace for Millennials, a Fortune Best Medium Workplace, a Fortune Best Workplace in Technology, and a top rated product on TrustRadius and G2.
Go behind-the-scenes on our careers site ( and @pagerduty on Instagram.
Additional Information
PagerDuty is an equal opportunity employer. PagerDuty does not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, parental status, veteran status, or disability status. Your privacy is important to us. By submitting an application, you confirm that you have read and understand PagerDuty's Privacy Policy ( .
PagerDuty is committed to providing reasonable accommodations for qualified individuals with disabilities in our job application process. Should you require accommodation, please email View email address on click.appcast.io and we will work with you to meet your accessibility needs.
PagerDuty uses the E-Verify employment verification program.
$95k - $171k
...Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts,...SuggestedPermanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$71.6k - $119.4k
...support application teams. Our services provide applications with reliability, security, and better customer experiences. About the Job:... ...automation, troubleshoot issues, and work closely with senior engineers to learn and apply best practices. You’ll gain exposure to a...SuggestedFull timeTemporary workInternshipLocal areaWork from home$104.9k - $174.7k
...Management. You can learn more about LexisNexis Risk at the link below, About the Role: We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...SuggestedFull timeWork at officeLocal areaRemote workWork from homeFlexible hours- ...work in the United States. However, we are not able to sponsor work visas now or in the future. “High-Fiver” Wanted: Software Engineer II Are you the type of person who always goes the extra mile? Are you passionate about your craft and always striving to improve?...SuggestedFull timeWork at officeVisa sponsorshipWork visa
- ...every worker, everywhere. SUMMARY We are seeking a Software Engineer II with expertise in Rails and React to join our growing Atlanta... ...6-8 weeks. RESPONSIBILITIES Build, design, and evolve reliable, scalable full-stack services and applications, including API...SuggestedFull timeWork experience placement
- ...success. Summary We're looking for a Software Developer II for Miovision Technologies US, LLC to architect and maintain AWS... ...or foreign education equivalent in Computer Science or Computer Engineering ~+ 3 years' experience in a software development role. ~...Remote workFlexible hours
$114.89k - $142.3k
...headquartered in Atlanta, Georgia, and serves customers in more than 35 countries worldwide. NCR Voyix Corporation Atlanta, GA Software Engineer II (f/t) Job Description: Collaborate with other developers to design, develop, test, deploy, maintain, and enhance new software...Part timeRemote workWorldwideFlexible hours- Software Engineer II Specializing in AkesoLIMS at Tempus Passionate about precision medicine and advancing the healthcare industry? Recent... ...laboratory workflows into technical solutions, and deliver reliable, scalable applications that improve efficiency and usability for...
$82.2k - $113.1k
...evolve. Job Description We are seeking a Full Stack Software Engineer with strong backend experience in .NET Core and frontend experience... ...and code quality. In this role as a Software Engineer II, you will: Design, develop, and maintain backend services and RESTful...For contractorsLocal areaRemote workWorldwideWork visaFlexible hours- ...customers in more than 35 countries worldwide. Title: Software Engineer II Grade: P2 Location: Atlanta, GA NCR Voyix is looking for a... ...with stakeholders and engineering teams to deliver scalable, reliable solutions. Contribute to product and platform roadmaps with consideration...WorldwideFlexible hours
$130k - $138k
...re looking for a curious, detail-oriented Software Development Engineer II who enjoys digging into existing systems and making them... ...understanding user needs, and turning complex issues into clean, reliable code, you'll feel right at home here. This is a great opportunity...Flexible hours- Software Engineer - II/ .Net Developer Client, a leader in payment technology, is seeking an experienced Software Engineer II to join our... ...cloud-based microservices in AWS, ensuring scalability, reliability, and security. Work closely with engineering leads and managers...
- Apply today to join Coreforce, where your Engineering expertise makes a real impact. Join Our Team as a Software Engineer II Company: Coreforce. Location: Atlanta (Hybrid... ...in Java with Spring Boot, ensuring reliability, performance, and clean design. Develop frontend...Full timeFlexible hours
$123.6k - $200.1k
...AI Machine Learning Engineer II (Full Time) – United States Note: This posting advertises potential opportunities. While the role... ...generative AI models. Optimize performance, scalability, and reliability. Engage in red‑teaming to validate robustness and...Full timeTemporary workFlexible hours$100k - $120k
...OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying...Full timeTemporary workWork experience placementFlexible hours- ...solving and decision-making abilities and the highest degree of professionalism. We are seeking an experienced AWS solution design engineer/architect to join our infrastructure cloud team. The infrastructure cloud team is responsible for internal services that provide...
- ...Site Reliability Engineer (SRE) When you join Atlanticus, you become a member of a fast-growing, mission-focused company that is committed to aid in meeting the financial needs of middle-class Americans. With a culture of collaboration and a one-team mindset, we encourage...Work at office
$60 - $68 per hour
...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with long-term potential and is located in Atlanta, GA (Onsite). Please review the job description below and contact me ASAP if you are interested...Contract workLocal areaImmediate start- ...and we're looking for new team members who want to be a part of this journey! We're looking for a proactive, hands-on Site Reliability Engineer who thrives in building and scaling cloud infrastructure in fast-moving startup environments. You're someone who enjoys owning...Work experience placementFlexible hours
- ...We are currently looking for a Senior Software Engineer to be a part of the Site Reliability Engineering (SRE) team in Atlanta, GA . The SRE team is an innovative team devoted to providing a Docker-based Platform as a Service and assisting a growing number of teams...Contract workWork at officeLocal area
- ...Lead Engineer, Site Reliability Engineering Team As a lead engineer with Retail, Site Reliability Engineering team, you will be at the forefront of Cloud and Big Data technology. In this role you will establish yourself as a technical leader by exposing yourself to...
$121.4k - $218.6k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office$127k - $249k
...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas... ...workloads. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This...Local areaRemote workWorldwideFlexible hours- ...availability. • Automation Experience with Build/deployment, Software Configuration/Continuous Integration/Continuous Delivery/Release Engineering related tasks in JavaEE/C++ Environments. • Experience in automating manual processes using Python, Ruby, Unix Shell (bash,...Immediate start
$130k - $150k
...recruiter to learn more. Base pay range $130,000.00/yr - $150,000.00/yr Overview: We are seeking a highly skilled Site Reliability Engineer (SRE) to join our team and help build and maintain scalable, reliable, and efficient systems. The ideal candidate will...Full timeRemote work- ...Site Reliability Engineer At Intercontinental Exchange (NYSE:ICE), we engineer technology, exchanges and clearing houses that connect companies around the world to global capital and derivative markets. With a leading-edge approach to developing technology platforms...
$123.4k - $222.53k
...! Ready grow your career as part of the Uncarrier journey at T-Mobile? Our team is searching for our next Sr. Site Reliability Engineer to strengthen the reliability and resilience of the systems powering T-Mobile's payment platforms, enabling faster, safer...Full timeTemporary workPart timeWork experience placementLocal areaFlexible hours- ...Site Reliability Engineer At Acuity, you will join an Agile team focused on building and supporting advanced platforms and applications that drive our business forward. We are seeking a Site Reliability Engineer (SRE) to help define and raise the reliability bar for...
$120k - $175k
...of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's possible...Full timeRemote workWork visaFlexible hours- ...Join to apply for the Site Reliability Engineer role at Motion Recruitment Join to apply for the Site Reliability Engineer role at Motion Recruitment Get AI-powered advice on this job and more exclusive features. Every year, nearly 200 million travelers...Contract workWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer II. Be the first to apply!
- site reliability engineer Atlanta, GA
- site reliability engineer remote Atlanta, GA
- site reliability engineer sre Atlanta, GA
- site safety Atlanta, GA
- website coordinator Atlanta, GA
- on-site clinical research associate (traveling/remote) Atlanta, GA
- site services specialist Atlanta, GA
- on site coordinator Atlanta, GA
- construction site safety Atlanta, GA
- junior website developer Atlanta, GA


