Site Reliability Engineer
GiveCampus
GiveCampus is the world's leading fundraising platform for non-profit educational institutions. Trusted by millions of donors and 1,300+ colleges, universities, and K-12 schools, our mission is to help advance the quality, the affordability, and the accessibility of education. At our current pace, we will facilitate $100 billion in charitable giving over the next decade–enough money to send more than 1 million students to college, tuition-free.
GiveCampus is backed by leading investors including Y Combinator, but we’re also practitioners of Sustainable Growth: we’ve made the Inc. 5000 list of America's fastest-growing private companies each of the last five years and we’ve been profitable nine of the last 10. In 2025, we celebrated a $140 million growth investment that included a major liquidity event for GiveCampus employees–the second in less than three years.
Our purpose-driven team of 130+ is located in 30+ states across the US: team members work from anywhere they choose. We have a beautiful 12,000sf office in Washington, DC that is available for people to use whenever they want, and we regularly organize team meet-ups, visit partner institutions, and host retreats in various locations.
While we operate at meaningful scale, we’re still small relative to the commercial and social good opportunities in front of us. Every GiveCampus employee plays a meaningful role in shaping what comes next, and we're growing the team in support of our ambitious plans–including a $100 million investment in AI product development. If you believe in the transformative power of education and want to join a fast-growing, mission-driven company, you’ll fit right in.
Location: This is a remote-first role based in the U.S. While we embrace flexible, distributed work, we also value in-person connection. Team members are expected to attend multiple company-wide and team-specific onsites throughout the year.
About the role
GiveCampus is looking for a hands-on Site Reliability Engineer to help improve the reliability, performance, and operational maturity of our platform.
With our migration to AWS complete, this role will focus on operating and strengthening our production environment: improving observability, automating infrastructure and operational work, responding to incidents, and partnering with product engineers to build resilient systems.
You will own well-scoped reliability projects and contribute to larger cross-functional initiatives. You should be comfortable working independently on straightforward problems, asking for guidance when needed, explaining technical tradeoffs, and keeping teammates informed at important milestones.
What you'll do
- Operate, maintain, and improve production infrastructure in AWS.
- Build and maintain infrastructure as code using Terraform.
- Support workloads running on Kubernetes and Amazon EKS.
- Improve dashboards, alerts, and service-level indicators using New Relic or comparable observability platforms.
- Investigate production issues, identify root causes, and implement durable fixes.
- Participate in the shared 24/7 on-call rotation and contribute to effective incident response.
- Participate in blameless postmortems and complete follow-up actions that reduce the likelihood or impact of repeat incidents.
- Partner with product engineers to troubleshoot performance and reliability issues throughout the application stack.
- Improve application resilience using established patterns such as timeouts, retries, queuing, backpressure, and idempotency.
- Maintain and improve CI/CD pipelines and deployment workflows using tools such as GitHub Actions and CircleCI.
- Automate repetitive operational tasks and identify opportunities to reduce engineering toil.
- Create and maintain runbooks, system diagrams, troubleshooting guides, and production documentation.
- Contribute to capacity planning, performance testing, database reliability, and production-readiness reviews.
- Apply established security, access-control, logging, and compliance practices to infrastructure work.
- Own small-to-medium reliability improvements from technical design and work breakdown through delivery.
- Communicate progress, risks, tradeoffs, and blockers clearly while incorporating feedback from engineering partners.
What we're looking for
- Approximately 5+ years of related experience in software engineering, infrastructure, systems engineering, SRE, Platform Engineering, DevOps, or equivalent practical experience.
- Hands-on experience operating production workloads in AWS.
- Experience building or maintaining infrastructure using Terraform or a similar infrastructure-as-code tool.
- Experience with New Relic, Datadog, or another modern observability platform.
- Experience troubleshooting production incidents and participating in an on-call rotation.
- Experience building or maintaining CI/CD pipelines.
- Software development or scripting experience, with the ability to read, debug, and make targeted changes to application or automation code.
- Working knowledge of Linux, networking, distributed systems, and relational databases.
- Ability to articulate root causes, explain technical tradeoffs, and translate findings into practical solutions.
- Ability to manage a well-scoped project with general direction and provide timely updates at key milestones.
- Strong written and verbal communication skills and a collaborative approach to working across engineering disciplines.
- A habit of automating repetitive work and improving the reliability of the systems you support.
Bonus points
- Experience with Ruby or Ruby on Rails.
- PostgreSQL administration or performance-tuning experience.
- Experience with Kubernetes and Amazon EKS.
- Experience with Redis, OpenSearch, or Amazon RDS.
- Experience operating enterprise SaaS products at scale.
- Familiarity with SLOs, SLIs, error budgets, capacity modeling, or load testing.
- Experience with payments, fintech, or other regulated systems.
- Experience supporting SOC 2 or similar security and compliance programs.
Ready to apply?
Be sure to keep an eye on your spam and promotions boxes in case our emails end up there!
At GiveCampus, we value diversity and we pledge to foster an environment of support, inclusivity, and learning, both on the job and throughout the application process. In this spirit, we encourage candidates of all backgrounds to apply.
GiveCampus is an Equal Opportunity Employer. Applicants and employees are not discriminated against because of race, color, creed, sex, sexual orientation, gender identity or expression, age, religion, national origin, citizenship status, disability, ancestry, marital status, veteran status, medical condition or any protected category prohibited by local, state or federal laws.
If you feel like you don't meet all of the requirements for this role, please apply anyways. We know confidence gaps and imposter syndrome often get in the way of connecting with incredible people, and we don't want them to prevent us from meeting you.
$7.5k
...investment management. We have become a multibillion-dollar asset manager, and we have ambitious goals for the future. As a Site Reliability Engineer (SRE), you will work at the intersection of production operations and software development as you improve, manage, and...SuggestedLocal areaRemote work- ...Senior Site Reliability Engineer Company: CyberArk Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: AWS, Kubernetes, Terraform, CloudFormation, Ansible, CloudWatch, Grafana, Datadog, OpenSearch, PagerDuty Requirements: Senior SRE...SuggestedFull timeRemote work
- ...Job Title: Site Reliability Engineer (Azure Government & Infrastructure) Pay Type : SALARIED EXEMPT Location: Remote Citizenship Requirement: U.S. Citizen (Required) Summary of Position Role/Responsibilities The Site Reliability Engineer (SRE) for...SuggestedFull timeRemote workMonday to Friday
- ...Site Reliability Engineer OXIO is the first NeoTelco. We arebuilding the world’s largest, most accessible, and insightful Telecom network. Our platform empowers anyone to spin up their own carrier from a browser, scaling and supporting you as you scale your network...SuggestedRemote work
- ...The Site Reliability Engineer (SRE) / Subject Matter Expert (SME) – Computer Systems Engineer/Architect will provide senior-level reach-back expertise to support the reliability, scalability, performance, and operational resilience of the GEOMAP platform in secure cloud...SuggestedFull timeContract workFor contractorsFor subcontractorRemote work
- ...Site Reliability Engineer Company: Crunchafi Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Azure, AKS, Azure Kubernetes Service, Terraform, Bicep, ARM templates, GitHub Actions, Azure DevOps, Kubernetes, Docker, App Insights...Full timeRemote work
- ...Site Reliability Engineer Company: Quzara Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: Azure, Terraform, Bicep, Ansible, Azure Monitor, Azure Automation, Azure Policy, Azure Site Recovery, TLS/SSL Requirements: 4+ years in SRE...Full timeRemote work
$160k - $180k
...big impact. See Arkestro in action at arkestro.com. About the Role Arkestro is hiring for a Senior SRE Engineer to manage our performance and reliability for our software platform and infrastructure. The right candidate will own and develop our infrastructural...Local areaRemote work- ...About The Role: We're looking for a Senior Site Reliability Engineer to help us mature and scale the infrastructure behind our multi-cloud SaaS platform. Most of our footprint runs on Microsoft Azure, built from the ground up around cloud architecture principles:...Remote workFlexible hours
$145k - $193k
...entertainment, we want to talk to you. About the Role & Team The SRE team at PENN Entertainment is looking for a Senior Site Reliability Engineer to help build and operate the infrastructure behind a large-scale sports betting and media platform. You'll own critical...Remote work$140k - $180k
...-making, and accelerated growth in the AI-driven world. Learn more at Opportunity We’re looking for a Senior Site Reliability Engineer to help build and scale a high-impact SRE function. You’ll be a technical leader on a team responsible for improving system...Work experience placementLocal areaRemote workVisa sponsorshipWork visa- ...Site Reliability Engineer Company: GitLab Work Type: Remote Employment: Full Time Location: CA, US Seniority: Senior Level Technologies: Terraform, Ansible, Kubernetes, Go, Ruby, Jsonnet, Prometheus, ELK, Grafana Requirements: Senior-level SRE with strong Terraform/IaC...Full timeRemote work
- ...Engineering, Product, Design, and Marketing Engineering Compensation ~ Zone 1 Base Pay: $214K – $260K Superhuman offers... ...role will be responsible for building software to ensure the reliability of our back-end systems, working with engineers who develop them...WorldwideHome officeFlexible hours
$141.8k - $195k
...work, grow fast, and bring their full selves to the herd. Why You’ll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in the U.S. Cribl provides...Temporary workRemote work- ...grow, adapt and use your skills consistently. Our customers rely on us in the moments that matter. Engineering delivers on that promise. The Senior Site Reliability Engineer is responsible for ensuring our SaaS products are fast, stable and optimized for our customers...Work experience placementRemote workFlexible hours
- ...encourage you to apply. The Role As a Senior Platform Engineer, you are a champion for DevOps and SRE culture and industry... ...met. \n What You Will Be Doing Improving production reliability and system resilience within an SRE scoped team Championing...Remote workFlexible hours
$135k - $170k
...Symmetrio is recruiting a Site Reliability Engineer for its customer, a rapidly growing international healthcare SaaS company aggressively expanding in the United States. As a member of our customer’s Site Reliability / Cloud Platform team, the role has two connected responsibilities...Full timeRemote work- ...Site Reliability Engineer Company: Milestone Systems Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Golang, Python, Linux, Shell scripting, Kubernetes, Docker, Terraform, CI/CD, GitOps, ArgoCD, Spinnaker, Prometheus, Datadog,...Full timeRemote work
- ...and performance our customers have come to expect, and help raise the reliability bar as we grow. What you would do: Design, build, and operate the shared platform foundations engineers ship on every day: GCP infrastructure, Kubernetes, networking, routing,...Remote workWorldwideFlexible hours
$147k - $168k
...Inc. as one of the most innovative and fastest-growing technology companies in the country. Role Summary As a Site Reliability Engineer at Filevine, you will improve the reliability, scalability, and operational maturity of the Filevine platform. You’ll...Full timeTemporary workWork experience placementWork at officeRemote work2 days per week3 days per week$125k - $250k
...we are reimagining how developers build reliable, scalable, event-driven applications without... ...possible Partner closely with engineering teams to improve system resiliency and scalability... ...For 5+ years of experience in Site Reliability Engineering, DevOps,...Full timeImmediate startRemote workFlexible hours- ...About the Role We are seeking a Senior Site Reliability Engineer to join our cloud engineering team. You will own the reliability, scalability, and observability of our critical financial SaaS applications and infrastructure, working across cloud platforms to ensure...Remote work
$186.82k - $224.18k
...made, and we are not tiptoeing into it. We are rebuilding our engineering culture around a simple belief: AI changes everything. How... ...Is Babylist is looking for a Senior Software Engineer, Site Reliability to join our Platform team. In this position, you will play...Work at officeLocal areaImmediate startRemote workFlexible hoursShift work$180k - $230k
...Acceleration Job Description We're looking for a Senior SRE to own the reliability, scalability, and observability of our production systems. You'll work closely with platform and data engineering to keep high-throughput, data-intensive services running at the...Work at officeLocal areaImmediate startRemote work3 days per week- ...Senior Site Reliability Engineer Company: Sphera Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Terraform, ARM templates, Kubernetes, Azure, SonarCloud, CheckPoint, Hadoop, Kafka, Presto, NewRelic, CI/CD, Linux, Windows, Redis...Full timeRemote work
$169k - $215k
...solving challenging problems in real-time video, scalability, reliability, data, machine learning, and user experience. We’re a... ...improvement. The Role We are seeking a remote Site Reliability Engineer who will elevate our infrastructure resilience and optimize...Temporary workWork at officeRemote work- ...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...through efficient, data-driven solutions. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling...Work at officeRemote work
$90k - $100k
...Site Reliability Engineer The Opportunity We are looking for a highly capable engineer to join our Platform and Site Reliability engineering team. You will be responsible for building, maintaining and operating the infrastructure platform on which all Ookla services...Remote workFlexible hours- ...remote within the U.S., with a preference for candidates located in the Mountain or Pacific Time zones. Overview: The Site Reliability Engineer (SRE) is responsible for the reliability, performance, and scalability of Precisely's infrastructure platforms across...Work experience placementRemote work
$153k - $210k
...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you passionate about building resilient, highly available cloud platforms that enable engineering teams to move quickly and confidently? Do you enjoy automating...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre United States
- site reliability engineer United States
- site reliability engineering manager United States
- site reliability engineer remote United States
- lead site reliability engineer United States
- IT site lead United States
- site safety United States
- site merchandiser United States
- website content developer United States
- site leader United States


