Site Reliability Engineer
Pacifica Continental
Job Title
Our engineering team has built the largest private Medicare marketplace in the country. We passionately focus on the continuous improvement of the systems we build.
We have spent many years growing and fostering a DevOps culture by bridging the divide between our Software and Infrastructure Engineering departments. We want the cross-functional teams that we are building to include Site Reliability Engineers. We operate in a complex, multi-tenant, hybrid cloud and on-premises infrastructure that spans both the Windows and Linux OS. We strive for security, reliability, and automation in line with DevOps and Site Reliability Engineering principles. If you are passionate about learning and improvement through metrics and automation, and passionate about engendering that mindset in others, we want to hear from you.
About the Role
Maintains shared cloud resources in use by numerous software engineering teams within our business unit. We aim to enable software engineering teams to build cloud native applications that adhere to security and regulatory requirements with limited handholding by our cloud engineers. We do still have a fair number of applications hosted in on-premise data centers, which we aim to support migrating to the cloud.
Requirements
Hands-on Engineering
5+ years of hands-on experience with a majority of the following technologies, along with a willingness to become proficient in the remaining areas:
- Windows and Linux Servers
- VMware
- Cloud platforms, preferably with Azure
- Active Directory
- Secrets management with Consul and Vault or similar systems
- Configuration management tools like Salt, Ansible and Terraform
- Firewalls and load balancers such as F5
- Web servers, including IIS and NGINX
- Database Server Infrastructure like Microsoft SQL Server and PostgreSQL
- Application Performance Monitoring with tools like New Relic
- Infrastructure monitoring with tools like Sensu, SolarWinds, Nagios, or Azure App Insights
- CI/CD tools like TeamCity, Octopus Deploy, Concourse, Azure DevOps, or GitHub Actions
- Log Aggregation tools like SumoLogic or Splunk
- Network theory and protocols such as DNS, DHCP, proxy servers, and firewalls
- Security operations with tools for SAST, DAST, RAST, and WAF
- Infrastructure as Code or automation experience.
Proficiency, high-comfort, and familiarity with:
- One or more programming languages, such as C#, JavaScript, Python or Go
- One or more scripting languages, such as PowerShell and BASH
- Command line tools such as (git, netcat, npm, terraform, etc.)
Responsibilities
- Make improvements to internal processes to reduce lead time and increase deployment frequency
- Identify improvements to the quality, security, and performance of our infrastructure
- Increase the velocity with which teams deliver, leveraging expertise from various functional disciplines
- Identify how to remediate production incidents more quickly and safely while reducing the frequency of outages
- Actively engage with other teams and departments to collaborate on best practices and implementation strategy
- Adhere to and advocate for best practices, including Infrastructure as Code, monitoring, high availability, disaster recovery, security, and DevOps methodologies
- Create SLIs, SLOs, and SLAs
- Contribute to capacity planning, advise and consult with teams who will be load/stress testing
- Keep up with industry innovations, recommending new tools or practices when appropriate
- Actively mentor peers, developing their expertise and inspiring others to innovate
- Provide timely assistance and remediation solutions during critical situations and production incident
- Document and share "lessons learned" from production, including root cause analysis
- Explore new ways of improving communication between other Site Reliability Engineers and with other teams
- Write and maintain architectural, stakeholder, and policy documentation
$76k - $127k
...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II The BizOps team at Mastercard is looking for a Site Reliability Engineer (SRE) who thrives on solving complex...SuggestedFull timePart timeWorldwideFlexible hoursEarly shift$96k - $163k
...services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Overview The BizOps team is looking for a Senior Site Reliability Engineer who can help us solve problems and...SuggestedFull timePart timeWorldwideFlexible hoursShift work$76k - $127k
...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that...SuggestedFull timePart timeWorldwideFlexible hours$96k - $163k
...services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that...SuggestedFull timePart timeWorldwideFlexible hours- ...and performance our customers have come to expect, and help raise the reliability bar as we grow. What you would do: Design, build, and operate the shared platform foundations engineers ship on every day: GCP infrastructure, Kubernetes, networking, routing,...SuggestedRemote workWorldwideFlexible hours
$147k - $168k
...Inc. as one of the most innovative and fastest-growing technology companies in the country. Role Summary As a Site Reliability Engineer at Filevine, you will improve the reliability, scalability, and operational maturity of the Filevine platform. You’ll...Full timeTemporary workWork experience placementWork at officeRemote work2 days per week3 days per week- ...Site Reliability Engineer Company: Milestone Systems Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Golang, Python, Linux, Shell scripting, Kubernetes, Docker, Terraform, CI/CD, GitOps, ArgoCD, Spinnaker, Prometheus, Datadog,...Full timeRemote work
- ...in Cupertino, California, invites an experienced CDN Solutions Engineer to join the Content Delivery Network Solutions team. You will... ...and collaborate with engineering groups across Apple to ensure reliable delivery at scale. The ideal candidate has 4+ years in CDNs and...
- ...Oracle Cloud Infrastructure (OCI) seeks a Senior Principal Engineer to lead the design and implementation of reliability validation for OCI control plane services, focusing on a high-performance, low-level systems approach. You will mentor engineers, define validation...
- ...provider is seeking a Senior Staff Technical Program Manager for Reliability to enhance the reliability and performance of their multi-... ...role involves leading programs in partnership with senior engineering leaders, requiring over 10 years of experience in cloud infrastructure...
- ...Motor Company seeks a Director of Cloud SRE to lead a team of engineering leaders and engineers, federating core SRE principles across... ...You will guide cross-domain collaboration, drive AI-enabled reliability, and ensure CI/CD integration while expanding reliability into...Remote work
- ...We are seeking a Senior Database Reliability Engineer (DBRE) to design, operate, and improve reliable, scalable, secure, and highly available... ...and data platforms. The role combines database engineering, site reliability engineering, Linux systems administration, and infrastructure...
- ...ServiceNow in Santa Clara, CA, seeks a Staff Software Engineer – SRE & AIOps to drive infrastructure automation, resilience, and toil... ...for global engineering teams. Embedded within the Site Reliability & Database Engineering organization, you will architect SRE tooling...
- ...Lambda Inc. in San Francisco is seeking a Storage Engineer to own the reliability, performance, and capacity of our production storage fleet across multiple data centers, using a software-defined data plane. You will build monitoring, dashboards, and alerting for storage...
- ...Jobtailor is seeking an experienced Platform Architect/Lead to shape major cloud platform decisions and drive reliability across enterprise-scale workloads. You will own design, implementation, and ongoing improvements for critical systems, with heavy emphasis on observability...
$276.1k - $311.4k
...Vehicle Software SRE team from the ground up — defining its charter, hiring its founding engineers, establishing the operating model, and creating the technical strategy that makes reliability a first-class property of the software running on our vehicles. You'll work in a...Permanent employmentFull timeWork at officeWork from home- ...automated detection, drain/cordon/taint, workload rescheduling. Feed the AIOps substrate The remediation-actuator and workflow engine land here — you make the control plane safe for automated action. Your CRDs are the schema the platform's predictors and...Local area
- ...Job Description The Opportunity: Versant's Sports & Entertainment Digital Products division is seeking a Senior Site Reliability Engineer to help drive the reliability, scalability, and usability of internal developer platforms, tooling, and engineering workflows...Local areaRemote workWorldwide
$186.82k - $224.18k
...made, and we are not tiptoeing into it. We are rebuilding our engineering culture around a simple belief: AI changes everything. How... ...Is Babylist is looking for a Senior Software Engineer, Site Reliability to join our Platform team. In this position, you will play...Work at officeLocal areaImmediate startRemote workFlexible hoursShift work- ...ByteDance’s Infrastructure Engineering team in Seattle designs, builds, and operates global infrastructure spanning public and private... ...storage. Join a fast-paced, collaborative team focused on reliability, scalability, and continuous optimization, driving improvements...
- ...Senior Site Reliability Engineer Company: Sphera Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Terraform, ARM templates, Kubernetes, Azure, SonarCloud, CheckPoint, Hadoop, Kafka, Presto, NewRelic, CI/CD, Linux, Windows, Redis...Full timeRemote work
$125k - $250k
...we are reimagining how developers build reliable, scalable, event-driven applications without... ...possible Partner closely with engineering teams to improve system resiliency and scalability... ...For 5+ years of experience in Site Reliability Engineering, DevOps,...Full timeImmediate startRemote workFlexible hours- ...About the Role We are seeking a Senior Site Reliability Engineer to join our cloud engineering team. You will own the reliability, scalability, and observability of our critical financial SaaS applications and infrastructure, working across cloud platforms to ensure...Remote work
- ...Site Reliability Engineer Platform and software · shared across customers Reports to: Director, Site Reliability Location: Remote (US) Department: Cloud Platform Engineering / SRE/Reliability Position summary The Site Reliability Engineer (SRE) owns...Remote workNight shift
- ...Eyes on glass. Hands on the pipeline. Real ownership from day one. This isn't a watch-and-wait monitoring seat. Our client needs engineers who can read a Kibana query at 3am, know the difference between a blip and a breach, and act on it, on a FedRAMP-authorised cloud...Hourly payFor contractorsRemote workShift workNight shiftWeekend work
- ...Apple Service Engineering (ASE) seeks a senior SRE software engineer to own the architectural direction of Kubernetes internals powering... ...You will define controllers and namespace management, raise reliability, and contribute to upstream Kubernetes. The role includes...
- ...Invisible Technologies is looking for a Principal Software Engineer (SRE/DevOps) to work remotely. The ideal candidate will possess dual expertise in application engineering and infrastructure, contributing to a variety of technical initiatives. This role includes...Remote work
- ...About the Role: We are looking for a Senior Site Reliability Engineer (SRE) to help modernize large-scale infrastructure and improve the reliability, scalability, and operational excellence of critical production systems. In this role, you will lead OS modernization...Remote work
- ...production paths and high-volume event processing in a fast-moving startup environment. We are looking for a Platform/Infrastructure engineer with 4+ years in SRE or DevOps, strong GCP experience, and hands-on work with serverless systems, IAM, and observability. #J-188...
- ...Site Reliability Engineer Company: GitLab Work Type: Remote Employment: Full Time Location: CA, US Seniority: Senior Level Technologies: Terraform, Ansible, Kubernetes, Go, Ruby, Jsonnet, Prometheus, ELK, Grafana Requirements: Senior-level SRE with strong Terraform/IaC...Full timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre United States
- site reliability engineering manager United States
- site reliability engineer United States
- site reliability engineer remote United States
- site recruiter United States
- site services specialist United States
- junior website developer United States
- official site United States
- on site coordinator United States
- site leader United States

