Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

Pacifica Continental

Job Title

Our engineering team has built the largest private Medicare marketplace in the country. We passionately focus on the continuous improvement of the systems we build.

We have spent many years growing and fostering a DevOps culture by bridging the divide between our Software and Infrastructure Engineering departments. We want the cross-functional teams that we are building to include Site Reliability Engineers. We operate in a complex, multi-tenant, hybrid cloud and on-premises infrastructure that spans both the Windows and Linux OS. We strive for security, reliability, and automation in line with DevOps and Site Reliability Engineering principles. If you are passionate about learning and improvement through metrics and automation, and passionate about engendering that mindset in others, we want to hear from you.

About the Role

Maintains shared cloud resources in use by numerous software engineering teams within our business unit. We aim to enable software engineering teams to build cloud native applications that adhere to security and regulatory requirements with limited handholding by our cloud engineers. We do still have a fair number of applications hosted in on-premise data centers, which we aim to support migrating to the cloud.

Requirements

Hands-on Engineering

5+ years of hands-on experience with a majority of the following technologies, along with a willingness to become proficient in the remaining areas:

  • Windows and Linux Servers
  • VMware
  • Cloud platforms, preferably with Azure
  • Active Directory
  • Secrets management with Consul and Vault or similar systems
  • Configuration management tools like Salt, Ansible and Terraform
  • Firewalls and load balancers such as F5
  • Web servers, including IIS and NGINX
  • Database Server Infrastructure like Microsoft SQL Server and PostgreSQL
  • Application Performance Monitoring with tools like New Relic
  • Infrastructure monitoring with tools like Sensu, SolarWinds, Nagios, or Azure App Insights
  • CI/CD tools like TeamCity, Octopus Deploy, Concourse, Azure DevOps, or GitHub Actions
  • Log Aggregation tools like SumoLogic or Splunk
  • Network theory and protocols such as DNS, DHCP, proxy servers, and firewalls
  • Security operations with tools for SAST, DAST, RAST, and WAF
  • Infrastructure as Code or automation experience.

Proficiency, high-comfort, and familiarity with:

  • One or more programming languages, such as C#, JavaScript, Python or Go
  • One or more scripting languages, such as PowerShell and BASH
  • Command line tools such as (git, netcat, npm, terraform, etc.)
Responsibilities
  • Make improvements to internal processes to reduce lead time and increase deployment frequency
  • Identify improvements to the quality, security, and performance of our infrastructure
  • Increase the velocity with which teams deliver, leveraging expertise from various functional disciplines
  • Identify how to remediate production incidents more quickly and safely while reducing the frequency of outages
  • Actively engage with other teams and departments to collaborate on best practices and implementation strategy
  • Adhere to and advocate for best practices, including Infrastructure as Code, monitoring, high availability, disaster recovery, security, and DevOps methodologies
  • Create SLIs, SLOs, and SLAs
  • Contribute to capacity planning, advise and consult with teams who will be load/stress testing
  • Keep up with industry innovations, recommending new tools or practices when appropriate
  • Actively mentor peers, developing their expertise and inspiring others to innovate
  • Provide timely assistance and remediation solutions during critical situations and production incident
  • Document and share "lessons learned" from production, including root cause analysis
  • Explore new ways of improving communication between other Site Reliability Engineers and with other teams
  • Write and maintain architectural, stakeholder, and policy documentation
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in United States vacancy
  • $76k - $127k

     ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II The BizOps team at Mastercard is looking for a Site Reliability Engineer (SRE) who thrives on solving complex... 
    Suggested
    Full time
    Part time
    Worldwide
    Flexible hours
    Early shift

    Mastercard

    O Fallon, MO
    19 hours ago
  • $96k - $163k

     ...services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Overview The BizOps team is looking for a Senior Site Reliability Engineer who can help us solve problems and... 
    Suggested
    Full time
    Part time
    Worldwide
    Flexible hours
    Shift work

    Mastercard

    O Fallon, MO
    19 hours ago
  • $76k - $127k

     ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that... 
    Suggested
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    19 hours ago
  • $96k - $163k

     ...services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that... 
    Suggested
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    19 hours ago
  •  ...and performance our customers have come to expect, and help raise the reliability bar as we grow. What you would do: Design, build, and operate the shared platform foundations engineers ship on every day: GCP infrastructure, Kubernetes, networking, routing,... 
    Suggested
    Remote work
    Worldwide
    Flexible hours

    Sanity

    United States
    2 days ago
  • $147k - $168k

     ...Inc. as one of the most innovative and fastest-growing technology companies in the country. Role Summary As a Site Reliability Engineer at Filevine, you will improve the reliability, scalability, and operational maturity of the Filevine platform. You’ll... 
    Full time
    Temporary work
    Work experience placement
    Work at office
    Remote work
    2 days per week
    3 days per week

    Filevine

    United States
    2 days ago
  •  ...Site Reliability Engineer Company: Milestone Systems Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Golang, Python, Linux, Shell scripting, Kubernetes, Docker, Terraform, CI/CD, GitOps, ArgoCD, Spinnaker, Prometheus, Datadog,... 
    Full time
    Remote work

    Milestone Systems Inc

    United States
    2 days ago
  •  ...in Cupertino, California, invites an experienced CDN Solutions Engineer to join the Content Delivery Network Solutions team. You will...  ...and collaborate with engineering groups across Apple to ensure reliable delivery at scale. The ideal candidate has 4+ years in CDNs and... 

    Apple

    Cupertino, CA
    1 day ago
  •  ...Oracle Cloud Infrastructure (OCI) seeks a Senior Principal Engineer to lead the design and implementation of reliability validation for OCI control plane services, focusing on a high-performance, low-level systems approach. You will mentor engineers, define validation... 

    Oracle

    Santa Clara, CA
    4 days ago
  •  ...provider is seeking a Senior Staff Technical Program Manager for Reliability to enhance the reliability and performance of their multi-...  ...role involves leading programs in partnership with senior engineering leaders, requiring over 10 years of experience in cloud infrastructure... 

    Menlo Ventures

    Bellevue, WA
    19 hours ago
  •  ...Motor Company seeks a Director of Cloud SRE to lead a team of engineering leaders and engineers, federating core SRE principles across...  ...You will guide cross-domain collaboration, drive AI-enabled reliability, and ensure CI/CD integration while expanding reliability into... 
    Remote work

    Ford Motor Company

    United States
    1 day ago
  •  ...We are seeking a Senior Database Reliability Engineer (DBRE) to design, operate, and improve reliable, scalable, secure, and highly available...  ...and data platforms. The role combines database engineering, site reliability engineering, Linux systems administration, and infrastructure... 

    Neshent Technologies

    Los Gatos, CA
    4 days ago
  •  ...ServiceNow in Santa Clara, CA, seeks a Staff Software Engineer – SRE & AIOps to drive infrastructure automation, resilience, and toil...  ...for global engineering teams. Embedded within the Site Reliability & Database Engineering organization, you will architect SRE tooling... 

    ServiceNow

    Santa Clara, CA
    4 days ago
  •  ...Lambda Inc. in San Francisco is seeking a Storage Engineer to own the reliability, performance, and capacity of our production storage fleet across multiple data centers, using a software-defined data plane. You will build monitoring, dashboards, and alerting for storage... 

    Lambda

    San Francisco, CA
    1 day ago
  •  ...Jobtailor is seeking an experienced Platform Architect/Lead to shape major cloud platform decisions and drive reliability across enterprise-scale workloads. You will own design, implementation, and ongoing improvements for critical systems, with heavy emphasis on observability... 

    Jobtailor

    Florida, NY
    4 days ago
  • $276.1k - $311.4k

     ...Vehicle Software SRE team from the ground up — defining its charter, hiring its founding engineers, establishing the operating model, and creating the technical strategy that makes reliability a first-class property of the software running on our vehicles. You'll work in a... 
    Permanent employment
    Full time
    Work at office
    Work from home

    Lindus Health

    Sunnyvale, CA
    3 days ago
  •  ...automated detection, drain/cordon/taint, workload rescheduling. Feed the AIOps substrate The remediation-actuator and workflow engine land here — you make the control plane safe for automated action. Your CRDs are the schema the platform's predictors and... 
    Local area

    Bitdeer Technologies Group

    San Jose, CA
    4 days ago
  •  ...Job Description The Opportunity: Versant's Sports & Entertainment Digital Products division is seeking a Senior Site Reliability Engineer to help drive the reliability, scalability, and usability of internal developer platforms, tooling, and engineering workflows... 
    Local area
    Remote work
    Worldwide

    Versant

    United States
    2 days ago
  • $186.82k - $224.18k

     ...made, and we are not tiptoeing into it. We are rebuilding our engineering culture around a simple belief: AI changes everything. How...  ...Is Babylist is looking for a Senior Software Engineer, Site Reliability to join our Platform team. In this position, you will play... 
    Work at office
    Local area
    Immediate start
    Remote work
    Flexible hours
    Shift work

    Babylist

    United States
    2 days ago
  •  ...ByteDance’s Infrastructure Engineering team in Seattle designs, builds, and operates global infrastructure spanning public and private...  ...storage. Join a fast-paced, collaborative team focused on reliability, scalability, and continuous optimization, driving improvements... 

    ByteDance

    Seattle, WA
    19 hours ago
  •  ...Senior Site Reliability Engineer Company: Sphera Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Terraform, ARM templates, Kubernetes, Azure, SonarCloud, CheckPoint, Hadoop, Kafka, Presto, NewRelic, CI/CD, Linux, Windows, Redis... 
    Full time
    Remote work

    Sphera

    United States
    2 days ago
  • $125k - $250k

     ...we are reimagining how developers build reliable, scalable, event-driven applications without...  ...possible Partner closely with engineering teams to improve system resiliency and scalability...  ...For 5+ years of experience in Site Reliability Engineering, DevOps,... 
    Full time
    Immediate start
    Remote work
    Flexible hours

    Orkes

    United States
    2 days ago
  •  ...About the Role We are seeking a Senior Site Reliability Engineer to join our cloud engineering team. You will own the reliability, scalability, and observability of our critical financial SaaS applications and infrastructure, working across cloud platforms to ensure... 
    Remote work

    MeridianLink

    United States
    2 days ago
  •  ...Site Reliability Engineer Platform and software · shared across customers Reports to: Director, Site Reliability Location: Remote (US) Department: Cloud Platform Engineering / SRE/Reliability Position summary The Site Reliability Engineer (SRE) owns... 
    Remote work
    Night shift

    STN Inc

    United States
    2 days ago
  •  ...Eyes on glass. Hands on the pipeline. Real ownership from day one. This isn't a watch-and-wait monitoring seat. Our client needs engineers who can read a Kibana query at 3am, know the difference between a blip and a breach, and act on it, on a FedRAMP-authorised cloud... 
    Hourly pay
    For contractors
    Remote work
    Shift work
    Night shift
    Weekend work

    C-Serv

    United States
    2 days ago
  •  ...Apple Service Engineering (ASE) seeks a senior SRE software engineer to own the architectural direction of Kubernetes internals powering...  ...You will define controllers and namespace management, raise reliability, and contribute to upstream Kubernetes. The role includes... 

    Socket

    San Francisco, CA
    2 days ago
  •  ...Invisible Technologies is looking for a Principal Software Engineer (SRE/DevOps) to work remotely. The ideal candidate will possess dual expertise in application engineering and infrastructure, contributing to a variety of technical initiatives. This role includes... 
    Remote work

    Invisible Technologies Inc. Defunct

    United States
    4 days ago
  •  ...About the Role: We are looking for a Senior Site Reliability Engineer (SRE) to help modernize large-scale infrastructure and improve the reliability, scalability, and operational excellence of critical production systems. In this role, you will lead OS modernization... 
    Remote work

    Halo Media

    United States
    2 days ago
  •  ...production paths and high-volume event processing in a fast-moving startup environment. We are looking for a Platform/Infrastructure engineer with 4+ years in SRE or DevOps, strong GCP experience, and hands-on work with serverless systems, IAM, and observability. #J-188... 

    BAM VENTURES LLC

    New York, NY
    1 day ago
  •  ...Site Reliability Engineer Company: GitLab Work Type: Remote Employment: Full Time Location: CA, US Seniority: Senior Level Technologies: Terraform, Ansible, Kubernetes, Go, Ruby, Jsonnet, Prometheus, ELK, Grafana Requirements: Senior-level SRE with strong Terraform/IaC... 
    Full time
    Remote work

    GitLab

    United States
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!