Site Reliability Engineer
Insight Global
Job Description
SAP NS2 is seeking a Site Reliability Engineer - OpenSearch to help ensure the highest levels of availability, performance, scalability, and Quality of Service (QoS) for mission-critical cloud services. This role will focus on the reliability, operations, automation, and continuous improvement of distributed search and analytics platforms built on OpenSearch, while working in a diverse, globally distributed team environment.
The ideal candidate brings deep experience in site reliability engineering, DevOps, cloud operations, automation, observability, and distributed systems, with proven hands-on expertise architecting, building, deploying, operating, and optimizing high-performance OpenSearch clusters and platforms from the ground up in production environments.
General Responsibilities
Provision, build, deploy, monitor, operate, and support cloud services in a globally distributed team environment
Architect, build, deploy, and maintain?high-performance OpenSearch clusters and platforms from the ground up
Administer and optimize OpenSearch environments for?high availability, resiliency, scalability, security, and performance
Monitor and troubleshoot cluster health, node performance, indexing throughput, search latency, shard allocation, replication, and storage utilization
Analyze and resolve operational issues, platform instability, and production incidents across infrastructure, platform, and application layers
Conduct incident response, root cause analysis, and post-incident remediation to drive continuous improvement
Maintain the integrity and security of servers, systems, and OpenSearch platform infrastructure
Support platform lifecycle activities including installation, configuration, upgrades, patching, hotfixes, backup, restore, and disaster recovery
Develop and maintain monitoring policies, alerting standards, operational runbooks, and support procedures
Automate testing, deployment, scaling, recovery, and operational workflows for OpenSearch and related cloud services
Ensure proper resource allocation and capacity planning across compute, memory, storage, and network resources
Partner with product development and engineering teams to design and enhance service reliability and operational readiness
Develop and implement testing strategies and document results for platform changes and operational improvements
Support log ingestion, index management, retention policies, lifecycle management, and search performance tuning
Work in a diverse environment and cross-train with other global team members
Participate in an on-call rotation and support weekend or after-hours operational needs as required
We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment regardless of their race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please send a request to View email address on click.appcast.io learn more about how we collect, keep, and process your private information, please review Insight Global's Workforce Privacy Policy:
Skills and Requirements
8+ years of hands-on experience in SRE or similar DevOps role
Experience deploying, and operating OpenSearch in a Kubernetes-based environment
Experience with Elasticsearch (as back up)
Expertise with?Git
Strong background as a Backend SRE / Platform Engineer supporting large-scale distributed systems
Ability to build and scale OpenSearch infrastructure from initial deployment through production operations ("0 to 100")
Experience with Kubernetes administration, containerized workloads, and cloud-native architectures
Experience infrastructure automation, CI/CD, and SRE best practices
- ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that... ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to...SuggestedFull time
- ...Senior Site Reliability & Cloud Systems Engineer This is a hybrid role - 2 days remote and 3 days in the Malvern, PA office. We are seeking a highly experienced Senior Site Reliability & Cloud Systems Engineer to architect, build, automate, and operate scalable,...SuggestedWork at officeRemote work
- ...Site Reliability Engineer Hybrid - Malvern, PA needs at least 8 years experience within the US The Site Reliability Engineer (SRE) is responsible for improving the reliability, resiliency, observability, and operational excellence of Client's Cash & Money Movement...SuggestedWork at office
$112.5k - $187.5k
...NoticePersonal Information We CollectYour Privacy ChoicesTeam OverviewAt TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential...SuggestedFull timeTemporary workWork experience placementWork at officeFlexible hours2 days per week- Position Summary:The Platform Engineer, Specialist supports the enterprise workload automation and job scheduling environment, ensuring... ...initiatives that enhance workload automation efficiency, reliability, and operational effectiveness.Qualifications:Bachelor's degree...SuggestedFull time
- Director, Enterprise Platform Engineering (Mac & Windows Endpoints)This position serves as the senior leader accountable for the strategy... ...PCI compliance enforcement, endpoint telemetry, and health & reliability engineering.Mature the platform‑as‑a‑product operating model—...Full time
$120.38k - $192.6k
...Senior Software Engineer, Tech LeadThe Role at a Glance We build the customer-facing web applications our policyholders use to manage their coverage — the digital front door to their relationship with us. You'll lead a small team of mid-level and junior engineers building...Work experience placement- ...with combinations of Matlab, Python, Perl, Java, C++, R, or similar platforms.Qualifications Required Skills:Bachelor's degree in engineering (aerospace, mechanical, electrical), physics, mathematics, computer science, data science, image science or similar technical...For contractors
- Company DescriptionSonsoft , Inc. is a USA based corporation duly organized under the laws of the Commonwealth of Georgia. Sonsoft Inc. is growing at a steady pace specializing in the fields of Software Development, Software Consultancy and Information Technology Enabled...Full timeH1bLocal area
- Company DescriptionArtech is the 10th Largest IT Staffing Company in the US, according to Staffing Industry Analysts' 2012 annual report. Artech provides technical expertise to fill gaps in clients' immediate skill-sets availability, deliver emerging technology skill-sets...Immediate start
- ...Engages third parties for the maintenance of supported devices. Qualifications A bachelor's degree in computer science, Computer Engineering, Information Systems, or other related field required. Two to four years related hands-on experience, or the equivalent...
- ...We're Hiring: SRE Production Support Engineer Malvern, PA · Onsite Contract Experience: 5+ yrs Skills: Shell, Bash, AWS,... ...Collaborate with development teams to improve application reliability. Support production deployments and release activities....Permanent employmentContract work
- Job RequirementsBachelor's degree in Computer Science, Information Technology, or a related field, or equivalent experience.Proven experience with observability tools such as Prometheus, Grafana, and Splunk.Hands-on experience with Kubernetes and container orchestration...
- Job TitleOver 12-15 years of overall experience needed.A solid foundation in computer science, with strong competencies in data structures, algorithms, and software design.Large systems software design and development experience.Experience performing in-depth troubleshooting...Remote work
- ...DevOps EngineerLocations- Newtown Square, PA/ Washington DC Metro Area1. Proven experience as a DevOps Engineer or in a similar role.2. Strong knowledge of CI/CD tools (e.g., Jenkins, GitLab CI, CircleCI). Experience in setting up, managing, and optimizing CI/CD pipelines...
- Cloud Development Operations EngineerPossesses a strong background in infrastructure as code, specifically Terraform, and has deep expertise and knowledge in AWS.Must also have deep Ansible experience. Also have experience with Gardner Kubernetes.Fine-tune and improve ...
- ...DevOps Engineer T4Newtown Square, PA (Onsite Travel Required) Job Type: Contract Must be a U.S. CitizenAbout the RoleWe are seeking... ...Qualifications5+ years of experience in DevOps, Cloud Engineering, or Site Reliability Engineering (SRE). Strong expertise in AWS (EC2, S3, IAM,...Contract work
$150.29k - $225.43k
...accurate payments. The Staff Software Engineer drives the design and implementation of... ...issues in production systems, ensuring reliability and uptime Drive adoption of engineering... ...in platform engineering, DevOps, or site reliability Contributions to open source...Full timeRemote work- ...Description Job Description Description: The Senior Software Engineer is responsible for leading the design and construction of high... ...(HIPAA, HITRUST). Optimize system performance and reliability for cloud deployment (AWS). Perform other duties as assigned...Local area
$71.2k - $88.8k
..., visit The Team You'll Join On the Tamarac Product Solutions Engineering team, we consult directly with key clients to optimize their usage... ...journey with us. Please visit our benefits page on our career site to learn more. Our Commitment to Inclusion & Belonging...- ...carefully selected external vendor data provider solutions, ensuring reliable, performant, and well‑governed use of data.Collaborate with... ...systems, data pipelines, and compute‑intensive workloadsDrive engineering best practices around resilience, observability, performance,...Full timeWork experience placement
- Job Title Responsibilities: ~ Automation engineering, specifically Out of Region recovery automation. Qualifications: Developing infrastructure as code Familiarity with Git, Ansible/AWX, Python Scripting and Terraform.
$85.16k - $131.91k
...responsibilities, you will play an important role in supporting team success and platform development.In this role, you will:Partner with engineers and architects to deliver effective technical solutionsCollaborate with stakeholders to understand requirements and improve user...Full time- SAP S4HANA Project ManagerAerospace & Defense/ Discrete Manufacturing experienceThe Project Manager (Commerce & OMS) will be responsible for leading the end-to-end delivery of the digital commerce and order fulfillment strategy, ensuring high-scale performance and seamless...
$102.02k - $153.03k
Microsoft Azure Cloud EngineerWe are seeking an experienced Azure Cloud Engineer to join our Cloud Competence Center. Your responsibilities will... ...position requires a minimum of two days per week on-site in our office in Wayne. Remote working arrangements are not possible...Work at officeLocal areaRemote workFlexible hours2 days per week- SAP Consultant SAP BASIS HANA Consultant Diverse Lynx LLC is an Equal Employment Opportunity employer. All qualified applicants will receive due consideration for employment without any discrimination. All applicants will be evaluated solely on the basis of their ability...
- ...network requirements to make sure the SAP security architecture framework can meet customer requirementsWork with SAP delivery and engineering teams to address customer-specific requirementsPresent the value proposition for SAP private cloud on Hyperscale clouds like...Permanent employmentFull timeWorldwideFlexible hours
- The Lead Senior Network Operations Center (NOC) Engineer (Tier 3) serves as the highest technical escalation point within the global... ...and observability platforms to enhance network visibility and reliability.Partners with Network Engineering teams to validate new designs...Full timeNight shift
$90k - $198.5k
...National Security Services Inc. (SAP NS2) is an independent U.S. subsidiary of SAP. At SAP NS2, we leverage best-in-breed technologies engineered by SAP to protect the lives, assets and information of Americans. Weoffer SAP solutions with specialized levels of security and...Permanent employmentFull timeWorldwideFlexible hours$102.8k - $214.4k
A leading technology firm in Maryland is seeking a Cloud Delivery Specialist to join their Enterprise Cloud Services. The role involves serving as a trusted technical advisor while collaborating with various teams to enhance customer investment in SAP solutions. Ideal candidates...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- on-site clinical research associate (traveling/remote) Newtown Square, PA
- junior site reliability engineer
- site reliability engineer remote
- lead site reliability engineer
- site reliability engineer
- site reliability engineer sre
- site reliability engineering manager
- junior website developer
- website auditor
- site development


