Site Reliability Engineer
Mindlance
Primary Skill Required for the Role Systems Engineer Level Required for Primary Skill dvanced (6-9 years experience) dditional Skills Requested for Role (No Value) dditional Details for Role Site Reliability Engineer
Note: Candidates with sub vendors cannot be considered.
Summary:
Site Reliability Engineer is responsible for supporting reliability driven development and operations through enablement and enhancement of tools, process, best practices, framework, training, and technology. SREs will be part of development lifecycle with more focus on shift left phases including Requirement, Design/Architecture, Coding, and testing and to make sure performance, resiliency, reliability, scalability, and availability are collected, factored in the design, developed, and tested before it goes into production and provide support/guidance for post-production operations.
Responsibilities:
• Facilitate Non-Functional Requirement & SLO Collection and Documentation
• Architecture and Design review for Reliability and Performance identify gaps to be added to the backlog.
• Conduct Single User profiling for resource (CPU, Memory, IO, Network) usage optimization
• Create automation framework/dashboard for faster root cause analysis required for issues arises from Non-functional testing
• Be escalation point for issues that Performance testing team can't figure out or need more support
• Takes ownership and proactively identifies issues and opportunities to improve the systems observability and/or reliability that they are involved in and acts accordingly by doing the work and/or providing a plan to be executed.
• Identify and configure Observability tools for the given application
• Develop dashboards for SLO/SLIs through configured observability tools
• Create alerts with thresholds for application/Infra issues through configured observability tools
• Enable proactive and predictive monitoring and alerting
• Crosstrain developers in the use of Observability Tools Splunk, Dynatrace, Grafana for issue resolution and application tracing
• Develop automated framework (dashboards and reports) for observability to include automated reports for SLOs and Error Budgets.
• Develop automated dashboards/alerts for new functionality/modules
• Develop or support self-healing solutions/framework for repeated prod issues
• Facilitate blameless RCA with the team for Production Issues to develop recommendations for improvement
• Escalation-point for complex issue resolution encountered in Production - escalation path: on-call -'Tech Lead/SME -? TA (Technical Architect)/SRE
• Establish a list of common Stability/Resilience recommendations each team should have in place for all deployments and assist teams with adopting. For existing applications and ensure new development includes the recommendations.
• In liaison with Technical Architects, develop automated framework for production readiness checklist to ensure deployments are configured for stability best practices that reduces risk and that error budgets are in a state to accept the risk of upcoming changes.
• Facilitate integrating Non-Functional testing into CICD pipeline
• Recommend best deployment strategy and tools
Qualifications / Requirements:
• Proficient in SRE Principles and Practices
• Hands on SRE experience with influencing skills on providing solutions and taking decisions
• Proficient in observability tools including Splunk, Dynatrace, Prometheus/Grafana
• Proficient in problem identification and solving
• 5+ years of experience in software, infrastructure, or platform engineering with a combination of the following:
• Scalability, resiliency, reliability analysis of application design/architecture
• Performance Testing/Tuning/Monitoring, maximizing system uptime and availability, ensuring functional and performance SLAs.
• Experience in Agile Methodologies and processes.
• Experience representing technical viewpoints to diverse audiences and making prudent & timely technical risk decisions
• Experience developing automation
• Experience in understanding end to end application component architecture
EEO:
"Mindlance is an Equal Opportunity Employer and does not discriminate in employment on the basis of - Minority/Gender/Disability/Religion/LGBTQI/Age/Veterans."
- ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team,, you will solve complex...SuggestedWork at office
$170k
...hackajob is collaborating with Barclays to connect them with exceptional professionals for this role. Join us as a Senior Site Reliability Engineer for CIAM at Barclays, where you will bring to life a new digital platform capability, transforming and modernizing our...SuggestedHourly payTemporary workWork at office- ...professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of... ...professionals and position yourself among the top echelon in site reliability. As an Associate Site Reliability Engineer at JPMorgan Chase...SuggestedWorldwide
- ...ears, and hands on the ground at a government customer site, ensuring the reliability and performance of Twenty's mission-critical platform running... ...of deep technical ownership and customer-facing engineering: you'll define how we measure reliability, lead incident...SuggestedFull timeWork at officeRemote workFlexible hours
$140k - $170k
...for local candidates** Want to work in technology in the financial industry? Our client is seeking a highly motivated Site Reliability Engineer responsible for ensuring reliability, scalability, and performance of large-scale systems and applications. The role blends...SuggestedFull timeLocal areaWorldwideVisa sponsorshipWork visa- ...education and experience required: Bachelor's degree in Electrical Engineering, Electronic Engineering, Computer Science or related field of study plus 7 years of experience in the job offered or as Site Reliability Engineer, Technical Lead, SharePoint Administrator, Software...Full time
- ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Commercial & Investment Bank, Production management team, you will solve complex and broad...Shift work
- Role Description Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and services. Our SRE teams solve reliability, security, and usability at scale for...Full timeWork at office
$54k - $150k
Role Description As Senior Site Reliability Engineer for Remote Build, you'll own the operational excellence and infrastructure strategy that makes Build's platform reliable, performant, and safe for customers. You'll report to the Engineering Manager and work closely with...Full timeLocal areaRemote workHome officeFlexible hours- Role Description As a Senior Site Reliability Engineer you will champion all things pertaining to reliability at Okta for Auth0. Working closely with the Product Engineers, Quality Engineers, Platform Engineers and Architecture teams, your primary focus will be on ensuring...Full timeRemote work
- Role Description We are looking for a talented and driven Sr. Site Reliability Engineering (SRE) to support our engineering team, which manages the infrastructure and services that power our Waystar products. This role is ideal for an experienced engineer who thrives in...Full timeLive outFlexible hours
- ...Istio) ~Defining and monitoring Service-Level Objectives (SLOs) and Service-Level Agreements (SLAs) to ensure that systems meet reliability and performance targets ~Monitoring Tools like New Relic, Prometheus, Grafana, and/or Datadog ~OpenTelemetry knowledge for...Full timeRemote work
- Role Description As a Site Reliability Engineer on the Central AI team, you will help Health Catalyst engineer teams adopt AI responsibly and effectively. You bring deep experience solutioning and implementing AI systems, and you use that expertise to evaluate architectures...Full time
- Role Description We’re looking for a Senior Site Reliability Engineer who takes ownership seriously — someone who designs for reliability, ships the automation, and stands behind it in production. You’ll work across cloud-native infrastructure on systems that process millions...Full time
- Role Description We are looking for a Site Reliability Engineer (SRE) to join our world-class team. This isn't just an operational "maintenance" role; you will be software-engineering the engine that powers tens of thousands of daily builds, ensuring our platform is as...Full time
- Role Description Versant's Sports & Entertainment Digital Products division is seeking a Senior Site Reliability Engineer to help drive the reliability, scalability, and usability of internal developer platforms, tooling, and engineering workflows across a portfolio of...Full timeLocal areaRemote workWorldwide
$135k - $170k
Role Description Climavision is seeking a Senior Site Reliability Engineer to contribute towards reliability, operational excellence, and production resilience for our customer-facing platform and weather data services. This role is focused on ensuring our systems consistently...Full timeTemporary workFlexible hours$137.9k - $221.4k
...for someone to lead development aspects of the Infrastructure engineering team at ServiceTitan. You must have a strong background in... ...leadership and strong architectural thought process. Our Site Reliability and Infrastructure Engineering team is an investment by Cloud...Full timeImmediate startFlexible hours- Role Description En Experis Argentina nos encontramos en la búsqueda de nuestro Site Reliability Engineer (SRE) para importante Compañía del rubro de Telecomunicaciones. Condiciones de contratación: ~Jornada de trabajo: Full Time 9am – 6pm. ~Modalidad de trabajo:...Full timeRemote work
- Role Description We are looking for a Site Reliability Engineer (SRE) who is passionate about infrastructure reliability, automation, and building scalable production systems. ~Own and improve production infrastructure reliability and stability ~Prepare, execute, and...Full timeRemote work
- Role Description Stack AV Site Reliability Engineers are responsible for enabling and ensuring our production systems meet their service-level objectives. Through the implementation of centralized observability and automation, the SRE team constantly ensures the health...Full time
- Role Description We are seeking a Site Reliability Engineer (SRE) with deep expertise in monitoring, observability, and reliability engineering to support systems running across on-premises infrastructure and Google Cloud Platform (GCP). This role is primarily responsible...Long term contractFull timeRemote workFlexible hours
$140k - $165k
...and optimizing development velocity without compromising on reliability. Qualifications ~Experience in an SRE role with an... ...plan you will receive if you were to be hired as a Senior Site Reliability Engineer at Flock Safety. The First 30 Days ~Onboarding. Make a...Full timeWork at officeWork from homeHome officeFlexible hours- Role Description We’re looking for a Senior Platform Engineer to design, build, and operate the core services that power Optura’s AI... ...systems end-to-end, from model and agent orchestration to routing, reliability, and observability. You will partner closely with product and...Full timeRemote work
- ...through intelligent automation and modern engineering. We are seeking a Senior SRE Engineer... ...efficient delivery, observability, and reliability across Sleek’s products and internal operations... ...~6+ years of progressive experience in Site Reliability Engineering (SRE). ~6+...Full timeRemote workFlexible hours
$54k - $150k
Role Description As Senior Site Reliability Engineer for Remote Build, you'll own the operational excellence and infrastructure strategy that makes Build's platform reliable, performant, and safe for customers. You'll report to the Engineering Manager and work closely with...Full timeLocal areaImmediate startRemote workHome officeFlexible hours- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Enterprise technology, Corporate technology team , you hold a leadership...
- ...professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of... ...professionals and position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer at JPMorgan Chase...
- ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III - DevOps Engineer at JPMorgan Chase within the Commercial and Investment Bank, you will solve complex and broad business...
- ...and shape the future of technology at a globally recognized firm, driven by pride in ownership. As a Senior Manager of Site Reliability Engineering at JPMorgan Chase within the Corporate Investment Bank, Markets team, you are the non-functional requirement owner and...Bank staffShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer United States
- site reliability engineer remote United States
- site reliability engineering manager United States
- site reliability engineer sre United States
- lead site reliability engineer United States
- IT site lead United States
- website coordinator United States
- after school site coordinator United States
- on site coordinator United States
- site merchandiser United States
















