SRE
$175k - $225kBenton Partners
Senior Site Reliability Engineer - Platform
Chicago New York
We are looking for a Site Reliability Engineer, to join our growing Platform Engineering team, who can cultivate our SRE philosophy, processes, and technologies from the ground up. This role entails driving standards and fostering adoption across our technology teams, whilst closely partnering with our DevOps and Cloud teams.
With a hands-on approach, you'll work across both cloud and on-premises hosting platforms, ensuring the reliability and scalability of our trading systems and production environments. This is a chance to play a pivotal role in transforming our operational capabilities and enhancing performance across a wide array of environments and platforms.
Key Responsibilities:
- Develop and promote our SRE philosophy, establishing best practices and processes that will be instrumental in scaling our infrastructure.
- Implement and scale end-to-end observability and monitoring solutions using Prometheus, Grafana, Loki, and Tempo, ensuring high visibility into application performance and infrastructure health.
- Participate in on-call rotation with approximately 1 week per month of on-call time shared equally across members of the team
- Review and define standards for application reliability requirements within our Kubernetes environment, ensuring application configuration is optimized for performance, cost and reliability.
- Develop automation and tooling to improve efficiency and reliability of deployment pipelines, system health checks, and recovery procedures.
- Collaborate with development teams to enhance service stability, scalability, and fault tolerance through SRE best practices like blameless post-mortems and service level objectives (SLOs).
To be considered a good fit, you must have:
- 8+ years of experience in SRE or similar roles within complex, distributed systems environments.
- A Bachelor's degree in engineering, computer science, information systems, or equivalent experience
- SME with key SRE technologies such as Prometheus, Grafana, Loki, Tempo, and Open Telemetry.
- Extensive knowledge of container orchestration using Kubernetes and containerization with Docker.
- Hands-on experience with both cloud (AWS preferred) and on-premises hosting platforms.
- Proven ability to script in languages like Python, Bash, or Go, to automate routine tasks and deployment pipelines.
- Strong understanding of CI/CD principles, agile methodologies, and DevOps culture.
- High level of initiative, passion for reliability engineering, detail orientation, and follow-through capabilities.
- Exceptional interpersonal and communication skills, with the ability to explain complex technical concepts to a diverse audience.
With respect to NY, CA, and IL based applicants, the starting base pay range for this role is between USD 175000 and USD 225000 annually. The actual base pay is dependent upon several factors, including, but not limited to, relevant experience, business needs and market demands. This role may also be eligible for bonus compensation and employee benefits.
$137.4k - $233.6k
...that help detect and address issues before they impact the business.Foster a broader community of practiceHelp build and sustain an SRE community of practice by collaborating to identify common improvement areas and define standards and governance.Communicate effectively...SuggestedFull timeH1bWorldwideFlexible hours$177.6k - $257.4k
The application window is expected to close on: 09/29/2026Meet the TeamAs a Site Reliability Engineering (SRE) Technical Leader on the Intersight Team you will play a key role in ensuring the reliability, scalability, and security of our cloud platforms. The broader team...SuggestedFull timeTemporary workLocal areaImmediate startFlexible hours$111.61k - $131.3k
...support, product/project management, or application developmentPreferred Skills/ExperienceStrongexpertiseinSiteReliabilityEngineering(SRE),DevOps,ProductionSupport,PlatformEngineering,andDistributedSystemsOperations.Experienceleadingtechnicalteams,incidentresponseefforts...SuggestedFull timeWork experience placementLocal area3 days per week$140k - $170k
...from you: BA/BS, in a related technical field; or the equivalent in education and work experience 8+ years of experience in DevOps, SRE, platform engineering, or similar roles supporting application teams running production services Strong CI/CD experience (Jenkins and...SuggestedFull timeWork experience placementFlexible hours- ...experienced Developer & Infrastructure Experts to evaluate AI-powered workflows across software development, cloud infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands, configurations, and workflows and provide practical feedback based on...SuggestedRemote jobFor contractors
$165k - $225k
...trends, forecast capacity needs, and optimize resource allocation for various workloads. Requirements Experience: 5+ years in SRE, DevOps, or infrastructure engineering roles with proven experience operating production infrastructure at scale. Kubernetes Infrastructure...Remote workFlexible hours- ...Job Description Job Description Site Reliability Engineer (SRE) – Application Support Diversified Services Network, Inc. (DSN) is seeking a full-time Site Reliability Engineer (SRE) – Application Support to join our team in their choice of our Chicago, IL or Peoria...Full timeWork at officeShift workWeekend work
- ...Site Reliability Engineer (SRE) Immediate need for a talented Site Reliability Engineer (SRE). This is a 12+ months contract opportunity with long-term potential and is in Chicago, IL (Hybrid). Key Requirements and Technology Experience: ~ Must have skills:...Contract workLocal areaImmediate start
$63 - $90 per hour
...Job Summary Our client, a leader financial services provider, is seeking a Senior DevOps Engineer / Site Reliability Engineer (SRE) to support enterprise backup, cyber recovery, and platform resiliency initiatives. The ideal candidate brings experience in backup and...Permanent employmentContract workLocal area$106k - $130.6k
...About the Team & Role We are looking for a highly motivated and high-potential mid-level Site Reliability Engineer Manager (SRE) to join our team and help drive meaningful business impact while accelerating your growth as a reliability leader. This is an exciting...Full timeFlexible hours- ...around language models, autonomous agents, machine learning, and agentic/AI-augmented systems• Ability to work closely with software, SRE, and cloud engineers• Software development experience is a plus• CISSP, OSCP, GIAC, and/or AWS Certified Security Specialty a...Full timeImmediate start
- ...problems, rather than execute one off remediations.What We're Looking For2+ years of experience in a DevOps, Platform Engineering, or SRE role.Working knowledge of Infrastructure as Code tools (Terraform/OpenTofu, Helm).A solid DevOps/SDLC mindset, with willingness to...Full time
$168.48k - $272.95k
...risk, governance, privacy, and data protection standards across all engineering activities.Partner with infrastructure, operations, SRE, and end-user support teams to ensure stable production operations and smooth lifecycle support for delivered capabilities.Provides leadership...Full timeWork at officeRemote workRelocationVisa sponsorshipRelocation package- ...Engineering, or a related field, or equivalent practical experience.4+ years of experience in DevOps, Site Reliability Engineering (SRE), Cloud Engineering, or a related role supporting production environments.Hands-on experience with Docker and containerization technologies...Full timeWork at office
$96.8k - $145.2k
...compliance automation.Experience supporting event-driven and microservices-based architectures.Knowledge of service reliability engineering (SRE) practices.Experience with AI-powered developer productivity and operational tooling.This role will have a Hybrid work schedule, with...Full timeTemporary workWork at office3 days per week- ...performance of critical applications and infrastructure.Requirements:7+ years of experience in DevOps, Site Reliability Engineering (SRE), Platform Engineering, or related infrastructure-focused roles.Proven experience designing, deploying, and operating large-scale production...Full time
$164k
...and Infrastructure experienceStrong development experience in Ruby on Rails and Go Lang.Previous professional experience in a DevOps/SRE/ProdEng role supporting customer-facing web services.Previous success in enabling effective partnerships and strong communication with...Full timeWork at officeLocal areaRemote work$250k - $350k
...at driving cultural change; Evangelizing and implementing transformational initiatives where the target state included agile, DevOps, SRE, cloud adoption at scale. Able to consolidate operations capabilities in large, complex enterprise organizations to deliver improved...Full timeLocal area- ...technologies.Experience using AI-assisted engineering tools with appropriate validation, governance, and human oversight.Working knowledge of SRE practices including service-level objectives, post-incident improvement, runbooks, and toil reduction.Strong collaboration,...Full timeWork at office
$100k - $120k
...incident response.Works proactively to prevent incidents and reduce their impact on our platform.Partners with the larger Cloud Operations, SRE, Engineering teams, and the business-at-large to advance our SaaS platforms.Other duties as assigned.QualificationsBachelor's degree...Full timeTemporary workWork experience placementFlexible hours$182.33k - $235.95k
..., or equity, option and futures trading platforms.Familiarity with the FIX protocol and OEMS or venue integrations.Observability and SRE practice — OpenTelemetry, Prometheus, Grafana, or Datadog, and working with SLOs and error budgets.Containerization (Docker/Kubernetes...Full timeContract workWork at office$106k - $117k
...degree in computer science, engineering, or a related field.Experience:5+ years of experience in an infrastructure, DevOps, platform, or SRE role supporting production systems.3+ years of experience supporting Linux servers in a production environment.5+ years of experience...Full timeWork experience placement$145k - $175k
...Establish and document standards, patterns, and runbooks that reduce tribal knowledge and make systems easier for the team to support.Embed SRE principles and engineering practices into how systems are built and how the team works day to day, favoring measurable reliability,...Full timeWork at officeImmediate startShift work$130k - $190k
...vulnerability remediation, and compliance (SOC 2, ISO 27001, etc.).Mentor and provide technical leadership to mid-level and senior DevOps/SRE engineers; conduct design and code reviews.Drive cost-optimization initiatives across cloud spend without compromising reliability or...Work at officeLocal areaVisa sponsorship3 days per week$100k - $130k
...plus) ~ Contribute to infrastructure-as-code practices using tools such as Terraform, CDK, or Pulumi ~ Collaborate with DevOps/SRE to define deployment pipelines and reliability standards for integration services ~ Construct evidence-based business cases for...Work experience placementLocal areaRemote workFlexible hours$101k - $203k
...clients. You will design and implement scalable DevOps pipelines, cloud infrastructure automation, and site reliability engineering (SRE) best practices to optimize application delivery and performance. In this client-facing role, you will work with cross-functional teams...Full timeWork experience placementInternshipLocal area$103.5k - $172.5k
...You possess mid-level development proficiency in Java or Python and experience managing data layers like Oracle, Postgres, or BigQuery.SRE DNA: A profound understanding of SRE principles, specifically in defining and defending SLIs, SLOs, and SLAs to maintain system...Full timeWorldwide$75 - $80 per hour
...from you: • BA/BS, in a related technical field; or the equivalent in education and work experience • 8+ years of experience in DevOps, SRE, platform engineering, or similar roles supporting application teams running production services • Strong CI/CD experience (Jenkins...Contract workTemporary workWork experience placementWork at officeLocal area$90k - $150k
...globally distributed databases.Experience designing multi-region and multi-cloud architectures.Knowledge of Site Reliability Engineering (SRE), SLOs, SLIs, error budgets, and operational readiness practices.Architecture, Cloud, Kubernetes, Security, or SRE certifications....Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week$62 - $80 per hour
...position is remote within the U.S., with alignment to Eastern Time business hours preferred.Required Skills & ExperienceSignificant SRE, DevOps, or cloud infrastructure engineering experience managing production AWS environments at scale, combined with strong Linux systems...Full timeTemporary workRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE. Be the first to apply!

