Site Reliability Engineer
$100k - $115kAnalytic Partners Inc
Denver, Colorado, United States / Dallas, Texas, United States / Miami, Florida, United StatesIT – Cloud & Corporate Technology Team /Regular Employee /HybridAnalytic Partners is a global leader in commercial measurement and optimization, turning data into expertise for the world’s largest brands for almost 25 years. Our holistic approach to decisioning is powered by our industry-leading platform and team of experts, who help leaders make better decisions, faster – unlocking business growth and creating powerful customer connections.With clients in 50+ countries and global offices across New York City, Miami, Dallas, Dublin, London, Paris, Singapore, Shanghai, Munich, Poznan, Sydney, Melbourne, Charlottesville and Denver, we’re growing fast. And we’re looking for top talent to join us in shaping the future of analytics. To learn more about what we do, visit analyticpartners.com – and see why we’re recognized as a Leader in the industry by independent research firms Forrester and Gartner.What You’ll Be DoingOwn the Internal Developer Platform (IDP) as a product, treating engineering teams as customers and optimizing for reliability, usability, and delivery velocity.Define and execute a platform roadmap aligned with business priorities, developer needs, and long-term scalability.Design, build, and evolve paved roads for application delivery, including CI/CD pipelines, infrastructure templates, service scaffolding, and standardized deployment patterns.Build self-service capabilities that enable teams to provision, deploy, observe, and operate services with minimal friction.Create and maintain reusable platform abstractions across AWS and Azure that standardize security, reliability, networking, and observability.Reduce developer cognitive load by abstracting unnecessary complexity while enforcing clear guardrails for security, cost, and compliance.Partner closely with application, product, and security teams to embed reliability, scalability, and security by design.Establish and evolve platform standards for logging, monitoring, alerting, tracing, and incident response workloads.Define, measure, and manage SLIs, SLOs, and error budgets for shared platform services.Drive the reduction of operational toil through automation, standardization, and platform-first solutions.Ensure shared platform services meet high standards for availability, performance, resilience, and scalability.Own system-to-system integration and messaging patterns used across the platform.Lead capacity planning, demand forecasting, and performance tuning for platform services.Plan and execute zero-downtime upgrades, migrations, and releases of platform components.Lead platform-level incident response workflows, post-incident reviews, and drive systemic improvements rather than one-off fixes.Evaluate incoming platform requests and translate them into scalable, productized capabilities.Mentor engineers and drive platform adoption through documentation, enablement, and technical evangelism.Participate in a 24x7 on-call rotation as an escalation point for platform reliability and availability issues.Operate effectively in ambiguous problem spaces, making sound architectural and product decisions with limited guidance.What We Look For In You:Bachelor’s degree in Computer Science or equivalent practical experience.4+ years of experience in Platform Engineering, Site Reliability Engineering, DevOps, or Systems Engineering roles.Strong expertise in Linux and Windows operating systems.Advanced automation and scripting skills using Python, Bash, and/or PowerShell.Deep, hands-on experience designing and operating AWS and Azure platforms at scale.Strong experience building and operating CI/CD platforms (Jenkins, GitHub Actions or equivalent).Strong experience with Infrastructure as Code and configuration management (Terraform, CloudFormation, ARM, or similar).Production experience with containerized and orchestration platforms such as Docker and Kubernetes.In-depth experience with the HashiCorp ecosystem (Nomad, Consul, Vault).Strong understanding of distributed systems, cloud-native architectures, and reliability patterns.Experience designing and operating observability platforms (e.g., Splunk, Sumo Logic, or similar).Familiarity with security and compliance practices, including vulnerability scanning and enterprise security tooling.Strong understanding of the software delivery lifecycle, release engineering, and platform lifecycle management.Experience working in Agile / DevOps environments with a strong product mindset.Demonstrated ability to influence without authority, set standards, and drive adoption across teams.Excellent communication skills, able to translate platform capabilities into clear developer value.Strong problem-solving skills with a bias toward durable, scalable solutions over short-term fixes.A mindset of continuous improvement, curiosity, and learning.Comfortable supporting a global, follow-the-sun operation when needed.How We Measure Success:Strong developer adoption and satisfaction with the platform (DX).Reduced deployment friction, lead time, and operational toil.Platform reliability and performance meeting or exceeding defined SLOs.Consistent, high-quality service delivery across engineering teams.Reduced incident frequency and severity driven by systemic platform improvements.Increased standardization, automation, and self-service adoption across the organization.$100,000 - $115,000 a yearOur differentiator is – Our People! We hire the brightest talent and develop them into leaders. We foster a culture of PEOPLE, PASSION and GROWTH. People: We value our people, customers, and partnersPassion: We love what we doGrowth: Unlimited growth means unlimited potentialAP is a customer-focused, team-oriented organization where innovation and results are rewarded, and individuals can chart the course of their own careers.As a woman founded and led company, this has meant supporting a meritocracy where everyone has opportunities to achieve their best and ensure we foster an environment of diversity, equity, and inclusion. In practice this means we will not only work to recruit a diverse workforce, but also maximize the full potential of all of our people. You can read more about our commitment to DEI Here
- ...Talent Acquisition Team will reach out to help you navigate our interview process.Lantern is seeking an experienced Senior Site Reliability Engineer to champion the reliability, availability, and performance of our Azure-based healthcare platform. In this pivotal role,...Suggested
$138.4k - $173k
...infrastructure as well as help improve the reliability, quality of services and overall... ...recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability... ...about our locations by visiting our site.Compensation & BenefitsThe base salary that...SuggestedFull timeFlexible hours- ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that... ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to...SuggestedFull time
- Qualifications: 8+ years of Software Engineering experience, or equivalent demonstrated through... ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform... ...vendor resources Willingness to work on-site at stated location in the job openingDepartment...SuggestedContract workFor contractorsWork experience placement
$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...SuggestedFull timeWork at officeLocal areaRemote workWork from home$197.3k - $313.7k
...ensure you are not duplicating efforts. Job Category Software Engineering Job Details About Salesforce Salesforce is the #1 AI CRM,... ...and you are the future of Salesforce. Job Title: Director, Site Reliability Engineering Location: New York, NY; San Francisco, CA;...Full timeImmediate start$174k - $253k
...systems by pushing for changes that improve reliability and velocity.Practice sustainable... ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical... ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you...$147k - $211k
...product or system development code.Review code developed by other engineers and provide feedback to ensure best practices (e.g., style... ..., and troubleshooting large-scale distributed systems. Site Reliability Engineering (SRE) is what you get when you treat operations...- ...Administrator / SRE in Dallas to own production Java environments, middleware, and cloud automation. You will optimize performance, drive reliability, and mentor teammates while aligning with enterprise security and AI-enabled integrations. You will work across Java apps, IBM...
- ...Site Reliability Engineer- W2 Role* Technical proficiency: Strong Proficiency in Java, Strong understanding of Database concepts (Oracle, SQL, Dynamo DB etc.) Industry standard SRE Tools like Prometheus, Grafana, Data Dog Etc Good to have skills: Cloud Concepts / AWS,...
- ...The Depository Trust & Clearing Corporation (DTCC) is seeking a Senior Application Support Engineer (SRE) to enhance reliability, scalability, and performance of mission-critical applications. You will apply SRE principles across engineering, infrastructure, and operations...
$207k - $301k
Develop strong, influential relationships with multiple stakeholders across the Site Reliability Engineering and Developer organizations.Serve as an expert on particular fields of knowledge related to rate limiting or sharding.Develop plans and lead projects on evolving...$114k - $148k
...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based...Full timeTemporary workWork experience placementRemote work- Site Reliability Engineer - Vice PresidentSite Reliability Engineering (SRE) is an engineering discipline that combines software and systems engineering to build and run scalable, massively distributed, fault-tolerant systems. At Goldman Sachs, SRE is responsible for improving...
$141.8k - $195k
...their best work, grow fast, and bring their full selves to the herd. Why You’ll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in the U.S. Cribl...Temporary workRemote work- ...This RoleAs a Senior Application Support Engineer, you will help power DTCC's global... ...markets infrastructure by ensuring the reliability, availability, and performance of Institutional... ...processing and settlement.Leveraging Site Reliability Engineering (SRE) principles...Flexible hoursAfternoon shift
$262k - $365k
Lead a team of Software/Systems Engineers on projects for users and be directly responsible for uptime.Own end-to-end availability... ...qualifications:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software and systems engineering...- We are Compliance Engineering, a global team of more than 300 engineers and scientists who work on the most complex, mission-critical... ...evolving systems by pushing for changes that improve capacity and reliability.Practicing sustainable incident management in a blameless...
$207k - $301k
Lead a team of Software/Systems Engineers on projects for users and be directly responsible for uptime.Own end-to-end availability... ...qualifications:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software and systems engineering...- Compliance EngineeringWe are Compliance Engineering, a global team of more than 500 engineers and scientists who work on the most complex... ...systems by pushing for changes that improve capacity and reliability.Practicing sustainable incident management in a blameless postmortem...
$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week$128.47k - $192.71k
...specific processes. • Leads a variety of types of business process design initiatives. • Assesses potential implications of re-engineering for multiple functions or departments. • Demonstrates mastery of re-engineering concepts, methods, and tools. Growth and Agility...Full timePart timeWorldwideFlexible hours- ...Pay Rate: $40/Hr. W2 Experience: 3-5 Years Overview We are seeking a remote Junior SRE/DevOps Engineer role. The ideal candidate has foundational knowledge of Site Reliability Engineering (SRE) and Kubernetes, and is enthusiastic about growing in a DevOps‑driven environment...Long term contractContract workInternshipRemote work
- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
$40 per hour
...A technology solutions provider is seeking a remote Junior SRE/DevOps Engineer. The ideal candidate should have foundational knowledge of Site Reliability Engineering (SRE) and Kubernetes. Responsibilities include gaining experience in a DevOps-driven environment. Applicants...Long term contractInternshipRemote work$141.3k - $237.4k
...AT&T, you won’t just imagine the future, you’ll build it.We are seeking a highly skilled and hands-on Lead Software Engineer to join Software Reliability Engineering (SRE) Onboarding and automation team. This role will drive innovation through automation, enhancement of...Full timeTemporary workWork at officeLocal areaRelocation- ...Job Description Job Description Forhyre is looking for engineers who can bring unique perspectives and innovative ideas to all areas... ...evangelize cloud best practices while building a culture of reliability and observability Engage in and improve the end to end lifecycle...
- ...maintaining, auditing, and assisting with Reliability Excellence (RX) and Process Safety... ...futureMust have Bachelor's Degree in an Engineering field from an accredited institution or... ...support RX related enhancement projectsLead site in achieving RX goals and maintenance system...Contract workFor contractorsLocal areaMonday to Friday
- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
$110k - $170k
...more information, visit . Follow Shield AI on LinkedIn, X, Instagram, and YouTube. Job Description:We are seeking a skilled Systems Engineer to join the Hivemind Solutions Systems Engineering team. In this role, the candidate will support the design, development,...Full timeTemporary workPart timeWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer Dallas, TX
- site reliability engineer sre Dallas, TX
- site services specialist Dallas, TX
- construction site safety Dallas, TX
- site leader Dallas, TX
- official site Dallas, TX
- website content developer Dallas, TX
- on site coordinator Dallas, TX
- IT site lead Dallas, TX
- site safety Dallas, TX

