Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

$100k - $115k

Analytic Partners Inc

Denver, Colorado, United States / Dallas, Texas, United States / Miami, Florida, United StatesIT – Cloud & Corporate Technology Team /Regular Employee /HybridAnalytic Partners is a global leader in commercial measurement and optimization, turning data into expertise for the world’s largest brands for almost 25 years. Our holistic approach to decisioning is powered by our industry-leading platform and team of experts, who help leaders make better decisions, faster – unlocking business growth and creating powerful customer connections.With clients in 50+ countries and global offices across New York City, Miami, Dallas, Dublin, London, Paris, Singapore, Shanghai, Munich, Poznan, Sydney, Melbourne, Charlottesville and Denver, we’re growing fast. And we’re looking for top talent to join us in shaping the future of analytics. To learn more about what we do, visit analyticpartners.com – and see why we’re recognized as a Leader in the industry by independent research firms Forrester and Gartner.What You’ll Be DoingOwn the Internal Developer Platform (IDP) as a product, treating engineering teams as customers and optimizing for reliability, usability, and delivery velocity.Define and execute a platform roadmap aligned with business priorities, developer needs, and long-term scalability.Design, build, and evolve paved roads for application delivery, including CI/CD pipelines, infrastructure templates, service scaffolding, and standardized deployment patterns.Build self-service capabilities that enable teams to provision, deploy, observe, and operate services with minimal friction.Create and maintain reusable platform abstractions across AWS and Azure that standardize security, reliability, networking, and observability.Reduce developer cognitive load by abstracting unnecessary complexity while enforcing clear guardrails for security, cost, and compliance.Partner closely with application, product, and security teams to embed reliability, scalability, and security by design.Establish and evolve platform standards for logging, monitoring, alerting, tracing, and incident response workloads.Define, measure, and manage SLIs, SLOs, and error budgets for shared platform services.Drive the reduction of operational toil through automation, standardization, and platform-first solutions.Ensure shared platform services meet high standards for availability, performance, resilience, and scalability.Own system-to-system integration and messaging patterns used across the platform.Lead capacity planning, demand forecasting, and performance tuning for platform services.Plan and execute zero-downtime upgrades, migrations, and releases of platform components.Lead platform-level incident response workflows, post-incident reviews, and drive systemic improvements rather than one-off fixes.Evaluate incoming platform requests and translate them into scalable, productized capabilities.Mentor engineers and drive platform adoption through documentation, enablement, and technical evangelism.Participate in a 24x7 on-call rotation as an escalation point for platform reliability and availability issues.Operate effectively in ambiguous problem spaces, making sound architectural and product decisions with limited guidance.What We Look For In You:Bachelor’s degree in Computer Science or equivalent practical experience.4+ years of experience in Platform Engineering, Site Reliability Engineering, DevOps, or Systems Engineering roles.Strong expertise in Linux and Windows operating systems.Advanced automation and scripting skills using Python, Bash, and/or PowerShell.Deep, hands-on experience designing and operating AWS and Azure platforms at scale.Strong experience building and operating CI/CD platforms (Jenkins, GitHub Actions or equivalent).Strong experience with Infrastructure as Code and configuration management (Terraform, CloudFormation, ARM, or similar).Production experience with containerized and orchestration platforms such as Docker and Kubernetes.In-depth experience with the HashiCorp ecosystem (Nomad, Consul, Vault).Strong understanding of distributed systems, cloud-native architectures, and reliability patterns.Experience designing and operating observability platforms (e.g., Splunk, Sumo Logic, or similar).Familiarity with security and compliance practices, including vulnerability scanning and enterprise security tooling.Strong understanding of the software delivery lifecycle, release engineering, and platform lifecycle management.Experience working in Agile / DevOps environments with a strong product mindset.Demonstrated ability to influence without authority, set standards, and drive adoption across teams.Excellent communication skills, able to translate platform capabilities into clear developer value.Strong problem-solving skills with a bias toward durable, scalable solutions over short-term fixes.A mindset of continuous improvement, curiosity, and learning.Comfortable supporting a global, follow-the-sun operation when needed.How We Measure Success:Strong developer adoption and satisfaction with the platform (DX).Reduced deployment friction, lead time, and operational toil.Platform reliability and performance meeting or exceeding defined SLOs.Consistent, high-quality service delivery across engineering teams.Reduced incident frequency and severity driven by systemic platform improvements.Increased standardization, automation, and self-service adoption across the organization.$100,000 - $115,000 a yearOur differentiator is – Our People! We hire the brightest talent and develop them into leaders. We foster a culture of PEOPLE, PASSION and GROWTH. People: We value our people, customers, and partnersPassion: We love what we doGrowth: Unlimited growth means unlimited potentialAP is a customer-focused, team-oriented organization where innovation and results are rewarded, and individuals can chart the course of their own careers.As a woman founded and led company, this has meant supporting a meritocracy where everyone has opportunities to achieve their best and ensure we foster an environment of diversity, equity, and inclusion. In practice this means we will not only work to recruit a diverse workforce, but also maximize the full potential of all of our people. You can read more about our commitment to DEI Here

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Dallas, TX vacancy
  •  ...Talent Acquisition Team will reach out to help you navigate our interview process.Lantern is seeking an experienced Senior Site Reliability Engineer to champion the reliability, availability, and performance of our Azure-based healthcare platform. In this pivotal role,... 
    Suggested

    Lantern

    Dallas, TX
    15 hours ago
  • $138.4k - $173k

     ...infrastructure as well as help improve the reliability, quality of services and overall...  ...recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability...  ...about our locations by visiting our site.Compensation & BenefitsThe base salary that... 
    Suggested
    Full time
    Flexible hours

    AppFolio

    Dallas, TX
    2 days ago
  •  ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that...  ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to... 
    Suggested
    Full time

    Vanguard

    Dallas, TX
    15 hours ago
  • Qualifications: 8+ years of Software Engineering experience, or equivalent demonstrated through...  ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform...  ...vendor resources Willingness to work on-site at stated location in the job openingDepartment... 
    Suggested
    Contract work
    For contractors
    Work experience placement

    Cedent Consulting

    Dallas, TX
    2 days ago
  • $104.9k - $174.7k

     ...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory... 
    Suggested
    Full time
    Work at office
    Local area
    Remote work
    Work from home

    RELX Group

    Dallas, TX
    2 days ago
  • $197.3k - $313.7k

     ...ensure you are not duplicating efforts. Job Category Software Engineering Job Details About Salesforce Salesforce is the #1 AI CRM,...  ...and you are the future of Salesforce. Job Title: Director, Site Reliability Engineering Location: New York, NY; San Francisco, CA;... 
    Full time
    Immediate start

    Salesforce

    Dallas, TX
    1 day ago
  • $174k - $253k

     ...systems by pushing for changes that improve reliability and velocity.Practice sustainable...  ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical...  ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you... 

    Google

    Sunnyvale, TX
    14 hours ago
  • $147k - $211k

     ...product or system development code.Review code developed by other engineers and provide feedback to ensure best practices (e.g., style...  ..., and troubleshooting large-scale distributed systems. Site Reliability Engineering (SRE) is what you get when you treat operations... 

    Google

    Sunnyvale, TX
    14 hours ago
  •  ...Administrator / SRE in Dallas to own production Java environments, middleware, and cloud automation. You will optimize performance, drive reliability, and mentor teammates while aligning with enterprise security and AI-enabled integrations. You will work across Java apps, IBM... 

    Motion Recruitment Partners LLC

    Dallas, TX
    2 days ago
  •  ...Site Reliability Engineer- W2 Role* Technical proficiency: Strong Proficiency in Java, Strong understanding of Database concepts (Oracle, SQL, Dynamo DB etc.) Industry standard SRE Tools like Prometheus, Grafana, Data Dog Etc Good to have skills: Cloud Concepts / AWS,... 

    RSA Tech Group

    Dallas, TX
    2 days ago
  •  ...The Depository Trust & Clearing Corporation (DTCC) is seeking a Senior Application Support Engineer (SRE) to enhance reliability, scalability, and performance of mission-critical applications. You will apply SRE principles across engineering, infrastructure, and operations... 

    The Depository Trust & Clearing Corporation

    Dallas, TX
    2 days ago
  • $207k - $301k

    Develop strong, influential relationships with multiple stakeholders across the Site Reliability Engineering and Developer organizations.Serve as an expert on particular fields of knowledge related to rate limiting or sharding.Develop plans and lead projects on evolving... 

    Google

    Sunnyvale, TX
    14 hours ago
  • $114k - $148k

     ...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based... 
    Full time
    Temporary work
    Work experience placement
    Remote work

    GrabJobs

    Irving, TX
    18 hours ago
  • Site Reliability Engineer - Vice PresidentSite Reliability Engineering (SRE) is an engineering discipline that combines software and systems engineering to build and run scalable, massively distributed, fault-tolerant systems. At Goldman Sachs, SRE is responsible for improving... 

    Goldman Sachs

    Dallas, TX
    4 days ago
  • $141.8k - $195k

     ...their best work, grow fast, and bring their full selves to the herd. Why You’ll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in the U.S. Cribl... 
    Temporary work
    Remote work

    GrabJobs

    Garland, TX
    1 day ago
  •  ...This RoleAs a Senior Application Support Engineer, you will help power DTCC's global...  ...markets infrastructure by ensuring the reliability, availability, and performance of Institutional...  ...processing and settlement.Leveraging Site Reliability Engineering (SRE) principles... 
    Flexible hours
    Afternoon shift

    DTCC- The Depository Trust & Clearing Corporation

    Dallas, TX
    1 day ago
  • $262k - $365k

    Lead a team of Software/Systems Engineers on projects for users and be directly responsible for uptime.Own end-to-end availability...  ...qualifications:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software and systems engineering... 

    Google

    Sunnyvale, TX
    4 days ago
  • We are Compliance Engineering, a global team of more than 300 engineers and scientists who work on the most complex, mission-critical...  ...evolving systems by pushing for changes that improve capacity and reliability.Practicing sustainable incident management in a blameless... 

    Goldman Sachs

    Dallas, TX
    2 days ago
  • $207k - $301k

    Lead a team of Software/Systems Engineers on projects for users and be directly responsible for uptime.Own end-to-end availability...  ...qualifications:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software and systems engineering... 

    Google

    Sunnyvale, TX
    14 hours ago
  • Compliance EngineeringWe are Compliance Engineering, a global team of more than 500 engineers and scientists who work on the most complex...  ...systems by pushing for changes that improve capacity and reliability.Practicing sustainable incident management in a blameless postmortem... 

    Goldman Sachs

    Dallas, TX
    3 days ago
  • $113.1k - $232.3k

    Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity... 
    Work at office
    Local area
    Visa sponsorship
    Flexible hours
    3 days per week

    Deloitte

    Dallas, TX
    1 day ago
  • $128.47k - $192.71k

     ...specific processes. • Leads a variety of types of business process design initiatives. • Assesses potential implications of re-engineering for multiple functions or departments. • Demonstrates mastery of re-engineering concepts, methods, and tools. Growth and Agility... 
    Full time
    Part time
    Worldwide
    Flexible hours

    Caterpillar Inc.

    Irving, TX
    11 hours ago
  •  ...Pay Rate: $40/Hr. W2 Experience: 3-5 Years Overview We are seeking a remote Junior SRE/DevOps Engineer role. The ideal candidate has foundational knowledge of Site Reliability Engineering (SRE) and Kubernetes, and is enthusiastic about growing in a DevOps‑driven environment... 
    Long term contract
    Contract work
    Internship
    Remote work

    BayOne Solutions

    Richardson, TX
    4 days ago
  •  ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless... 

    AmNet Services

    Irving, TX
    2 days ago
  • $40 per hour

     ...A technology solutions provider is seeking a remote Junior SRE/DevOps Engineer. The ideal candidate should have foundational knowledge of Site Reliability Engineering (SRE) and Kubernetes. Responsibilities include gaining experience in a DevOps-driven environment. Applicants... 
    Long term contract
    Internship
    Remote work

    BayOne Solutions

    Richardson, TX
    4 days ago
  • $141.3k - $237.4k

     ...AT&T, you won’t just imagine the future, you’ll build it.We are seeking a highly skilled and hands-on Lead Software Engineer to join Software Reliability Engineering (SRE) Onboarding and automation team. This role will drive innovation through automation, enhancement of... 
    Full time
    Temporary work
    Work at office
    Local area
    Relocation

    AT&T

    Dallas, TX
    4 days ago
  •  ...Job Description Job Description Forhyre is looking for engineers who can bring unique perspectives and innovative ideas to all areas...  ...evangelize cloud best practices while building a culture of reliability and observability Engage in and improve the end to end lifecycle... 

    Forhyre

    Dallas, TX
    16 days ago
  •  ...maintaining, auditing, and assisting with Reliability Excellence (RX) and Process Safety...  ...futureMust have Bachelor's Degree in an Engineering field from an accredited institution or...  ...support RX related enhancement projectsLead site in achieving RX goals and maintenance system... 
    Contract work
    For contractors
    Local area
    Monday to Friday

    Sherwin-Williams

    Garland, TX
    2 days ago
  •  ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless... 

    AmNet Services

    Irving, TX
    2 days ago
  • $110k - $170k

     ...more information, visit . Follow Shield AI on LinkedIn, X, Instagram, and YouTube. Job Description:We are seeking a skilled Systems Engineer to join the Hivemind Solutions Systems Engineering team. In this role, the candidate will support the design, development,... 
    Full time
    Temporary work
    Part time
    Worldwide

    Shield AI

    Dallas, TX
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!