Site Reliability Engineer
3B Staffing LLC
Location: Irving, TX
Duration: 6+ Month Contract to hire
Interview: Onsite. (mandatory)
Term: Hybrid (2 days in office)
Responsibilities
Duration: 6+ Month Contract to hire
Interview: Onsite. (mandatory)
Term: Hybrid (2 days in office)
Responsibilities
- Lead architecture and development teams to ensure applications are highly available, reliable, and performant at a global scale.
- Partner with the architecture team to ensure operability, measurability, and manageability are integrated into business features and enablers.
- Collaborate with product owners and managers to establish service level objectives (SLOs) for applications and define consequences if objectives are not met.
- Work with development team members to identify monitoring gaps, improve application performance, and assist with troubleshooting issues.
- Drive Root Cause Analysis (RCA) of production issues and other failures within the product software, pipeline, or other DevOps support processes or technology.
- Design, build, and advocate for automated solutions to optimize application/service/platform uptime with minimal human intervention.
- Participate in an on-call rotation to support troubleshooting and communication efforts outside of normal business hours.
- Create and implement standards and best practices, driving adoption across development teams and external vendors as applicable.
- Ensure compliance with all company policies and procedures.
- Bachelor of Computer Science or related Engineering field required.
- Master's Degree preferred.
- 5-7 years of hands-on SRE experience.
- 1-2 years of leading and mentoring others.
- Hands-on experience supporting Linux production environments, hands-on administration on Spark, and hands-on experience with MS Azure Cloud technologies.
- 3-5 years hands-on experience with scripting with bash, perl, ruby, or python required.
- 3-5 years experience with Docker Datacenter required.
- 2-4 years of hands-on administration experience on Machine learning platforms required.
- Minimum of 1 year of experience in Mesos, Kubernetes, OpenShift and/or Deis or other such container/platform-as-a-service orchestrator required.
- Minimum of 1 year of hands-on experience on CICD tools & Technologies required.
- Minimum of 1 year of lead experience of site reliability engineering team required.
- Proven leadership skills and the ability to guide and mentor a team.
- Strong collaboration and communication skills.
- A proactive approach to problem-solving and continuous improvement.
- Passion for automation and operational excellence.
- Deep expertise in cloud technologies and software development, with a strong technical background.
- Experience with Java
- Proficiency in SQL and Powershell.
- Expertise in defining, implementing, and evaluating Service Level Objectives (SLOs) and Service Level Indicators (SLIs), and associated consequences.
- Strong skills in performing Root Cause Analysis (RCA) and Problem Management.
- Extensive experience in cloud native applications Azure/AWS (monitoring, networking, containerization, infrastructure).
- Proficiency in containerization technologies such as Azure Kubernetes Service, Kubernetes (open source), and Docker.
- Knowledge of metrics and monitoring tools like Azure Application Insights and Azure Monitor.
- Familiarity with networking technologies relevant to Azure and AWS, including Azure DNS, Virtual Networks, Azure API Manager, Azure Application Gateway, Akamai WAF/CDN, AWS Route 53, AWS VPC, AWS API Gateway, and AWS CloudFront.
- Strong experience with Terraform for infrastructure as code.
- Ability to establish and maintain a culture of learning through the development and sharing of skills, knowledge, processes, and tools; combat traditional silos that create "us and them" environments.
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Irving, TX vacancy
- ...Senior Site Reliability Engineer Our client is seeking a Senior Site Reliability Engineer for a month 6-month contract in Irving, TX. Will be working on an onsite schedule. Contract Duration: 6 Months Required Skills & Experience ~ Bachelors/4 Year Degree ~5+...SuggestedFull timeContract workTemporary workWork experience placementFlexible hours
- ...generative AI and cloud-native platforms to advanced release engineering practices, our teams are redefining how financial technology... ...AI-driven solutions that accelerate development and improve reliability. Your work will directly influence how GM Financial leverages...SuggestedFull timeH1bWork at officeRemote workVisa sponsorshipFlexible hours2 days per week3 days per week
$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...SuggestedFull timeWork at officeLocal areaRemote workWork from home$138.4k - $173k
...infrastructure as well as help improve the reliability, quality of services and overall... ...recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability... ...about our locations by visiting our site.Compensation & BenefitsThe base salary that...SuggestedFull timeFlexible hours- ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that... ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to...SuggestedFull time
- ...cloud-native platforms to advanced release engineering practices, our teams are redefining how... ...alert behavior preferred Exposure to reliability engineering concepts such as SLOs/SLIs and... ...office#LI-KC1#GMFjobsAbout The Role:The Site Reliability Engineer under the general...Work experience placementH1bWork at officeRemote workVisa sponsorshipFlexible hoursShift work2 days per week
- Qualifications: 8+ years of Software Engineering experience, or equivalent demonstrated through... ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform... ...vendor resources Willingness to work on-site at stated location in the job openingDepartment...Contract workFor contractorsWork experience placement
- ...improving platform infrastructure and applications with high reliability, resiliency, performance & quality, and faster time-to-market... ...documentation, including runbooks/playbooks; and, Using Chaos Engineering to test the robustness of the systems and applications....
- ...Senior Site Reliability Engineer This role will require someone onsite at our client office in Cleveland, OH, Pittsburgh, PA, or Dallas, TX. Systemone is looking for a Site Reliability Engineer who will work within the Production support and Performance Management...Work at officeFlexible hoursShift workWeekend work
- ...Site Reliability Engineer TXSE is building the next-generation exchange infrastructure to support transparent, efficient, and resilient capital markets. With SEC approval and $275MM in funding, we are currently hiring a Site Reliability Engineer to help with a greenfield...Currently hiring
- ...Senior Site Reliability Engineer (Permanent Role) Cleveland, OH, Pittsburgh, PA, or Dallas, TX Your future duties and responsibilities . Monitoring distribution systems and notifying them of any potential issues. . Assisting with troubleshooting on call....Permanent employmentFull timeLocal areaFlexible hoursShift workWeekend work
$120.6k - $150.9k
...About the Role We are looking for a highly motivated, high-potential Staff Site Reliability Engineer (SRE) to join our team as a technical leader and drive transformative impact across WEX’s platform reliability and operational excellence. This is a particularly exciting...Full timeFlexible hours- ...and continuously improving the platforms that power TI's digital integration, automation and DevOps capabilities. As an IT Site Reliability Engineer within the Enterprise Platforms team, you will serve as the primary technical platform owner for TI's Apigee Edge private...Local area
- Site Reliability Engineer - Vice PresidentSite Reliability Engineering (SRE) is an engineering discipline that combines software and systems engineering to build and run scalable, massively distributed, fault-tolerant systems. At Goldman Sachs, SRE is responsible for improving...
- Compliance EngineeringWe are Compliance Engineering, a global team of more than 500 engineers and scientists who work on the most complex... ...systems by pushing for changes that improve capacity and reliability.Practicing sustainable incident management in a blameless postmortem...
$110.11k - $204.49k
...employees feel respected, valued and have an opportunity to contribute to the company’s success. As a Software Engineering Manager within PNC's Site Reliability organization, you will be based in one of these Technology Hub locations: Pittsburgh, PA, Cleveland, OH,...Permanent employmentFull timeTemporary workPart timeWork experience placementWork at officeAfternoon shift$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
$159k - $305k
...Wells Fargo is seeking a Senior Lead Platform Reliability Engineer to join the CTO Platform organization. This role is designed for highly experienced... ...Reliability Engineering (PRE) team, you will apply modern Site Reliability Engineering (SRE) practices to improve the...Full timeWork experience placement- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
- ...McKesson seeks a seasoned Release Train Engineer (RTE) in Irving, TX to lead Agile Release Trains delivering strategic initiatives. You will coordinate PI planning, manage risks, and provide executive reporting to align delivery with enterprise priorities. Required 12...
- Compliance Engineering, Site Reliability Engineering, Vice President, Dallas location_on Dallas, TX, United States We are Compliance Engineering, a global team of more than 300 engineers and scientists who work on the most complex, mission-critical problems. We build and...Full timeTemporary workWork at office
- Required Technical Skills • Linux bash, Python, pip, conda, Node/npm, rpm, GNU tools (g++,make,configure)Desirable Technical Skills • Public cloud platform experience• Bash script development• Python development, package management, package building and testing• Source ...
- ...asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing... ...- Manage, administer, and maintain all internet and intranet sites - Research/analyze data processing functions, methods and procedures...
$107.48k - $143.31k
...employees. What You'll Be Doing Lead and apply regional reliability engineering strategies to improve equipment performance, uptime, and... ...management skills with the ability to support multiple sites remotely. Willingness to travel to the plant and corporate...Temporary workRemote workFlexible hours- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...Work experience placement
$77.4k - $135.4k
...apply strong SQL and Python expertise to enhance performance, reliability, and data accessibility while supporting advanced analytics and... ...solutions. Partner with application developers, data engineers, BI teams, and analytics partners to deliver integrated solutions...Full time$124.36k - $146.3k
...each other. Job DescriptionResponsibilitiesAs a senior-level Reliability Engineer specializing in observability, this role partners closely with... ...Models.Partner with Product Owners, Application Engineering, Site Reliability Engineering (SRE), and Operations Teams to ensure...Full timeWork experience placementLocal area3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
Related searches
- junior website developer Irving, TX
- website content developer Irving, TX
- site leader Irving, TX
- site recruiter Irving, TX
- historic site Irving, TX
- on-site clinical research associate (traveling/remote) Irving, TX
- official site Irving, TX
- site services specialist Irving, TX
- site safety Irving, TX
- construction site safety Irving, TX


