Site Reliability Engineer
System One
Senior Site Reliability Engineer (SRE)
Location: Pittsburgh, PA / Cleveland, OH / Dallas, TX
FTE
We are seeking an experienced Senior Site Reliability Engineer (SRE) to support production operations, application reliability, performance management, and continuous improvement initiatives.
The selected candidate will work closely with production support and engineering teams to ensure critical internal and external applications maintain appropriate levels of availability, reliability, and uptime .
This role requires strong experience in production support, incident management, monitoring, troubleshooting, log analysis, automation identification, infrastructure technologies, databases, and application servers . The SRE will also provide technical leadership and collaborate with geographically distributed teams. Key Skills
- Site Reliability Engineering (SRE)
- Production Support / Application Support
- Incident & Problem Management
- Linux
- Windows Server
- Oracle / PL/SQL / DB2
- Dynatrace / DT Managed
- GlassBox / ITCAM / TrueSight / OEM
- Tomcat / Apache / WebSphere (WAS) / IIS
- REST & SOAP Web Services
- Log Analysis & Troubleshooting
- AIOps / NLP
- Monitoring & Performance Management
- Automation
- Root Cause Analysis
- Business Analytics
- Agile
- Technical Leadership
- Client-Facing Production Support
- Monitor distributed systems and proactively identify potential production issues.
- Support troubleshooting and participate in on-call activities.
- Manage, track, and coordinate production incidents and application outages.
- Lead incident-analysis and problem-management meetings.
- Identify opportunities for operational and production-support automation.
- Monitor applications and related infrastructure to maintain system reliability.
- Coordinate follow-up activities through incident resolution and closure.
- Troubleshoot complex application issues using system and application logs.
- Participate in critical incident calls and contribute technical expertise toward resolution.
- Perform root cause analysis and recommend corrective actions.
- Research and reproduce user issues to validate solutions.
- Resolve technical problems that cannot be handled by junior team members.
- Provide technical guidance and solutions to the production-support team.
- Introduce process improvements and innovative solutions for operational challenges.
- Develop and maintain SOPs, operational procedures, and knowledge documentation.
- Collaborate with offshore and geographically distributed teams.
- Work with client technical teams, SMEs, and leadership.
- Support extended or weekend hours when required during critical production events.
- Participate in overlapping business-hour shifts for critical meetings and activities.
- 5+ years of overall IT experience.
- 2–3 years of business analytics and technical leadership experience.
- Strong experience with production/application support in a client-facing environment.
- Strong understanding of Site Reliability Engineering and production operations .
- Hands-on experience troubleshooting production applications and analyzing log files .
- Strong knowledge of system-management, monitoring, and support analytics tools.
- Experience with incident management, root cause analysis, and problem resolution .
- Strong understanding of AIOps and NLP concepts .
- Experience identifying opportunities for automation and process improvement .
- Strong problem-solving and analytical capabilities.
- Ability to recommend efficient and cost-effective technical solutions.
- Experience working with geographically distributed/onshore-offshore teams.
- Excellent client-facing verbal and written communication skills.
Strong knowledge of:
- Oracle
- PL/SQL
- DB2
Experience developing and consuming:
- REST APIs
- SOAP Web Services
Application Servers / Web Servers
Strong knowledge of:
- Tomcat
- Apache
- WebSphere (WAS)
- IIS
- Extensive experience with Linux
- Good understanding of Windows Server
- Linux and Windows server configuration and troubleshooting
Experience with monitoring tools such as:
- Dynatrace
- Dynatrace Managed / DT Managed
- GlassBox
- ITCAM / ITCAMS
- TrueSight
- Oracle Enterprise Manager (OEM)
- Agile methodology
- SOP and technical documentation
- Performance management
- System reliability and availability
- Production incident coordination
- Technical research and solution evaluation
- Process improvement
- Automation opportunity identification
- Strong stakeholder and client communication
#M1
#DI-CB2
System One, and its subsidiaries including Joulé, ALTA IT Services, TeamPeople, and Mountain Ltd., are leaders in delivering outsourced services and workforce solutions across North America. We help clients get work done more efficiently and economically, without compromising quality. System One not only serves as a valued partner for our clients, but we offer eligible employees health and welfare benefits coverage options including medical, dental, vision, spending accounts, life insurance, voluntary plans, as well as participation in a 401(k) plan.
System One is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender identity, age, national origin, disability, family care or medical leave status, genetic information, veteran status, marital status, or any other characteristic protected by applicable federal, state, or local law.
- ...: Lovelace is the only provider of enterprise-scale context engines capable of analyzing trillions of real-time data points to create... ...: ~ Lovelace AI is seeking a highly skilled and motivated Site Reliability Engineer (SRE) to join our growing team. As an SRE at...SuggestedFull time
- ...Job Description Job Description Job Title: Senior Site Reliability Engineer Job Category: Infrastructure/Cloud Job Type: Permanent Full Time Location: Pittsburgh, Pennsylvania, United States Position Description This role will require someone onsite...SuggestedPermanent employmentFull timeWork at officeFlexible hoursShift workWeekend work
$70.8k - $156.7k
Senior Site Reliability Engineer - Local to Cleveland, Pittsburgh, or Dallas Position Description This role will require someone onsite at our client office in Cleveland, OH, Pittsburgh, PA, or Dallas, TX. Love technology? We do too. CGI is looking for a Site...SuggestedWork at officeLocal areaFlexible hoursShift workWeekend work$182.8k - $247.3k
...mission to develop education for our half a billion (and growing!) learners around the world.About the role...As a Senior Site Reliability Engineer, you will work closely with both product and platform engineering teams to ensure Duolingo’s sophisticated distributed systems...SuggestedWork experience placement$51 - $61 per hour
...onsite at the project, significantly reducing and/or eliminating the demands to travel. Key Responsibilities:As a Release Train Engineer, you will be responsible for facilitating Agile Release Train events and processes including communicating with stakeholders...SuggestedHourly payLive inWork at officeLocal areaImmediate startFlexible hoursShift work$93.61k - $119.72k
...funcionen mejor. This role will support several ATS site locations within the Midwest and East. The role will be... ...procedures. Partners with internal/external customer for engineered solutions to improve reliability and throughput. Identifies opportunities for Capital...Work at officeRemote workHome office- ...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas... ...crucial workloads. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This...Full timeRemote workWorldwide
$140k - $200k
...people around the globe work on Speechify in a 100% distributed setting – Speechify has no office. These include frontend and backend engineers, AI research scientists, and others from Amazon, Microsoft, and Google, leading PhD programs like Stanford, high growth startups...Remote jobFull timeWork at office$100.22k - $111.18k
...Basic Qualifications Requires a Bachelor's degree in Systems Engineering, or a related Science, Engineering, Technology or Mathematics... ...Options: This position is located in Pittsburgh PA. On site work is required and Hybrid/Flex work schedule is permitted. Please...For subcontractorSecond jobWork at officeFlexible hours- ...Who We Are Engineered to outperform, Teraswitch is on a mission to provide high-performance infrastructure services for critical workloads... ..., Support, and other internal stakeholders to deliver reliable, secure, and scalable software. We’re not looking for a...Full time
$140k - $200k
...– Speechify has no office. These include frontend and backend engineers, AI research scientists, and others from Amazon, Microsoft, and... ...→ testing → release → maintenance. Ensure quality, reliability, and consistency across releases. Identify, diagnose, and resolve...Full timeWork at office- ...Description XDIN subsidiary of ALTEN Group, includes 500 employees dedicated to the automotive engineering development. ALTEN is a Leader in Engineering & Information Technology system, and operates in over 21 countries (Europe, North America, Asia, Africa and Middle...Full timeTemporary work
- ...AI and advanced autonomous systems designed to enhance safety, reliability, and efficiency of modern operations. Stack's autonomous... ...perception problems for the product. As a Senior Software Engineer, you will develop foundational ML architecture and systematic...Full time
- ...Description: Senior Software Engineer working with our Cloud Infrastructure team to develop and maintain distributed services for... ...requirements. Lead technical design efforts to ensure scalability, reliability, and efficient use of infrastructure resources. Collaborate...Full time
- ...What We Do Gecko Robotics is helping the world’s most important organizations ensure the availability, reliability, and sustainability of critical infrastructure. Gecko's complete and connected solutions combine wall-climbing robots, industry-leading sensors, and an AI...Full timeWork at officeLocal areaWork from homeFlexible hours
- ...Senior Software Engineer Pittsburgh, Pennsylvania, United States This role may require up to 10% travel Onsite role. About the Client We are a leader in enterprise readiness. Our mission is to establish readiness as a real-time condition that is continuously...Full timeWork experience placementWork at office
- ...across machine learning and robotics, cloud platforms, mapping, sensors and compute systems, test operations, systems and safety engineering – all dedicated to making a real, positive impact on the driving experience for millions of people. As a Ford Motor Company subsidiary...Permanent employmentFull timeImmediate startVisa sponsorship
- ...based company specializing in wireless power, low power electronics design, and wireless communication solutions seeks a Software Engineer to support numerous application and product lines. Overview Work with the engineering team designing software for various...Full timeWork at office
$70k - $300k
...robots interact with the real world. We are building risk-aware, reliable, and field-ready AI systems that address the most complex... ...hardware and software is critical. We’re looking for a Software Engineer - Mission Workflows to maintain and develop robot user workflows...Permanent employmentFull timeFlexible hours- ...Us: Lovelace is the only provider of enterprise-scale context engines capable of analyzing trillions of real-time data points to... ...Conduct thorough testing, debugging, and optimization to ensure the reliability and performance of software applications. Continuous...Full time
$85 - $95 per hour
...Job Description Job Description Genesis10 is currently seeking a AVP / Head of Enterprise Site Reliability Engineering (SRE) with our consumer finance lender firm client in their Pittsburgh, PA location. This is a Right to hire position. Summary: An experienced...Hourly payPermanent employmentFull timeContract workImmediate startRemote work- ...Title: Lead Design Engineer Location: Onsite, Pittsburgh, PA 15213 Type: 6-month contract to hire Hours: Standard business hours Overview: Join a cutting-edge lab to discover novel therapeutics that are seeking a highly motivated Lead Engineer...Contract work
- ...AI and advanced autonomous systems designed to enhance safety, reliability, and efficiency of modern operations. Stack's autonomous... ...power all of Autonomy development. We work hand in hand with AV engineers to provide cutting edge solutions to all their data needs, working...Full time
$210k - $280k
...Swan because our mission matters. We’re doing practical, high-impact work in AI safety, that sense of purpose is a core reason why engineers and researchers choose us. The Role We’re looking for a software engineer who wants to build things end to end and doesn’t...Full timeWork at officeVisa sponsorshipFlexible hours- ...AI and advanced autonomous systems designed to enhance safety, reliability, and efficiency of modern operations. Stack's autonomous... ...evangelizing best practices and frameworks among Machine Learning Engineers (MLEs) across the company. Responsibilities: Analyze ML...Full time
- ...demands, the front line gets what it needs to succeed. Job Description We are seeking a skilled and dedicated Software Engineering Lead to join our Engineering team. As the Software Engineering Lead at Air, you will be essential to crafting and implementing...Full timeWork at office
- ...and advanced autonomous systems designed to enhance safety, reliability, and efficiency of modern operations. Stack's autonomous technology... ...You will collaborate with various teams in Autonomy, Systems Engineering and Operations, fostering a culture of safety, principled...Full time
- ...Lead Enterprise Systems Engineer Under supervision of the Chief Technology Officer, the Lead Enterprise Systems Engineer works to analyze... ...and gain efficiency Work as part of a team Work with on-site equipment What You Bring to the Team Other duties as...For contractorsWork experience placementWork at officeNight shiftWeekend work
$140k - $200k
...people around the globe work on Speechify in a 100% distributed setting - Speechify has no office. These include frontend and backend engineers, AI research scientists, and others from Amazon, Microsoft, and Google, leading PhD programs like Stanford, high growth startups...Full timeWork at officeRemote work- ...AI and advanced autonomous systems designed to enhance safety, reliability, and efficiency of modern operations. Stack's autonomous... ...the Role: Join a high-caliber team of experienced software engineers responsible for delivering some of Stack AV's most foundational...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- junior website developer Pittsburgh, PA
- on site coordinator Pittsburgh, PA
- construction site safety Pittsburgh, PA
- site services specialist Pittsburgh, PA
- website content developer Pittsburgh, PA
- website coordinator Pittsburgh, PA
- on-site clinical research associate (traveling/remote) Pittsburgh, PA
- historic site Pittsburgh, PA
- official site Pittsburgh, PA
- site leader Pittsburgh, PA




