Site Reliability Engineer
National Oilwell Varco
As a Site Reliability Engineer, you will be responsible for: Operational Excellence & Incident Management- Maintain and monitor production systems for availability, latency, and performance.- Lead incident response efforts, including communication, resolution, and postmortem documentation.- Design and implement health checks, alerting systems, and automated remediation workflows.- Drive root cause analysis and implement permanent resolutions for recurring issues.Observability & Insights- Set up and maintain full observability stacks (logging, metrics, tracing) using tools like Prometheus, Grafana, Datadog, OpenTelemetry, or ELK.- Analyze telemetry and logs to identify trends, anomalies, and opportunities for improvement.- Conduct post-incident reviews and use insights to inform future engineering investments.Performance & Systems Optimization- Tune and optimize distributed systems, including AKKA.NET actors, for performance and resource efficiency.- Work with developers to evolve architecture and improve system throughput, latency, and stability.- Optimize PostgreSQL performance, queries, and maintenance strategies.CI/CD & Automation- Design and maintain modern CI/CD pipelines using GitHub Actions, Azure Pipelines, or GitLab CI.- Automate deployment, testing, and rollback processes to reduce friction and increase deployment frequency.- Standardize infrastructure as code practices across environments.We’d love to talk to you if you have:- 5+ years of experience in SRE, DevOps, or Infrastructure Engineering roles.- Expertise in Kubernetes and container orchestration at scale.- Strong experience with AKKA.NET or similar actor-based frameworks.- Proficiency with scripting and automation (Bash, PowerShell, Python).- Experience with observability tools (Phobos,Datadog, Prometheus, Grafana, OpenTelemetry, ELK).- Hands-on experience with cloud platforms (AWS, Azure, or GCP).- Strong PostgreSQL knowledge—performance tuning, query optimization, maintenance.- Proven ability to lead incident management and drive postmortem processes.- A builder’s mindset with high standards for operational excellence and technical ownership.Preferred Tools & Ecosystem Experience- CI/CD: GitHub Actions, Azure Pipelines, GitLab CI- Infrastructure: Kubernetes, Docker, Terraform- Monitoring: Phobos (AKKA.NET), Datadog, Prometheus- Source Control: GitHub, GitLab, Azure DevOps- Programming: C#, Python, Bash, PowerShellEvery day, the oil and gas industry’s best minds put more than 150 years of experience to work to help our customers achieve lasting success.We Power the Industry that Powers the WorldThroughout every region in the world and across every area of drilling and production, our family of companies has provided the technical expertise, advanced equipment, and operational support necessary for success—now and in the future.Global FamilyWe are a global family of thousands of individuals, working as one team to create a lasting impact for ourselves, our customers, and the communities where we live and work. Purposeful InnovationThrough purposeful business innovation, product creation, and service delivery, we are driven to power the industry that powers the world better.Service Above AllThis drives us to anticipate our customers’ needs and work with them to deliver the finest products and services on time and on budget.CorporateOur family of companies is supported by our global Corporate teams, providing expert knowledge from functions including Human Resources, Information Technology, Compliance, Finance, QHSE, Marketing and Legal centers of expertise. We are structured to provide guidance and service above all to all our business operations.Full timePosting Date: 2026-06-23
- The Senior Site Reliability Engineer is responsible for improving the reliability, availability, scalability, and operational excellence of our critical infrastructure platforms and services. This role partners closely with Engineering, Security, and Infrastructure teams...SuggestedFull timeWork at officeLocal area
- Reliability Engineering Design, implement, and operate scalable, resilient, and highly available systems on Google Cloud Platform. Improve service... ...Skills, and Abilities Three or more years of experience in Site Reliability Engineering, platform engineering, DevOps, cloud...SuggestedRemote work
- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Corporate Technology, Risk Technology team, you will solve complex and broad business problems...Suggested
- ...Title:Site Reliability Engineer Location: Houston, TX 77002 (Hybrid: 3 days onsite / 2 days remote) Duration: Contract to Hire Work Requirements:U.S.Citizen, GC Holders,or Authorized to Work in the U.S. Job Description The Site Reliability Engineer is a founding...SuggestedContract workRemote workFlexible hours
$136.2k - $214.01k
...outcomes Visionary in future focused problem-solving Exceptional in execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to...SuggestedFull timeFlexible hours- ...Site Reliability Engineer The SRE manages and maintains infrastructure which supports cloud-based prototype applications as they transition into a production environment. The SRE assists in defining and measuring service level agreements and non-functional requirements...
- ...on one unified cloud. One cloud for compute, inference, and agents. Role Overview We are seeking a skilled Site Reliability Engineer to join the GMI Global Infrastructure team. This role is hands-on and critical to ensuring the stability, efficiency, and...
- ...As a Site Reliability Engineer, you will be responsible for: Operational Excellence & Incident Management Maintain and monitor production systems for availability, latency, and performance. Lead incident response efforts, including communication, resolution...Permanent employment
- ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Corporate Technology, Corporate Know Your Customer (KYC) team, you will solve complex and...
$55k - $151.47k
...ApplicableSpecialismIFS - Internal Firm Services - OtherManagement LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability, scalability, and performance of our...Full timeH1b- As an Entry-Level DevOps Site Reliability Engineer, you will join a team responsible for continuous improvement and support of customer facing products. Responsibilities will include collecting system requirements; improving existing tools and processes through scripting...Work from home2 days per week
- ...ENGINEERLocation: HOUSTON, TXFLSA Class: EXEMPTResponsible to: Directo of Software EngineeringPosition Summary: DevOps / Site Reliability Engineer to implement and evolve the infrastructure, deployment pipelines, and reliability posture of our systems. You'll work closely...Full timeLocal area
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate & Investment Bank (CIB) Management and Support Functions Digital & Platform...
- As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate Know Your Customer (KYC) Technology group, you hold a leadership role in your team, demonstrate strong knowledge across multiple technical domains, and advise others on the technical and business...
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate & Investment Bank (CIB) Management and Support Functions Digital &...
- ...contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology Team, you draw upon your advanced knowledge to identify new...
$213.1k - $300k
Lead a team of engineers to maintain service uptime while managing global on-call rotations... ...improve operational practices to drive reliability, maintainability, and stakeholder alignment... ...or in a Manager, Software Engineer, Site Reliability Engineering-related occupation...Full timeWork at office- ...JPMorgan Chase is seeking a Lead Site Reliability Engineer to shape the future of reliability for a globally recognized firm within the Corporate & Investment Bank domain. The role emphasizes leadership across multiple technical domains and mentoring peers. You will...
$61k - $101k
...formal training or certification in software engineering concepts, along with 5+ years of applied... .... We need deep expertise in reliability, scalability, performance, security, enterprise... ...architecture, toil reduction, and other site reliability practices, with the ability...Full time- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate & Investment Bank (CIB) Management and Support Functions Digital &...
- ...Talentify is seeking a Site Reliability Engineer in Houston to own reliability, observability, and performance across distributed systems. You will lead incident response, design health checks, and drive postmortems while collaborating with developers to evolve architecture...
- ...Please extend your support for this role. Local candidate will get 1st preference. Job Title: SRE Engineer Location: Houston, TX and Jersey City, NJ - 3 Days Onsite Role FTE role with Mphasis Client: Mphasis H1B transfer will work...Work experience placementH1bLocal area
$113k - $141.53k
...leader in global energy. Senior Solutions Engineer - Systems Integration serves as a... ...functionally in the field, ensuring safe, reliable, and performant operation across diverse... ...Willingness to travel to factories and project sites (25%).Preferred QualificationsMaster’s degree...Full timeFor contractorsLocal areaRemote workWorldwideFlexible hours$94k - $112.63k
...global energy. AES is seeking a Solutions Engineer - Systems Integration to contribute to... ...interface correctly and operate safely, reliably, and as intended across project environments... ...to factories, laboratories, and project sites to support inspections, testing, commissioning...Full timeFor contractorsLocal areaWorldwide- Position: Software Engineer- Flight & Ground Systems Location: Houston, TX Remote Status: On-Site Job Id: 866 # of Openings... ...requirements and translate them into reliable software solutions.Self-motivated and...Permanent employmentFull timeRemote workRelocation package
$76k - $155.7k
...Required: Up to 10%Type of Travel: Continental US* * *The Opportunity:CACI is looking for an experienced Flight Software Systems Engineer to support NASA’s Moon Base flight software development at the Johnson Space Center. The Moon Base is humanity’s first permanent lunar...Permanent employmentContract workFor contractorsWork experience placementFlexible hours- ...We are seeking two Release Engineers to support Azure-based cloud and release engineering activities. This role is responsible for the... ...engineer will collaborate with cross-functional teams to deliver reliable and scalable cloud solutions and ensure smooth transitions of...
- Lead ETRM Systems Developer:On behalf of our Energy client, Procom is searching for a Lead ETRM Systems Developer for a permanent role. This position is onsite at our client’s Houston, Texas office. Lead ETRM Systems Developer - Job Description:This role involves supporting...Permanent employmentWork at officeImmediate start
$51 - $61 per hour
...onsite at the project, significantly reducing and/or eliminating the demands to travel. Key Responsibilities: As a Release Train Engineer, you will be responsible for facilitating Agile Release Train events and processes including communicating with stakeholders escalating...Hourly payLive inWork at officeLocal areaImmediate startFlexible hoursShift work- ...Release Train Engineer 4 Months- Contract To Hire Pay- $65-$70 W2 Onsite Houston, TX Job Description The Release Train... ...sure all team activity, dashboards, and metrics are visible and reliable. Coach teams and Scrum Masters on agile and Scrum practices,...Contract workWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre Houston, TX
- site reliability engineer Houston, TX
- official site Houston, TX
- remote website tester Houston, TX
- site services specialist Houston, TX
- construction site safety Houston, TX
- IT site lead Houston, TX
- site recruiter Houston, TX
- site leader Houston, TX
- site safety Houston, TX



