Staff Site Reliability Engineer, AQI Data SRE
$207k - $300kDefine the technical goal, architectural roadmap, and reliability strategy for AQI Data SRE across the SAGE/SPARK production stack.Partner with dev leadership to lead system design reviews, drive backend simplification and resource isolation, and ensure production readiness for flagship Ads launches.Architect and lead the implementation of critical infrastructure projects, including data-pipeline resilience, automated rollback/restart platforms, and consolidated observability.Mentor and grow engineers on the team, and push for software engineering excellence, AI-first tooling, and SRE best practices; drive the operational maturity handover that helps partner dev teams mature.Lead the response to complex production incidents, correlate production signals with customer and business impact, and drive high-impact post-incident architectural improvements and blameless postmortems and participate in a healthy tier 2 oncall rotation for core shared SAGE/SPARK infrastructure.Minimum qualifications:Bachelor’s degree in Computer Science, a related field, or equivalent practical experience.8 years of experience with software development in one or more programming languages.3 years of experience leading projects.3 years of experience designing, analyzing, and troubleshooting distributed systems.Experience with large-scale data processing.Preferred qualifications:Master's degree in Computer Science or Engineering.Experience in distributed, real-time systems and high-throughput data pipelines.Knowledge or experience with agentic flows/development and with AI-assisted production/reliability tooling.Proven track record of driving cross-functional architectural alignment and influencing technical decisions across large engineering organizations.Excellent collaboration skills, with a track record of building trusted, blameless partnerships with development teams.Excellent non-abstract large systems design skills, with experience building scalable, self-service reliability and observability.Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating work through automation. On the SRE team, you’ll have the opportunity to manage the complex challenges of scale which are unique to Google Cloud, while using your expertise in coding, algorithms, complexity analysis and large-scale system design. SRE's culture of intellectual curiosity, problem solving and openness is key to its success. Our organization brings together people with a wide variety of backgrounds, experiences and perspectives. We encourage them to collaborate, think big and take risks in a blame-free environment. We promote self-direction to work on meaningful projects, while we also strive to create an environment that provides the support and mentorship needed to learn and grow.Ads Quality Infrastructure (AQI) Data Site Reliability Engineering (SRE) keeps Search Ads and Ads on Google Experiences (SAGE/SPARK) data processing reliable, so those teams can deliver useful ads that create advertiser value. We own the core infrastructure behind the Google Ads business across Google-owned surfaces, including Google Search, Google Maps, and Feed Ads, and have a portion of Google business.AQI SRE owns: serving, data extraction and copy, and the platforms that keep those systems fast and healthy. We partner with development teams to co-build and support this infrastructure.Behind everything our users see online is the architecture built by the Technical Infrastructure team to keep it running. From developing and maintaining our data centers to building the next generation of Google platforms, we make Google's product portfolio possible. We're proud to be our engineers' engineers and love voiding warranties by taking things apart so we can rebuild them. We keep our networks up and running, ensuring our users have the best and fastest experience possible.Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $207000 - $300000 (USD) + 20% bonus target + equity + benefitsLearn more about benefits at Google.Bachelor’s degree in Computer Science, a related field, or equivalent practical experience.8 years of experience with software development in one or more programming languages.3 years of experience leading projects.3 years of experience designing, analyzing, and troubleshooting distributed systems.Experience with large-scale data processing.
- ...the only provider of enterprise-scale context engines capable of analyzing trillions of real-time data points to create knowledge graphs that are... ...AI is seeking a highly skilled and motivated Site Reliability Engineer (SRE) to join our growing team. As an SRE at Lovelace...SuggestedFull time
$85 - $95 per hour
...Job Description Job Description Genesis10 is currently seeking a AVP / Head of Enterprise Site Reliability Engineering (SRE) with our consumer finance lender firm client in their Pittsburgh, PA location. This is a Right to hire position. Summary: An experienced...SuggestedHourly payPermanent employmentFull timeContract workImmediate startRemote work- ...Senior Site Reliability Engineer (SRE) Location: Pittsburgh, PA / Cleveland, OH / Dallas, TX FTE Position Overview We are seeking an experienced Senior Site Reliability Engineer (SRE) to support production operations, application reliability, performance management...SuggestedFull timeLocal areaShift workWeekend work
$148k - $249k
...Qualifications: - 5+ years software engineering or systems/performance engineering experience... ...activities and social events both on-site, off-site & virtually. As we grow, this... ...If you would like more information about how your data is processed, please contact us.SuggestedFull timeWork at officeWork from homeFlexible hours$70.8k - $156.7k
Senior Site Reliability Engineer - Local to Cleveland, Pittsburgh, or Dallas Position Description This role will require someone onsite at our client office in Cleveland, OH, Pittsburgh, PA, or Dallas, TX. Love technology? We do too. CGI is looking for a Site...SuggestedWork at officeLocal areaFlexible hoursShift workWeekend work- ...Job Description Job Description Job Title: Senior Site Reliability Engineer Job Category: Infrastructure/Cloud Job Type: Permanent Full Time Location: Pittsburgh, Pennsylvania, United States Position Description This role will require someone onsite...Permanent employmentFull timeWork at officeFlexible hoursShift workWeekend work
- ...are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and... ...Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong... ...of the market. We have redefined the data platform for the AI era, enabling builders...Full timeRemote workWorldwide
- Company Description Randstad Technologies Job Description Java 1.7 Oracle SQL Web sphere Application Server Additional Information All your information will be kept confidential according to EEO guidelines.Full time
$58k - $123.8k
...You will collaborate with engineering, platform, SRE, and security partners to automate... ...workflows, strengthen reliability and controls, and enable... ...-functional teams (DevOps, Data Engineering, Product Owners... ...here to be directed to our site that is dedicated to veterans...Work at officeLocal areaRemote work- ...Description Job Description Job Title: Software Engineer (Python) Duration : Contract to Hire... ...will collaborate with engineering, platform, SRE, and security partners to automate delivery workflows, strengthen reliability and controls, and enable teams to ship...Contract work
$172k - $229k
...hidden within petabytes of multimodal sensor data. Our next-generation autonomous driving... ...multimodal data mining framework, is the engine that powers this discovery. As a... ...production components. Drive Production Reliability : Establish patterns for graceful degradation...Full timeWork at officeRemote work$140k - $200k
...- Speechify has no office. These include frontend and backend engineers, AI research scientists, and others from Amazon, Microsoft, and... ...their own companies. Overview We're looking to hire for our Data side of our AI team at Speechify. This role is responsible for...Full timeWork at officeShift work- ...Python Data Engineer Strong Python and SQL skills, ETL development experience with Microsoft SSIS, and familiarity with Agile,... ...with relational databases and large-scale datasets to ensure reliable data integration. Develop and manage workflows in...Contract workLocal area
$147.9k - $220k
...Cloud Infrastructure / Site Reliability Engineer As a Cloud Infrastructure / Site Reliability Engineer... ...2 and 3 support for NetApp's Cloud Data Service solutions. Analysis, and Infrastructure... ...SaaS/IaaS environments. Implement SRE best practices for effective resolution...Odd jobLocal area- ...of Vice President, AI / Machine Learning Engineer to join our AI Hub team. This role is... ...high-quality delivery across architecture, reliability, and responsible AI adoption. To be... ...Bachelor's degree in Computer Science, Data Engineering, or a related field or the equivalent...WorldwideFlexible hours
- ...Software Engineer – Interoperability & Data Platforms We are seeking a highly skilled Software Engineer – Interoperability to design, build,... ...Optimize BDM jobs for performance, scalability, and reliability. Support metadata management, lineage, and operational...
- ...Job Description Job Description Site Reliability Maintenance Engineer This is an exciting new position meant to be a key player in our newly created... ...engineering standards based upon equipment failure data and lessons learned · Is able to work independently on...
- ...Big Data DeveloperLocation: Strongsville OH / Pittsburgh PA / Dallas TXDuration: Full TimeJob Description:Bigdata Hadoop & ecosystem, Scala/Python, PySpark.Oracle PL/SQL, CI/CD, Excellent communication, Data Lake, Informatica, Teradata.Roles & ResponsibilitiesStrong experience...
- ...Datacenter Engineer The engineer in this position should have a wealth of experience engineering... ...include the complete engineering of data centers including, but not limited to,... ...little direction read and understand site survey data, able to think through the information...For contractorsWork at officeLocal area
- ...Release Train Engineer Pittsburgh, PA (Hybrid, Onsite 3 days/week) 7 months Contract to Hire Supporting Enterprise Compliance Engineering Group. Majority of the work is related to Financial Crimes - Actimize, anti-money laundering (AML) and smaller crimes sanctions...Contract workWork at office3 days per week
$182.8k - $247.3k
...experiments (300+ at a time!) with our massive user base to make data-driven decisions, and educating our users and employees... ...!) learners around the world.About the role...As a Senior Site Reliability Engineer, you will work closely with both product and platform...Work experience placement- ...Overview: Role Summary We are seeking an experienced Release Train Engineer (RTE) to lead Agile Release Train execution in a SAFe environment. The role involves coordinating multiple Agile teams, managing dependencies, and driving program-level delivery excellence...
- ...We are seeking an experienced AI Engineer to join our AI Enablement team, focused on... ...built on LLMs that showcases your skill at reliably automating complex tasks. Experience... ...interact with backend systems, models, and data processing to increase internal team velocity...Full timeWork at office
- ...AI-enabled Applications and best-in-class data to more rapidly imagine, develop, and... ...We are seeking an experienced AI Engineer to join our Agentic AI team as we scale our... ...more effective agent, focusing on planning, reliable execution over longer time horizon tasks,...Full timeWork at office
$18 - $50 per hour
...3.0 GPA Perks: Employee discounts at our top customer sites Networking with our global leaders Mentorship from senior... ...: Siemens Digital Industries Software is seeking an AI & Engineering Data Intern to support emerging Artificial Intelligence and...Remote jobHourly payFull timeInternshipLocal area$179.2k - $268.8k
...systems, test operations, systems and safety engineering – all dedicated to redefining the... ...teams across the company to deploy software reliably to autonomous vehicles.This role sits at... ...to have: Experience in a DevOps, SRE, or infrastructure role in autonomous vehicles...Permanent employmentFull timeWork at officeImmediate startVisa sponsorship- ...Description Job Title: DevOps Engineer Duration : Contract to Hire... ...cross-functional teams to ensure reliable and compliant production... ...Collaborate with development, QA, SRE, and security teams to improve... ...DevOps, Cloud Engineering, or Site Reliability Engineering (SRE)...Contract work
- ...Data Engineer – Azure Databricks Contract-to-Hire Pittsburgh, PA – Onsite Job ID J0726-0569 Visa : USC, GC, EAD (No Sponsorship... ..., scalability, and cost efficiency. Ensure data quality, reliability, governance, and regulatory compliance through validation,...Full timeContract workTemporary workLocal area
- ...Senior AWS Data Engineer Full-time Hybrid – Lafayette, LA | Knoxville, TN | Birmingham, AL | Columbia, SC Position Overview Seeking a Senior AWS Data Engineer to develop, enhance, and maintain enterprise risk analytics applications in a financial services...Full timeLocal area
$70.8k - $156.7k
Senior Data Engineer Position Description This role will require someone at our client site 5 days a week in Pittsburgh, PA, We are seeking a Data Engineer with 6+ years... .... Experience ensuring data quality, reliability, and compliance in regulated environments...Hourly payLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer, AQI Data SRE. Be the first to apply!
- engineering aide Pittsburgh, PA
- technology administrator Pittsburgh, PA
- senior staff systems engineer Pittsburgh, PA
- staff engineer Pittsburgh, PA
- assistant engineer Pittsburgh, PA
- aws data engineer Pittsburgh, PA
- data engineer machine learning Pittsburgh, PA
- data science developer Pittsburgh, PA
- senior data engineer Pittsburgh, PA
- finance data engineer Pittsburgh, PA




