Senior AI Site Reliability Engineer, AI.x
The Charles Schwab Corporation
Your OpportunityAt Schwab, you will build a rewarding career while making a difference in the lives of our millions of clients. Here, innovative thinking meets creative problem solving as we work together to challenge the status quo. Joining Schwab means joining a company committed to transforming the financial industry and putting clients at the center of everything we do.Schwab’s AI Strategy & Transformation team, known as AI.x, is the central hub for Artificial Intelligence at Schwab. We are an integrated product, engineering, strategy and risk team, all based in San Francisco. We help set the enterprise vision for AI, invest in the most promising opportunities, and accelerate delivery across the company. We also build the core platform that powers AI at scale and explore next-generation GenAI efforts that will redefine how we serve our clients. As a Senior Engineer on AI.x, you will play a key role in bringing these priorities to life by designing and delivering innovative AI solutions.This role is an opportunity to join a high-profile team shaping Schwab’s future with AI, to build solutions that matter to millions of clients, and to grow your career in one of the most exciting areas of technology today.As a Senior AI Site Reliability Engineer you will support reliability efforts for cutting-edge GenAI applications that enhance the client experience and create value. You will work closely with architects and engineers to ensure scalability, reliability and security of solutions that build towards an enterprise strategy. You will lead automation-first initiatives, build robust CI/CD pipelines for one-touch deployments, and implement comprehensive observability frameworks to minimize MTTD and MTTR. This role requires participation in on-call rotations to ensure 24/7 reliability of critical AI systems. Above all, you will apply the rigor, discipline, and technical depth to help shape the next generation of AI at Schwab.Roles & Responsibilities:Lead automation-first initiatives to eliminate toil and manual interventions, defining and executing the strategic roadmap for reliability, observability, and self-healing systems across AI.x platformsDesign and implement robust CI/CD pipelines enabling one-touch deployments with automated testing, validation, and rollback capabilities to accelerate delivery velocity and reduce deployment riskImplement comprehensive observability frameworks for real-time monitoring of AI services, including metrics, logs, and traces, with intelligent alerting and automated diagnostics to minimize MTTD and MTTRParticipate in on-call rotation providing 24/7 support for production AI systems, ensuring rapid incident response, root cause analysis, and resolution with measurable SLO targetsEstablish and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), error budgets, and incident response runbooks to drive continuous reliability improvementsChampion Infrastructure-as-Code (IaC) practices and automate environment provisioning, configuration management, and deployment processes to ensure consistency, repeatability, and operational efficiencyCollaborate seamlessly with AI Engineering teams to integrate SRE practices early in the development lifecycle, promoting a culture of reliability and shared responsibilityProactively identify and resolve reliability, performance, and scalability issues through data-driven analysis, capacity planning, and system optimizationImplement and maintain monitoring, alerting, and incident response frameworks to ensure system health and reliability, maximizing production availabilityChampion reliability, monitoring, observability, and operational best practices for AI systems and data pipelines, establishing patterns and standards for the organizationWhat you haveRequired Qualifications8+ years of software engineering experience, with 4+ years as a hands-on Site Reliability Engineer in startups and/or large organizations.Bachelor’s degree in Computer Science or related field, or equivalent experience.5+ years building complex products from scratch, running them in production, and ensuring operational reliability.3+ years working with containers and cloud-native applications, operationalizing them in the public cloud with infrastructure as code and CI/CD pipelines.3+ years of experience working in high-availability hybrid-cloud environments.Preferred QualificationsStrong computer science fundamentals and experience across the tech stack.Experience with proprietary or open-source LLMs (e.g., Gemini, Claude, OpenAI), deploying LLM-powered applications to production and maintaining availability.Strong written and verbal communication skills to clearly convey ideas and feedback.Strong understanding of observability, incident management and reliability engineering principles.Mindset of continuous learning and improvement, adept at both giving and receiving feedback.Ability to troubleshoot complex problems with ambiguous or incomplete data in distributed systems.Curiosity about new technologies and processes, proactively sharing knowledge and seeking improvement.Experience with Terraform and Google Cloud Platform.In addition to the salary range, this role is eligible for bonus or incentive opportunities.Job SummaryRequisition ID: 2026-122527Posted Date: 1 week ago(7/29/2026 4:07 PM)Category: Engineering & Software DevelopmentSalary Range: USD $170,000.00 - $220,000.00 / YearApplication deadline: 8/11/2026Position Type: Full time
- ...takes that seriously!The RoleThe Senior SRE at 2K is a hands-on... ...while partnering with network engineers, systems architects, and game... ...technical direction, influencing reliability from architecture review... ...cloud scale.Experience with AI and Agentic Development.Cloud...Senior
- ...work better. Our software solutions harness the power of AI and shape the future of digitalization.We believe that our... ...thrive. Are you ready to join our team and make an impact?As a Senior Site Reliability Engineer at TeamViewer, you’ll be a key player in ensuring the...SeniorTemporary workCasual workWorldwide
$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of... ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager... ...redefined the data platform for the AI era, enabling builders to create, transform...SeniorWork at officeLocal areaRemote workWorldwideFlexible hours$152k - $241.5k
...intelligence.We’re looking for a Senior SRE to join our Compute Farm... ...You’ll harness the power of AI to deliver groundbreaking... ...lifecycle management, fleet reliability/auto-healing, E2E observability... ...Perl, or Ruby.Mentored other engineers and influenced technical direction...SeniorFull time- ...for the selected candidate for this role to work on site in the specified location(s).As a Senior Reliability Engineer, you will help shape the reliability, scalability,... ...scale. Leveraging modern observability practices, AI/ML-enabled operational capabilities, and proactive...SeniorFull timeWork at office
$165k - $241.4k
...ones they don’t own. Powered by AI and an unmatched set of cloud,... ....We’re looking for talented engineers with a software or operations... ...development teams to ensure the reliability, performance and security of... ...Please see the Cisco careers site to discover more benefits and...SeniorFull timeTemporary workWork at officeLocal areaFlexible hours1 day per week$127k - $249k
...zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support,... ...Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure... ...redefined the data platform for the AI era, enabling builders to create, transform...SeniorLocal areaRemote workWorldwideFlexible hours$127k - $249k
We are looking for an experienced Senior or Staff Engineer for our SRE, InfraSec team, to guide the security of our cloud-based infrastructure. As... ...of the market. We have redefined the data platform for the AI era, enabling builders to create, transform, and disrupt industries...SeniorLocal areaRemote workWorldwideFlexible hours$152k - $195k
...Senior Site Reliability Engineer Austin, TX (Hybrid) SecurityScorecard is the global leader in cybersecurity ratings, with over 12 million companies... ...systems. You will also own the infrastructure behind our AI tooling — building MCP servers and defining safe,...Senior- ...Apptronik is a human-centered robotics company developing AI-powered robots to support humanity in every facet of life. Our... ...the better. JOB SUMMARY We are seeking an experienced Site Reliability Engineer to own and maintain the deployment of our cloud-based...SeniorFull timeLocal area
$141k - $208k
...analytics, data warehousing, observability, and AI workloads. The company’s sustained,... ...are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and...SeniorLocal areaRemote workHome officeFlexible hours$121.4k - $218.6k
...delivery challenges? Join our critical AI Hardware SRE Team! The AI Hardware... ...ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings... ...them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for:...SeniorWork experience placementWork at office- ...missions of the United States and allied partners. We build AI systems that determine how logistics decisions are made — not... ...advantage. About the Role Gallatin is looking for a Site Reliability Engineer to keep our production systems running with the reliability...SeniorFull timeLocal area
- ...digital space - ready to join us? What’s the position? We are looking for a Senior Site Reliability Engineer who combines deep infrastructure expertise with a forward-thinking approach to AI-driven operations. In this role you will maintain and improve the reliability...SeniorRemote workFlexible hoursNight shift
$81.1k - $187k
...according to terms for reliability and functionality. ~... ...Gains basic knowledge of site reliability trends and... ...and escalate issues to senior team members. Collects... ...skilled Site Reliability Engineer to design, build,... ...-saving care. And with AI embedded across our products...SeniorTemporary workImmediate startFlexible hoursShift work- ...the selected candidate for this role to work on site in the specified location(s).We are seeking a Kafka Site Reliability Engineer to help build, operate, and continuously... ...application of automation, predictive analytics, and AI-driven operational insights, you will help...SeniorFull timeWork at office
- ...Description Job Description Sr. Software Engineer - Site Reliability About ShipperHQ: ShipperHQ is a... ...Position Overview: We’re seeking a Senior Site Reliability Engineer to join our... ..., solutions-oriented approach ~ AI fluency: You leverage a diverse AI toolkit...SeniorFull timeWork at office
$184k - $287.5k
...over Converged Ethernet) we make powerful ML/AI platforms possible. We believe in our... ...team!We seek experienced software embedded engineers to help support our groundbreaking, innovative... ...networking technologies such as Spectrum-X and work with customers on their technical...SeniorFull timeRemote work- ...-15382** Onsite in Austin or Southlake 4 x weekly** Your Opportunity At Client, you’... ...our core technology infrastructure. As a Senior AI Developer, this role will be a leader in... ...coding assistants used across day-to-day engineering workflows in the SDLC—implementation, refactoring...Senior
- ...times per week.The RoleWe are looking for a Senior Software Engineer to join the team that is revolutionizing... ...core focus of this role is Automation, AI Enablement, and Modernization — you will... ...in accessibility programming (WCAG 2.x compliance) to meet regulatory standards...SeniorFull timeLocal areaWork from homeRelocation package
$140k - $215k
...with the world’s most advanced AI-native platform. We work on... ...our Core Platform and Embedded Reliability charters: building the... ...embedding directly with product engineering teams and their leadership to... ...deployment processes.At the Senior Engineer level, your influence...SeniorFull timeWork experience placementWork at officeLocal area2 days per week3 days per week$168k - $270.25k
...tapping into the unlimited potential of AI to define the next era of computing.... ...Experience (NVEX) Solutions Engineering team is looking for a senior Computer or Software Engineer. This person... ...like InfiniBand, NVLink, and Spectrum-X that link GPUs and AI compute infrastructure...SeniorFull timeWeekend work$140k - $224.25k
The NVIDIA Experience (NVEX) Solutions Engineering team is looking for a senior Computer or Software Engineer who is... ...-breaking network technology used in AI clusters. Our team of software... ...for InfiniBand, NVLink, and Spectrum-X network systems that interconnect GPUs...SeniorFull timeWeekend work- ...leadership in cloud, data and AI with unmatched industry... ..., Operations, Industry X and Song, together with... ...applied AI and data engineering. We help the world’s... ....You Are:You are a Senior AI engineering Leader who... ...for LLMOps, security, reliability, and cost governance across...SeniorFull timeWork experience placementLive inWork at officeLocal area
- ...coolest projects you can imagine.The Work:The Senior SAP APO Consultant is responsible for... ...technology and leadership in cloud, data and AI with unmatched industry experience, functional... ..., Technology, Operations, Industry X and Song, together with our culture of shared...SeniorFull timeWork experience placementLive inWork at officeLocal areaShift work
- ...Tank is financial AI + payments for the... ...front-end software engineer. Proven track record... ...building highly scalable, reliable systems. Strong... ...challenge with a senior engineer to assess... ...min, virtual). On-Site Interviews: Meet... ...2 non-technical (4 x 60 min, in person in...SeniorWork at officeFlexible hours
$98.58k - $138.02k
...Silicon Valley Region / Denver, COProduct Engineering - DevOps /Full Time /HybridRestaurant36... ..., TX; Irvine, CA; or Akron, OH. The Site Reliability Engineer II will be responsible for supporting... ....We may use artificial intelligence (AI) tools to support parts of the hiring...Full timeWork at office- ...the elite technical and product engine of the Accenture Google... ...technology: the move to Agentic AI and Product-Led Operating Models... ...Platform (GCP) Agentic AI Delivery Senior Engineer, you are a new breed of... ...Technology, Operations, Industry X and Song, together with our...SeniorFull timeWork experience placementLive inWork at officeLocal areaShift work
$196k - $269.5k
Senior Principal AI Agent EngineerThe Software Engineering team delivers next-generation software application enhancements and new products for a changing world. Working at the cutting edge, we design and develop software for platforms, peripherals, applications and diagnostics...Senior- ...engagement. A Cybersecurity Forward Deployed Engineer is a production engineer who works... ...their security and engineering teams—to make AI systems secure, governed, and resilient in... ...Consulting, Technology, Operations, Industry X and Song, together with our culture of...SeniorFull timeWork experience placementLive inWork at officeLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior AI Site Reliability Engineer, AI.x. Be the first to apply!
- ai engineer Austin, TX
- machine learning ai engineer Austin, TX
- senior ai engineer Austin, TX
- ai ml engineer Austin, TX
- ai prompt engineer Austin, TX
- ai engineer remote Austin, TX
- ai developer Austin, TX
- site reliability engineer remote Austin, TX
- site reliability engineer sre Austin, TX
- site reliability engineer Austin, TX


