Senior Kafka SRE Engineer
The Charles Schwab Corporation
Your OpportunityYour OpportunityAt Schwab, you’re empowered to make an impact on your career. Here, innovative thought meets creative problem solving, helping us challenge the status quo and transform the finance industry together. We believe in the importance of in-office collaboration and fully intend for the selected candidate for this role to work on site in the specified location(s).We are seeking a Kafka Site Reliability Engineer to help build, operate, and continuously improve Schwab's enterprise streaming platform ecosystem. This role combines deep expertise in Confluent Kafka technologies with modern Site Reliability Engineering practices to deliver highly available, secure, and resilient streaming services that support critical business capabilities across the firm.As a member of the team, you will drive operational excellence through automation, observability, Infrastructure as Code, and AIOps-driven capabilities. You will play a key role in designing, deploying, and supporting Confluent Kafka environments across on-premises and cloud platforms while helping engineering teams deliver reliable real-time data solutions. Success in this role requires the ability to proactively identify risks, solve complex technical challenges, improve platform performance, and enhance system reliability through data-driven decision-making and continuous improvement.You will contribute to the evolution of Kafka infrastructure supporting capabilities such as Schema Registry, Kafka Connect, ksqlDB, Cluster Linking, and cloud-native deployments. Through the application of automation, predictive analytics, and AI-driven operational insights, you will help reduce operational toil, strengthen platform resiliency, improve incident response, and accelerate issue resolution.The ideal candidate thrives in highly distributed environments and enjoys partnering with engineers, architects, and infrastructure teams to improve platform health, streamline deployment processes, increase observability, and establish best practices across the streaming ecosystem. This role offers the opportunity to influence the future of event streaming at Schwab while developing expertise in emerging AIOps, cloud, automation, and reliability engineering capabilities.What you haveRequired Qualifications5-7 years of experience supporting and administering enterprise-scale production Confluent Kafka platforms, including Confluent Platform and Confluent Cloud.5-7 years of experience developing Python automation, operational tooling, observability dashboards, and alerting solutions.Hands-on experience applying AIOps concepts, including anomaly detection, event correlation, alert reduction, predictive analytics, automated remediation, and AI-driven operational insights within production environments.Experience leveraging predictive monitoring and intelligent automation to improve platform reliability, incident management, and operational efficiency.Hands-on experience deploying and operating Kafka platforms on Kubernetes, preferably Google Kubernetes Engine (GKE), using Helm charts and cloud-native operational practices.Strong experience using Terraform and Ansible to automate infrastructure provisioning, configuration management, and platform lifecycle activities.Deep expertise with the Confluent Kafka ecosystem, including ZooKeeper, KRaft, Schema Registry, Kafka Connect, ksqlDB, Cluster Linking, MirrorMaker, and role-based access controls (RBAC).3+ years of experience working with public cloud technologies, with Google Cloud Platform (GCP) preferred.Strong understanding of Kafka architecture and client internals, including producers, consumers, partitions, replication, serialization, consumer groups, performance tuning, and exactly-once processing.Strong Linux administration, troubleshooting, performance tuning, and networking experience supporting high-throughput distributed systems.Experience supporting large-scale distributed systems, highly available platforms, fault-tolerant architectures, and production-critical workloads.Experience performing incident response, root-cause analysis, problem resolution, and post-incident continuous improvement activities.Experience implementing and maintaining observability solutions, operational metrics, service-level objectives (SLOs), dashboards, and actionable monitoring controls.Bachelor's degree in Computer Science, Engineering, Information Technology, or a related discipline.Preferred QualificationsConfluent certifications such as CCDAK or CCAAK.Google Cloud certifications such as Professional Cloud Architect or Professional Cloud DevOps Engineer.Experience operating Confluent for Kubernetes (CFK) and Confluent Cloud APIs.Experience implementing or supporting AIOps platforms for predictive incident detection, automated root-cause analysis, and self-healing infrastructure capabilities.Experience with observability and monitoring technologies such as Grafana, InfluxDB, BigQuery, Prometheus, Splunk, Datadog, or comparable monitoring platforms.Experience with CI/CD technologies such as GitHub Actions, Cloud Build, or similar automation frameworks.Experience supporting additional messaging and streaming technologies such as RabbitMQ, IBM MQ, Solace, or Google Pub/Sub.Understanding of modern Site Reliability Engineering practices, including SLIs, SLOs, error budgets, reliability engineering, and operational excellence methodologies.Strong communication, collaboration, and relationship-building skills with the ability to effectively partner across technical and business teams.Demonstrated ability to adapt to changing priorities, drive initiatives independently, and maintain a strong sense of ownership and accountability.In addition to the salary range, this role is eligible for bonus or incentive opportunities.Job SummaryRequisition ID: 2026-124407Posted Date: 3 hours ago(9/28/2026 5:34 PM)Category: Engineering & Software DevelopmentSalary Range: USD $120,000.00 - $155,000.00 / YearApplication deadline: 10/12/2026Position Type: Full time
- ...Opportunity: We are looking for a skilled engineer with disciplines that incorporate aspects... .... What you’ll do: • Evangelize SRE mindset and solve problems through systematization... ...Brokers: Solace, RabbitMQ, IBM MQ, Kafka. • Knowledge of Splunk, AppDynamics, or...Senior
$168k - $200k
...Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team.... ...ideal for someone who combines a strong SRE mindset with deep cloud... ...like Snowflake, S3, Delta Lake, and Kafka. Champion Event-Driven Architectures...SeniorRemote work- ...days per week TOP 3 MUST HAVE SKILLS: Experience with Confluent Kafka Platform Experience with both On-Prem and Cloud Must have... ...AI agents, custom skills, or equivalent frameworks to automate engineering and operational workflows at scale • Leverage AI-powered tools...Suggested
- ...experts can help you find the best job for you. Role: SRE DevOps Engineer Location: Austin TX Duration: 6 months Required... ..., Kustomize, Flux, Crossplane, CRDs, Python, Github, Kafka, Linux, Trino • Strong experience with Python scripting...SuggestedPermanent employmentContract workRemote work
- .... Our infra has to match. The role We're looking for a Senior SRE to own the reliability, scalability, and operational posture... ...using AI-assisted development workflows Partner closely with engineering on reliability reviews and architecture decisions ~5-8...Senior
- ...and best in class outcomesVisionary in future focused problem-solvingExceptional in execution and impactThe RoleAs a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various servicesand applications that come together to deliver...SeniorFull timeFlexible hours
- ...commercialization, and mass production to change the world for the better. JOB SUMMARY We are seeking an experienced Site Reliability Engineer to own and maintain the deployment of our cloud-based infrastructure to customer sites. In this role, you will work closely with...SeniorFull timeLocal area
- ...our core technology infrastructure. As a Senior AI Developer, this role will be a leader... ...assistants used across day-to-day engineering workflows in the SDLC—implementation, refactoring... ...of messaging technologies (Rabbit MQ, Kafka, or equivalent) • Experience in...Senior
- ...this role to work on site in the specified location(s). As a Senior Site Reliability Engineer within the CET SAvE organization, you will play a critical... ..., scalability, and resilienceImplement and evolve SRE best practices (SLOs, error budgets, incident reduction strategies...SeniorFull timeWork at office
- ...thrive. Are you ready to join our team and make an impact?As a Senior Site Reliability Engineer at TeamViewer, you’ll be a key player in ensuring the... ...Engineering, IT, or equivalent practical experience.5+ years in SRE, DevOps, or software development roles.Strong experience...SeniorTemporary workCasual workWorldwide
$121.4k - $218.6k
...Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages... ...Cloud's products and services. Our SRE teams solve reliability, security, and usability... ...and resource optimization. As a Senior Site Reliability Engineer, you will be...SeniorWork experience placementWork at office$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...SeniorTemporary workImmediate startFlexible hoursShift work- ...enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service... .... Your Impact You will be the most senior technical individual contributor on the team... ...experience. 7+ years of experience in SRE, cloud operations, and systems...Senior
$110.7k - $171.8k
...Description We are looking for a talented Sr. Site Reliability Engineer (SRE) to join our Observability team. In this role, you will work... ...to incidents, and implementing fixes under the guidance of senior team members. You'll also contribute to building and enhancing...SeniorWork at officeLocal area$168k - $270.25k
...impact on the world.NVIDIA is seeking an experienced software engineer to join the Cloud Foundations Automation team. Our team builds... ...distributed systems using databases, Redis, and messaging platforms (Kafka, NATS, SQS).Experience designing, building, and operating...SeniorFull timeRemote work- ...Senior Java Developer Austin, TX Requirement: ~ Bachelor's degree in Computer Science or equivalent is required ~8 or more... ...) ~ Proficiency in data processing technologies (e.g Kafka, Spark, Flink ) ~ Experience crafting scalable micro services...Senior
- ...Employment Type: Full-time, In-Office Department: Engineering Salary: Top of market salary + equity We're hiring a Senior Software Engineer on the growth team with... ...(Next.js) Django / FastAPI Postgres Kafka (we build on an event driven architecture)...SeniorFull timeWork at office
$140k - $224.25k
NVIDIA’s DGX Cloud organization is seeking a Senior Data Engineer to become part of its data team! We develop the reliable data foundation that... ...lakehouse or distributed-compute platform.Experience with Kafka or another streaming platform, change-data capture, event development...SeniorFull timeRemote work$110.7k - $171.8k
...asynchronous, workflow-driven automation; familiarity with orchestration engines such as Temporal is a strong plus - Experience with... ...- Building microservices that connect to MySQL, MongoDB, and Kafka - Strong understanding and hands-on expertise with Microservices...SeniorFull timeWork experience placementWork at officeLocal area$190k - $210k
...Senior Devops EngineerAustin, TXOne of the fastest growing tech companies... ...looking for a Senior Devops Engineer to own the infrastructure... ...in a Production Engineering, SRE, or Platform Engineering role... ...and queues - PostgreSQL, Redis, Kafka, or similarProficiency in at least...SeniorFlexible hours- ...sequencing and prioritization. • Leading a team and coaching them on engineering excellence and practices. • Fostering a culture of engineering... ...• 3+ years of experience in messaging technologies (RabbitMQ, Kafka or equivalent). • Experience in Test Driven Development, QA...SeniorFor contractors
- ...Job Summary The Senior Integration Software Engineer will design, build, and scale cloud-native enterprise... ...engineering, cloud engineering, DevOps, SRE, or related technical disciplines ~... ...technologies such as GCP Pub/Sub, Kafka, or similar platforms ~ Strong Infrastructure...Senior
- ...Job Description Job Description Job Title: Senior Cloud Engineer (Kubernetes / Helm / AWS) Location: Remote (U.S.; Austin, TX time zone... ..., and namespace isolation. Manage in-cluster PostgreSQL, Kafka, Grafana, and Elasticsearch via Helm or Operators....SeniorContract workRemote work
- Request ID: 109004-1 Title: Senior Java Developer Location: Hybrid: Austin, TX, Onsite Duration: 6 months Pay Range: $50 - $58/Hour on... ...RESTful services. • In-depth experience with messaging: SQS, SNS, and Kafka. • Experience in service performance profiling and optimization....Senior
- ...across the System Development Life Cycle. You will work with .NET engineering, event-driven architecture, cloud-native platforms, and... ...scalable, resilient, and event-driven applications using .NET, Kafka, and cloud-native architecture patterns.Build and modernize high...SeniorWork experience placement
- ...Senior Integration Software Engineer Location: Temple, TX or Austin, TX (Hybrid) Schedule: Hybrid, minimum... ..., cloud engineering, DevOps, SRE, or related technical disciplines.... ...driven technologies such as GCP Pub/Sub, Kafka, or similar platforms. Strong Infrastructure...Senior
$70.32 - $78.13 per hour
...the industry. Are you a skilled engineer passionate about combining software systems... ...the Site Reliability Engineering (SRE) mindset, build groundbreaking tools, and... ...Brokers, such as Solace, RabbitMQ, IBM MQ, or Kafka. Experience with Splunk, AppDynamics...Hourly payTemporary work- Senior Principal Software Engineer (ServiceNow Process Architect)Be a part of a team that’s ensuring Dell Technologies' product integrity and customer satisfaction. Our IT Software Engineer team turns business requirements into technology solutions by designing, coding...Senior
$85 - $90 per hour
...hr Date Posted: 04/01/2025 Hiring Organization: Rose International Position Number: 480571 Job Title: Senior Principal Software Engineer Work Model: Onsite Employment Type: Temporary Min Hourly Rate($): 85.00 Max Hourly Rate($): 90.00...SeniorHourly payTemporary workFlexible hours- ...Procore Technologies seeks a Senior Principal Software Engineer to lead the Workflows platform, shaping long‑term technical direction and reference patterns. You will actively code and design while guiding multiple teams toward scalable, AI‑enabled construction SaaS platform...SeniorImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Kafka SRE Engineer. Be the first to apply!
- site reliability engineer Austin, TX
- site reliability engineer sre Austin, TX
- senior living director Austin, TX
- senior manager customer operations Austin, TX
- senior support engineer Austin, TX
- senior product manager mobile Austin, TX
- senior java developer Austin, TX
- senior software engineer ruby on rails Austin, TX
- sr finance manager Austin, TX
- sr marketing manager Austin, TX



