Site Reliability Engineer
GrabJobs
About Airwallex Airwallex is the only unified payments and financial platform for global businesses. Powered by our unique combination of proprietary infrastructure and software, we empower over 200,000 businesses worldwide – including Brex, Rippling, Navan, Qantas, SHEIN and many more – with fully integrated solutions to manage everything from business accounts, payments, spend management and treasury, to embedded finance at a global scale. Proudly founded in Melbourne, we have a team of over 2,000 of the brightest and most innovative people in tech across 26 offices around the globe. Valued at US$8 billion and backed by world-leading investors including T. Rowe Price, Visa, Mastercard, Robinhood Ventures, Sequoia, Salesforce Ventures, DST Global, and Lone Pine Capital, Airwallex is leading the charge in building the global payments and financial platform of the future. If you’re ready to do the most ambitious work of your career, join us. Attributes We Value We hire successful builders with founder-like energy who want real impact, accelerated learning, and true ownership. You bring strong role-related expertise and sharp thinking, and you’re motivated by our mission and operating principles . You move fast with good judgment, dig deep with curiosity, and make decisions from first principles, balancing speed and rigor. You're humble and collaborative; turn zero‑to‑one ideas into real products, and you “get stuff done” end-to-end. You use AI to work smarter and solve problems faster. Here, you’ll tackle complex, high‑visibility problems with exceptional teammates and grow your career as we build the future of global banking. If that sounds like you, let’s build what’s next. About the team Airwallex's Database team sits within the Infrastructure division and is responsible for the reliability, performance, security, and automation of Airwallex's database infrastructure — primarily Postgres (Google CloudSQL), Redis (Google Memorystore), Clickhouse and many more. The team's mission is to make databases invisible: product engineers should be able to provision, scale, and operate databases safely without needing to become database experts themselves. Today the team has 4 engineers across Singapore and China. We are expanding to Seattle to build a truly global team and to tackle the next generation of database automation — including AI-powered database operations, self-service platforms, and unified observability. You will join a team that values engineering rigor, automation-first thinking, and a bias toward building tools over performing manual toil. What you'll do As a Senior/Staff Software Engineer on the Database SRE team, you will design and build the platforms, automation, and AI-driven tooling that power Airwallex's database infrastructure. This is not a traditional DBA role — you are an engineer who builds systems that automate database operations at scale. You'll work across Postgres (CloudSQL) and Redis (Memorystore) on GCP, tackling challenges at the intersection of database reliability engineering, platform automation, and AI. Your work will directly impact the availability, performance, and security of the data layer that underpins Airwallex's global payments platform. This role is based in Seattle. Responsibilities: Build a unified database observability platform: Design and implement an "eagle-eye" view across all database instances, providing real-time visibility into availability, security posture, reliability metrics, and latency — enabling proactive issue detection before incidents occur. Build MCP (Model Context Protocol) for safe AI agent access: Design and implement secure interfaces that allow AI agents to safely query and interact with production databases, enabling the next generation of AI-powered internal tooling while maintaining strict security and access control. Build an Agentic AI DBA: Develop AI-powered automation that handles routine DBA tasks — performance tuning, capacity planning, schema review, incident triage, and remediation — reducing manual toil and enabling the team to scale operations without scaling headcount. Build self-service database platform: Create tooling that enables product engineering teams to provision, configure, scale, and manage Postgres and Redis instances through self-service workflows, with built-in guardrails for security, compliance, and best practices. Drive database reliability and operational excellence: Establish and enforce database best practices across the organization — connection pooling, query optimization, backup/restore strategies, failover procedures, and change management. Participate in on-call rotations for database incidents. Collaborate with engineering teams: Partner with product engineers to diagnose database performance issues, review schema designs, and provide guidance on data modeling and access patterns for Postgres and Redis. Who you are We're looking for people who meet the minimum requirements for this role. The preferred qualifications are great to have, but are not mandatory: 5+ years of professional software engineering experience with strong proficiency in backend languages such as Python, Go, Java, or Kotlin. Deep experience with PostgreSQL — performance tuning, query optimization, replication, backup/recovery, schema design, and operational best practices. Experience with Redis — data modeling, clustering, persistence strategies, and operational management. Infrastructure and cloud expertise: Hands-on experience with GCP (or equivalent cloud provider), infrastructure-as-code (Terraform), Kubernetes, and Docker. Demonstrated track record of building automation, tooling, or platforms that replaced manual operational work — not just operating databases, but engineering solutions that make operations scalable. Strong system design skills, with the ability to architect reliable, scalable infrastructure services. Preferred qualifications: Experience with Google CloudSQL and/or Google Memorystore specifically. Experience building internal developer platforms, database-as-a-service systems, or self-service infrastructure tooling. Experience with AI/ML technologies, LLM integration, or building AI-powered automation — particularly applied to infrastructure or operations use cases. Familiarity with MCP (Model Context Protocol) or similar patterns for safe programmatic access to production systems. Experience with observability tools (Grafana, Prometheus, Datadog) and building monitoring/alerting systems for database infrastructure. Experience in fintech, payments, or other regulated industries where data integrity, security, and compliance are critical. Applicant Safety Policy: Fraud and Third-Party Recruiters To protect you from recruitment scams, please be aware that Airwallex will not ask for bank details, sensitive ID numbers (i.e. passport), or any form of payment during the application or interview process. All official communication will come from an @ airwallex.com email address. Please apply only through careers.airwallex.com or our official LinkedIn page. Airwallex does not accept unsolicited resumes from search firms/recruiters. Airwallex will not pay any fees to search firms/recruiters if a candidate is submitted by a search firm/recruiter unless an agreement has been entered into with respect to specific open position(s). Search firms/recruiters submitting resumes to Airwallex on an unsolicited basis shall be deemed to accept this condition, regardless of any other provision to the contrary. Equal opportunity Airwallex is proud to be an equal opportunity employer. We value diversity and anyone seeking employment at Airwallex is considered based on merit, qualifications, competence and talent. We don’t regard color, religion, race, national origin, sexual orientation, ancestry, citizenship, sex, marital or family status, disability, gender, or any other legally protected status when making our hiring decisions. If you have a disability or special need that requires accommodation, please let us know.
$160k - $240k
...passionate about building unified IT solutions that simplify the way IT organizations work. We are currently looking for a Senior Site Reliability Engineer to join our SRE team in the Platform Engineering organization and help us scale our products to millions of end-users. We...SuggestedPermanent employmentFull timeRemote workWork from homeRelocationFlexible hours$114k - $148k
...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based...SuggestedFull timeTemporary workWork experience placementRemote work- ...scale, we invite you to bring your talents to Zscaler to help shape the future of cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or Bellevue WA) to join our Zero Trust Exchange team. This is a remote role based in San...SuggestedInternshipWork at officeLocal areaRemote workWorldwide
$100k - $115k
...Internal Developer Platform (IDP) as a product, treating engineering teams as customers and optimizing for reliability, usability, and delivery velocity.Define and... ....4+ years of experience in Platform Engineering, Site Reliability Engineering, DevOps, or Systems Engineering...SuggestedTemporary work$117k - $209.33k
Job Requisition ID #26WD99276Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products.As part of a new SRE team supporting...SuggestedFull timeFor contractorsRemote work- ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that... ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to...Full time
$138.4k - $173k
...infrastructure as well as help improve the reliability, quality of services and overall... ...recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability... ...about our locations by visiting our site.Compensation & BenefitsThe base salary that...Full timeFlexible hours$96.8k - $145.2k
...If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now.We are currently seeking a Site Reliability Engineer (Onsite Hybrid) to join our team in Plano, Texas (US-TX), United States (US).Job Responsibilities Include: Own and manage...Temporary workWork at officeRemote workFlexible hours- Qualifications: 8+ years of Software Engineering experience, or equivalent demonstrated through... ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform... ...vendor resources Willingness to work on-site at stated location in the job openingDepartment...Contract workFor contractorsWork experience placement
- ...Talent Acquisition Team will reach out to help you navigate our interview process.Lantern is seeking an experienced Senior Site Reliability Engineer to champion the reliability, availability, and performance of our Azure-based healthcare platform. In this pivotal role,...
$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...Full timeWork at officeLocal areaRemote workWork from home- ...: Meghana GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8 to... ...are seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will have...Remote work
$152.6k - $191.5k
...is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include composing... ...improvement.Position Summary:The Senior GCP Site Reliability Engineer acts as an advanced senior individual...Full timeWork at officeDay shift- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the IP, you will solve complex and broad business problems with simple and straightforward solutions...
- ...we are dedicated to connecting talented professionals with your ideal opportunities. We are currently seeking a qualified Site Reliability Engineer (AI & Agentic Systems) to join our client’s organization and contribute to their ongoing success. Job summaryThis role demands...
$174k - $253k
...systems by pushing for changes that improve reliability and velocity.Practice sustainable... ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical... ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you...$147k - $211k
...product or system development code.Review code developed by other engineers and provide feedback to ensure best practices (e.g., style... ..., and troubleshooting large-scale distributed systems. Site Reliability Engineering (SRE) is what you get when you treat operations...$141k
...be a part of our journey! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability...Local areaRemote workHome officeFlexible hours- ...The Depository Trust & Clearing Corporation (DTCC) is seeking a Senior Application Support Engineer (SRE) to enhance reliability, scalability, and performance of mission-critical applications. You will apply SRE principles across engineering, infrastructure, and operations...
- ...improving platform infrastructure and applications with high reliability, resiliency, performance & quality, and faster time-to-market... ...documentation, including runbooks/playbooks; and, Using Chaos Engineering to test the robustness of the systems and applications....
- ...Site Reliability Engineer II Our client, a leading organization in the financial services industry, is seeking a Site Reliability Engineer II to join their team. As a Site Reliability Engineer II, you will be part of the Infrastructure Support Department supporting...Weekly pay
- .... You will lead end-to-end product vision across production and sub-productions, translating reliability needs into a measurable strategy. You will partner with SRE engineering, platform teams, and stakeholders to deliver customer-centric outcomes and data-driven decisions...
$96.8k - $145.2k
...Site Reliability Engineer (Onsite Hybrid) NTT DATA strives to hire exceptional, innovative and passionate individuals who want to grow with us. If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now. We are currently seeking...Temporary workFlexible hours- Location: Plano, TX (Hybrid)3 days onsite 2 days remote look for nearby Candidates Must have Skills: Need SRE mindset Preferred coming from development background AWS Splunk App Dynamics (good Monitoring background ) Job responsibilities ...Remote work
$50 - $53 per hour
...area onsite at the project, significantly reducing and/or eliminating the demands to travel. Key Responsibilities: Site Reliability Engineers are expected to be able to drive technology triage efforts to completion by assisting with restoral steps, identifying...Hourly payLive inWork at officeLocal areaFlexible hours3 days per week- ...Administrator / SRE in Dallas to own production Java environments, middleware, and cloud automation. You will optimize performance, drive reliability, and mentor teammates while aligning with enterprise security and AI-enabled integrations. You will work across Java apps, IBM...
- ...Site Reliability Engineer (SRE) Our client, a IT Services and Consulting company, is looking for a Site Reliability Engineer (SRE) for their Plano, TX/Hybrid location. Requirements: Years of experience required: 11+ Mandatory skills: Azure DevOps (ADO), GitHub...
- ...Site Reliability Engineer III Location: Plano, Texas (Hybrid) Duration: 18 months Role Overview We are seeking a Site Reliability Engineer to support the production and operations of critical applications. This role focuses on establishing and improving monitoring...Work experience placement
- ...Site Reliability Engineer- W2 Role* Technical proficiency: Strong Proficiency in Java, Strong understanding of Database concepts (Oracle, SQL, Dynamo DB etc.) Industry standard SRE Tools like Prometheus, Grafana, Data Dog Etc Good to have skills: Cloud Concepts / AWS,...
$69.62k - $99.45k
...'ll be enabling progress and enhancing lives by providing reliable, high-speed connectivity solutions that keep the world connected... ..., Optimum is for you! Job Summary As a Site Reliability Engineer II, you will be a primary driver in the long-term management...Permanent employmentLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site services specialist Garland, TX
- construction site safety Garland, TX
- site leader Garland, TX
- official site Garland, TX
- IT site lead Garland, TX
- site safety Garland, TX
- on-site clinical research associate (traveling/remote) Garland, TX
- site reliability engineering manager
- junior site reliability engineer
- site reliability engineer

