Site Reliability Engineer
GrabJobs
The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, networking, load balancing (including our public-facing edge and internal service mesh), and observability and alerting systems. The Fleet Management team provides the core runtime environment that empowers our developers to build and ship products to delight our customers. We manage the end-to-end lifecycle of our Kubernetes fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper). As our infrastructure scales to support new use cases and products, we are spearheading a migration from Terraform-based Infrastructure as Code (IaC) to an Operator-driven lifecycle management model. This role can be based out of our Austin, Boston, Los Angeles, New York City, Raleigh, or San Francisco offices, remotely in the United States region, or our European office in Dublin. Responsibilities Contribute to developing and maintaining a scalable and secure runtime environment on top of Kubernetes that supports product needs across MongoDB Provide internal support for our Kubernetes ecosystem, partnering with engineering teams to help them solve domain-specific problems Participate in a 24/7 on-call rotation to resolve critical issues Prioritize blameless post-mortems and dedicate engineering time to systemic fixes, ensuring you aren’t paged for the same issue twice You may be a good fit if you Have 6+ years of experience in software development and operating distributed systems Are proficient in Go, Python, or a similar language, with a strong commitment to code quality and testing practices (writing unit, integration, and E2E tests) Have deep experience using and extending containerization technologies, preferably Kubernetes Have a solid understanding of Linux operating system internals and networking concepts (e.g., filesystems, TCP/IP, DNS, TLS) Possess a customer focused mindset, treating internal developers as your primary users Have strong operational ownership, including a track record of debugging complex production issues and driving them to resolution Prefer automation over manual processes ("allergic to ops work") We are a small team of software engineers with a strong bias toward building software solutions to eliminate toil Strong candidates may also have experience with Designing and implementing secure, multi-tenant runtime environments from first principles Proficiency with Kubernetes ecosystem tools such as Helm, Kustomize, Gatekeeper, Kyverno, and CRDs/Operators, CRI, CSI Expertise in cloud infrastructure platforms, including AWS, GCP, or Azure Proficiency in provisioning infrastructure using tools like Terraform, Crossplane, and AWS Controllers for Kubernetes (ACK) Advanced Linux systems internals and networking concepts specifically relevant to containers, such as namespaces and cgroups About MongoDB MongoDB is built for change, empowering our customers and our people to innovate at the speed of the market. We have redefined the database for the AI era, enabling innovators to create, transform, and disrupt industries with software. MongoDB’s unified database platform, the most widely available, globally distributed database on the market, helps organizations modernize legacy workloads, embrace innovation, and unleash AI. Our cloud-native platform, MongoDB Atlas, is the only globally distributed, multi-cloud database and is available across AWS, Google Cloud, and Microsoft Azure. With offices worldwide and over 60,000 customers, including 75% of the Fortune 100 and AI-native startups, relying on MongoDB for their most important applications, we’re powering the next era of software. Our compass at MongoDB is our Leadership Commitment, guiding how and why we make decisions, show up for each other, and win. It’s what makes us MongoDB. To drive the personal growth and business impact of our employees, we’re committed to developing a supportive and enriching culture for everyone. From employee affinity groups, to fertility assistance and a generous parental leave policy , we value our employees’ wellbeing and want to support them along every step of their professional and personal journeys. Learn more about what it’s like to work at MongoDB , and help us make an impact on the world! MongoDB is committed to providing any necessary accommodations for individuals with disabilities within our application and interview process. To request an accommodation due to a disability, please inform your recruiter. MongoDB, Inc. provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type and makes all hiring decisions without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws. Req ID: 426182
- ...sector specific focus The ability to work with cross-functional teamsRequired Skill and Experience• Strong experience in Site Reliability Engineering / Production Engineering.• Hands-on expertise with: IBM MQ (queue managers, clustering, channels, DLQ management).•...SuggestedFull timeTemporary workRelocation
$152.6k - $191.5k
...is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include composing... ...improvement.Position Summary:The Senior Azure Site Reliability Engineer acts as an advanced senior...SuggestedFull timeWork at officeDay shift- ...Join Our Team as a Site Reliability Engineer (SRE)! About Us At Energy Worldnet, Inc. (EWN), we deliver innovative technology solutions that empower our clients and support the future of the energy industry. Our SaaS platform helps organizations operate more safely, efficiently...SuggestedTemporary workCasual workH1bWork at officeRemote workHome officeMonday to FridayFlexible hoursShift work
- ...Site Reliability Engineer III Location: Chandler, Arizona (Hybrid) Role Overview We are seeking a skilled Cloud Site Reliability Engineer (SRE) with expertise in Infrastructure-as-Code (IaC) and large-scale VMware environments. This role is critical in the enterprise...Suggested
- ...customers globally. As part of the ongoing investment in the reliability and modernization of these systems, client is making a multi-... ...of change without impact to reliability. Integration Engineering, ensuring that client is consuming the right infrastructure and...SuggestedWork experience placement3 days per week
$56.49 - $66.49 per hour
...resiliency/HA Minimum of a 4-year degree in computer science or equivalent experience 8-10 years infrastructure or software engineering / development experience Desired Qualifications Experience with web programming languages, database design especially SQL...Hourly payContract workTemporary workWork experience placementLocal areaRemote workShift workWeekend work$70.23 per hour
...Job Title: Site Reliability Engineer (SRE) / Cloud Automation Engineer Location: Chandler, AZ (Hybrid, 3 days onsite per week required). Local candidates only. Duration: Contract - 12 months Pay Range: $70.23/hr (W2) Job ID: 407394 About BCforward BCforward is a leading...Contract workLocal area3 days per week$131.5k - $175.5k
...functionality of Tenable cloud products and ensuring they’re reliable and highly available in cloud environments Responsible for responding... ...with peers on complex projects Collaboration with cloud engineers in understanding new cloud technologies, assessing impact to security...Work experience placementH1bLocal areaRemote workFlexible hours- ...scale, we invite you to bring your talents to Zscaler to help shape the future of cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (Virginia) to join our Zero Trust Exchange team. This is an onsite role based in Crystal City, Virginia...InternshipWork at officeLocal areaRemote workNight shift
- ...within the Infrastructure division and is responsible for the reliability, performance, security, and automation of Airwallex's database... ...more. The team's mission is to make databases invisible: product engineers should be able to provision, scale, and operate databases...Worldwide
$130k - $180k
...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and... ...an in-house AI R&D team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to...Temporary workWork at officeImmediate startRemote workFlexible hours- ...Our client, a leading organization in the technology and financial services industry, is seeking a Site Reliability Engineer III to join their team. As a Site Reliability Engineer III, you will be part of the Infrastructure Support Department supporting the Application...Weekly payTemporary workFlexible hours
- ...solutions comply with applicable standards, promoting design, engineering, and organizational practices, and advocating and advancing... ..., and managing stakeholders.Overview:Seeking a seasoned Site Reliability Engineering (SRE) Leader to drive the reliability, scalability...Full timeWork at officeDay shift
$141k - $208k
...be a part of our journey! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability...Local areaRemote workHome officeFlexible hours- ...Senior Site Reliability Engineer (Enterprise Platform) Location: Remote - US - Open to Europe if happy to overlap with EST Compensation: Competitive We are a high-growth software company supporting the development of a premier open-source, EVM-compatible public ledger...Contract workCurrently hiringRemote work
$150k - $200k
...our CEO's funding announcement: . The Reliability team owns the availability, performance,... ...enforcing reliability standards across engineering Designing incident response processes and... ...strong ownership of production systems. As a Site Reliability Engineer on the Reliability...Remote workVisa sponsorshipWork visaFlexible hours- ...Partner with software developers, platform engineers, and IT staff to improve system design,... ...requirements, service quality, reliability, security, and compliance needs. Drive continuous... ...Required: 8+ years of experience in Site Reliability Engineering, DevOps, Platform...Work at officeRemote work
$141.8k - $195k
...their best work, grow fast, and bring their full selves to the herd. Why You’ll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in the U.S. Cribl...Temporary workRemote work- Job Title:Lead Information Security Site Reliability EngineerWells Fargo is back in the office collaborating for fabulous outcomes!This role... ....About this Role: We are seeking a Lead Site Reliability Engineer (SRE) to lead a team of SRE engineers to help mature our platform...Full timeWork experience placementWork at officeWork from homeVisa sponsorship2 days per week3 days per week
$112k - $137k
...The selected colleague will work at an MUFG office or client sites four days per week and work remotely one day. A member of... ...MUFG is seeking a highly motivated Certified Sr. Cloud Site Reliability Engineer to build a robust, scalable, and reliable web application environment...Full timeWork at officeLocal areaRemote work- ...Senior Site Reliability Engineer Come join a growing bank at the heart of the innovation, technology, green tech and life sciences space. We continue to expand our global footprint and our banking technology is at the core of everything we do. As a Senior Site Reliability...
$114k - $148k
...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based...Full timeTemporary workWork experience placementRemote work$115k - $145k
...600,000+ successful outcomes, EarWell ® is a proven, non-invasive treatment option for families. We are looking for a Site Reliability Engineer to join our growing team. The ideal candidate has a strong technical background in software development and systems operations...Temporary workWork at officeWork visa2 days per week$66.73 per hour
*Description* We are seeking a Site Reliability Engineer (SRE) to help establish and scale our client's Google Cloud Platform (GCP) SRE practice. This is a unique opportunity to join a team during its formative stage and play a key role in building the operational foundation...Contract workTemporary work- ...Job Summary: Platform Engineering & Operations - Administer, monitor, and tune Oracle Enterprise... ...optimization initiatives. Reliability & Automation (SRE Practices) - Define and implement Site Reliability Engineering (SRE) principles (SLIs...
- ...Join us!Position Summary:We are seeking a highly skilled Software Engineer with expert-level Golang experience, strong system design... ...candidate is a strong software engineer who can design and build reliable distributed systems in Go, develop software that integrates deeply...Full timeWork at officeFlexible hoursShift workDay shift
$122k - $200k
...opportunities to learn, grow, and make an impact. Join us!Job Description:This role is responsible for defining and leading the engineering approach for complex features to deliver significant business outcomes. Key responsibilities of the role include delivering complex...Full timeWork at officeDay shift- ...Fargo is seeking a Lead Systems Operations Engineer in technology as part of Commercial and... ...of engineering, service management, and reliability. You will be embedded early in the... ...and stabilized at scale.This role embeds Site Reliability Engineering and production engineering...Full timeWork experience placement
- ...Hardware-In-The-Loop Systems Engineer - Level 6RELOCATION ASSISTANCE: Relocation assistance may be availableCLEARANCE REQUIRED FOR START: NoCLEARANCE TYPE: NoneTRAVEL: Yes, 10% of the TimeDescription At Northrop Grumman, our employees have incredible opportunities to...Relocation package
$79.3k - $137.6k
...Our employees are not only part of history, they're making history.Northrop Grumman's Space Systems Sector is seeking a Software Engineer - Level 2 to join our team. This position can be located in either Redondo Beach, CA, Aurora, CO, or Gilbert, AZ.This position is...Full timeRelocation packageShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre Chandler, AZ
- site reliability engineer Chandler, AZ
- on-site clinical research associate (traveling/remote) Chandler, AZ
- website coordinator Chandler, AZ
- junior website developer Chandler, AZ
- site leader Chandler, AZ
- historic site Chandler, AZ
- website content developer Chandler, AZ
- construction site safety Chandler, AZ
- official site Chandler, AZ


