Senior Site Reliability Engineer
$4,000 - $5,000 per monthEnumerate
Senior Site Reliability Engineer
Latin America - Remote
We're looking for a Senior Site Reliability Engineer who can own the architecture, governance, and cost efficiency of our cloud and platform infrastructure. In this role you'll design and evolve our production environments, define standards and best practices, and partner with engineering and IT teams to build scalable, reliable systems that are easy to operate and cost-effective to run.
You'll be a hands-on technical leader: designing reference architectures, building CI/CD and automation pipelines, leading incident response practices, and setting guardrails around security, reliability, and cost management across our platforms.
This role is a remote contractor role. We are seeking candidates located in LATAM.
Key Responsibilities
Architecture & Infrastructure Ownership
- Design, implement, and evolve cloud infrastructure architectures for high availability, reliability, security, and scale.
- Define and maintain reference architectures and patterns for services, applications, and environments across the organization.
- Develop workflow processes and standards for building, deploying, and maintaining applications within a distributed architecture.
- Lead infrastructure modernization initiatives (e.g., containerization, Kubernetes adoption, infrastructure as code, platform consolidation).
Governance, Standards & Cost Management
- Establish and enforce governance standards for infrastructure, CI/CD, observability, and operational practices.
- Define and maintain policies for environment management, access control, configuration management, and change management.
- Implement cost management practices (e.g., tagging, budget alerts, rightsizing, reservations/committed use, auto-scaling policies) to optimize cloud spend.
- Partner with product and engineering leadership to balance performance, reliability, and cost-efficiency across environments.
- Use DORA metrics and industry benchmarks to drive continuous improvement in delivery and operational performance.
CI/CD, Automation & Operations
- Design, implement, and maintain CI/CD pipelines for multiple applications and environments using tools such as Git, Azure DevOps, GitLab, or Jenkins.
- Develop and manage automation pipelines for deployment, configuration, and infrastructure management.
- Build and maintain monitoring, alerting, and logging systems to ensure visibility, high availability, and performance of applications and services.
- Manage cloud infrastructure resources and services to ensure reliability, security, and scalability.
Incident Management & Reliability
- Lead incident response efforts, including triage, root cause analysis, and post-incident reviews.
- Contribute to and maintain incident response processes, runbooks, and on-call practices.
- Partner with engineering teams to design resilient systems and reduce mean time to recovery (MTTR).
Leadership, Mentorship & Cross-Functional Collaboration
- Collaborate with software engineering, QA, product, and IT teams to determine the best way to tackle complex infrastructure, security, and delivery challenges.
- Mentor engineers in DevOps and platform practices, tools, and standards across the organization.
- Lead departmental initiatives related to DevOps, platform engineering, and infrastructure disciplines; present plans and progress to stakeholders.
- Drive new department initiatives based on organizational needs and your expertise in modern technologies and industry trends.
- Stay current on emerging technologies, tools, and best practices; evaluate their potential application within our tech stack.
Required Experience
- BS or MS in Computer Science, Engineering, or a related technical field, or equivalent practical experience.
- 6+ years of experience with container orchestration services (Kubernetes preferred).
- 6+ years of experience administering and deploying CI/CD tooling (e.g., Git, Azure DevOps, Jira, GitLab, Jenkins).
- 6+ years of experience managing scalable applications in one or more major cloud providers.
- 8+ years of significant experience with both Windows and Linux operating system environments.
- 7+ years of experience with scripting and automation using tools such as PowerShell, Bash, or Python.
- 4+ years of experience with infrastructure-as-code and orchestration platforms (e.g., Terraform, ARM/Bicep, CloudFormation, Ansible, etc.).
- Demonstrated expertise designing architectures for scalable, reliable, and secure tech stacks in distributed systems.
- Demonstrated expertise implementing workflow processes for operating and maintaining applications in distributed architectures.
Qualifications & Skills
- Strong experience working in agile-leaning software development environments and across varying application stacks.
- Deep understanding of best practices and IT operations in distributed, cloud-native architectures.
- Experience defining and implementing governance and guardrails around infrastructure, CI/CD, and security.
- Strong grasp of cloud cost management and optimization techniques (e.g., usage analysis, rightsizing, scaling policies).
- Excellent problem-solving, troubleshooting, and incident management skills.
- Excellent oral and written communication skills; capable of presenting complex technical concepts to technical and non-technical audiences.
- Process-oriented with strong documentation skills and attention to detail.
- Ability to translate loosely defined product or platform requirements into robust, scalable technical solutions.
Total monthly compensation:
$4,000 - $5,000 USD
- About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure...Senior
- ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that... ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to...SeniorFull time
$170k - $220k
Who We're Looking ForWe’re looking for a hands-on, high-agency Site Reliability Engineer to help shape and scale the reliability layer of our stack. You'll own the release pipeline end-to-end — managing daily releases, weekly deploys, and hotfixes — while also automating...Senior$65 - $75 per hour
DescriptionKforce has a client seeking a remote Senior Site Reliability Engineer to be a l be a leading member of the team working with a diverse range of technologies. You will enjoy working in a friendly environment and benefit from our investment in staff. The role...SeniorRemote work$168k - $270.25k
NVIDIA is looking for a Senior Site Reliability Engineer (SRE) to join its GeForce Now (GFN) team. SRE at NVIDIA ensures that our internal and external-facing GPU cloud gaming services have reliability and uptime as promised to the users and at the same time enables developers...SeniorFull time- ...TechMContact: Meghana GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8... ...are seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will...SeniorRemote work
- IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion...SeniorWork at officeImmediate start
$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...SeniorCurrently hiringRemote work$168k - $270.25k
...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance...SeniorFull time$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...SeniorFull timeWork at officeLocal areaRemote workWork from home$152.6k - $191.5k
...responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include... ...and continuous improvement.Position Summary:The Senior GCP Site Reliability Engineer acts as an advanced senior...SeniorFull timeWork at officeDay shift- ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and... ...and networking teams to improve service reliability and deployment workflowsDeploy and... ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering...SeniorWork at officeLocal areaWork from homeFlexible hours
- Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work with...SeniorFlexible hours
- ...professionalism. We are seeking an experienced AWS solution design engineer/architect to join our infrastructure cloud team. The... ...product features efficiently and confidently them into production.As Senior SRE, you will be responsible for providing leadership, design and...Senior
- ...candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS...SeniorTemporary work
$267k - $356k
...day is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-... ...workloads in the industry, which means reliability and performance aren't just goals—they're... ...defined storage across new and existing sites using tools such as Ansible, Jenkins etc...SeniorWork experience placementWork at officeLocal areaWork from homeFlexible hours- Job Description:About the Role: We are looking for a Senior SRE to join our Platform Engineering team as the operations owner of our observability platforms. You’ll be responsible for the reliability, scalability, and continued evolution of the tools that give our engineering...SeniorFull time
- ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering... ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or...SeniorWork at officeLocal areaWork from homeFlexible hours
- LeanData helps the world’s fastest-growing companies automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly to the SVP of Engineering, this role is...SeniorFull timeWork at office2 days per week
$119.8k - $234.7k
...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...EngineeringDiscipline: Site Reliability EngineeringCompany: MicrosoftOverviewMicrosoft... ...’s most demanding workloads. As a Senior Site Reliability Engineer, you will lead reliability...SeniorOngoing contractLocal area3 days per week$101k - $161k
...several prestigious awards, such as Best Engineering Team, Best Company for Diversity,... ...DescriptionWho You'll Work WithWe’re looking for Site Reliability Engineers to join our growing Arista’s... ...: EngineeringExperience level: Mid-Senior LevelIndustry: Computer NetworkingSenior$80k - $140k
Job DescriptionRBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience, and operational excellence of critical applications and...SeniorFull timeFlexible hoursShift work- Job Description:Note: Fidelity will not provide immigration sponsorship for this positionThe RoleOur Site Reliability Engineering group within Enterprise Infrastructure combines Operations Excellence with the Development Experience to deliver services at high scale, high...SeniorFull time
$152.5k - $205k
...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate the secure, scalable platform infrastructure behind...SeniorFlexible hours$90k - $180k
...nutritionals and branded generic medicines. Our 115,000 colleagues serve people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.We...SeniorRemote work$158.5k - $172k
...the exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate,... ...environment. This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire...SeniorFull timeTemporary workWork at officeFlexible hours3 days per week- ...the team takes that seriously!The RoleThe Senior SRE at 2K is a hands-on technical leader... ...regions while partnering with network engineers, systems architects, and game studio developers... ...technical direction, influencing reliability from architecture review through production...Senior
$104.9k - $174.7k
Are you passionate about improving reliability, scalability, and resilience in complex database... ....Own prioritization of reliability engineering tasks within team backlogs.Lead incident... ...a Service (IaaS).Background in DevOps, site reliability engineering practices, or related...SeniorFull timeLocal area$98k - $176k
...joy of everyday life. We bring that vision to life through our values and culture. Learn more about Target here. As a Senior Site Reliability Engineer within Digital Enablement, you specialize in building and supporting the platforms and tools that enable teams to deliver...SeniorFull timeTemporary workWork experience placementFlexible hours$15k
...benefits packages, technology talks by our experts, a beautiful modern office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to meet our growing needs, and you will leverage...SeniorWork at officeLocal areaRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer remote United States
- site reliability engineer sre United States
- site reliability engineering manager United States
- site reliability engineer United States
- lead site reliability engineer United States
- senior groundskeeper United States
- senior maintenance supervisor United States
- senior operations associate United States
- senior safety specialist United States
- lcb senior living United States
