SRE Engineer
Jobgether
Sre Engineer
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an SRE Engineer based in Brazil.
As an SRE Engineer, you will play a key role in ensuring the reliability, scalability, performance, and resilience of critical digital environments. You will work across cloud infrastructure, Kubernetes, observability, automation, and incident management to keep systems stable and highly available. The role combines proactive engineering with hands-on troubleshooting, helping teams identify risks before they become production issues. You will establish and monitor reliability metrics such as SLIs, SLOs, SLAs, MTTR, MTTD, and error budgets to continuously improve operational performance. You'll collaborate closely with multidisciplinary teams to embed reliability practices throughout the software development lifecycle. Automation and Infrastructure as Code will be central to reducing manual work and creating more efficient, consistent operations. This is an opportunity to contribute to a culture of continuous improvement while working on modern, cloud-based and distributed technology environments.
Accountabilities:
- Define, monitor, and continuously improve reliability indicators, including SLIs, SLOs, SLAs, MTTR, MTTD, and error budgets.
- Implement and evolve observability, monitoring, alerting, and APM solutions across applications and infrastructure.
- Monitor latency, traffic, errors, saturation, availability, and overall system performance.
- Prevent, investigate, and resolve incidents, minimizing their impact on users and business operations.
- Conduct root cause analyses and define corrective and preventive actions to avoid recurring incidents.
- Identify operational risks, bottlenecks, single points of failure, and opportunities to strengthen system resilience.
- Support the design and evolution of highly available, scalable, resilient, and fault-tolerant solutions.
- Automate operational activities and reduce repetitive manual tasks through scripting, automation, and Infrastructure as Code.
- Operate and continuously improve Kubernetes and Docker environments.
- Support capacity planning, business continuity, disaster recovery, and cloud cost optimization initiatives.
- Participate in deployments and contribute to application stabilization following releases.
- Collaborate with engineering, development, infrastructure, and other technical teams to incorporate reliability from the earliest stages of solution design.
- Create and maintain operational dashboards, alerts, procedures, runbooks, and technical documentation.
- Promote a culture centered on reliability, observability, automation, prevention, and continuous improvement.
Requirements:
- Proven professional experience as a Site Reliability Engineer, SRE, or in an equivalent reliability/platform engineering role.
- Practical experience with cloud environments, using one or more of GCP, AWS, or Azure.
- Hands-on knowledge of Kubernetes and Docker.
- Experience implementing and managing observability, monitoring, alerting, and APM solutions.
- Strong understanding of SRE concepts and metrics, including SLI, SLO, SLA, MTTR, MTTD, and error budgets.
- Experience managing, investigating, troubleshooting, and resolving production incidents.
- Knowledge of application and infrastructure troubleshooting in complex environments.
- Experience administering Linux environments.
- Understanding of networking, security, performance, scalability, and high availability.
- Experience with automation and Infrastructure as Code practices.
- Hands-on experience with CI/CD pipelines and modern software delivery practices.
- Strong communication and collaboration skills, with the ability to work effectively across multidisciplinary teams.
- Analytical, proactive, collaborative, and prevention-oriented mindset.
- Experience with GKE, EKS, or AKS is a plus.
- Knowledge of Dynatrace, Datadog, Grafana, Prometheus, or comparable observability platforms is a plus.
- Experience with ELK Stack, Elasticsearch, and Kibana is desirable.
- Knowledge of Terraform and Ansible is desirable.
- Experience supporting critical systems and distributed architectures is an advantage.
- Experience in financial institutions or other regulated environments is a plus.
- Experience with cloud capacity management and cost optimization is desirable.
- Knowledge of disaster recovery and business continuity practices is beneficial.
- Cloud, Kubernetes, or SRE certifications are considered a plus.
Benefits:
- Meal and food allowance.
- Home office allowance.
- Medical insurance.
- Dental insurance.
- Life insurance.
- Birthday Day Off.
- TotalPass / Wellhub access.
- Health and wellness support through the Boon Saúde app.
- Discounts and partnerships with a variety of establishments.
- Partnerships with educational institutions and other services.
- Welcome kit.
- Structured onboarding program.
- Access to continuous learning and professional development initiatives.
- Dedicated learning and knowledge-sharing programs.
- Employee support and engagement initiatives.
- Fully remote work opportunity.
$117.2k - $176.7k
...U.S. government background investigation and clearance required for this role.Overview of the Role:Join our Site Reliability Engineering (SRE) team, where you'll work alongside Infrastructure and Research & Development (R&D) partners to keep Salesforce cloud services...SuggestedFull timeWork experience placement- ...intend for the selected candidate for this role to work on site in the specified location(s).We are seeking a Kafka Site Reliability Engineer to help build, operate, and continuously improve Schwab's enterprise streaming platform ecosystem. This role combines deep...SuggestedFull timeWork at office
- Job ID: 18699182Reference Number: 23-00128Title: SRE EngineerLocation: Iselin, NJ, 08830Posted Date: 2023-01-17Company: HAN Staffing Role: CICS System programmer - scripting skills (anyone - but python preferred) - drive meeting/calls (not pure PM) (Tech coordination...Suggested
- ...every day. If you're ready to grow, lead and make a difference, come join our team and help shape the future of convenience.The SRE RunOps Engineer 2 is responsible for ensuring the reliability, availability, and performance of the 7NOW delivery platform and associated...SuggestedHourly payWork experience placement
- ...innovation, empowering airlines, hoteliers, agencies and other partners to retail, distribute and fulfill travel worldwide.SRE Software Systems Engineer IV - Data Intelligence and AI OperationsWe are seeking a SRE Software Systems Engineer to join our global Data...SuggestedFull timeWorldwideFlexible hoursWeekend work
- #CareersJC 1483593Qualifications· Strong experience supporting production systems hosted on AWS, including EC2, VPC, ALB/NLB, RDS, Lambda, and EKS.· Hands-on experience with incident management and 24/7 production support models.· Proficiency with monitoring and observability...
- Job ID: 19257033Reference Number: 23-00565Title: SRE Devops EngineerLocation: Iselin, NJ, 08830Posted Date: 2023-03-16Contact: Shyam... ...hanstaffing.comContact Phone: (***) ***-****Company: HAN StaffingSRE engineer- Summit Minneapolis Charlotte Dallas Chandler Atlanta - Hybrid...
- ...management - Mandatory.? Triaging incidents.? Strong debugging mindset.? Ability to automate.? Troubleshooting skills? Reliability-first engineering approach.SRI Tech Solutions is an equal opportunity employer and does not discriminate on the basis of race, color, gender,...Full time
$207k - $300k
...outlook, architectural roadmap, and reliability strategy for Home SRE.Minimum qualifications:Bachelor’s degree in Computer Science, a... ...qualifications:Master's degree in Computer Science or Engineering, or a related field.Experience architecting complex client-to-cloud...Worldwide$140k - $170k
We are looking for a Senior Site Reliability Engineer to work as part of a lean, product‑focused engineering organization. This role is about... ...education and work experience 8+ years of experience in DevOps, SRE, platform engineering, or similar roles supporting application...Full timeWork experience placementFlexible hours$1,000 per month
...financial data safe. We're also building the AI infrastructure our engineers use every day. When this team does its job well, engineers ship... ...audits find real controls instead of gaps.As a Senior DevOps / SRE Engineer on this team, you'll own reliability and deployments...Temporary workWork at officeImmediate startRemote workFlexible hours- ...scale, where your contributions directly impact the stability and performance of critical financial services.As an Infrastructure Engineer III at JPMorganChase within Enterprise Technology (Infrastructure Platforms), you apply strong knowledge of software, applications...Shift work
- ...System|UNIX Technical/Domain Skill 3Foundational|Service Management|ITIL Technical/Domain Skill 4Technology|DevOps|Site Reliability Engineering(SRE) Work LocationAlpharetta, GA, New York, NYCountryUSAState / Region / ProvinceGeorgia, New YorkCompanyITL USA Interest...Full timeTemporary workRelocationShift work
- ...SRE Engineer Hi All, We have an immediate requirement on "SRE Engineer" @ Remote till Covid. So please share your suitable resumes to ****@*****.*** (***) ***-**** Ext 205. Role: SRE Engineer Location: Remote till Covid Type: C2C Job Description: ~9+ Yrs...Immediate startRemote work
- ...Site Reliability Engineer (SRE) Location: Remote Contract Length: 12 months w/ high likeliness of conversion or extension to FTE Contract Type: W2 Overview: The selected candidate will be responsible for the administration and engineering work concerning...Contract workCasual workRemote work
$31 - $42 per hour
DescriptionKforce has a client that is seeking a SRE AI Ops Engineer in Miami, FL (Florida).Summary:We are seeking an SRE AIOps Engineer to design, build, and support AI-driven automation solutions while providing advanced application and SRE operational support. This role...- ...Job Title Required qualifications: GCP data engineer skillset working with previous experience in cloud functions, workflow, big query, dataplex, python, dbt, sql and data warehousing for enterprise level system Knowledge/experience of data modelling Knowledge...Remote work
- Site Reliability Engineer - Vice PresidentSite Reliability Engineering (SRE) is an engineering discipline that combines software and systems engineering to build and run scalable, massively distributed, fault-tolerant systems. At Goldman Sachs, SRE is responsible for improving...
- Site Reliability Engineering (SRE) is an engineering discipline that combines software development and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. At Goldman Sachs, SRE is responsible for the availability and reliability...
- ...SRE / DevOps Engineer Canada / Remote 6+ Months Contract Position Requirements Collaborate closely with Development teams to improve services and operational targets. Implement and maintain CI/CD practices using tools like Sonar. Manage application integration...Contract workRemote work
- ...Duration: Long Term Contract Pay Rate: $40/Hr. W2 Experience: 3-5 Years Overview We are seeking a remote Junior SRE/DevOps Engineer role. The ideal candidate has foundational knowledge of Site Reliability Engineering (SRE) and Kubernetes, and is enthusiastic about growing...Long term contractContract workInternshipRemote work
- ...SRE DevOps EngineerLocation: Newport Beach, CADuration: Full TimeJob Description:We are seeking experienced DevOps Engineers with strong Site Reliability Engineering (SRE) capabilities who can work independently, think critically, and contribute immediately to our technical...Immediate start
- ...SRE DevOps EngineerTekfortune is a fast-growing consulting firm specialized in permanent, contract & project-based staffing services... ...experts can help you find the best job for you. Role: SRE DevOps Engineer Location: Austin TX Duration: 6 months Required Skills:...Permanent employmentContract workRemote work
$120k - $130k
...Tittle : SRE DevOps Engineer Location: Across USA any Location The pay range for this role is $120k - $130k per annum including any bonuses or variable pay. Job Description # Experience working on Google Cloud ( GCS, BigQuery ). # Experience...Remote work- ...DevOps Engineer Key Responsibilities Technical Skills • Strong hands-on experience with AWS cloud services • Expertise in Terraform... ...• Experience in Linux system administration DevOps / SRE Expertise • Strong understanding of DevOps principles and SRE...Remote work
- ...SRE Devops EngineerRootshell Enterprise Technologies Inc. is a recognized provider of professional IT Consulting services in the US. We are actively seeking SRE Devops Engineer Fulltime Role for one of our direct client.Role: SRE Devops EngineerLocation: Santa Clara,...Full timeLocal areaRemote work
- ...SRE/DevOps EngineerWe are looking for a highly technical, hands-on engineer to join a critical engineering organization supporting enterprise infrastructure and platform modernization initiatives. This role involves solving complex technical problems, automating infrastructure...
- ...SRE / DevOps EngineerJob Location: Mclean, VADuration: 3-month assignment with possibility of extensionJob Description:We are seeking a Site Reliability Engineer (SRE) with strong expertise in the client ecosystem and deep knowledge of cloud-native infrastructure. The...
- ...SRE/DevOps EngineerVersana is an industry-backed data and technology company on a mission to make the syndicated loan market better... ...source of deal information.Versana is seeking a motivated SRE/DevOps Engineer with strong observability experience to join our growing...Work experience placementLocal area
$100k - $170k
...and enhance our overall quality of life.We are seeking a DevOps Engineer who is eager to have an immediate impact in establishing and building... ...systemsWhat you bring to this role:5+ years of experience in SRE, DevOps, or Platform Engineering rolesProven experience...Full timeWork at officeImmediate startVisa sponsorshipNight shift
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE Engineer. Be the first to apply!
- site reliability engineer United States
- site reliability engineering manager United States
- site reliability engineer sre United States
- site reliability engineer remote United States
- lead site reliability engineer United States
- site reliability engineer
- site reliability engineering manager
- junior site reliability engineer
- site reliability engineer sre
- site reliability engineer remote

