Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

OutSystems

There are NO limits to your career: come shape the future and be part of a truly unique global culture at OutSystems!

Hybrid Onsite in Menlo Park, CA

Site Reliability Engineering (SRE) is a discipline that incorporates aspects of software engineering and applies them to infrastructure and operations problems. The main goals of SRE are to create scalable and highly reliable systems. Our SREs ensure our production systems' reliability, performance, and scalability while enabling rapid development and deployment of new features and services.

SREs at OutSystems work closely with development teams, acting as an extension of the team, in adopting the reliability tenets with the shared goal of meeting Service Level Objectives (SLOs) and thus delivering a smooth and frictionless Customer Experience.

Site Reliability Engineer Role
As an SRE at OutSystems here are your key responsibilities and duties:

Lead and onboard services and teams to the reliability tenets;

Establish and maintain Service Level Objectives (SLOs) and Service Level Agreements (SLAs);

Design and implement scalable, reliable, and secure infrastructure, while ensuring cloud-native best practices;

Collaborate with software development teams to ensure systems are resilient (observable, fault-tolerant, recoverable, scalable) and performant;

Implement monitoring, alerting, logging, and tracing solutions to detect and respond to incidents;

Lead incident response efforts, ensuring quick resolution and minimal downtime, and conduct RCA/post-mortems;

Automate every operational task, with a special focus on fast incident detection & recovery;

Programming in Python supported by Gen AI tooling to accelerate development of mission critical automation and tools.

Foster a culture of continuous improvement and knowledge sharing;

Communicate effectively with stakeholders, providing updates on system reliability and performance;

Participate in on-call rotation to provide 24/7 support for production systems.

Site Reliability Engineering Performance Indicators
The main KPIs that aid in understanding the impact and success of the SRE function at OutSystems are:

SLA and Service Level Objectives (SLO) compliance;

SLO Coverage and Detection Ratio;

MTTA - Mean time to acknowledge;

MTTR - Mean time to resolve.

Qualifications and Skills
To illustrate the desired profile for a Site Reliability Engineer. Nevertheless, the selection of candidates will always vary depending on specific knowledge of the field and prior experience.

Qualifications
BS/MS in Computer Science or Equivalent

6+ years of experience in Site Reliability Engineering, managing infrastructure and services at scale

History of end-to-end project delivery

Experience managing Hadoop and Kubernetes infrastructure and related services, or equivalent experience

Advanced knowledge of Linux, Networking, and Containers

Proficiency in at least one high-level programming language (Python, GoLang etc.).

Strong troubleshooting and debugging skills.

Fluency in English and excellent communication skills.

An understanding or hands-on experience with Prompt engineering in software development;

Familiarity with AI Native IDEs or AI Assistants such as Cursor, GitHub CoPilot, and Claude.

Soft Skills
Communication - able to communicate effectively (in English) both orally and written showing empathy for the other person;

Collaboration - Proactive collaboration and presentation skills to effectively communicate ideas and represent the deliverables and needs of the SRE team with leadership.

Humbleness - accepts mistakes and acts accordingly, with a humble attitude, apologizing for them and mitigating them ASAP to avoid higher impact.

Accountability - takes ownership of problems and makes sure to see them through. Even if he does not have all the necessary knowledge to move on alone, can involve the right people to reach closure.

Negotiation Skills - has tough and politically complex conversations with colleagues and customers, defusing disagreements and leading towards a mutual agreement and understanding of all parties involved.

Process Oriented - is organized and able to properly follow defined processes, whilst being able to properly challenge inefficient processes and suggest improvements.

Problem-solving - Has a top-down approach to problems, breaking them into smaller pieces and solving them by starting with a wider scope and narrowing it down as the analysis progresses. Has critical thinking, so can analyze information objectively and make a reasoned judgment.

Technical Skills
Experience in any of the following is valued, but not fully required:

Ability to establish, monitor, and improve Service Level Objectives (SLOs), Indicators (SLIs), and Agreements (SLAs) in line with business needs.

Containerization technologies and orchestration platforms, mainly Kubernetes and EKS
(CKA, CKAD, CKS certifications are valued);

Experience with automation and Infrastructure as Code (IaC) tools, such as AWS CloudFormation, Terraform, Puppet, Chef, Spacelift, etc;

Experience with Python, Go, Bash/Shell scripting, or other automation tools/languages;

Familiarity with AWS services like EC2, RDS, ELB, CloudFront, Lambda, etc;

Proficiency in monitoring and troubleshooting complex distributed systems;

Experience with Grafana, ELK stack, Prometheus, or others;

Strong understanding of designing resilient and fault-tolerant systems;

Expertise in debugging complex distributed systems.

More about OutSystems

OutSystems is a leading AI Development Platform built for the enterprise. Global organizations trust OutSystems to rapidly build mission-critical apps and agents, modernize legacy processes with agentic systems, and govern their entire AI portfolio across complex regulatory environments, all on one unified platform.

As the future becomes agentic, our customers need us now more than ever. While AI has opened the door to extraordinary possibilities, most large organizations find themselves stuck on one side of the "enterprise gap" because AI by itself doesn't solve their complex use cases and business challenges. OutSystems bridges the "enterprise gap" by combining the speed of generative AI with a deterministic, enterprise-grade framework. We provide the tools for teams of any size to deliver high-quality, reliable AI solutions that drive real business impact.

We are looking for passionate, talented, and motivated people to join us as we empower organizations to build, deploy, and scale the next generation of enterprise software. While we are leading the charge into the agentic era, our mission is broader: we are the platform enterprise leaders trust to evolve their entire business, accelerating innovation through secure, governed human-AI collaboration.


OutSystems is a global company, with more than 900k developer community members, 1,700 employees, more than 600 partners, and thousands of active customers in over 75 countries and across 21 industries. Founded in 2001, OutSystems now has offices in the United States, United Kingdom, the Netherlands, Portugal, Germany, the UAE, Japan, Hong Kong, Malaysia, Australia, India, and Singapore, and includes a thriving, worldwide community of remote employees.

Our customers are some of the world's most recognizable brands across diverse industries- such as Toyota, Heineken, Bosch, KeyBank, and UCLA-who trust OutSystems to deliver ROI and transformational impact.


Consistently recognized as a leader by top analyst firms Gartner, IDC and Forrester, OutSystems continues to shape the future of enterprise software development in the agentic era. We are proud to be named a leader in more than 100 categories on G2, including #1 in Customer Satisfaction in Enterprise Low Code Development, and most recently as a leader in AI Agent Building in the G2 Spring 2026 Reports.


Working at OutSystems

Our culture is built on our core values of Trust, Customer Success, Innovation, and Alignment. We operate as one global OutSystems team, taking ownership to pursue our vision of being the AI platform enterprise leaders trust to build, secure, and evolve their most critical applications and systems.

What do we have to offer you?
  • A company at the vanguard of the agentic revolution, where we don't just react to AI innovation-we architect it. Joining OutSystems means stepping onto a high-growth rocket ship that combines the fearless agility of a startup with the sophisticated, global foundation of an enterprise powerhouse.
  • Real growth opportunities. We don't just talk about development; we invest in it through structured programs designed to scale your expertise. Whether you are aiming for vertical progression, exploring lateral moves into new domains, or mastering specialized AI skills through our Professional Development Fund and Internal Mobility Program, we provide the resources to get you there.
  • A global collective of world-class talent, where you'll collaborate with enterprise software legends and sought-after thought leaders. At OutSystems, our industry experts aren't just visionaries-they are accessible, approachable mentors who are deeply invested in your growth as we architect the agentic future together.

OutSystems nurtures an inclusive culture where talented individuals from all backgrounds are empowered to learn, experiment and make an impact. . We believe that driving our next phase of growth requires the radical creativity that only comes from diverse perspectives. We are committed to building a team as global and diverse as the organizations we serve, ensuring every individual can perform to their full potential. As an equal opportunity employer, all qualified applicants receive equal consideration regardless of race, origin, religion, sex, sexual orientation, gender identity, disability, veteran status, or any other protected status.
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in San Francisco, CA vacancy
  •  ...US Corp. is seeking a Lead Site Reliability Engineer to spearhead our mission of delivering highly available and performant systems. With an average of over 12 years of industry experience, the successful candidate will bridge the gap between software development and systems... 
    Senior

    Axiom Pursuits

    San Francisco, CA
    8 hours ago
  •  ...About the Role We're looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You'll partner with engineers and data scientists to build, automate, and maintain... 
    Senior

    Alembic Limited

    San Francisco, CA
    14 hours ago
  • $210k - $240k

     ...Join to apply for the Senior Site Reliability Engineer role at Alembic Technologies This range is provided by Alembic Technologies. Your actual pay will be based on your skills and experience — talk with your recruiter to learn more. Base pay range $210,000.00/yr - $2... 
    Senior
    Full time

    Alembic Technologies

    San Francisco, CA
    3 days ago
  •  ...Senior Engineering Role at Salesforce Salesforce is the #1 AI CRM, where humans with agents drive customer success together. Here,...  ...Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with... 
    Senior
    Worldwide
    Weekend work

    Salesforce

    San Francisco, CA
    1 day ago
  • $175k - $250k

     ...000.00/yr - $250,000.00/yr Job Title: Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On-Site only. Must live within commuting distance...  ...scalability, performance, and reliability across environments. What You’ll Do Design... 
    Senior
    Full time
    Remote work
    Relocation
    Relocation package

    The Recruiting Guy

    San Francisco, CA
    3 days ago
  •  ...of healthcare, we'd love to meet you. Apply now to join our growing team. About the Role Plenful is hiring a Senior Site Reliability Engineer (SRE) to keep our production systems reliable, performant, and scalable as we grow. This role is centered on operating... 
    Senior
    Full time
    Work at office
    Remote work
    Flexible hours
    2 days per week

    Plenful

    San Francisco, CA
    1 day ago
  • $117k - $209.33k

     ...Job Requisition ID # 26WD99273 Position Overview Want to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products. As part of a... 
    Senior
    For contractors

    Autodesk

    San Francisco, CA
    2 days ago
  • $164k - $205k

     ...ensuring high availability and performance Design intelligent alerting and observability systems Collaborate with engineering teams to embed reliability into the development lifecycle, shifting left on operational concerns Automate incident response workflows and... 
    Senior
    Work experience placement
    Summer holiday
    Live out
    Work at office
    Local area
    Flexible hours
    Shift work
    2 days per week

    SupportFinity

    San Francisco, CA
    8 hours ago
  • $160k - $250k

     ...DevOps And Systems Engineer Hive is the leading provider of cloud-based AI solutions to understand, search, and generate content...  ...machine learning models, we also need to grow our DevOps and Site Reliability team to maintain the reliability of our enterprise SaaS offering... 
    Senior

    Hive

    San Francisco, CA
    1 day ago
  •  ...Airbyte Infrastructure And Reliability Engineer Airbyte is the data and action layer for AI agents. We give agents fast, accurate, authenticated access to business data across hundreds of sources, so they can discover the entities that matter, reason over real-time... 
    Senior
    Work at office
    Local area
    Flexible hours

    Airbyte

    San Francisco, CA
    2 days ago
  • $185.5k - $232k

     ...Senior Site Reliability Engineer New York, NY; Boston, MA; San Francisco, CA About Formation Bio Formation Bio is a tech and AI driven pharma company differentiated by radically more efficient drug development. Advancements in AI and drug discovery are creating... 
    Senior
    Work experience placement
    Work at office
    Local area
    Relocation
    3 days per week

    Formation Bio (Formerly TrailSpark)

    San Francisco, CA
    1 day ago
  • $181k - $225k

     ...Senior Site Reliability Engineer Los Angeles, CA Altruist is transforming the multi-trillion dollar wealth management industry by building an AI platform for wealth professionals. We partner with financial advisors nationwide, empowering them to grow, optimize time... 
    Senior
    Work at office
    Immediate start
    3 days per week

    Altruist

    San Francisco, CA
    4 days ago
  • $189k - $283.6k

     ...the SRE team, you will proactively and reactively improve the reliability of Block's platform and critical infrastructure. You are metrics...  ...accountability ~ A strong desire to perform and grow as an engineer ~5+ years of software development experience... 
    Senior
    Full time
    Relocation package
    Flexible hours
    Shift work

    Block Inc

    San Francisco, CA
    8 hours ago
  • $81.1k - $187k

     ...Site Reliability Engineer 3 We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving... 
    Senior
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle

    San Francisco, CA
    2 days ago
  • Job Title At U.S. Bank, we're on a journey to do our best. Helping the customers and businesses we serve to make better and smarter financial decisions and enabling the communities we support to grow and succeed. We believe it takes all of us to bring our shared ambition...
    Senior
    Temporary work
    Work experience placement

    Phenom People

    San Francisco, CA
    1 day ago
  • $153k - $191.3k

     ...hardware design, manufacturing, data processing, and software engineering, our office is a truly inspiring mix of experts from a...  ...deployments across operating environments, to guarantee the reliability, scalability, and availability of our services. To do this, you... 
    Senior
    Full time
    Temporary work
    For contractors
    Work at office
    Local area
    Remote work
    Home office
    3 days per week

    Planet Labs PBC

    San Francisco, CA
    14 hours ago
  • $220k - $235k

     ...are seeking a strategic, high‑output Staff/Senior Staff SRE to define the future of our cloud platform and champion engineering excellence across Ironclad. In this role,...  ...leadership and strategic direction for the Site Reliability Engineering team and our broader Cloud... 
    Senior
    Full time
    Work at office

    Ironclad Inc

    San Francisco, CA
    8 hours ago
  • $181k - $263k

     ...and supporting deployments of global products, and providing first line operational support. We are looking for a Senior Staff Site Reliability Engineer who will set the technical direction for reliability engineering across LiveRamp's global infrastructure. This is a... 
    Senior
    Full time
    Work from home
    Worldwide
    Flexible hours
    Night shift

    LiveRamp

    San Francisco, CA
    3 days ago
  •  ...management. We have become a multibillion‑dollar asset manager, and we have ambitious goals for the future. As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to meet our growing needs, and you will leverage... 
    Senior
    Local area

    The Voleon Group

    Berkeley, CA
    1 day ago
  • $232.34k - $290.42k

     ...same: to make access to data as simple and reliable as electricity. With Fivetran, customer...  ..., canonical and ready to query, with no engineering or maintenance required. We're proud...  ...integrate our teams, systems, and career sites. About the Role Fivetran and dbt Labs... 
    Senior
    Full time
    Work at office
    Remote work

    dbt Labs

    Oakland, CA
    3 days ago
  • $250k

     ...across Europe, while now significantly expanding its footprint in the United States. The company is looking for a Senior / Staff Site Reliability Engineer to support and scale large-scale HPC and cloud environments powering GPU-intensive workloads. The role involves... 
    Senior
    Full time
    Remote work
    San Francisco, CA
    more than 2 months ago
  • $300k

     ...thousands of H100s, H200s, and B200s, ready for experimentation, full-scale model training, or inference. As a Platform Engineer/Senior Site Reliability Engineer, you’ll own the reliability, performance, and automation of this GPU-powered infrastructure, ensuring... 
    Senior
    Permanent employment
    San Francisco, CA
    more than 2 months ago
  • $165k - $227k

     ...opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Engineering Opportunity We are looking for an experienced Senior Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly... 
    Senior
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    22 days ago
  • $173k - $230k

     ...employment Visa sponsorship. Role Summary The Principal Site Reliability Engineer applies software engineering and systems engineering...  ...~ Establishes enterprise technical direction, develops senior technical leaders, and demonstrates impact well beyond systems... 
    Hourly pay
    Work at office
    Immediate start
    Visa sponsorship
    Work visa
    Flexible hours

    Early Warning Services

    San Francisco, CA
    2 days ago
  •  ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems...  ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes... 
    Senior

    Saviynt

    San Francisco, CA
    14 days ago
  •  ...Site Reliability Engineer Specter's mission is to help automate the physical world. Today, we build video sensors with state-of-the-art AI agents that answer any question, anywhere in their environments. Our systems can automatically detect and reason about any physical... 
    Remote work

    Specter Services LLC

    San Francisco, CA
    3 days ago
  •  ...access to life-saving treatment. What We Look for in a Great Engineer You have the intensity and technical mastery to own mission...  ...high-velocity feature release while maintaining the highest reliability. DevX Support: Support Developer Experience (DevX) work to... 
    Work at office

    LATENT

    San Francisco, CA
    2 days ago
  •  ...JOB DESCRIPTION Project Outline: We are looking for a Site Reliability Engineer with experience in incident response. In this role, you will help Shipt understand where we can improve stability and reliability. There will be a focus on the intersection of systems... 

    BayOne Solutions

    San Francisco, CA
    2 days ago
  •  ...Zof AI is seeking a Site Reliability Engineer to run the infrastructure that lets fleets of sandboxed agents execute customer code safely and cheaply...  ...constraints they personally own. Engineering · Mid to Senior · Full-time · On-site · San Francisco, CA... 
    Full time

    Zof AI

    San Francisco, CA
    7 hours ago
  •  ...Site Reliability Engineer Job Location: San Francisco, CA or Charlotte, NC. Job Type: Contract Work with local API development squads, platform teams, product owners, scrum masters, and architects. The SRE ensures that both our internally critical and our externally... 
    Contract work
    Local area

    InterSources

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!