Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior SRE

$145k - $170k

Banyan Software

Senior SRE

Banyan Software is the best permanent home for software businesses that serve specialized industries, their employees, and their customers. With a buy-grow-and-hold-for-life approach and a permanent capital base, Banyan acquires and grows companies worldwide, honoring founder legacies and helping portfolio companies modernize through shared AI expertise and operational discipline. Founded in 2016, Banyan operates more than 120 portfolio companies across North and South America, Europe, and APAC, and has appeared on the Inc. 5000 list for six consecutive years. The Banyan Software Foundation, endowed with $100 million in Banyan stock, leverages technology to build a greener and more equitable world.

Job Title: Senior SRE (Site Reliability Engineer) – Modernized Application Operations

Remote: US/Canada

Overview

We are seeking a highly experienced and hands-on SRE to own the operational excellence of the modernized SaaS applications produced by the Banyan AI Factory. This is not a role focused on building the factory itself; instead, you will run the reliability of the modernized applications the factory delivers to our Operating Companies (OpCos).

You will join a team that provides 24x7 coverage with rotating on-call responsibilities, serving as Tier 1 Site Reliability Engineering (SRE) for our OpCos' distributed applications. Day to day this will include: automated deployments, cloud service integration, application performance and availability monitoring/observability, and security incident response across our two target clouds — Amazon Web Services (AWS) and Microsoft Azure. The ideal candidate has a track record of keeping secure, highly available production systems running at scale.

Key Responsibilities

  • 24x7 Operations & On-Call: Operate as part of a team providing round-the-clock coverage of OpCo containerized applications, participating in a rotating on-call schedule to ensure continuous availability and rapid response.
  • Tier 1 SRE & Operations: Serve as Tier 1 SRE for the modernized applications, managing day-to-day cloud integrations across our two target clouds — AWS and Azure — to keep production systems healthy, performant, and secure.
  • Performance & Availability Monitoring/Observability: Implement and maintain robust application observability tooling (monitoring, logging, tracing) to track performance and availability, proactively detect degradation, and drive down mean-time-to-detect and mean-time-to-resolve.
  • Disaster Recovery and Service Restoration: Develop, maintain, test, and execute disaster recovery and business continuity procedures. Ensure the timely recovery and restoration of services following geographic disruptions, cyber incidents, infrastructure failures, or other disaster events.
  • Security Incident Response: Respond to security incidents and operational events affecting OpCo SaaS platforms, executing established runbooks, coordinating remediation
  • Automation & Infrastructure-as-Code : Use Infrastructure-as-Code (Terraform) and CI/CD pipelines (e.g., GitHub Actions, GitLab CI) to manage, deploy, and automate the operational environments of modernized applications, reducing toil and improving consistency.
  • AI Agents & DevSecOps Scale: Build scale in our DevSecOps practice by designing, building, and operating AI agents that automate SRE tasks and incident response, reducing toil and accelerating detection, triage, and remediation.
  • Hands-on Problem Solving: Serve as a technical escalation point for operational challenges, applying strong analytical skills to resolve infrastructure, network, and automation issues across distributed, multi-tenant SaaS environments while navigating technical ambiguity.

Required Qualifications & Experience

  • Experience: 5–7 years of progressive experience in Software Engineering, and/or Site Reliability Engineering, with a focus on operating distributed systems.
  • Automation Coding Experience: Deep expertise in Python, Javascript, or Go. Building automation and integrations between tools. This may be with AI assistance, but you must have a deep understanding of the code and scripting principals such as: authentication, parallelization, triggering, APIs, data transformation, etc.
  • Containerization: Deep expertise in container technologies (Docker/Kubernetes) supporting highly scalable and resilient distributed systems.
  • Infrastructure-as-Code with Terraform: Have experience working with modules at scale. This is a requirement for the role.
  • Cloud Native Services: hands-on experience operating production workloads on Amazon Web Services (AWS) (e.g., EC2, Lambda, EKS, S3, RDS) and / or Microsoft Azure (e.g., Container Apps, AKS, Container Storage).
  • CI/CD & Automation: Deep history of hands-on work with CI/CD platforms (GitHub Actions, GitLab CI) and embedding DevSecOps practices directly into operational workflows.
  • Operations, Monitoring & Observability: Experience with application level logging, troubleshooting, and tracing tools, with a proven track record operating highly available production systems.
  • AI-Fluent Engineering: Experience with AI-assisted engineering tools such as Claude Code or similar
  • Application Performance Management (APM): Familiarity with APM tooling and practices (e.g., Datadog, New Relic, Dynatrace, or similar) to instrument, profile, and optimize application performance in production.
  • Incident & Security Response: Demonstrated experience participating in on-call rotations, responding to production and security incidents, and executing disaster recovery procedures.
  • Communication & Collaboration: Exceptional communication, presentation, and collaboration skills, with a proven ability to coordinate across teams.
  • Education: Bachelor's degree in Computer Science or a related technical field.

Preferred Skills (A Plus)

Familiarity with advanced cloud security tools like Wiz, Prisma Cloud, and Checkov.

The expected base salary for this position is approximately USD $145,000 - $170,000 for US-based candidates and CAD $120,000 - $145,000 for Canada-based candidates, excluding annual bonus and equity (when applicable). Salary is based on a number of factors, including market conditions, location, job-related skills and experience, and may vary accordingly.

Diversity, Equity, Inclusion & Equal Employment Opportunity at Banyan: Banyan affirms that inequality is detrimental to our Global Teams, associates, our Operating Companies, and the communities we serve. As a collective, our goal is to impact lasting change through our actions. Together, we unite for equality and equity. Banyan is committed to equal employment opportunities regardless of any protected characteristic, including race, color, genetic information, creed, national origin, religion, sex, affectional or sexual orientation, gender identity or expression, lawful alien status, ancestry, age, marital status, or protected veteran status and will not discriminate against anyone on the basis of a disability. We support an inclusive workplace where associates excel based on personal merit, qualifications, experience, ability, and job performance.

Please Note : Banyan Software does not accept unsolicited resumes or applications submitted via email, LinkedIn, or other direct channels. All candidates must apply through our official Careers site to be considered for employment. Applications submitted outside of our applicant tracking system will not be reviewed.

Recruitment Notice Banyan Software may use artificial intelligence (AI) tools to assist in screening and/or assessing applicants during the recruitment process. All hiring decisions are made by our team. Personal information submitted through your application will be collected and used for recruitment purposes in accordance with applicable privacy laws. Contact us at any time with questions about our process or to request accommodation.

Beware of Recruitment Scams

We have been made aware of individuals fraudulently posing as members of our Talent Acquisition team and extending fake job offers. These scams may involve requests for personal information or payment for equipment.

Protect yourself by following these steps:

  • Verify that all communications from our recruiting team come from an @banyansoftware.com email address.
  • Remember, employers will never request payment or banking information during the hiring process.
  • If you receive a suspicious message, do not respond — instead, forward it to View email address on click.appcast.io and/or report it to the platform where you received it.

Your safety and security are important to us. Thank you for staying vigilant.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Senior SRE in United States vacancy
  •  ...Palo Alto Networks is seeking a Senior Principal Engineer/Architect to serve as the technical authority for global SRE and Platform Engineering initiatives across the US and India. You will architect AI-driven, self-healing platform capabilities and partner with product... 
    Senior

    Jobleads-US

    Santa Clara, CA
    3 days ago
  • $129k - $231k

     ...is in active transition from human-executed operations to engineering-led, agent-assisted reliability. Role Summary: As the Senior Principal SRE Engineering Lead, you are the senior-most engineering authority for the reliability of the supported production estate. You... 
    Senior
    Contract work
    Flexible hours
    Shift work

    Eli Lilly and Company

    Indianapolis, IN
    2 days ago
  •  ...Wonder is hiring for a Staff-level SRE/DevOps role in New York to own a production platform end-to-end. You will architect resilient, scalable services, manage multi-region AWS deployments, and drive reliability through SLOs, tracing and cost-aware observability. You... 
    Senior

    Jobleads-US

    New York, NY
    3 hours ago
  •  ...Job Title: Senior Site Reliability Engineer (SRE) Work Location: Southlake, TX 76092 Contract duration: 12 months Interview Mode- In-person Interview Job Details: Must Have Skills:: SRE Grafana Python Nice to have skills AI Cloud... 
    Senior
    Contract work

    eTeam

    Southlake, TX
    4 days ago
  •  ...Senior Site Reliability Engineer (SRE) Our client is a global technology consulting and digital solutions company that enables enterprises across industries to reimagine business models, accelerate innovation, and maximize growth by harnessing digital technologies.... 
    Senior
    Local area

    E-Solutions

    Los Angeles, CA
    2 days ago
  • $150k - $250k

     ...Job Description: Senior Site Reliability Engineer New York, NY Seeking a Senior Site Reliability Engineer to design, build, and...  ...across on-prem and cloud environments. This role will help establish SRE best practices and ensure the stability, scalability, and... 
    Senior

    Berkeley Square IT

    New York, NY
    2 days ago
  •  ...reliable, high-performing, scalable software to our global customer base. The ideal candidate would be excited to be an early member of the SRE team. We are looking for someone who is passionate about heavily influencing our SRE roadmap and excited to roll up their sleeves and... 
    Senior

    1872 Consulting

    Westwood, MA
    3 days ago
  • Senior Director - Observability | SRE About the RoleThe Senior Director - Observability and SRE, is a strategic leader accountable for ensuring the reliability, availability, and performance of the enterprise technology ecosystem. This role oversees Observability, Site... 
    Senior

    GAP Inc.

    Coppell, TX
    a month ago
  •  ...Verisk is seeking a Principal DevOps Engineer/SRE for its Boston hybrid site to lead design of a scalable, self‑service DevOps platform. You will own pipelines, infrastructure as code, and security standards while partnering with engineering teams to drive reliability.... 
    Senior

    Jobleads-US

    Boston, MA
    14 hours ago
  • Kong Inc. is seeking a Senior Site Reliability Engineer for Managed Gateways in Washington. You will own production reliability for a cloud...  ...spanning AWS, GCP, and Azure, leading a high-performing SRE team and shaping the enterprise deployment experience. You will... 
    Senior

    Cacheflow

    Seattle, WA
    1 day ago
  •  ...Senior Operations Analyst (SRE) CGI's Advantage Cloud Operations is an SRE-driven operating model, anchored on an Operations Control Plane that unifies telemetry, event management, automation, and IT Service Management (ITSM). The Senior Operations Analyst is a senior... 
    Senior
    Work at office

    CGI

    Fairfax, VA
    3 days ago
  •  ...Role: Senior Site Reliability Engineer (SRE) Cloud & Kubernetes Location: Atlanta, GA (Onsite) Contract Role Summary: Lead the reliability, scalability, security, and operational excellence of customer-facing platforms across Azure, GCP, and Kubernetes... 
    Senior
    Contract work

    Noblesoft Technologies

    Atlanta, GA
    2 days ago
  • $185k - $200k

    Back to All JobsSenior Site Reliability Engineer (SRE) Dayton, OH (Remote) full time Top Secret (TS) $185,000 - $200,000 Job Description Position Overview Metronome is seeking a Senior Site Reliability Engineer (SRE) to support AFRL's Google Cloud Platform environment... 
    Senior
    Full time
    Remote work

    Metronome LLC

    Dayton, OH
    1 day ago
  • Oracle seeks a senior Oracle database engineer to own end-to-end management of the database stack, emphasizing security, resiliency, scale, and performance. You will collaborate with SRE and development teams to define and deliver premier capabilities across on-premises... 
    Senior
    Shift work
    Weekend work

    Ll Oefentherapie

    Seattle, WA
    1 day ago
  • $152k - $241.5k

    NVIDIA DGX Cloud builds and operates large-scale GPU infrastructure for AI workloads. We are looking for Software Engineers with SRE or Production Engineering experience who have worked hands-on with bare-metal NVIDIA systems. This team builds the software and operational... 
    Senior
    Permanent employment
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  •  ...— 5,000+ transactions before go-live — through auto-remediation, capacity planning, and actionable dashboards. Build specialized SRE agents using Cursor AI • Design and ship AI agents for incident triage, log analysis, and root-cause investigation (to name a few... 
    Senior
    Remote work

    Accelerant Inc

    United States
    1 day ago
  •  ...the scripts, hooks, and automation that make them stick. Develop the guardrails and developer-facing tooling that reduce toil for SRE and product engineering alike, measured by build times, merge friction, and interrupt volume. Set the technical bar through code reviews... 
    Senior

    Zoox

    Foster, CA
    15 days ago
  • Tatari seeks a Senior Site Reliability Engineer to design, scale, and automate AWS infrastructure. Responsibilities managing AWS infrastructure...  ...AI developer tools for automation Requirements 4-6 years in SRE, DevOps, or systems engineering 3+ years with AWS and Kubernetes... 
    Senior
    Remote work

    Tatari

    Poland, NY
    2 days ago
  •  ...who will join our team to help move us forward and achieve our mission. The RoleAs a Software Engineering - Site Reliability Engineer (SRE) on the Connectivity team, you will drive the reliability, scalability, and performance of critical software systems while solving... 
    Senior
    Full time
    Work experience placement
    Local area
    Work from home
    Relocation
    Relocation package

    General Motors

    Warren, MI
    3 days ago
  • $61k - $101k

     ...reliability engineering domains. Preferred knowledge includes SRE concepts such as SLOs/SLIs, error budgets, incident management,...  ...hackajob and J.P. Morgan to find exceptional professionals for this Senior Lead Software Engineer role within our AI/ML Data Platforms... 
    Senior
    Full time

    J.P. Morgan

    Washington DC
    9 days ago
  •  ...Senior Software Engineer (SRE) Location: Onsite in Miami · Reports to : Engineering Lead· Department: Engineering  About eMed eMed is a leading digital health company specializing in cardio metabolic health through managed GLP-1 programs. Focused on delivering... 
    Senior
    Full time
    Temporary work

    eMed, LLC

    Miami, FL
    21 days ago
  • $175k - $229k

     ...must have both the best technology and the best access to that technology to win. Requirements: ~5 or more years of DevOps or SRE experience deploying and operating commercial SaaS platforms on public cloud infrastructure, AWS preferred. ~ Expert knowledge with... 
    Senior

    Instrumental Inc

    Palo Alto, CA
    4 days ago
  • $167k - $196.5k

     ...engineering operational documentation for supported products. Build and maintain product operational documentation and establish SRE practices. Optimize performance and cost of systems, right‑size Kubernetes containers. Collaborate with SRE and engineering teams... 
    Senior
    Remote work
    Flexible hours

    C0035 LiveRamp, Inc.

    San Francisco, CA
    3 days ago
  • $170k - $277k

     ...drives great outcomes. Job Summary Your Career We are looking for a visionary Senior Principal Engineer/Architect to serve as the technical authority for our global SRE and Platform Engineering initiatives across the US and India. This isn't just a "keep the... 
    Senior
    Full time
    Work at office
    Visa sponsorship
    Work visa
    Flexible hours

    Palo Alto Networks, Inc.

    Santa Clara, CA
    3 days ago
  • $149.4k - $202k

     ...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing... 
    Senior
    Remote work

    Noctua Technology

    Washington DC
    5 days ago
  •  ...collaboratively in teams and build meaningful relationships to achieve common goals. Preferred Qualifications ~10+ years in an SRE or production support role with AWS Cloud, Databricks, Snowflake or similar Technologies. ~ AWS, Snowflake or Databricks certifications... 
    Senior

    Next Frontier Capital

    Jersey City, NJ
    2 days ago
  •  ...Senior Site Reliability Engineer At Swile, we believe that good products can help reduce friction in daily professional life and boost...  ...and Brazil. Your role as a Senior Site Reliability Engineer (SRE) centers around creatively solving problems, ensuring a balance... 
    Senior
    Remote work

    Swile

    United States
    1 day ago
  •  ...managing and operating applications — including AI/ML-driven approaches to observability and reliability. What you’ll do: • Evangelize SRE mindset and solve problems through systematization. • Identify opportunities to build innovative tools and solve unique operations... 
    Senior

    Mindlance

    Austin, TX
    3 days ago
  •  ...TENEX.AI is seeking a Principal DevOps Engineer to own the Azure platform, CI/CD pipelines, and SRE practices at scale, ensuring high availability and security for petabytes of security data. You will collaborate with Software, AI/ML, and Security Operations to drive automation... 

    Jobleads-US

    Overland Park, KS
    4 days ago
  • $70k - $150k

     ...Address:4100 Gordon Baker RoadJob Family Group:TechnologyWe are seeking a highly motivated and technically strong Senior Lead, Site Reliability Engineering (SRE) to provide technical leadership within the DCOE Reliability Engineering organization. This role is critical to... 
    Senior
    Full time
    Contract work
    Part time

    BMO Bank

    Toronto, OH
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior SRE. Be the first to apply!