Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Software Engineer, Site Reliability Engineering

$153k - $210k

Ridge Line Services

Senior Software Engineer, Site Reliability Engineering

Are you passionate about building resilient, highly available cloud platforms that enable engineering teams to move quickly and confidently? Do you enjoy automating complex operational challenges, improving observability, and eliminating manual toil through thoughtful engineering? Are you excited by the opportunity to support mission-critical production systems while collaborating with talented engineers in a fast-moving, innovative environment? If so, we invite you to be a part of our innovative team.

As a Site Reliability Engineer, you'll help ensure the reliability, scalability, and operational excellence of Ridgeline's mission-critical SaaS platform. You'll partner closely with product and platform engineers to improve service reliability, accelerate engineering velocity through automation, and build systems that are easier to operate from day one. Our team of engineers are building with cutting-edge technologies—like Claude Code and Cursor—in a fast-moving, creative, progressive work environment. You'll play a key role in advancing our observability, release engineering, incident response, and automation capabilities while contributing measurable improvements to platform stability and developer productivity.

At Ridgeline, how we work matters as much as what we build. Ridgeliners act like owners, choose growth over comfort, and communicate with transparency. We assume positive intent, bias toward action, and bring solutions—not just problems. We celebrate wins, learn from setbacks, and thrive in a resilient, collaborative, high-performing culture. If this excites you, we'd love to meet you!

You must be work authorized in the United States without the need for employer sponsorship.

The impact you will have
  • Improve the reliability, availability, and performance of Ridgeline's mission-critical production SaaS platform.
  • Build automation that measurably increases engineering velocity while reducing operational toil.
  • Own and improve production observability through metrics, structured logging, distributed tracing, dashboards, and actionable alerting.
  • Design and enhance CI/CD pipelines, deployment automation, progressive delivery strategies, and rollback mechanisms.
  • Define and improve Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budget practices to proactively manage reliability.
  • Identify capacity constraints and reliability risks before they impact customers.
  • Participate in an on-call rotation, triaging production issues, coordinating incident response, and driving issues to resolution with very infrequent after-hours support.
  • Lead blameless postmortems and implement long-term improvements that strengthen platform resilience.
  • Partner with software engineers on infrastructure design reviews to build highly operable, scalable services.
  • Develop Infrastructure as Code solutions using Terraform and AWS best practices.
  • Collaborate across a distributed engineering organization while fostering a culture of ownership, transparency, learning, and continuous improvement.
What we look for
  • 3–6 years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or a related discipline.
  • At least 2 years supporting mission-critical production SaaS workloads running on AWS.
  • Experience operating production systems where uptime, performance, and reliability are business critical.
  • Hands-on experience with AWS services including EC2, ECS or EKS, RDS, S3, IAM, CloudWatch, and managed database or messaging services.
  • Strong understanding of observability, including monitoring, alerting, distributed tracing, and production diagnostics.
  • Experience designing or significantly improving CI/CD pipelines using tools such as GitHub Actions, CircleCI, Buildkite, or similar platforms.
  • Experience with deployment strategies including blue/green, canary, or progressive rollouts.
  • Proficiency in Python, Go, Bash, or another scripting language used for automation and tooling.
  • Experience implementing Infrastructure as Code using Terraform.
  • Comfortable participating in an on-call rotation and leading incident response with composure.
  • Excellent communication skills with the ability to explain technical concepts to both technical and non-technical stakeholders.
  • Demonstrated ability to make measurable improvements to platform reliability, operational efficiency, or developer productivity.
  • Strong analytical and troubleshooting skills with a passion for solving complex technical challenges.
  • A collaborative mindset with a desire to learn, mentor others, and contribute to a positive engineering culture.
Bonus
  • Experience with Kubernetes and Helm.
  • Familiarity with chaos engineering or fault injection practices.
  • Experience building or contributing to SLO and error budget programs.
  • Working knowledge of Kotlin, Node.js, or TypeScript.
  • Experience supporting highly distributed cloud-native applications.
  • Bachelor's degree in Computer Science, Information Systems, or a related technical discipline.

About Ridgeline

Ridgeline is the industry cloud platform for investment management. It was founded by visionary tech entrepreneur Dave Duffield (co-founder of both PeopleSoft and Workday) to apply his successful formula of solving operational business challenges with bold innovation and human connectivity to the unique needs of the investment management industry.

Ridgeline started with a clean sheet of paper and a deep bench of experts bound by a set of core values and motivated to revolutionize an industry underserved by its current tech offerings. We are building a new, modern platform in the public cloud, purpose-built for the investment management industry and we are prioritizing security, agility, and usability to empower business like never before.

With a growing campus in Reno and offices in New York, Lake Tahoe, and the Bay Area, Ridgeline is proud to have built a fast-growing, people-first company that has been recognized by Fast Company as a "Best Workplace for Innovators," by The Software Report as a "Top 100 Software Company," and by Forbes as one of "America's Best Startup Employers."

Ridgeline is proud to be a community-minded, discrimination-free equal opportunity workplace.

Ridgeline processes the information you submit in connection with your application in accordance with the Ridgeline Applicant Privacy Statement. Please review the Ridgeline Applicant Privacy Statement in full to understand our privacy practices and contact us with any questions.

Compensation and Benefits

The cash compensation amount for this role is targeted at $153,000 - $210,000. Final compensation amounts are determined by multiple factors, including candidate location, candidate experience and expertise, and may vary from the amount listed above.

As an employee at Ridgeline, you'll have many opportunities for advancement in your career and can make a true impact on the product.

In addition to the base salary, 100% of Ridgeline employees can participate in our Company Stock Plan subject to the applicable Stock Option Agreement. We also offer rich benefits that reflect the kind of organization we want to be: one in which our employees feel valued and are inspired to bring their best selves to work. These include unlimited vacation, educational and wellness reimbursements, and $0 cost employee insurance plans. Please check out our Careers page for a more comprehensive overview of our perks and benefits.

Vacancy posted 21 hours ago
Similar jobs that could be interesting for youBased on the Senior Software Engineer, Site Reliability Engineering in New York, NY vacancy
  • $158.5k - $172k

     ...they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you...  ...high-impact position driving continuous reliability, deep system optimization, and automation...  ...fast, secure, and friction-free software delivery workflows.Secure and Standardize... 
    Senior
    Full time
    Work at office
    3 days per week

    GrubHub

    New York, NY
    4 days ago
  • $141k - $216.6k

     ...and justice issues with our ecosystem of devices and cloud software. Like our products, we work better together. We connect...  ...building a safer, more connected world.Position OverviewAs a Site Reliability Engineer, you'll own the reliability, observability, and... 
    Senior
    Work experience placement
    Work at office

    Axon

    New York, NY
    3 days ago
  • $139k - $257.55k

     ...Creative Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning,...  ...team blends startup energy with the resources of a large software company.What you'll doThis is a role for engineers who... 
    Senior
    Full time
    Temporary work
    Local area
    Remote work
    Worldwide

    Adobe Systems

    New York, NY
    1 day ago
  • $165k - $241.4k

     ...effective.We’re looking for talented engineers with a software or operations background, experienced...  ...application development teams to ensure the reliability, performance and security of our...  .... Please see the Cisco careers site to discover more benefits and perks.... 
    Senior
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    1 day per week

    CISCO Systems

    New York, NY
    4 days ago
  • $160k - $180k

     ...Socure is seeking a Site Reliability Engineer in New York to enhance our identity trust infrastructure. In this role, you will take full ownership of AWS and Kubernetes platforms, ensuring high reliability and operability. The ideal candidate will possess extensive experience... 
    Senior

    Socure Inc

    New York, NY
    16 hours ago
  •  ...globally recognized firm, driven by pride in ownership.As a Senior Manager of Site Reliability Engineering at JPMorgan Chase within the Corporate Investment...  ..., monitoring, instrumentation, and automation of the software in your area. You act in a blameless, data-driven... 
    Senior
    Bank staff
    Shift work

    JP Morgan Chase

    New York, NY
    2 days ago
  •  ...Karsun Solutions, LLC is seeking a Site Reliability Manager to lead a multi-disciplinary team responsible for reliability, security, and platform lifecycle across AWS-based services. The role emphasizes collaboration, observability, and continuous improvement in a client... 
    Senior

    Karsun Solutions

    New York, NY
    16 hours ago
  •  ...Karsun Solutions in the DMV area is seeking a Site Reliability Manager to ensure reliability, scalability, and performance of our systems. You will lead a team focusing on Application Reliability, DevSecOps, and Platform Lifecycle Management. The ideal candidate has 1... 
    Senior

    Karsun Solutions

    New York, NY
    16 hours ago
  •  ...We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and...  ...Spaces Direct enables you to build highly available, software-defined storage by pooling local disks (SSDs, NVMe drives,... 
    Senior
    Local area

    2T Consulting

    New York, NY
    1 day ago
  • $104.9k - $174.7k

     ...Management. You can learn more about LexisNexis Risk at the link below, About the Role: We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory... 
    Senior
    Full time
    Work at office
    Local area
    Remote work
    Flexible hours

    LexisNexis Risk Solutions

    New York, NY
    16 hours ago
  •  ...raising the bar. This is the place. The role As a Senior Site Reliability Engineer you'll join the founding SRE team at our new NYC...  ...pressure and preventing repeat incidents Comfortable writing software and building automation to solve reliability problems... 
    Senior
    Work at office

    Legora

    New York, NY
    4 days ago
  • $225k - $325k

     ...to-day Ensure the scalability, reliability, and observability of our systems to maintain...  ...environment. Lead a range of engineering projects, from developing proprietary...  ...to implementing open source and vendor software for orchestration and automation tooling... 
    Senior
    Hourly pay

    D. E. Shaw & Co.

    New York, NY
    2 days ago
  •  ...build safer, more resilient organizations. The Role: As a Senior Site Reliability Engineer (SRE) at Dune Security, you will play a critical role in...  ...mechanisms and prevent bot attacks. Establish best practices for software reliability, incident response, and fault tolerance. Lead... 
    Senior
    Full time
    Work at office

    Dune Security

    New York, NY
    3 days ago
  • $86k - $148k

     ...and make a difference. Position Summary We’re looking for a Senior Engineer to lead complex initiatives and elevate our managed services...  ...incidents, implement monitoring solutions, and improve system reliability. Security-First Mindset: Experienced in aligning engineering... 
    Senior
    Work at office
    Immediate start
    Remote work
    Flexible hours

    GrabJobs

    New York, NY
    4 days ago
  • $140k - $170k

     ...Senior Site Reliability Engineer New York About us @Symphony Secure. Connected. Intelligent. Symphony is an AI-powered communication and...  ...to provide production insight into running and operating software at-scale in a globally distributed and highly available... 
    Senior
    Local area

    Symphony Service Corp

    New York, NY
    1 day ago
  • $115k - $160k

     ...professionals for this role. Embark on a transformative journey as a Senior Site Reliability Engineer - AVP - Credit Trade Floor. At Barclays, our vision is...  ...of preventative maintenance tasks on hardware and software and utilisation of monitoring tools/metrics to identify,... 
    Senior
    Hourly pay
    Work at office

    Barclays

    New York, NY
    2 days ago
  • $153k - $210k

     ...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you passionate about building resilient, highly available cloud platforms that enable engineering teams to move quickly and confidently? Do you enjoy automating... 
    Senior

    Synthesia

    New York, NY
    16 hours ago
  • $220k - $235k

     ...Staff/Senior Staff Site Reliability Engineer Ironclad is the leading AI contracting platform that transforms agreements into assets. Contracts move faster, insights surface instantly, and agents push work forward, all with you in control. Whether you're buying or selling... 
    Senior
    Full time
    Contract work
    Work at office

    Ironclad Inc

    New York, NY
    2 days ago
  • $400k

     ...in financial markets, the organization combines innovation, engineering excellence, and data-driven insights to support complex trading operations worldwide. This opportunity is for a Senior Site Reliability Engineer to join a high-performance infrastructure... 
    Senior
    Permanent employment
    Worldwide
    New York, NY
    a month ago
  • $500 per month

     ...accounts. Our global team is a diverse group of experienced engineers, traders, and brokerage professionals who are working to...  ...significant impact, we encourage you to apply. Your Role: As a Site Reliability Engineer at Alpaca, you'll help keep our brokerage platform... 
    Senior
    Home office

    Alpaca

    New York, NY
    4 days ago
  •  ...growing its team rapidly, and they are looking for a Senior DevOps Engineer / Site Reliability Engineer who can join. If you’re passionate about...  ...minimize errors. Deployment: Use configuration management software to automatically deploy updates and fixes into the... 
    Senior

    The Greene Group

    New York, NY
    2 days ago
  • $123k - $165k

    Job Posting Title:Site Reliability Engineer IIReq ID:10143234Job Description:Department/Group OverviewOur engineering fleet is a horizontal set...  ..., alerting, and operational workflows.Collaborate with software engineering teams to implement SRE best practices, including... 
    Full time

    Hulu

    New York, NY
    5 days ago
  • $138.1k - $198.2k

     ...technology that simply works.  The SRE Engineering Enablement Team supports our CI...  ...day-to-day work. We support the entire software development lifecycle (SDLC), including...  ...engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter... 
    Permanent employment
    Full time
    Temporary work
    Work experience placement
    Local area
    Remote work
    Flexible hours

    CISCO Systems

    New York, NY
    4 days ago
  • $200k - $250k

    Hudson River Trading (HRT) is seeking a Senior Site Reliability Engineer focused on storage to join our growing Enterprise SRE team. This team is responsible...  ...this stack, and are the principal drivers of growth for software and infrastructure practice within our larger Enterprise... 
    Work at office
    Local area
    Immediate start

    Hudson River Trading

    New York, NY
    1 day ago
  • $120k - $200k

     ...PermContact: Kunal DaveContact Email: ****@*****.*** Reliability Engineer(SRE) ResponsibilitiesGlobal Architecture & Disaster Recovery...  ...practices (e.g., Chaos Engineering, resilience testing, automated recovery)SkillsBilingual Mandarin Site Reliability Engineer(SRE)
    Overseas

    Comrise

    New York, NY
    2 days ago
  •  ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial & Investment...  ...engineering best practices within your teamCollaborates with other software engineers and teams to design, develop, test, and... 
    Shift work

    JP Morgan Chase

    New York, NY
    1 day ago
  • $110k - $120k

     ...technology.Job DescriptionJob Title: Site Reliability Engineer (SRE) / L3 Support EngineerGetting to...  ..., scale, and technology.Kick off your software engineering career on our Quality & Automation...  ..., DevOps, Platform Engineering, or a senior production support role.Experience... 
    Ongoing contract
    Full time
    Casual work
    Remote work
    Flexible hours

    SS&C Technologies

    New York, NY
    1 day ago
  • $150k - $160k

    Front-End & AdTech Site Reliability Engineer (SRE)Haymarket Media, Inc. is seeking a Front-End & AdTech Site Reliability Engineer (SRE) to join...  ...Hands-on experience managing Cloudflare (including Workers for senior roles).Comfort with GCP (Cloud Run, Storage, GKE) and... 
    Work at office
    Local area

    Haymarket Media Group

    New York, NY
    5 days ago
  • $131k - $164k

     ...OverviewWe are seeking a highly skilled Staff Site Reliability Engineer with deep technical expertise across...  ...team. This role is a hands-on senior engineering position responsible for designing...  ...who not only want to help build the software company of the future, but who want to... 
    Work at office
    Local area
    Visa sponsorship
    Flexible hours

    Diligent

    New York, NY
    5 days ago
  • $194k - $267k

     ...all in on this mission. If you are too, let's talk.Position Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    New York, NY
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Software Engineer, Site Reliability Engineering. Be the first to apply!