Senior Software Engineer, Site Reliability Engineering
$153k - $210kRidgeline
Senior Software Engineer, Site Reliability Engineering
Reno, NV; San Ramon, CA; NYC - Hybrid
Are you passionate about building resilient, highly available cloud platforms that enable engineering teams to move quickly and confidently? Do you enjoy automating complex operational challenges, improving observability, and eliminating manual toil through thoughtful engineering? Are you excited by the opportunity to support mission‑critical production systems while collaborating with talented engineers in a fast‑moving, innovative environment? If so, we invite you to be a part of our innovative team.
As a Site Reliability Engineer, you'll help ensure the reliability, scalability, and operational excellence of Ridgeline's mission‑critical SaaS platform. You'll partner closely with product and platform engineers to improve service reliability, accelerate engineering velocity through automation, and build systems that are easier to operate from day one. Our team of engineers are building with cutting‑edge technologies—like Claude Code and Cursor—in a fast‑moving, creative, progressive work environment. You'll play a key role in advancing our observability, release engineering, incident response, and automation capabilities while contributing measurable improvements to platform stability and developer productivity.
At Ridgeline, how we work matters as much as what we build. Ridgeliners act like owners, choose growth over comfort, and communicate with transparency. We assume positive intent, bias toward action, and bring solutions—not just problems. We celebrate wins, learn from setbacks, and thrive in a resilient, collaborative, high‑performing culture. If this excites you, we'd love to meet you!
You must be work authorized in the United States without the need for employer sponsorship.
Impact you will have
- Improve the reliability, availability, and performance of Ridgeline's mission‑critical production SaaS platform.
- Build automation that measurably increases engineering velocity while reducing operational toil.
- Own and improve production observability through metrics, structured logging, distributed tracing, dashboards, and actionable alerting.
- Design and enhance CI/CD pipelines, deployment automation, progressive delivery strategies, and rollback mechanisms.
- Define and improve Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budget practices to proactively manage reliability.
- Identify capacity constraints and reliability risks before they impact customers.
- Participate in an on‑call rotation, triaging production issues, coordinating incident response, and driving issues to resolution with very infrequent after‑hours support.
- Lead blameless postmortems and implement long‑term improvements that strengthen platform resilience.
- Partner with software engineers on infrastructure design reviews to build highly operable, scalable services.
- Develop Infrastructure as Code solutions using Terraform and AWS best practices.
- Collaborate across a distributed engineering organization while fostering a culture of ownership, transparency, learning, and continuous improvement.
What we look for
- 3–6 years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or a related discipline.
- At least 2 years supporting mission‑critical production SaaS workloads running on AWS.
- Experience operating production systems where uptime, performance, and reliability are business critical.
- Hands‑on experience with AWS services including EC2, ECS or EKS, RDS, S3, IAM, CloudWatch, and managed database or messaging services.
- Strong understanding of observability, including monitoring, alerting, distributed tracing, and production diagnostics.
- Experience designing or significantly improving CI/CD pipelines using tools such as GitHub Actions, CircleCI, Buildkite, or similar platforms.
- Experience with deployment strategies including blue/green, canary, or progressive rollouts.
- Proficiency in Python, Go, Bash, or another scripting language used for automation and tooling.
- Experience implementing Infrastructure as Code using Terraform.
- Comfortable participating in an on‑call rotation and leading incident response with composure.
- Excellent communication skills with the ability to explain technical concepts to both technical and non‑technical stakeholders.
- Demonstrated ability to make measurable improvements to platform reliability, operational efficiency, or developer productivity.
- Strong analytical and troubleshooting skills with a passion for solving complex technical challenges.
- A collaborative mindset with a desire to learn, mentor others, and contribute to a positive engineering culture.
Bonus
- Experience with Kubernetes and Helm.
- Familiarity with chaos engineering or fault injection practices.
- Experience building or contributing to SLO and error budget programs.
- Working knowledge of Kotlin, Node.js, or TypeScript.
- Experience supporting highly distributed cloud‑native applications.
- Bachelor's degree in Computer Science, Information Systems, or a related technical discipline.
About Ridgeline
Ridgeline is the industry cloud platform for investment management. It was founded by visionary tech entrepreneur Dave Duffield (co‑founder of both PeopleSoft and Workday) to apply his successful formula of solving operational business challenges with bold innovation and human connectivity to the unique needs of the investment management industry.
Ridgeline started with a clean sheet of paper and a deep bench of experts bound by a set of core values and motivated to revolutionize an industry underserved by its current tech offerings. We are building a new, modern platform in the public cloud, purpose‑built for the investment management industry and we are prioritizing security, agility, and usability to empower business like never before.
With a growing campus in Reno and offices in New York, Lake Tahoe, and the Bay Area, Ridgeline is proud to have built a fast‑growing, people‑first company that has been recognized by Fast Company as a “Best Workplace for Innovators,” by The Software Report as a “Top 100 Software Company,” and by Forbes as one of “America’s Best Startup Employers.”
Ridgeline is proud to be a community‑minded, discrimination‑free equal opportunity workplace.
Ridgeline processes the information you submit in connection with your application in accordance with the Ridgeline Applicant Privacy Statement. Please review the Ridgeline Applicant Privacy Statement in full to understand our privacy practices and contact us with any questions.
Compensation and Benefits
The cash compensation amount for this role is targeted at $153,000 - $210,000. Final compensation amounts are determined by multiple factors, including candidate location, candidate experience and expertise, and may vary from the amount listed above.
As an employee at Ridgeline, you’ll have many opportunities for advancement in your career and can make a true impact on the product.
In addition to the base salary, 100% of Ridgeline employees can participate in our Company Stock Plan subject to the applicable Stock Option Agreement. We also offer rich benefits that reflect the kind of organization we want to be: one in which our employees feel valued and are inspired to bring their best selves to work. These include unlimited vacation, educational and wellness reimbursements, and $0 cost employee insurance plans. Please check out our Careers page for a more comprehensive overview of our perks and benefits.
#J-18808-Ljbffr$158.5k - $172k
...they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you... ...high-impact position driving continuous reliability, deep system optimization, and automation... ...fast, secure, and friction-free software delivery workflows.Secure and Standardize...SeniorFull timeTemporary workWork at officeFlexible hours3 days per week$139k - $257.55k
...Creative Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning,... ...team blends startup energy with the resources of a large software company.What you'll doThis is a role for engineers who...SeniorFull timeTemporary workLocal areaRemote workWorldwide$141k - $216.6k
...and justice issues with our ecosystem of devices and cloud software. Like our products, we work better together. We connect... ...building a safer, more connected world.Position OverviewAs a Site Reliability Engineer, you'll own the reliability, observability, and...SeniorWork experience placementWork at office$167.7k - $245.2k
...approximately 2 days per week on-site at Cisco offices in either... ...as intended, improving reliability and reducing risks. This... ...and control.As a Senior Site Reliability Engineer (SRE), you will build, operate... ...Code tools• Collaborate with software engineers and customers to...SeniorFull timeTemporary workLocal areaFlexible hours2 days per week$167.7k - $245.2k
...effective.We’re looking for talented engineers with a software or operations background, experienced... ...application development teams to ensure the reliability, performance and security of our... .... Please see the Cisco careers site to discover more benefits and perks....SeniorFull timeTemporary workWork at officeLocal areaFlexible hours1 day per week- ...globally recognized firm, driven by pride in ownership.As a Senior Manager of Site Reliability Engineering at JPMorgan Chase within the Corporate Investment... ..., monitoring, instrumentation, and automation of the software in your area. You act in a blameless, data-driven...SeniorBank staffShift work
- ...Job Description:- Our client is seeking a Senior Site Reliability Engineer (SRE) with 10 15 years of experience to support front-office trading systems in a production environment. This role focuses on troubleshooting complex trading infrastructure, managing observability...Senior
- Kong is seeking a Senior SRE for Managed Gateways to lead reliability engineering for our fastest-growing SaaS product. You will own the end-to-end operational lifecycle, drive incident response, and ensure platform resilience across AWS, GCP, and Azure. You will mentor...Senior
$400k
...in financial markets, the organization combines innovation, engineering excellence, and data-driven insights to support complex trading operations worldwide. This opportunity is for a Senior Site Reliability Engineer to join a high-performance infrastructure...SeniorFull timeWorldwide$150k - $170k
...Senior Site Reliability Engineer – Zip Co Join to apply for the Senior Site Reliability Engineer role at Zip Co At Zip, we build cloud‑native software applications that serve millions of customers and process billions of dollars in payments. We’re looking for...SeniorCasual workWork at officeRemote workFlexible hours$104.9k - $174.7k
...experience in SRE, DevOps, or infrastructure engineering roles We need strong production... ...with application teams to improve reliability, performance, and deployability We automate... ...data management. We are hiring a Senior Site Reliability Engineer to take hands-on ownership...SeniorFull timeWork at officeRemote work$500 per month
...accounts. Our global team is a diverse group of experienced engineers, traders, and brokerage professionals who are working to... ...significant impact, we encourage you to apply. Your Role: As a Site Reliability Engineer at Alpaca, you'll help keep our brokerage platform...SeniorHome office- ...growing its team rapidly, and they are looking for a Senior DevOps Engineer / Site Reliability Engineer who can join. If you’re passionate about... ...minimize errors. Deployment: Use configuration management software to automatically deploy updates and fixes into the...Senior
$160k - $180k
...Socure is seeking a Site Reliability Engineer in New York to enhance our identity trust infrastructure. In this role, you will take full ownership of AWS and Kubernetes platforms, ensuring high reliability and operability. The ideal candidate will possess extensive experience...Senior- ...build safer, more resilient organizations. The Role: As a Senior Site Reliability Engineer (SRE) at Dune Security, you will play a critical role in... ...mechanisms and prevent bot attacks. Establish best practices for software reliability, incident response, and fault tolerance. Lead...SeniorFull timeWork at office
$127k - $249k
...We are looking for an experienced Senior or Staff Engineer for our SRE, InfraSec team, to guide the security of our cloud-based infrastructure. As a Staff SRE, you will be very hands‑on technically while also mentoring a small team of SREs. The InfraSec team collaborates...SeniorLocal areaRemote workFlexible hours$140k - $215k
...Core Platform and Embedded Reliability charters: building the foundational... ...directly with product engineering teams and their leadership to... ...deployment processes.At the Senior Engineer level, your... ...source community; evangelize software engineering best practices,...SeniorFull timeWork experience placementWork at officeLocal area2 days per week3 days per week$156k - $262k
...research copilots, or internal knowledge tools, we're the missing link between LLMs and the real world. The Role: Senior Site Reliability Engineer ~ Managing Kubernetes clusters across multiple environments and regions ~ Owning infrastructure as...SeniorFull timeTemporary workWork at officeImmediate startRemote work- Innowise Group, located in Georgia, is seeking a skilled professional for a role specializing in Cloud technologies and containerization. You'll join a fast-growing team of IT experts, contributing to impactful projects across a variety of domains. The position requires...Senior
- Site Reliability Engineer - Equity Trading PlatformLocation: New York | Practice Area: Capital Markets - Technology & Engineering | Type: PermanentKeep critical equity trading platforms resilient, reliable, and ready for the markets.The RoleWe are seeking a highly motivated...Permanent employmentWork at officeWeekend workAfternoon shift
$123k - $165k
Job Posting Title:Site Reliability Engineer IIReq ID:10143234Job Description:Department/Group OverviewOur engineering fleet is a horizontal set... ..., alerting, and operational workflows.Collaborate with software engineering teams to implement SRE best practices, including...Full time$110k - $120k
...technology.Job DescriptionJob Title: Site Reliability Engineer (SRE) / L3 Support EngineerGetting to... ..., scale, and technology.Kick off your software engineering career on our Quality & Automation... ..., DevOps, Platform Engineering, or a senior production support role.Experience...Ongoing contractFull timeCasual workRemote workFlexible hours$200k - $250k
Hudson River Trading (HRT) is seeking a Senior Site Reliability Engineer focused on storage to join our growing Enterprise SRE team. This team is responsible... ...this stack, and are the principal drivers of growth for software and infrastructure practice within our larger Enterprise...Work at officeLocal areaImmediate start$120k - $200k
...PermContact: Kunal DaveContact Email: ****@*****.*** Reliability Engineer(SRE) ResponsibilitiesGlobal Architecture & Disaster Recovery... ...practices (e.g., Chaos Engineering, resilience testing, automated recovery)SkillsBilingual Mandarin Site Reliability Engineer(SRE)Overseas- ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial & Investment... ...engineering best practices within your teamCollaborates with other software engineers and teams to design, develop, test, and...Shift work
- ...developers on how to make things better. We collectively strive to build and maintain a rapid-feedback platform that enables our engineers to accomplish their own goals instead of creating friction.ResponsibilitiesEKS & Karpenter Management: Manage, upgrade, and autoscale...For contractors
$138.1k - $198.2k
...technology that simply works. The SRE Engineering Enablement Team supports our CI... ...day-to-day work. We support the entire software development lifecycle (SDLC), including... ...engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter...Permanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours$150k - $160k
Front-End & AdTech Site Reliability Engineer (SRE)Haymarket Media, Inc. is seeking a Front-End & AdTech Site Reliability Engineer (SRE) to join... ...Hands-on experience managing Cloudflare (including Workers for senior roles).Comfort with GCP (Cloud Run, Storage, GKE) and...Work at officeLocal area- Sr Site Reliability Engineer (Linux, UNIX, Reliability Engineering, Python, C, C++, Java, DevOps) in New York City C, C++, DevOps Engineer, Java... ..., platform management and capacity planning. • Develop software and systems architectural frameworks and tooling. •...SeniorPermanent employmentFull timeRemote work
$194k - $267k
...If you are too, let's talk.The TeamThe Site Reliability team is dedicated to architecting and... ...that maximize platform reliability and engineering velocity.The ideal candidate is... ...cloud environmentExperience debugging software using gdb, strace, ltrace, tcpdump, Wireshark...Local areaWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Software Engineer, Site Reliability Engineering. Be the first to apply!
- senior robotics software engineer New York, NY
- software system engineer New York, NY
- part time software developer New York, NY
- fall software engineering internship New York, NY
- security software engineer New York, NY
- intel software engineer New York, NY
- software developer fintech New York, NY
- new graduate software engineer New York, NY
- software development engineer aws New York, NY
- information technology software engineer New York, NY


