Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer, Site Reliability

$142k - $196.6k

Upstart

About Upstart

At Upstart, we're united by a mission that matters: to radically reduce the cost and complexity of borrowing for all Americans. Every day, we bring creativity, experimentation, and advanced AI to reshape access to credit, helping millions move forward financially with clarity and confidence.

As the leading AI lending marketplace, we partner with banks and credit unions to expand access to affordable credit through technology that's both radically intelligent and deeply human. Our platform runs over one million predictions per borrower using more than 3,000 signals, powering smarter, fairer decisions for millions of customers. But the numbers only hint at the impact. Every idea, every voice, and every contribution moves us closer to a world where credit never stands between people and their financial progress.

We're proudly digital-first, giving most Upstarters the flexibility to do their best work from wherever they thrive, alongside teammates across 80+ cities in the US and Canada. Digital-first doesn't mean distant. We're intentional about in-person connection through team onsites, planning sessions, and moments that spark creativity and trust. And whether you choose to work primarily from home or collaborate in-person from one of our offices in Columbus, Austin, the Bay Area, or New York City, you'll have the support to work in the way that works best for you.

If you're energized by tackling meaningful problems, excited to innovate with purpose, and motivated by work that truly matters, we'd love to hear from you.

The Team

Upstart's Site Reliability Engineering team enables engineers to operate reliable, resilient, and observable production systems at scale. We build the platforms, tooling, automation, and operational practices that help teams understand system health, respond effectively when things go wrong, and continuously improve the reliability of the services they own.

Our goal is to make reliability an integrated part of how software operates at Upstart. We provide shared observability and reliability capabilities, improve incident response and operational readiness, automate recurring operational work, and use system and customer data to identify where reliability investments will have the greatest impact.

SRE partners closely with product engineering, Cloud Platform, Delivery, Developer Platform, Security, and other infrastructure teams. SRE provides shared reliability capabilities and operational practices, while engineering teams remain accountable for the reliability and operation of the services they build.
The Role

As a Software Engineer on the Site Reliability Engineering team, you will build and operate systems that improve the reliability, resiliency, and observability of Upstart's production environment.

You will work across observability platforms, reliability tooling, incident response systems, operational automation, and resiliency capabilities. You will independently deliver well scoped engineering projects, contribute to technical design, and use production data and operational experience to improve systems used across engineering.

We are looking for engineers who are thoughtful and intentional about how AI changes software development and operations. You should be comfortable using AI throughout the engineering lifecycle, including understanding unfamiliar systems, investigating production behavior, planning implementation, accelerating development, validating changes, and automating repetitive work. We expect engineers to continually develop more effective ways of working with increasingly capable AI tools and to apply sound engineering judgment to where they provide the most leverage.
How you'll make an impact
  • Build and improve the tooling, services, and automation that help engineers understand and improve the reliability of production systems
  • Develop shared observability capabilities that make metrics, logs, traces, service health, and customer impact easier to understand and act on
  • Improve incident response and operational readiness through better tooling, automation, standards, and actionable production signals
  • Build resiliency capabilities that help teams identify failure modes, reduce operational risk, and recover effectively from infrastructure or application failures
  • Identify recurring operational toil and reliability problems and replace manual processes with durable software and automation
  • Use AI as an integrated part of software development and operational problem solving, while identifying opportunities for AI enabled capabilities that improve incident investigation, observability, reliability, and engineering efficiency
Minimum Qualifications
  • 3+ years of professional experience in software engineering, site reliability engineering, or a related engineering discipline
  • Strong software development skills in one or more general purpose programming languages such as Python, Go, JavaScript, or TypeScript
  • Experience designing, building, testing, and operating production software, internal tooling, or infrastructure
  • Experience with cloud infrastructure, distributed systems, observability, monitoring, or production operations
  • Experience participating in on call or incident response for production systems
  • Demonstrated ability to independently deliver well scoped engineering projects, navigate technical ambiguity, and collaborate effectively across engineering teams
  • Demonstrated experience using AI assisted development tools across multiple stages of the software engineering lifecycle, with an interest in continually evolving how you use these tools as their capabilities advance
Preferred Qualifications
  • Experience with Kubernetes, AWS, infrastructure as code, and cloud native production environments
  • Experience building internal reliability, observability, incident management, or operational automation tools
  • Experience with observability platforms such as Datadog, Sumo Logic, CloudWatch, or similar technologies
  • Experience with reliability practices such as service level objectives, capacity planning, resiliency testing, disaster recovery, or operational readiness
  • Experience operating distributed applications with complex dependencies and high availability requirements
  • Experience building AI enabled operational workflows, tools, or automation that extend beyond individual code generation
Position location This role is available in the following locations: Remote

Travel requirements As a digital first company, the majority of your work can be accomplished remotely. The majority of our employees can live and work anywhere in the U.S but are encouraged to to still spend high quality time in-person collaborating via regular onsites. The in-person sessions' cadence varies depending on the team and role; most teams meet once or twice per quarter for 2-4 consecutive days at a time.

#LI-REMOTE

#LI-Associate

At Upstart, your base pay is one part of your total compensation package. The anticipated base salary for this position is expected to be within the below range. Your actual base pay will depend on your geographic location-with our "digital first" philosophy, Upstart uses compensation regions that vary depending on location. Individual pay is also determined by job-related skills, experience, and relevant education or training. Your recruiter can share more about the specific salary range for your preferred location during the hiring process.

In addition, Upstart provides employees with target bonuses, equity compensation, and generous benefits packages (including medical, dental, vision, and 401k).

United States | Remote - Anticipated Base Salary Range

$142,000-$196,600 USD

What you'll love

At Upstart, our benefits are designed to support your health, financial well-being, family, and personal growth. Here's what you can expect:
  • Competitive compensation, including base pay, bonus opportunities, and annual equity grants that vest quarterly
  • Retirement benefits to help you plan for the future, including a 401(k) or Group Retirement Savings Plan with a company match of $2 for every $1 contributed, up to $15,000 annually (USD in the US, CAD in Canada)
  • Employee Stock Purchase Plan (ESPP) with discounted stock purchase options for eligible employees (US only)
  • Comprehensive health coverage designed to support you and your family, including medical, dental, vision, and wellness resources for US and supplemental health coverage for Canada.
  • Health Savings Account contributions from Upstart for eligible plans (US only)
  • Income protection benefits, including life insurance and disability coverage for added financial security
  • Paid time off, sick leave, and company holidays, in line with local requirements
  • Paid family and parental leave to support caregiving and major life moments (duration varies by country)
  • Family-centered benefits to support fertility, parenthood, and caregiving needs
  • Employee Assistance Program (EAP) offering mental health support and life-centered resources
  • Financial wellness resources, including access to financial planning tools and a financial concierge service (US Only)
  • Annual wellness allowance to support your physical and emotional well-being and personal development, based on what matters most to you
  • Annual productivity allowance to invest in relevant tools and resources you need to do your best work, no matter where you work from
  • Connection and community through team events, all-company updates, and employee resource groups (ERGs)
  • Onsite perks, including catered lunches and fully stocked micro-kitchens when working from one of our offices in the Bay Area, Austin, Columbus, and New York City
For roles based in Canada, please note that we are not currently able to hire in Quebec.

Upstart is a proud Equal Opportunity Employer. Just as we are dedicated to improving access to affordable credit for all, we are committed to inclusive and fair hiring practices.

If you require reasonable accommodation in completing an application, interviewing, completing any pre-employment testing, or otherwise participating in the employee selection process, please email

_privacy_policy
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Software Engineer, Site Reliability in United States vacancy
  •  ...Starline, and Google Lens. Before that, Clay led the product and design teams for Google Workspace. What you'll do As a Software Engineer on our Site Reliability team at Sierra, you will be responsible for defining and building the foundation of reliability, observability,... 
    Website
    Full time
    Flexible hours

    Sierra

    San Francisco, CA
    1 day ago
  • $170k - $240k

    Senior Software Engineer - Observability and Reliability New York City, NY Senior Software Engineer - Observability and Reliability About the Role We are growing...  ...Practices When you submit a job application on this site, Sigma processes your personal data for the purposes... 
    Website
    Full time
    Work at office
    Flexible hours

    Sigma Computing

    New York, NY
    17 hours ago
  • $120k - $190k

     ...Job Description Job Description Senior Software Engineer, Reliability Remote, US About Nametag Nametag is building the future of secure digital...  ..., Ann Arbor, Denver, New York City, and beyond. Off-sites: We bring the team together once per quarter for in-... 
    Website
    Full time
    Remote work
    Visa sponsorship
    Flexible hours

    Nametag

    Boston, MA
    24 days ago
  •  ...Robot Co-Adaptation is seeking a Robotics Software Engineer on the UT Main Campus to build and...  ...software stack for HERO deployments across sites. You will collaborate with researchers...  ..., develop ROS 2 packages, and ensure reliability, testing, and deployment readiness. #J... 
    Website

    Phase2 Technology

    Austin, TX
    3 days ago
  •  ...and Robot Co-Adaptation seeks a Robotics Software Engineer to build and deploy a common robotics...  ...infrastructure for HERO deployments across multiple sites. You will own software stack architecture, integrate components, and ensure reliability in long-running robot operations. You... 
    Website

    University of Texas

    Austin, TX
    17 hours ago
  • $149.4k - $202k

     ...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing... 
    Website
    Remote work

    Noctua Technology

    Washington DC
    17 hours ago
  • $125k - $160k

     ...actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars. SOFTWARE ENGINEER, SITE RELIABILITY ENGINEER (APPLICATION SOFTWARE) The application software team is the central nervous system of SpaceX. We build... 
    Website
    Temporary work
    Weekend work

    SpaceX

    Hawthorne, CA
    4 days ago
  • $90k - $150k

    Position Purpose: As a Software Engineer II on the Reliability Engineering Tooling team, you will build and own internal tooling that supports Developers...  ...will focus on automating manual processes, assessing site health through meaningful data, and owning the SOPs that... 
    Website
    Remote job
    Work experience placement
    Local area

    Home Depot

    New York, NY
    17 hours ago
  •  ...Owning the reliability of the alerting pipeline, the full-time Staff Software Engineer, Reliability will manage end-to-end alert evaluation and delivery, define SLOs and...  ...fixes Required Qualifications: Experience as a Site Reliability Engineer, DevOps engineer, or in a... 
    Website
    Full time
    Remote work

    Virtual Vocations Inc

    United States
    2 days ago
  •  ...Site Reliability Engineer (SRE) Location: North Little Rock AR (onsite) Duration: Contract Required/Desired Skills: • Strong web development skills with a strong focus in C#/.NET • Someone who currently works in a hybrid skillset of BOTH.Net development AND... 
    Website
    Contract work

    Software Technology Inc

    North Little Rock, AR
    3 days ago
  • $203.5k - $248.5k

     ...create it. Who you are Metropolis is seeking a Staff Software Engineer focused on Reliability to own reliability across the entire Metropolis platform...  ...office-first model, which requires employees to be on-site at least four days a week, fostering organic interactions... 
    Website
    Temporary work
    Work at office
    Local area

    Metropolis

    Washington DC
    9 days ago
  •  ...this role. JOB DESCRIPTION As a Lead Software Engineer at JPMorgan Chase within the Test...  ...problems. Leads initiatives to improve the reliability and stability of the applications and...  ...comprehensive health care coverage, on-site health and wellness centers, a... 
    Website
    Plano, TX
    18 days ago
  •  ...creating category-leading enterprise software that unleashes that power.To make that...  ...to be part of this journey?At UiPath's Site Reliability team, we build the platforms and systems...  ..., powered by AI.This is a software engineering role. You will not be the person who identifies... 
    Website
    Work at office
    Immediate start
    Remote work

    UiPath

    Bellevue, WA
    4 days ago
  • Site Reliability Engineer Company: ContainIQ Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Requirements: Remote, full-time role; prior SRE experience preferred. Role: Individual Contributor Category: Information Technology
    Website
    Full time
    Remote work

    ContainIQ

    United States
    1 day ago
  •  ...Site Reliability Engineer Company: GitLab Work Type: Remote Employment: Full Time Location: CA, US Seniority: Senior Level Technologies: Terraform, Ansible, Kubernetes, Go, Ruby, Jsonnet, Prometheus, ELK, Grafana Requirements: Senior-level SRE with strong Terraform/IaC... 
    Website
    Full time
    Remote work

    GitLab

    United States
    1 day ago
  • $180k - $225k

     ...sending money globally, providing secure, simple, and reliable ways to manage their money, ensuring true peace of...  ...borders.About the Role:Come join us as a Senior Software Developer on Remitly's Site Reliability Engineering (SRE) Team! We work across Remitly's engineering... 
    Website
    Full time
    Work at office
    Worldwide
    Flexible hours

    Remitly

    Seattle, WA
    2 days ago
  •  ...Site Reliability Engineer Company: Quzara Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: Azure, Terraform, Bicep, Ansible, Azure Monitor, Azure Automation, Azure Policy, Azure Site Recovery, TLS/SSL Requirements: 4+ years in SRE... 
    Website
    Full time
    Remote work

    Quzara LLC

    United States
    1 day ago
  •  ...Site Reliability Engineer Company: Milestone Systems Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Golang, Python, Linux, Shell scripting, Kubernetes, Docker, Terraform, CI/CD, GitOps, ArgoCD, Spinnaker, Prometheus, Datadog,... 
    Website
    Full time
    Remote work

    Milestone Systems Inc

    United States
    1 day ago
  •  ...Site Reliability Engineer Company: Crunchafi Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Azure, AKS, Azure Kubernetes Service, Terraform, Bicep, ARM templates, GitHub Actions, Azure DevOps, Kubernetes, Docker, App Insights... 
    Website
    Full time
    Remote work

    Crunchafi

    United States
    1 day ago
  •  ...Senior Site Reliability Engineer Company: CyberArk Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: AWS...  ...exposure preferred; US citizenship required. Education: Bachelors Role: Individual Contributor Category: Software Development... 
    Website
    Full time
    Remote work

    CyberArk

    United States
    1 day ago
  •  ...Job Title: Site Reliability Engineer Location: Dallas TX (HYBRID) Duration :Full Time Job Description: Skill: Site Reliability...  ...normal operations. • Keep abreast of current software and hardware technologies to enhance software development.... 
    Website
    Full time
    Work at office

    Syntricate Technologies

    Dallas, TX
    2 days ago
  • $180k - $270k

     ...critical supplies quickly and reliably. Today, Zipline operates on...  ...complexity scale. We create software that detects issues in live operations...  ..., and distributed site assets maintenance orchestration...  ...maintenance teams, service engineering, and flight operations to diagnose... 
    Website
    Full time
    Local area
    Immediate start

    Zipline

    United States
    more than 2 months ago
  • $230k - $390k

     ...Starline, and Google Lens. Before that, Clay led the product and design teams for Google Workspace. What you'll doAs a Software Engineer on our Site Reliability team at Sierra, you will be responsible for defining and building the foundation of reliability, observability,... 
    Website
    Full time
    Flexible hours

    Sierra

    San Francisco, CA
    4 days ago
  • $110.1k - $204.49k

     ...employees feel respected, valued and have an opportunity to contribute to the company’s success. As a Software Engineering Manager within PNC's Corporate Banking and Site Reliability Engineering Center, you will be based at one of the following IT Hubs: Phoenix, Arizona;... 
    Website
    Full time
    Temporary work
    Part time
    Work experience placement
    Work at office
    Afternoon shift

    PNC

    Phoenix, AZ
    2 days ago
  •  ...Working remotely, the full-time Site Reliability Engineer will apply a software engineering mindset to enhance the reliability and performance of production services on Heroku and AWS, collaborating with the engineering team to build automated, scalable solutions that... 
    Website
    Full time
    Remote work

    Virtual Vocations Inc

    United States
    2 days ago
  • $167.7k - $245.2k

     ...and security — partnering across engineering, security, compliance, and product...  ...global technology leader.As a Senior Software Engineer in Application Reliability, you will own the reliability of...  ...insurance. Please see the Cisco careers site to discover more benefits and... 
    Website
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    2 days ago
  •  ...Senior Site Reliability Engineer (Permanent Role) Cleveland, OH, Pittsburgh, PA, or Dallas, TX Your future duties and responsibilities . Monitoring distribution systems and notifying them of any potential issues. . Assisting with troubleshooting on call... 
    Website
    Permanent employment
    Flexible hours
    Shift work
    Weekend work

    System One Holdings, LLC

    Farmers Branch, TX
    3 days ago
  • $158k - $237k

    About The TeamThe Rubrik Engineering team is comprised of people who produce...  ...are driven to build efficient, reliable, and cost effective products....  ...to talk to you!About The Role:Site Reliability Engineers at Rubrik are systems/software engineers who ensure that Rubrik... 
    Website
    Local area

    Rubrik

    Palo Alto, CA
    15 hours ago
  •  ...Senior Site Reliability Engineer Cleveland, OH, Pittsburgh, PA, or Dallas, TX Your future duties and responsibilities: Monitoring distribution systems and notifying them of any potential issues. Assisting with troubleshooting on call. Managing and tracking... 
    Website
    Flexible hours
    Shift work
    Weekend work

    System One Holdings, LLC

    Cleveland, OH
    17 hours ago
  •  ...Manager, Site Reliability Engineering Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work... 
    Website

    Okta, Inc.

    Washington DC
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer, Site Reliability. Be the first to apply!