Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

$120k - $140k

Best Egg

Best Egg is a market-leading, tech-enabled financial platform helping people build financial confidence through a variety of installment lending solutions and financial health tools. We aim to help customers make smart financial decisions and stay on track, so they can be money confident no matter what life throws at them. We offer top-tier benefits and growth opportunities in a culture built on our core values: Put People First – We foster an inclusive, flexible, and fun workplace. Create Clarity – Open communication drives trust and results. Get Things Done – We focus, prioritize, and deliver with excellence. Deliver with Heart – We lead with kindness, humility, and strong teamwork. Listen to Our Customers – Their needs drive our innovation. Barclays has entered into an agreement to acquire Best Egg with closing expected to take place in Q2 2026. This acquisition will give us the resources and capital to continue on our mission and drive our strategy forward. With an aligned culture, lower cost of funds, and increased employee growth opportunities across a global brand, we are excited about the future of the Best Egg brand under the Barclays umbrella. We are looking for collaborative, innovative team players who like to solve problems. There will also be immense opportunities for those willing to dive in. If you're inspired by growth and want to make a real difference, Best Egg is the place for you. We’re proud to be an equal opportunity employer committed to building a diverse, inclusive team. The Job As Site Reliability Engineer, you will serve as a reliability subject matter expert who leads major incident recovery, drives observability and reliability improvements, mentors associate engineers, reduces operational toil, influences technical decisions, and improves resiliency standards. This role requires depth across production systems, telemetry, batch operations, automation, and incident response. You will be expected to guide technical direction for reliability improvements and help teams prevent recurring failures. Employees joining Best Egg's Information Technology organization can expect a culture centered on Continuous Delivery, Total Quality Management, Knowledge Sharing, Personal and Career Advancement, Empowerment, Innovation, and Collective Ownership. Duties & Responsibilities Lead technical recovery efforts for major incidents, coordinating triage, evidence review, restoration actions, and validation. Optimize observability strategy, alert quality, dashboard standards, and telemetry coverage across multiple services. Drive reliability initiatives that reduce recurring failures, noisy alerts, manual work, and operational risk. Mentor associate engineers on troubleshooting methods, RCA evidence, runbook quality, and production support judgment. Influence engineering decisions by identifying reliability risks, missing telemetry, supportability gaps, and resiliency patterns. Improve JAMS, GoAnywhere, Datadog, xMatters, and service support practices through automation and standards. Partner with leaders and technical teams to prioritize remediations based on customer impact, business impact, and operational exposure. Required Hands‑on familiarity with production support, monitoring, alerting, and incident response practices. Working knowledge of Datadog dashboards, monitors, logs, metrics, and APM concepts. Ability to troubleshoot application, infrastructure, batch, or file transfer issues using runbooks and telemetry. Exposure to AWS or cloud operations and scripting with Python, PowerShell, Bash, or similar tools. Clear communication skills during incidents, service requests, and post‑incident follow‑through. Strong experience leading production incident recovery and cross‑system reliability investigations. Ability to mentor engineers and influence technical decisions without direct authority. Recommended Datadog, AWS, ITIL, Linux, or automation certification. Experience with JAMS, GoAnywhere, xMatters, ServiceNow/Jira, or CI/CD environments. Exposure to AIOps, anomaly detection, operational automation, or reliability engineering. Familiarity with financial services controls, secure file transfer, or regulated operations. Skill Set Serves as a reliability SME across observability, incident response, batch operations, and operational platforms. Leads major incident recovery with calm command of telemetry, dependencies, impact, and restoration options. Reduces operational toil through automation, standards, better alerting, and durable remediation. Mentors engineers and improves the quality of technical support practices across the team. Influences design and readiness decisions that improve resiliency and operational resilience. Success Metrics in First 90 Days Lead a major incident or complex reliability investigation with clear recovery and follow‑through. Deliver an observability or automation improvement that measurably reduces alert noise, toil, or repeat issues. Mentor associate engineers through troubleshooting reviews, runbook improvements, or incident debriefs. Identify and influence remediation of a meaningful resiliency or supportability gap. Improve standards or patterns for telemetry, escalation, batch support, or operational validation. $120,000 - $140,000 a year Please note this job description is not designed to cover every activity, duty, or responsibility required for the job. Duties, responsibilities, and activities may change at any time with or without notice. Best Egg celebrates diversity and equal opportunity. We are committed to building a team that represents a variety of backgrounds, perspectives, and skills. The more inclusive we are, the better we will grow. Employee Benefits Pre‑tax and post‑tax retirement savings plans with a competitive company matching. Generous paid time‑off plans including vacation, personal/sick time, paid short‑term, long‑term disability leaves, paid parental leave, and paid company holidays. Multiple health care plans to choose from, including dental and vision options. Flexible Spending Plans for Health Care, Dependent Care, and Health Reimbursement Accounts. Company‑paid benefits such as life insurance, wellness platforms, employee assistance programs, and Health Advocate programs. Other great discounted benefits including identity theft protection, pet insurance, fitness center reimbursements, and many more. #J-18808-Ljbffr

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in New York, NY vacancy
  •  ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Commercial & Investment Bank, Production management team, you will solve complex and broad... 
    Suggested
    Shift work

    J.P. Morgan

    New York, NY
    7 days ago
  •  ...and shape the future of technology at a globally recognized firm, driven by pride in ownership. As a Senior Manager of Site Reliability Engineering at JPMorgan Chase within the Corporate Investment Bank, Markets team, you are the non-functional requirement owner and... 
    Suggested
    Bank staff
    Shift work

    J.P. Morgan

    New York, NY
    18 days ago
  •  ...paced, regulated financial environment. • Excellent communication and collaboration skills. Role Overview The System Engineer will be responsible for designing, implementing, and maintaining enterprise-level infrastructure solutions across Linux... 
    Suggested

    Q1 Technologies

    New York, NY
    4 days ago
  • $176.75k - $209.1k

     ...development and learning. It allows us to scale easily, enabling our engineers to maximize attention on new features and capabilities. A...  ...of customers all over the world. Peloton is looking for a Site Reliability Engineer with an operations focus to work with teams across... 
    Suggested
    Temporary work
    Local area

    Peloton

    New York, NY
    4 days ago
  • $400k

     ...Salary : Up to $400,000 Total Compensation Senior Site Reliability Engineer We are working with a leading trading technology firm building a high-performance infrastructure engineering team focused on reliability, automation, and large-scale platform resilience. This is... 
    Suggested

    Hamilton Barnes ?

    New York, NY
    2 days ago
  • $120k - $160k

     ...and benefits packages, technology talks by our experts, a beautiful modern office, daily catered lunches, and more. As a Site Reliability Engineer (SRE), you will work at the intersection of production operations and software development as you improve, manage, and monitor... 
    Work at office
    Local area

    The Voleon Group

    New York, NY
    2 days ago
  • $150k - $170k

     ...Senior Site Reliability Engineer – Zip Co Join to apply for the Senior Site Reliability Engineer role at Zip Co At Zip, we build cloud‑native software applications that serve millions of customers and process billions of dollars in payments. We’re looking for a seasoned... 
    Casual work
    Work at office
    Remote work
    Flexible hours

    ZIP

    New York, NY
    22 hours ago
  • $111k - $160k

     ...Join Mizuho as a Site Reliability Engineer! In this role you will play a crucial role in maintaining the reliability, scalability, and overall performance of our production systems. This position collaborates closely with development, operations, and product teams to automate... 
    Work at office
    Local area
    Remote work

    Mizuho

    New York, NY
    2 days ago
  •  ...Responsibilities Improve the reliability of mission-critical solutions, applications, and platforms Software development for enterprises...  ...Windows and Linux Years of Experience: 5 Years of Software Engineering Seniority level Mid-Senior level Employment type Full-time Job... 
    Full time
    Work experience placement

    InterEx Group

    New York, NY
    2 days ago
  • $89k - $178k

     ...industry. Learn more at What You’ll Do Build and maintain the reliability, scalability, and performance of our digital media...  ...prevent recurrence Required Experience & Skills 4+ years in Site Reliability Engineering, DevOps, or related operational roles with proven... 

    DoubleVerify

    New York, NY
    2 days ago
  •  ...Site Reliability Engineer Discover your future at Citi. Working at Citi is far more than just a job. A career with us means joining a team of more than 230,000 dedicated people from around the globe. At Citi, you’ll have the opportunity to grow your career, give back to... 

    Citi

    New York, NY
    22 hours ago
  •  ...public sector—co‑creating customized AI systems that they can run on their terms. The Role We are seeking highly experienced Site Reliability Engineers (SRE) to shape the reliability, scalability and performance of our platform and customer‑facing applications. You will... 
    Relocation package

    Mistral

    New York, NY
    2 days ago
  •  ...configure the monitoring and alerting metrics so the support engineers can proactively and timely validate, troubleshoot and...  ...monitoring high availability critical application compon ents.1+ Years in Site Reliability Engineering organization prefe #J-18808-Ljbffr... 
    Work experience placement

    PineQ Lab Technology

    New York, NY
    2 days ago
  • $164k - $205k

     ...ensuring high availability and performance Design intelligent alerting and observability systems Collaborate with engineering teams to embed reliability into the development lifecycle, shifting left on operational concerns Automate incident response workflows and build... 
    Work experience placement
    Summer holiday
    Work at office
    Local area
    Flexible hours
    Shift work
    2 days per week

    BetterUp

    New York, NY
    2 days ago
  • $111k - $218k

     ...The Site Reliability Engineering team designs and builds the global infrastructure on which we deploy our services, focusing on the flagship MongoDB Atlas platform. As our customers grow and globalize, our services must satisfy demands for low-latency requests around... 
    Local area
    Flexible hours

    MongoDB

    New York, NY
    2 days ago
  • $200k

     ...Role: Site Reliability Engineer (Capital Markets) Location: New York City Salary: Up to $200k No capital markets experience needed, you will be part of a small, fast-paced global team responsible for the deployment, maintenance, and continuous enhancement of a sophisticated... 
    Worldwide

    Autonomai Recruitment

    New York, NY
    2 days ago
  •  ...Site Reliability Engineer, Commodities Technology What you’ll do Ensure high availability and uptime of Commodities Technology services and applications Automate and streamline manual processes Contribute to root cause analysis and post-mortem reports for production incidents... 
    Work experience placement

    Triangle Workforce

    New York, NY
    2 days ago
  • $123k - $165k

     ...Department/Group Overview Our engineering fleet is a horizontal set of teams providing engineering...  .... Our specific team provides reliability engineering and operational support to backend...  ...and brands. We are seeking a Site Reliability Engineer who will contribute... 

    The Walt Disney Company

    New York, NY
    14 hours ago
  •  ...Site Reliability Engineer I, Abhishek, would like to share a job opportunity as Site Reliability Engineer in Jacksonville, FL, Cary, NC or New York, NY (Onsite) location for a Fulltime position. In case, if you are not comfortable with this location, please share your... 
    Full time
    Work visa

    Syntricate Technologies

    New York, NY
    2 days ago
  •  ...Senior Site Reliability Engineer (SRE / Infrastructure) Role Overview We’re hiring a Senior SRE to build and scale the infrastructure behind a high-growth, production system. You’ll ensure reliability, performance, and scalability as the platform grows from early traction... 

    The Cypress Group

    New York, NY
    2 days ago
  •  ...A leading quantitative trading firm is seeking a Senior Site Reliability Engineer to build and evolve the reliability, observability, and automation capabilities powering a highly performance-sensitive trading environment. Working at the intersection of software and infrastructure... 

    Acquire Me

    New York, NY
    2 days ago
  •  ...firm that has been operating at scale since 2014, with millions of global users and a reputation for rigorous engineering. The Role As Senior Site Reliability Engineer , you will own the infrastructure foundation that the entire engineering organization depends on. This... 

    Harrison Clarke

    New York, NY
    2 days ago
  •  ...Our client is looking for an Infrastructure Engineer to build and operate the foundational systems that power our data, analytics, and...  ...container orchestration, CI/CD, and monitoring keeping our platform reliable, scalable, and secure. If you are excited about AI and want to... 
    Flexible hours
    3 days per week

    The Phoenix Group

    New York, NY
    2 days ago
  • $150k - $200k

     ...Site Reliability Engineer at Triomics (W21) $150K - $200K AI Agents for Oncology EHRs Triomics is building the agentic AI layer for oncology electronic health records (EHRs). Cancer hospitals spend billions on highly trained staff manually reading unstructured patient... 
    Full time
    Work at office
    Remote work
    Day shift

    Triomics

    New York, NY
    22 hours ago
  •  ...About the job Senior Site Reliability Engineer About the Company Stellar is a decentralized, public blockchain that gives developers the tools to create experiences that are more like cash than crypto. The network is faster, cheaper, and far more energy-efficient... 

    TechChain Talent

    New York, NY
    2 days ago
  •  ...A leading quantitative trading firm is hiring a Site Reliability Engineer to help build and operate the critical systems powering its global trading infrastructure. This is a high-impact role focused on reliability, automation, scalability, and performance across mission... 

    Goliath Partners LP

    New York, NY
    2 days ago
  • $200k - $220k

     ...pivots, and the challenges of building in a high‑growth startup, we’d love to talk. This is more than a job—it’s a journey. Site Reliability Engineers (SREs) are responsible for the overall performance and reliability of ASAPP's infrastructure and products. The team owns... 
    Remote work

    ASAPP

    New York, NY
    2 days ago
  •  ...Quant Trading space, this firm combines cutting edge technology and world class engineering to execute high profit, high quality trades at speed. They are now searching for a Site Reliability Engineer (SRE) to join one of their elite equity teams in New York. The Role... 
    Contract work
    Remote work

    Thurn Partners

    New York, NY
    2 days ago
  • $100k - $250k

     ...systems) can be yours. What you’ll do Improve observability, reliability and availability by defining and measuring key metrics. Build...  ...function improvements. Educate, mentor and hold accountable the engineering team to improve the reliability of our systems and make... 
    Local area

    Kalshi

    New York, NY
    2 days ago
  •  ...New York City or Chicago (Hybrid) A technology-driven investment firm is expanding its Platform Engineering organization and is seeking an experienced Senior Site Reliability Engineer to help shape reliability practices across its infrastructure and production... 

    Mission Staffing

    New York, NY
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!