Senior Service Reliability Engineer
Fitch Group
Senior Service Reliability Engineer
Fitch Group is currently seeking a Senior Service Reliability Engineer to embed with Fitch Ratings development squads in Toronto and to partner with developers across Chicago, London, Manchester and New York. The role is based out of our Toronto office. As a leading, global financial information services provider, Fitch Group delivers vital credit and risk insights, robust data, and dynamic tools to champion more efficient, transparent financial markets. With over 100 years of experience and colleagues in over 30 countries, Fitch Group's culture of credibility, independence, and transparency is embedded throughout its structure, which includes Fitch Ratings, one of the world's top three credit ratings agencies, and Fitch Solutions, a leading provider of insights, data and analytics. With dual headquarters in London and New York, Fitch Group is owned by Hearst.
Fitch's Technology & Data Team is a dynamic department where innovation meets impact. Our team includes the Chief Data Office, Chief Software Office, Chief Technology Office, Emerging Technology, Shared Technology Services, Technology, Risk and the Executive Program Management Office (EPMO). Driven by our investment in cutting-edge technologies like AI and cloud solutions, we're home to a diverse range of roles and backgrounds united by a shared passion for leveraging modern technology to drive projects that matter to our organization and clients. We are also proud to be recognized by Built In as a "Best Place to Work in Technology" 3 years in a row. Whether you're an experienced professional or just starting your career, we offer an exciting and supportive environment where you can grow, innovate, and make a difference.
Fitch Group SRE provides Service Reliability Engineering expertise to Fitch's development organizations. This squad joins Core Engineering, Networking, and other SRE groups as part of Cloud Infrastructure & Platform Engineering (CI&PE), serving as subject matter experts in cloud technologies, systems engineering, infrastructure automation, and DevOps tooling across Fitch Group.
The role itself will be dedicated to ensuring excellence in Fitch Ratings services, with focus on new AI development.
How You'll Make an Impact:
- Lead the delivery of reliable, scalable, mission-critical services.
- Guide squads on Kubernetes and modern deployment patterns.
- Mentor associate engineers and set best practices for areas of expertise.
- Partner closely with Fitch Ratings Development Squads and Operations to design and advance service builds, automation, AI tooling, and operational excellence.
- Partner with Core Engineering to architect and govern GitHub Actions CI/CD with quality gates, canary/blue-green strategies, and AI-assisted redeploy checks.
- Own observability in Datadog—define SLIs/SLOs, dashboards, alerting, and MS Teams integrations—and reduce incidents via telemetry-driven automation and blameless postmortems.
- Champion AI-enabled operations using AWS Bedrock/SageMaker and Model Context Protocol (MCP) for log analysis, anomaly detection, incident triage, and workflow orchestration; establish adoption guardrails.
- Define and enforce cloud guardrails and security controls (SCPs/IAM boundaries, OPA policies, tagging, centralized logging with AWS Config/CloudTrail/Security Hub) in partnership with Security and Risk.
- Influence cross-functional roadmaps, lead complex release planning, and drive strategic platform initiatives across CI&PE serve as an escalation point and participate in the L3 on-call rotation.
You May be a Good Fit if:
- You have deep, hands-on experience in SRE, DevOps, or Platform Engineering across both AWS and Azure, with a strong track record operating Docker and Kubernetes in production environments.
- You're highly proficient in administering both Linux and Windows, and have practical, enterprise-level experience supporting IIS/.NET applications as well as Java Spring Boot services.
- You have built and maintained CI/CD pipelines (primarily GitHub Actions; Bamboo experience a plus) with DevSecOps principles baked in—integrating security scans, policy-as-code, and compliance gates—and script confidently in Python, PowerShell, or Bash.
- You have experience with cloud security best practices (IAM, secrets management, container/image scanning) and understand core infrastructure fundamentals (networking, storage, DNS) and APM/telemetry tooling.
What Would Make You Stand Out:
- Practical experience with agentic AI for operations—incident triage, runbooks, and change management—with clear guardrails, auditability, and human-in-the-loop controls.
- Supporting AI/ML workloads at scale: SageMaker endpoints, GPU node groups, autoscaling, and Kubernetes-based model serving.
- Policy-as-code (OPA) and compliance implementation across CIS, NIST, ISO 27001, with automated remediation integrated via CSPM tools (e.g., Wiz).
- Applying AI in CI/CD, observability, and incident response using AWS DevOps Agent, Claude Code, or others with Skills and Model Context Protocol (MCP).
- Hands-on Agile delivery experience, actively participating in stand-ups and sprint ceremonies.
Why Choose Fitch:
- Hybrid Work Environment: 2 to 3 days a week in office required based on your line of business and location
- A Culture of Learning & Mobility: Dedicated trainings, leadership development and mentorship programs designed to ensure that your time at Fitch will be a continuous learning opportunity
- Investing in Your Future: Retirement planning and tuition reimbursement programs that empower you to achieve your short and long-term goals
- Promoting Health & Wellbeing: Comprehensive healthcare offerings that enable physical, mental, financial, social, and occupational wellbeing
- Supportive Parenting Policies: Family-friendly policies, including a generous global parental leave plan, designed to help you balance career and family life effectively
- Inclusive Work Environment: A collaborative workplace where all voices are valued, with Employee Resource Groups that unite and empower our colleagues around the globe
- Dedication to Giving Back: Paid volunteer days, matched funding for donations and ample opportunities to volunteer in your community
Fitch is committed to providing global securities markets with objective, timely, independent and forward-looking credit opinions. To protect Fitch's credibility and reputation, our employees must take every precaution to avoid conflicts of interest or any appearance of a conflict of interest. Should you be successful in the recruitment process at Fitch Ratings you will be asked to declare any securities holdings and other potential conflicts prior to commencing employment. If you, or your immediate family, have any holdings that may conflict with your work responsibilities, you may be asked to divest yourself of them before beginning work.
Fitch is proud to be an Equal Opportunity and Affirmative Action Employer. We evaluate qualified applicants without regard to race, color, national origin, religion, sex, sexual orientation, gender identity, disability, protected veteran status, and other statuses protected by law.
FOR TORONTO ROLES ONLY: Expected base pay rates for the role will be between CAD 120,000 and CAD 150,000 per year. Actual salaries will be determined on an individualized basis and may vary based on factors including but not limited to education, training, experience, past performance, and other job-related factors. Base pay is one part of Fitch's total compensation package, which, depending on the position, may also include commission earnings, discretionary bonuses, long-term incentives, and other benefits sponsored by Fitch.
- ...Senior Database Reliability Engineer (DBRE)Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex...Senior
$153k - $210k
...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you passionate about building resilient, highly... ...with product and platform engineers to improve service reliability, accelerate engineering velocity through automation...SeniorFull time- ...technology company, we're a force for good in financial services. We're redefining how community banks and credit unions connect... ...(EIT) organization is expanding, and we are seeking a Senior Site Reliability Engineer to help drive a major architectural modernization. In...SeniorPermanent employmentFull timeH1bLocal areaRemote workShift work
- ...Senior Site Reliability Engineer (SRE) Our client is a global technology consulting and digital solutions company that enables enterprises across industries to reimagine business models, accelerate innovation, and maximize growth by harnessing digital technologies....SeniorLocal area
$200k - $240k
...problems and help health systems deliver better care, we'd love to meet you! About the role We're looking for a Senior Site Reliability Engineer to join our Infrastructure Engineering team and get their hands directly into the systems that keep our healthcare platform...SeniorWork at office3 days per week$182.8k - $247.3k
...learners around the world. About the role... As a Senior Site Reliability Engineer, you will work closely with both product and platform engineering... ...to release new features and become an authority on our services ✅ You have... Experience identifying and solving...SeniorWork experience placement$150k - $170k
...Senior Site Reliability Engineer – Zip Co Join to apply for the Senior Site Reliability Engineer role at Zip Co At Zip, we build cloud‑native... ...across cloud environments (Azure, Kubernetes, Service Mesh). Define, measure, and improve Service Level Objectives...SeniorCasual workWork at officeRemote workFlexible hours- ...Senior Site Reliability Engineer (SRE) Plenful is hiring a Senior Site Reliability Engineer (SRE) to keep our production systems reliable, performant... ...and implement SLIs, SLOs, and error budgets across core services. Own production system health: uptime, latency, and...SeniorFull timeWork at officeRemote workFlexible hours2 days per week
$185.5k - $232k
...Senior Site Reliability Engineer New York, NY; Boston, MA; San Francisco, CA About Formation Bio Formation Bio is a tech and AI driven pharma... ...infrastructure for product applications, containerized services, internal tools, data systems, ML pipelines, inference,...SeniorWork experience placementWork at officeLocal areaRelocation3 days per week$500 per month
...subsidiaries, Alpaca is a licensed financial services company, serving hundreds of financial... ...team is a diverse group of experienced engineers, traders, and brokerage professionals... ...you to apply. Your Role: As a Site Reliability Engineer at Alpaca, you'll help keep...SeniorHome office$189k - $283.6k
...people. Square makes commerce and financial services accessible to sellers. Cash App is the... ...proactively and reactively improve the reliability of Block's platform and critical... ...strong desire to perform and grow as an engineer ~5+ years of software development experience...SeniorFull timeLocal areaRemote workRelocation packageFlexible hoursShift work- ...raising the bar. This is the place. The role As a Senior Site Reliability Engineer you'll join the founding SRE team at our new NYC engineering hub, sitting within Foundations. You'll own critical services end-to-end and partner across engineering to raise the...SeniorWork at office
$139k - $257.55k
...Creative Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning,... ...as Behance's creative community and Content's advanced services. Our infrastructure spans multiple cloud platforms and is...SeniorTemporary workLocal areaRemote workWorldwide$191k - $226k
...than anyone else can. About the role: We are seeking a Senior Site Reliability Engineer to own the reliability, performance, and resilience of... ...workloads; define, measure, and uphold SLOs across our critical services Lead Incident Response: Serve in the on-call rotation...SeniorRemote workWork visaFlexible hours- ...Job Description Job Description A major financial services company in NYC is growing its team rapidly, and they are looking for a Senior DevOps Engineer / Site Reliability Engineer who can join. If you’re passionate about high-availability, reliability, automation...Senior
- The Sr Customer Service Engineer - is an intermediate position that performs tasks related to the repair of a variety of technology-based products typically associated in an end-user computing environment. Due to government contract requirements, U.S. Citizenship is required...SeniorContract workWorldwide
$132.6k - $174.04k
...SummaryTemporal customers are building some of the most critical, reliability-sensitive systems in production. They choose Temporal... ..., and operate durable workflows at scale.As a Senior Professional Services Engineer, you will help customers and partners move from design...Senior- ...Senior Site Reliability Engineer (SRE) Our client is seeking a Senior Site Reliability Engineer (SRE) with 10–15 years of experience to support front-office trading systems in a production environment. This role focuses on troubleshooting complex trading infrastructure...Senior
- ...Job Description Job Description Senior Database Reliability Engineer Role Type: Full-time Location: Remote About the Role We are looking for an experienced Database Reliability Engineer to architect, optimize, and maintain highly available PostgreSQL environments...SeniorRemote jobFull time
$104k - $178k
...Sr. Site Reliability Engineer I You will join the Site Reliability Engineering (SRE) team within DoubleVerify's Technology organization.... ...Monitoring and maintaining high-availability infrastructure and services across GCP, AWS, and on-premises environments. Responding...Senior- ...Responsibilities Lead data-center smart-hands, cabling, and hardware work Deliver complex installs and migrations Mentor junior field engineers Follow strict change-control and safety procedures What you’ll bring 5+ years field / data-center engineering Structured...SeniorFlexible hours
- ...curious, appreciates complexity, knows or wants to learn when to step back and when to dive deep. We call this role a Cloud Service Reliability Engineer. The Cloud Service Reliability Engineer will be responsible for effective design, execution, and maintenance of...
- ...threats and building digital government services, OTI is at the forefront of how the city... ...We are seeking a versatile and driven Senior Data Service Developer to join our innovative... ...and support of complex data engineering solutions for City agencies. The Senior...SeniorFull timeWork at officeShift workNight shiftWeekend workAfternoon shift
$75k - $150k
...threats and building digital government services, OTI is at the forefront of how the city... ...We are seeking a versatile and driven Senior Data Service Developer to join our innovative... ...and support of complex data engineering solutions for City agencies. The Senior...SeniorFull timeWork at officeShift workNight shiftWeekend workAfternoon shift$160k - $240k
Senior Software Engineer - BQL Reliability Engineering Location New York Business Area Engineering and CTO Ref # 10053945 Description... ..., resilience, and transparency of BQL and the services it depends on. You’ll work on engineering problems at...SeniorTemporary workFor contractorsWork experience placement- ...General arrangement drawings Technical manuals and other engineering documents Development and validation of equipment and functional... ...Support for Preventive Maintenance, Spare Parts, BOM, Reliability, and Maintenance Engineering activities Coordination with...Temporary workWork at office
$400k
...in financial markets, the organization combines innovation, engineering excellence, and data-driven insights to support complex trading operations worldwide. This opportunity is for a Senior Site Reliability Engineer to join a high-performance infrastructure...SeniorPermanent employmentWorldwide- ...Principal Site Reliability Engineer Location: New York, NY (Onsite) Job type: Contract Job Description: Job Requirements Must Have:... ...self-healing rules - Automated diagnostics and resolution service integration - Observability integration (AWS CloudWatch, Grafana...Contract work
$168k - $200k
...healthcare. What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the... ...teams to build a modern data platform that is secure, self-service, and production-grade. What You Will Do Operate and...Senior$174k - $252k
...in and improve the whole lifecycle of services - from inception and design through to... ...systems by pushing for changes that improve reliability and velocity.Practice sustainable... ...Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Service Reliability Engineer. Be the first to apply!
- field support engineer New York, NY
- client services engineer New York, NY
- managed services engineer New York, NY
- amazon web services engineer New York, NY
- IT field engineer New York, NY
- senior technical service engineer New York, NY
- electrical field service engineer New York, NY
- service virtualization engineer New York, NY
- position field engineer New York, NY
- field applications engineer New York, NY



