Senior Site Reliability Engineer
$140k - $210.9kFederal Reserve Bank of Boston
Senior Site Reliability Engineer
Federal Reserve Financial Services (FRFS) delivers a suite of payments services to financial institutions via FedLine® Solutions, FedNowSM, Fedwire®, National Settlement Service (NSS), FedCash®, FedACH® (Automated Clearing House), and Check Services. We are currently leading a strategic effort to transform FRFS to a national, enterprise-focused organization. Through our evolved structure, we will meet the needs of the marketplace for new products and services more quickly, seek to provide a more robust and unified customer experience across our financial service offerings, and create new career growth opportunities for FRFS staff.
The Federal Reserve has developed a new interbank 24x7x365 real-time gross settlement (RTGS) service with integrated clearing functionality, called the FedNow Service. This service enables financial institutions to provide their customers with the ability to send and receive payments any time, any day, and have full access to those funds within seconds. This position is a unique opportunity to be part of this mission-critical Federal Reserve initiative that is transforming the payments landscape in the United States.
The position will be primarily on-site with residency commutable to one of our offices required. Candidates may come from infrastructure/DevOps backgrounds or software engineering backgrounds (e.g., Java Python, Go) with strong interest in operating and improving reliability of distributed production systems.
Responsibilities
As a Senior Engineer of the SRE / Production Operations team for FedNow, you will operate the production environment for the program. You will architect, implement, and leverage solution monitoring and tooling to be used for capacity planning, utilization reporting, and scaling. The team uses open source and proprietary software to support Engineering, DevOps, and DevSecOps tools, services, and solutions. CI/CD and IaC Pipeline automation design and development. Resiliency, DR and BCP (including testing) The SRE / Production Operations team is part of the Technical Operations (TechOps) department and has the overall responsibility for the design, management and execution of operations required to support the ongoing technical and delivery needs of the FedNow Program, as well as the transition to production support and operations. It owns ongoing ITIL processes, and the implementation and driving of continuous improvement initiatives. The role applies both software engineering and system engineering practices to operate and improve large-scale distributed systems. You will work closely with Engineers and Architects of the FedNow program in order to maintain seamless automation across the entire platform. Proactively identify suspected gaps in system architecture and design experiments to expose them
Key Skills
Strong communication and collaboration skills Extensive knowledge and understanding of working in AWS environments & services EC2, EBS, EKS, RDS, Aurora, S3, Route 53, ELB, IAM, etc. Hashicorp Terraform, Consul, Vault, and Ansible Experience developing automation or operational tooling using scripting or programming languages such as Python, Java, Go, or similar languages. Experience working with cloud infrastructure platforms or distributed system environments Experience working in Linux environment and shell scripting Experience supporting infrastructure for large multi-services applications Experience working with continuous deployment in micro-services architectures Experience working with Docker, Containers, ECR and EKS. Observability - CloudWatch, OpenSearch, Dynatrace, Grafana, Prometheus Familiarity with Fault Injection tooling (i.e. AWS Fault Injection Simulator, Gremlin, ChaosToolkit, Chaos Monkey) Automation mindset to enable consistency and dependability in common actions
The salary range for this position is $140,000 - $210,900. The position and job description posted is for a Senior Site Reliability Engineer however, candidates will be placed in an appropriate level within the Site Reliability Engineer job family based on the extent of their experience.
The Federal Reserve Bank of Boston is committed to provide equal employment opportunities to all persons without regard to race, color, religion, national origin, sex, sexual orientation, gender identity, age, genetic information, disability, or military service.
All employees assigned to this position will be subject to FBI fingerprint/ criminal background and Patriot Act/ Office of Foreign Assets Control (OFAC) watch list checks at least once every five years.
For this job, any offer of employment is contingent upon successfully passing a two-phase security screening. The first phase consists of the satisfactory completion of a physical examination (including a drug screening), reference checks, and a security investigation consisting of credit and criminal history checks.
The second phase, which might not be complete until after you begin working at the Reserve Bank, is an additional risk-based security screening determined by the risk rating of the position. Depending upon the sensitivity of the position, this phase may include, and is not limited to, work and residency eligibility verification, and personal interviews with the candidate, references, and prior employers.
All applicants must have resided in the United States for at least three (3) years.
Full Time / Part Time
Full time
Regular / Temporary
Regular
Job Exempt (Yes / No)
Yes
Job Category
Information Technology Family Group
Work Shift
First (United States of America)
$160k - $200k
...data, ideally using promQLKey Responsibilities:Mentor and evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage & remediation processes as a player / coachPerform incident...SeniorTemporary workWork at officeLocal areaFlexible hours3 days per week$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours$134.25k - $214.8k
...upload. Every piece of digital evidence. Every chain of custody log that holds up in court. That's us.Axon's Platform team is the engine behind what hundreds of thousands of officers rely on every day. We're one of the world's largest blob storage customers, ingesting...SeniorWork experience placementWork at officeRemote work- Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that... ...maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering...SeniorLocal areaWorldwideFlexible hours
$166k - $220k
...failure. As such, it is critical that Anduril services are reliable and maintainable. This means that all services &... ...ground systems & Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build & operate Anduril...SeniorFull timeWork experience placementImmediate start- ...The Depository Trust & Clearing Corporation (DTCC) seeks a Senior Application Support Engineer to ensure reliability and performance of its critical trade processing platforms. You will apply SRE principles, drive automation, and partner with global teams to support AWS...Senior
- ...and AI agent a cryptographically secured identity, improving engineering velocity while maintaining security. We make trusted computing... ...problems that allow our customers to trust us for secure and reliable access to their infrastructure. Excellent security is table stakes...SeniorWork at officeLocal areaRemote workSleeping nights
- ...commuting distance of one of our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team, you will operate the... ...candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based...SeniorFull time
$141k - $208k
...be a part of our journey! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability...SeniorLocal areaRemote workHome officeFlexible hours$160k - $200k
...Senior Site Reliability Engineer This role is located in Somerville, MA – We are a hybrid work environment and are in the office 3+ days/per week. Tulip, the leader in AI-native frontline operations, is helping companies around the world equip their workforce with...SeniorTemporary workWork at officeLocal areaFlexible hours3 days per week$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner... ...indicators and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling...SeniorWork experience placementWork at office- ...foundation of success and bringing it to the digital space - ready to join us? What’s the position? We are looking for a Senior Site Reliability Engineer who combines deep infrastructure expertise with a forward-thinking approach to AI-driven operations. In this role you...SeniorRemote workFlexible hoursNight shift
- ...Information Technology group delivers secure, reliable technology solutions that enable... ...You Will Have in This RoleAs a Senior Application Support Engineer, you will help power DTCC's global... ...processing and settlement.Leveraging Site Reliability Engineering (SRE) principles...SeniorRemote workFlexible hours
$134.25k - $214.8k
...real change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability...SeniorWork at officeRemote workFlexible hours$130k - $150k
...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are... ...career mentoring and performance coaching from an assigned senior colleague. Additional leadership and collaboration opportunities...Work at officeWork from home3 days per week$138.1k - $198.2k
...more intuitive with technology that simply works. The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments... ...Our customers are all engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter of our engineering...Permanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours$160k - $200k
Role Overview We are looking for a Control System Engineer/Site Reliability Engineer (SRE) to integrate and maintain the hardware and software systems that enable QuEra’s quantum controls and software stack. You’ll work closely with software engineers, physicists, hardware...Local areaRemote work$84.9k - $209.5k
...Job Description As a Principal Site Reliability Engineer (IC4), you will be responsible for designing, building, and operating highly available... ...Lead critical production incident response and act as a senior escalation point during major service events. Drive root...Temporary workFlexible hours- ...mission-critical industries, helping partners move more quickly and reliably from algorithm to silicon. Our platform accelerates deployment... .... The Roles We are looking for an experienced software engineer to help us build a new generation of transpilation tools...SeniorFull timeRemote workRelocation packageFlexible hours
$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours- ...scale, we invite you to bring your talents to Zscaler to help shape the future of cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or Bellevue WA) to join our Zero Trust Exchange team. This is a remote role based in San...InternshipWork at officeLocal areaRemote workWorldwide
$114k - $148k
...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based...Full timeTemporary workWork experience placementRemote work$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office- ...SRE Support Engineer - Observability While this position is not currently open, we are interviewing strong candidates for upcoming opportunities... ...support across Slack and tickets, improving monitoring reliability, and reducing incident impact through better triage,...Remote work
$74.1k - $148.3k
...systems. Facilitate service capacity planning and demand forecasting, software performance analysis, and system tuning. As a Site Reliability Engineer, you will solve interesting technical challenges by defining, designing, deploying, and solving key Oracle Cloud services,...Temporary workImmediate startFlexible hours- ...Job Title: Site Reliability Engineer Location: Remote with Quarterly visits to Chennai, Tamil Nadu, India Duration: Full-Time bout BigRio: BigRio is a remote-based, technology consulting firm headquartered in Boston, MA. We deliver software solutions...Full timeRemote work
- ...ISEE is seeking an experienced Senior Software Engineer to join our team. The ideal candidate has several years of work experience, and have worked on complex, performance-critical code-bases. Role responsibilities include: - Support full software development life...SeniorFull timeWork experience placement
$108k - $209k
...seeking an experienced, creative, and talented Principal / Senior Software Engineer. The ideal candidate will have a strong background in software... .... Leverage AWS cloud infrastructure to build scalable, reliable, and efficient applications and AI-powered services. Uphold...Senior$130k - $140k
...mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionSite Reliability EngineerLocation(s): Waltham, MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to ensuring systems are reliable, scalable, and efficient...SeniorOngoing contractFull timeTemporary workWork experience placement$138k - $252k
...scaled autonomy Solicit and incorporate feedback from end users of the APIs and implementations, and collaborate with adjacent engineering teams to help make the product vision a reality Develop and improve our APIs for commanding and controlling teams of...SeniorFull timeWork experience placementLocal areaRelocation packageFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer remote Boston, MA
- site reliability engineer sre Boston, MA
- site reliability engineer Boston, MA
- senior operations associate Boston, MA
- senior safety specialist Boston, MA
- senior technology project manager Boston, MA
- remote senior business analyst Boston, MA
- senior manager clinical operations Boston, MA
- senior supervisor Boston, MA
- senior leadership Boston, MA


