Senior Site Reliability Engineer

The Federal Reserve Bank
Boston, MA, United States
3 days ago
Apply on www.juju.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$140,000.0 - $210,900.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Systems Engineering Cloud Computing Continuous Delivery Continuous Integration Linux DevOps Distributed Systems Identity and Access Management
+21 more
Python (Programming Language) Open Source Technology Proprietary Software Reliability Engineering Ansible Prometheus Shell Script Software Engineering Scripting DevOps Tools - Open-source Delivery Pipeline Grafana Hashicorp Route53 Cloudwatch Terraform Dynatrace Docker Golang Programming Languages Microservices

Job description

As a Senior Engineer of the SRE / Production Operations team for FedNow, you will operate the production environment for the program.

You will architect, implement, and leverage solution monitoring and tooling to be used for capacity planning, utilization reporting, and scaling.

The team uses open source and proprietary software to support Engineering, DevOps, and DevSecOps tools, services, and solutions.

CI/CD and IaC Pipeline automation design and development.

Resiliency, DR and BCP (including testing)

The SRE / Production Operations team is part of the Technical Operations (TechOps) department and has the overall responsibility for the design, management and execution of operations required to support the ongoing technical and delivery needs of the FedNow Program, as well as the transition to production support and operations.

It owns ongoing ITIL processes, and the implementation and driving of continuous improvement initiatives.

The role applies both software engineering and system engineering practices to operate and improve large-scale distributed systems. You will work closely with Engineers and Architects of the FedNow program in order to maintain seamless automation across the entire platform.

Proactively identify suspected gaps in system architecture and design experiments to expose them

The ideal candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based highly available, high performing applications.

Requirements

Strong communication and collaboration skills

Extensive knowledge and understanding of working in AWS environments & services

EC2, EBS, EKS, RDS, Aurora, S3, Route 53, ELB, IAM, etc.

Hashicorp Terraform, Consul, Vault, and Ansible

Experience developing automation or operational tooling using scripting or programming languages such as Python, Java, Go, or similar languages.

Experience working with cloud infrastructure platforms or distributed system environments

Experience working in Linux environment and shell scripting

Experience supporting infrastructure for large multi-services applications

Experience working with continuous deployment in micro-services architectures

Experience working with Docker, Containers, ECR and EKS.

Observability - CloudWatch, OpenSearch, Dynatrace, Grafana, Prometheus

Familiarity with Fault Injection tooling

(i.e. AWS Fault Injection Simulator, Gremlin, ChaosToolkit, Chaos Monkey)

Automation mindset to enable consistency and dependability in common actions

About the company

Federal Reserve Bank of Boston

Federal Reserve Financial Services (FRFS) delivers a suite of payments services to financial institutions via FedLine® Solutions, FedNowSM, Fedwire®, National Settlement Service (NSS), FedCash®, FedACH® (Automated Clearing House), and Check Services. We are currently leading a strategic effort to transform FRFS to a national, enterprise-focused organization. Through our evolved structure, we will meet the needs of the marketplace for new products and services more quickly, seek to provide a more robust and unified customer experience across our financial service offerings, and create new career growth opportunities for FRFS staff.

The Federal Reserve has developed a new interbank 24x7x365 real-time gross settlement (RTGS) service with integrated clearing functionality, called the FedNow Service. This service enables financial institutions to provide their customers with the ability to send and receive payments any time, any day, and have full access to those funds within seconds. This position is a unique opportunity to be part of this mission-critical Federal Reserve initiative that is transforming the payments landscape in the United States.

The position will be primarily on-site with residency commutable to one of our offices required.

Candidates may come from infrastructure/DevOps backgrounds or software engineering backgrounds (e.g., Java Python, Go) with strong interest in operating and improving reliability of distributed production systems., OUR BANK has one of the most recognizable brands around the world. The Federal Reserve is the central bank of the United States-one of the world’s most influential, trusted and prestigious financial organizations. The Federal Reserve is charged with the important mission of promoting a strong economy and a stable financial system and fulfills this responsibility by formulating national monetary policy, supervising and regulating banks and bank holding companies, and providing financial services for banks and the U.S. government.

OUR PEOPLE are diverse in background and ideas, which allows for ongoing creativity and innovation. Ultimately, they are the ones who push our high-performance, exchange-driven culture forward.

Why Our People Choose Us

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.juju.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

4:23 min

Reviewing AWS infrastructure deployment configuration and planning

Devlin Duldulao · LIVE

Videos

See all

Related articles

See all