Rb
Senior Site Reliability Engineer
Boston, MA · Senior · Part-time
Sponsorship not specified$140k-$211kDetected 30 days ago
PythonJavaGoBashDistributed SystemsElasticsearchAWSCloud PlatformsDockerTerraformAnsibleCI/CDLinuxPrometheusGrafanaDevOpsSite Reliability EngineeringCommunicationCollaboration
About the role
- The Federal Reserve has developed a new interbank 24x7x365 real-time gross settlement (RTGS) service with integrated clearing functionality, called the FedNow Service.
- Candidates may come from infrastructure/DevOps backgrounds or software engineering backgrounds (e.g., Java Python, Go) with strong interest in operating and improving reliability of distributed production systems.
- As a Senior Engineer of the SRE / Production Operations team for FedNow, you will operate the production environment for the program.
Responsibilities
- You will architect, implement, and leverage solution monitoring and tooling to be used for capacity planning, utilization reporting, and scaling.
- The team uses open source and proprietary software to support Engineering, DevOps, and DevSecOps tools, services, and solutions.
- CI/CD and IaC Pipeline automation design and development.
- It owns ongoing ITIL processes, and the implementation and driving of continuous improvement initiatives.
- You will work closely with Engineers and Architects of the FedNow program in order to maintain seamless automation across the entire platform.
- Proactively identify suspected gaps in system architecture and design experiments to expose them
Skills
- Strong communication and collaboration skills
- Extensive knowledge and understanding of working in AWS environments & services
- EC2, EBS, EKS, RDS, Aurora, S3, Route 53, ELB, IAM, etc.
- Hashicorp Terraform, Consul, Vault, and Ansible
- Experience working with cloud infrastructure platforms or distributed system environments
- Experience working in Linux environment and shell scripting
- Experience supporting infrastructure for large multi-services applications
- Experience working with continuous deployment in micro-services architectures
- Experience working with Docker, Containers, ECR and EKS.
- Observability - CloudWatch, OpenSearch, Dynatrace, Grafana, Prometheus
- Familiarity with Fault Injection tooling
- (i.e. AWS Fault Injection Simulator, Gremlin, ChaosToolkit, Chaos Monkey)
Compensation
- The salary range for this position is $140,000 - $210,900.
This listing is sourced directly from Rb's careers page and normalized into a canonical job model.