Rb

Rb

Senior Site Reliability Engineer

Boston, MA · Senior · Part-time

Sponsorship not specified$140k-$211kDetected 30 days ago
PythonJavaGoBashDistributed SystemsElasticsearchAWSCloud PlatformsDockerTerraformAnsibleCI/CDLinuxPrometheusGrafanaDevOpsSite Reliability EngineeringCommunicationCollaboration

About the role

  • The Federal Reserve has developed a new interbank 24x7x365 real-time gross settlement (RTGS) service with integrated clearing functionality, called the FedNow Service.
  • Candidates may come from infrastructure/DevOps backgrounds or software engineering backgrounds (e.g., Java Python, Go) with strong interest in operating and improving reliability of distributed production systems.
  • As a Senior Engineer of the SRE / Production Operations team for FedNow, you will operate the production environment for the program.

Responsibilities

  • You will architect, implement, and leverage solution monitoring and tooling to be used for capacity planning, utilization reporting, and scaling.
  • The team uses open source and proprietary software to support Engineering, DevOps, and DevSecOps tools, services, and solutions.
  • CI/CD and IaC Pipeline automation design and development.
  • It owns ongoing ITIL processes, and the implementation and driving of continuous improvement initiatives.
  • You will work closely with Engineers and Architects of the FedNow program in order to maintain seamless automation across the entire platform.
  • Proactively identify suspected gaps in system architecture and design experiments to expose them

Skills

  • Strong communication and collaboration skills
  • Extensive knowledge and understanding of working in AWS environments & services
  • EC2, EBS, EKS, RDS, Aurora, S3, Route 53, ELB, IAM, etc.
  • Hashicorp Terraform, Consul, Vault, and Ansible
  • Experience working with cloud infrastructure platforms or distributed system environments
  • Experience working in Linux environment and shell scripting
  • Experience supporting infrastructure for large multi-services applications
  • Experience working with continuous deployment in micro-services architectures
  • Experience working with Docker, Containers, ECR and EKS.
  • Observability - CloudWatch, OpenSearch, Dynatrace, Grafana, Prometheus
  • Familiarity with Fault Injection tooling
  • (i.e. AWS Fault Injection Simulator, Gremlin, ChaosToolkit, Chaos Monkey)

Compensation

  • The salary range for this position is $140,000 - $210,900.

This listing is sourced directly from Rb's careers page and normalized into a canonical job model.