Twenty Twenty Therapeutics

Twenty Twenty Therapeutics

Forward Deployed Site Reliability Engineer (TS/SCI

Arlington, VA · Vp · Full-time

No sponsorshipDetected 79 days ago
TypeScriptAWSDockerTerraformPrometheusGrafana

About the role

  • ABOUT THE COMPANY America is under sustained cyber attack.
  • Our adversaries infiltrate our networks, steal our IP, and degrade the digital infrastructure that modern life runs on.
  • They've learned-correctly-that those attacks rarely produce consequences.

Responsibilities

  • Use error budgets to drive reliability conversations with the Arlington engineering team, translating operational data into prioritized engineering work.
  • Identify and eliminate toil: build automation for repetitive operational tasks within the constraints of the secure environment.
  • Conduct post-incident reviews, own root cause analysis, and drive durable fixes in partnership with the engineering team.
  • Maintain and continuously improve runbooks for operational procedures and emergency response protocols.
  • Manage containerized services (Docker, Docker Compose) across deployment lifecycle - configuration, updates, rollbacks.
  • Apply and validate Terraform-based infrastructure changes within the enclave, in coordination with the DSO engineer who owns IaC policy and guardrails.
  • Perform capacity planning and flag scaling requirements to the Arlington team before they become incidents.
  • Partner with the DevSecOps engineer on compliance, logging, and audit requirements specific to the customer environment.
  • Provide technical guidance and support to customer stakeholders on system behavior and troubleshooting procedures.
  • You own reliability outcomes, not just uptime dashboards - you define what "healthy" means and hold the system to it.

Requirements

  • You're as comfortable writing a runbook as you are deep in a production incident with limited tooling and no safety net.
  • You understand that in a restricted environment, you are the feedback loop - and you take that responsibility seriously.
  • 5+ years of professional experience in site reliability engineering, production operations, or a closely related infrastructure role.
  • Proven experience defining and tracking SLIs, SLOs, and error budgets in a production environment.
  • Hands-on experience with Docker, Docker Compose, and AWS (EC2, ECS, RDS, VPCs, security groups) in production deployments.
  • Experience with Terraform for infrastructure provisioning and configuration, working within DSO-provided policy guardrails.
  • Experience with the LGTM observability stack or equivalent (Grafana, Loki, Prometheus/Mimir, distributed tracing).

Nice to have

  • Experience with NATS or similar pub/sub messaging systems in production.
  • Background in cyber operations, intelligence systems, or signals environments.
  • AWS certifications (Solutions Architect, SysOps, or DevOps Engineer).
  • U.S. citizenship required
  • What's on the table:
  • 12 weeks for birthing parents, 4 for non-birthing parents, 6 weeks for adoptive, foster, or intended parents through surrogacy.
  • Take what you need.
  • 401(k) with pre-tax and Roth options.

Benefits

  • We consider all qualified applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, veteran status, disability, or any other protected status.
  • Medical, dental, and vision plan options.
  • Life / AD&D, disability coverage options.
  • Paid parental leave for eligible full-time employees.
  • Paid holidays and flexible PTO.
  • HSA/FSA options, dependent care FSA.
  • Commuter benefits.
  • Building fitness center.
  • Desk setup stipend.
  • Benefits vary by location, role, and eligibility.

Company info

  • America is under sustained cyber attack.
  • Twenty was founded to change that, by making our adversaries think twice before they attack us.
  • Our vision is American and allied primacy in cyberspace-a future where they cannot contest us, deterrence is assured, and the free world remains secure.
  • Founded in 2024, Twenty Technologies (www.twenty.io http://www.twenty.io) industrializes offensive cyber operations for the U.S. and its allies.
  • Headquartered in Arlington, Virginia, Twenty has raised $138M from Accel, Caffeinated Capital, Friends & Family Capital, Point72 Ventures, General Catalyst, and In-Q-Tel.
  • You'll be our eyes, ears, and hands on the ground at a government customer site, ensuring the reliability and performance of Twenty's mission-critical platform running in a restricted, air-gapped AWS environment.
  • This role sits at the intersection of deep technical ownership and customer-facing engineering: you'll define how we measure reliability, lead incident response in a constrained environment, and serve as the primary technical link between what's happening on-site and the engineering team back in Arlington.
  • You'll work closely with the DevSecOps engineer to ensure the platform operates within government security and compliance requirements, and with product engineers to translate operational reality into actionable feedback.
  • You'll report directly to the VP of Engineering.
  • If you thrive operating with autonomy in high-stakes environments and find satisfaction in making complex systems provably reliable, this role is for you.
  • You build trust naturally with external stakeholders, including government customers, and can translate complex technical situations into plain language under pressure.

Equal opportunity

  • equal opportunity employer.
  • If you need a reasonable accommodation during the hiring process, let us know and we will work with you.

Visa & Work Authorization

  • Must possess and be able to maintain a TS/SCI security clearance with appropriate polygraph

This listing is sourced directly from Twenty Twenty Therapeutics's careers page and normalized into a canonical job model.