Twenty Twenty Therapeutics
Forward Deployed Site Reliability Engineer (TS/SCI
Arlington, VA · Vp · Full-time
No sponsorshipDetected 79 days ago
TypeScriptAWSDockerTerraformPrometheusGrafana
About the role
- ABOUT THE COMPANY America is under sustained cyber attack.
- Our adversaries infiltrate our networks, steal our IP, and degrade the digital infrastructure that modern life runs on.
- They've learned-correctly-that those attacks rarely produce consequences.
Responsibilities
- Use error budgets to drive reliability conversations with the Arlington engineering team, translating operational data into prioritized engineering work.
- Identify and eliminate toil: build automation for repetitive operational tasks within the constraints of the secure environment.
- Conduct post-incident reviews, own root cause analysis, and drive durable fixes in partnership with the engineering team.
- Maintain and continuously improve runbooks for operational procedures and emergency response protocols.
- Manage containerized services (Docker, Docker Compose) across deployment lifecycle - configuration, updates, rollbacks.
- Apply and validate Terraform-based infrastructure changes within the enclave, in coordination with the DSO engineer who owns IaC policy and guardrails.
- Perform capacity planning and flag scaling requirements to the Arlington team before they become incidents.
- Partner with the DevSecOps engineer on compliance, logging, and audit requirements specific to the customer environment.
- Provide technical guidance and support to customer stakeholders on system behavior and troubleshooting procedures.
- You own reliability outcomes, not just uptime dashboards - you define what "healthy" means and hold the system to it.
Requirements
- You're as comfortable writing a runbook as you are deep in a production incident with limited tooling and no safety net.
- You understand that in a restricted environment, you are the feedback loop - and you take that responsibility seriously.
- 5+ years of professional experience in site reliability engineering, production operations, or a closely related infrastructure role.
- Proven experience defining and tracking SLIs, SLOs, and error budgets in a production environment.
- Hands-on experience with Docker, Docker Compose, and AWS (EC2, ECS, RDS, VPCs, security groups) in production deployments.
- Experience with Terraform for infrastructure provisioning and configuration, working within DSO-provided policy guardrails.
- Experience with the LGTM observability stack or equivalent (Grafana, Loki, Prometheus/Mimir, distributed tracing).
Nice to have
- Experience with NATS or similar pub/sub messaging systems in production.
- Background in cyber operations, intelligence systems, or signals environments.
- AWS certifications (Solutions Architect, SysOps, or DevOps Engineer).
- U.S. citizenship required
- What's on the table:
- 12 weeks for birthing parents, 4 for non-birthing parents, 6 weeks for adoptive, foster, or intended parents through surrogacy.
- Take what you need.
- 401(k) with pre-tax and Roth options.
Benefits
- We consider all qualified applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, veteran status, disability, or any other protected status.
- Medical, dental, and vision plan options.
- Life / AD&D, disability coverage options.
- Paid parental leave for eligible full-time employees.
- Paid holidays and flexible PTO.
- HSA/FSA options, dependent care FSA.
- Commuter benefits.
- Building fitness center.
- Desk setup stipend.
- Benefits vary by location, role, and eligibility.
Company info
- America is under sustained cyber attack.
- Twenty was founded to change that, by making our adversaries think twice before they attack us.
- Our vision is American and allied primacy in cyberspace-a future where they cannot contest us, deterrence is assured, and the free world remains secure.
- Founded in 2024, Twenty Technologies (www.twenty.io http://www.twenty.io) industrializes offensive cyber operations for the U.S. and its allies.
- Headquartered in Arlington, Virginia, Twenty has raised $138M from Accel, Caffeinated Capital, Friends & Family Capital, Point72 Ventures, General Catalyst, and In-Q-Tel.
- You'll be our eyes, ears, and hands on the ground at a government customer site, ensuring the reliability and performance of Twenty's mission-critical platform running in a restricted, air-gapped AWS environment.
- This role sits at the intersection of deep technical ownership and customer-facing engineering: you'll define how we measure reliability, lead incident response in a constrained environment, and serve as the primary technical link between what's happening on-site and the engineering team back in Arlington.
- You'll work closely with the DevSecOps engineer to ensure the platform operates within government security and compliance requirements, and with product engineers to translate operational reality into actionable feedback.
- You'll report directly to the VP of Engineering.
- If you thrive operating with autonomy in high-stakes environments and find satisfaction in making complex systems provably reliable, this role is for you.
- You build trust naturally with external stakeholders, including government customers, and can translate complex technical situations into plain language under pressure.
Equal opportunity
- equal opportunity employer.
- If you need a reasonable accommodation during the hiring process, let us know and we will work with you.
Visa & Work Authorization
- Must possess and be able to maintain a TS/SCI security clearance with appropriate polygraph
Apply directly at Twenty Twenty Therapeutics →Create a free account for alerts like thisView Twenty Twenty Therapeutics immigration profile
This listing is sourced directly from Twenty Twenty Therapeutics's careers page and normalized into a canonical job model.