ReflectionAI
Member of Technical Staff - Distributed Systems Engineer
New York · Staff+
H1B sponsorship availableDetected 132 days ago
GoRustDistributed SystemsKubernetesSite Reliability EngineeringgRPCA/B TestingSystems EngineeringResearch
About the role
- You will help define our future as a company, and help define the future of open foundational models.
Responsibilities
- Reflection is a research lab making intelligence open and accessible for everyone to use, customize, and build on.
- We build open models that let anyone control their intelligence and help shape the future of AI.
- Build and operate shared services that multiple teams rely on across research and production workflows.
- Maintain strong operational readiness with runbooks, incident playbooks, and capacity planning.
- Develop APIs, SDKs, and internal platforms that enable high-velocity experimentation and iteration.
- Idempotency, retries, backpressure, SLI/SLO design, tail latency optimization, service reliability engineering.
- Sponsorship support: We sponsor visas to help exceptional talent join our team and support long-term immigration pathways where applicable.
- Team building: We have regular off-sites, happy hours, and team celebrations.
- Joining Reflection means building from the ground up as part of a talent-dense team.
Compensation
- Salary and equity structured to recognize and retain our talent globally.
- Health & wellness: Comprehensive medical, dental, vision, and life, with an annual wellness allowance.
Benefits
- Top-tier compensation: Salary and equity structured to recognize and retain our talent globally.
- Stock options: Everyone who joins and contributes to Reflection's success gets to share in the upside through stock options.
- Health & wellness: Comprehensive medical, dental, vision, and life, with an annual wellness allowance.
- Life & family: 22 weeks paid parental leave for all new birthing and non-birthing parents, including adoptive and surrogate journeys.
- Vacation days: Unlimited paid time off in the U.S. and 30 days in the U.K.
Company info
- make intelligence open and accessible to all.
- Build and operate a company-wide foundations platform that accelerates every team by providing reliable, scalable developer infrastructure, SRE capabilities, and high-throughput data ingestion tooling enabling Reflection to move faster as we scale.
- Build and operate the core shared services that power our research, training, and production environments.
- These systems form the foundational platform that multiple teams depend on for model development, deployment, and evaluation, unifying data, compute, and workflow management across the stack while enabling rapid experimentation and reliable production systems.
- Define and uphold reliability targets through SLIs, SLOs, and healthy on-call practices.
- Ensure correctness and performance under load, addressing consistency, tail latency, and failure modes.
- Reduce operational burden through better tooling, standardization, and platform patterns that scale across teams.
- Container Abstractions: Containers-as-a-Service, Kubernetes abstraction layers, container orchestration, reproducible environments, multi-tenant isolation.
- Distributed Systems Architecture: Sharding, replication, coordination services, high-concurrency systems, concurrency control.
Visa & Work Authorization
- We sponsor visas to help exceptional talent join our team and support long-term immigration pathways where applicable.
Apply directly at ReflectionAI →Create a free account for alerts like thisView ReflectionAI immigration profile
This listing is sourced directly from ReflectionAI's careers page and normalized into a canonical job model.