Nebius
Senior Site Reliability Engineer (In-Office
New York City, New York, United States · Senior
Sponsorship not specifiedDetected 21 days ago
Distributed SystemsFull-Stack DevelopmentCloud PlatformsKubernetesTerraformCI/CDDevOpsSite Reliability EngineeringAPI DevelopmentMachine LearningData EngineeringLLMsRAGResearch
About the role
- Managing Kubernetes clusters across multiple environments and regions
- Owning infrastructure as code for all resources
- Maintaining and improving CI/CD pipelines and GitOps-based deployments
Responsibilities
- Working closely with a small engineering team - you'd own infra, not a slice of it
Requirements
- 5-8 years in a DevOps or SRE role, working in production environments
- Strong Kubernetes experience in a managed cloud environment
- Proficiency with infrastructure as code (Terraform or similar)
- Experience with GitOps-based deployment workflows
- Experience handling production incidents calmly and methodically
Nice to have
- Multi-region deployments <span data-ccp-props="{"134233117":false,"134233118":fal
Skills
- Nebius is leading a new era in cloud infrastructure for the global AI economy.
- Built by engineers, for engineers.
- About Tavily
- Managing cloud costs and capacity planning
This listing is sourced directly from Nebius's careers page and normalized into a canonical job model.