Nebius

Nebius

Senior Site Reliability Engineer (In-Office

New York City, New York, United States · Senior

Sponsorship not specifiedDetected 21 days ago
Distributed SystemsFull-Stack DevelopmentCloud PlatformsKubernetesTerraformCI/CDDevOpsSite Reliability EngineeringAPI DevelopmentMachine LearningData EngineeringLLMsRAGResearch

About the role

  • Managing Kubernetes clusters across multiple environments and regions
  • Owning infrastructure as code for all resources
  • Maintaining and improving CI/CD pipelines and GitOps-based deployments

Responsibilities

  • Working closely with a small engineering team - you'd own infra, not a slice of it

Requirements

  • 5-8 years in a DevOps or SRE role, working in production environments
  • Strong Kubernetes experience in a managed cloud environment
  • Proficiency with infrastructure as code (Terraform or similar)
  • Experience with GitOps-based deployment workflows
  • Experience handling production incidents calmly and methodically

Nice to have

  • Multi-region deployments <span data-ccp-props="{"134233117":false,"134233118":fal

Skills

  • Nebius is leading a new era in cloud infrastructure for the global AI economy.
  • Built by engineers, for engineers.
  • About Tavily
  • Managing cloud costs and capacity planning

This listing is sourced directly from Nebius's careers page and normalized into a canonical job model.