Zello

Zello

Senior Site Reliability Engineer, Database Infrastructure

Austin, Texas · Senior

Sponsorship not specifiedDetected 70 days ago
PythonGoBashMySQLMongoDBRedisElasticsearchCassandraAWSGCPAzureCloud PlatformsDockerKubernetesLinuxPrometheusDevOpsSite Reliability EngineeringCybersecurityIncident ResponseLogisticsEmbedded SystemsCommunicationCollaboration

About the role

  • Operated Zello's MySQL and MongoDB clusters to documented availability targets, with automated backups, regularly tested restores, and failover the on-call team trusts under real incident pressure.
  • Cut latency or capacity cost on at least one critical database workload through measurable performance work - index strategy, query tuning, schema changes, or sharding.
  • Extended our Observability coverage so incidents are diagnosed in minutes rather than hours, with dashboards and alerts the team actually uses.

Responsibilities

  • Design, deploy, and operate highly available MySQL and MongoDB clusters across our cloud environments
  • Participate in the Platform on-call rotation, lead incident response for data-tier issues, and write postmortems that drive durable change.
  • building agents, automation, and developer-enablement tooling that scale the team's reliability work
  • replication, sharding, backups, point-in-time recovery, and failover drills you've actually run, not just designed on paper.
  • Design, deploy, and operate highly available MySQL and MongoDB clusters across our cloud environments; replication, sharding, backups, point-in-time recovery, upgrades, and disaster recovery.

Requirements

  • All Zello personnel are required to comply with defined security, privacy, and compliance requirements applicable to their role along with requirements that are applicable to all Zello personnel.

Nice to have

  • You communicate clearly under pressure and after the fact.
  • You bring an opinion on managed vs. self-managed databases, and can defend the trade-off based on availability, cost, and operational burden.
  • 7+ years in SRE, DevOps, platform, infrastructure, or database reliability roles, with at least 3 years owning production databases.
  • BSc in Computer Science or equivalent practical experience.
  • ScyllaDB/Cassandra or Elasticsearch experience is a plus
  • You've shipped meaningful work on at least two of bare metal Linux, containerized workloads (Docker, Kubernetes, or similar), and a major cloud (GCP preferred
  • AWS or Azure equivalent is fine).

Skills

  • As Zello scales, the line between "database problem" and "platform problem" keeps blurring.
  • every channel, every message, every login depends on them.

Benefits

  • We have competitive pay, equity with significant upside, and intentionally design our benefits to encourage healthy and well-balanced employees, flexible schedules and time off.

Company info

  • We hire for potential, passion for our mission, and a knack for solving difficult problems over checking every qualification box.

This listing is sourced directly from Zello's careers page and normalized into a canonical job model.