HiveWatch
Senior Site Reliability Engineer
El Segundo, CA · Senior
Sponsorship not specified$145k-$173kDetected 7 days ago
TypeScriptPythonJavaRustKotlinDistributed SystemsFull-Stack DevelopmentObject-Oriented ProgrammingGitSQLPostgreSQLAWSCloud PlatformsDockerKubernetesTerraformHelmCI/CDGitHub ActionsLinuxPrometheusGrafanaDatadogDevOps
About the role
- You'll help ensure exceptional performance, reliability, and observability across our distributed environment.
Responsibilities
- Perform root cause analysis requiring deep code-level investigation and implement preventive measures
- Build automation and tooling to reduce operational toil and improve system reliability
- Maintain CI/CD pipelines, observability infrastructure, and database performance optimization
- Maintain on-call procedures and disaster recovery processes
- HiveWatch is seeking a Senior Site Reliability Engineer to join our Platform Team, where you'll build and operate mission-critical edge infrastructure that connects our SaaS platform to customer systems.
- Competitive compensation packages designed to reward top talent
Requirements
- 5+ years of software engineering experience with strong coding skills in production environments
- 3+ years of SRE, DevOps, or production operations experience
- Experience with Infrastructure as Code (Terraform, CloudFormation, or similar)
- Proficiency in at least one object oriented programming language in our tech stack (Java, Kotlin, Python)
- Hands-on experience with relational databases and SQL performance optimization
- Experience with monitoring and observability tools (Prometheus, Grafana, DataDog, or equivalent)
- Bachelor's degree in Computer Science, Engineering, or equivalent practical experience
Nice to have
- Strong experience with AWS architecture and services
- Experience in physical security, IoT, or edge computing environments
- Experience with advanced AWS services (Kinesis, Lambda, EKS, RDS)
- Experience with Terraform and Terragrunt specifically
- Background in high-availability, multi-tenant SaaS environments
- Experience participating in incident response and post-mortem processes
- Experience with edge computing and distributed system architectures
- Previous experience in a startup or high-growth environment (50-200 employees)
Skills
- Kotlin, Rust, TypeScript, Python
Compensation
- Eligible to participate in HiveWatch Equity Incentive Plan
Benefits
- At HiveWatch, we're passionate about taking care of our people - and it shows in the benefits we offer.
- Flexible paid time off so you can recharge when you need it
- Additional benefits include ClassPass credits and a discount on pet insurance
- Participate in a regular on-call rotation to provide 24/7 coverage for critical systems
Company info
- HiveWatch is a tech-forward, inclusive organization fostering the evolution of the physical security industry. We are a diverse team of forward thinkers who empower each other to find creative and collaborative solutions in an industry ripe for modernization. We are passionate about the problems we're solving for our customers and equally passionate about the company we're building.
- HiveWatch is here to help security teams pivot from chasing threats to preventing them. We protect organizations, people, and property through the intelligent orchestration of physical security programs. With better communication, more insights, and less "noise", we are modernizing what it means for businesses and their employees to truly feel safe.
- HiveWatch is seeking a Senior Site Reliability Engineer to join our Platform Team, where you'll build and operate mission-critical edge infrastructure that connects our SaaS platform to customer systems. You'll help ensure exceptional performance, reliability, and observability across our distributed environment.
- HiveWatch is a tech-forward, inclusive organization fostering the evolution of the physical security industry.
- We are a diverse team of forward thinkers who empower each other to find creative and collaborative solutions in an industry ripe for modernization.
- We are passionate about the problems we're solving for our customers and equally passionate about the company we're building.
- HiveWatch is here to help security teams pivot from chasing threats to preventing them.
- We protect organizations, people, and property through the intelligent orchestration of physical security programs.
- With better communication, more insights, and less "noise", we are modernizing what it means for businesses and their employees to truly feel safe.
- Improve the reliability of mission-critical systems including production monitoring, alerting, and capacity planning
- Debug and resolve complex production issues across the full stack, from infrastructure to application code
- Increase the resiliency, scalability, and maintainability of production environments
- Contribute to on-call runbooks, postmortems, and reliability best practices
- Languages: Kotlin, Rust, TypeScript, and Python
- Deployments: GitHub Actions, Terraform, Terragrunt, and Helm
- Infrastructure: AWS (Kinesis, Serverless, RDS, EKS), Kubernetes, Docker, Postgres, IoT Edge, Red Hat Enterprise Linux, Rocky Linux
Equal opportunity
- HiveWatch is an equal opportunity employer and we are committed to cultivating a work environment that supports, inspires, and respects all individuals.
- We execute our hiring practices so that they are merit-based and we do not discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity/expression, marital status, age, disability, medical condition, genetic information, national origin, ancestry, military or veteran status, or other protected characteristic.
Apply directly at HiveWatch →Create a free account for alerts like thisView HiveWatch immigration profile
This listing is sourced directly from HiveWatch's careers page and normalized into a canonical job model.