favorited

favorited

Member of Technical Staff - Infrastructure

Santa Monica, CA · Staff+ · Full-time

Sponsorship not specified$150k-$200kDetected 30 days ago
PythonGoGCPCloud PlatformsDockerKubernetesPrometheusGrafanaDatadogDevOpsSite Reliability EngineeringAgentic AIIncident ResponseTCP/IPCommunicationCollaboration

About the role

  • We are looking for a Member of Technical Staff to help ensure the reliability, scalability, and performance of the infrastructure that powers favorited's real-time platform.

Responsibilities

  • Design, implement, and maintain highly reliable and scalable infrastructure supporting real-time applications.
  • Build automation and tooling to improve system reliability, deployment processes, and operational efficiency.
  • Develop and maintain monitoring, logging, and alerting systems to ensure high availability and rapid incident response.
  • Partner closely with engineering teams to improve service reliability, performance, and observability.
  • Support incident response, root cause analysis, and postmortems, ensuring learnings are incorporated into system improvements.
  • Optimize infrastructure for performance, cost efficiency, and scalability.
  • Manage and scale containerized environments using Docker, Kubernetes, and related orchestration technologies.
  • You will play a key role in building and maintaining systems that support high-traffic applications used by a rapidly growing global audience.

Requirements

  • 6+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering roles.
  • Experience managing infrastructure for large-scale systems supporting millions of users.
  • Strong expertise with cloud infrastructure, ideally Google Cloud Platform (GCP).
  • Hands-on experience with Kubernetes, container orchestration, and distributed systems.
  • Experience implementing monitoring and observability systems (Dash0, Prometheus, Grafana, Datadog, or similar).
  • Strong scripting or programming experience in languages such as Python, Go, or TypeScript.
  • Strong collaboration skills and ability to work cross-functionally with engineering teams.

Nice to have

  • Experience supporting real-time streaming, gaming, or large-scale consumer applications.
  • Familiarity with event-driven architectures and large-scale data processing systems.
  • Experience optimizing infrastructure costs in high-growth environments.
  • Practical experience integrating AI agents into DevOps workflows.
  • 401(k) plan to invest in your future.
  • Paid company holidays for time to recharge.

Compensation

  • $150k - $200k base salary + options.
  • Competitive salary that values your expertise and contributions.

Benefits

  • Unlimited PTO to prioritize work-life balance.
  • Comprehensive health insurance to support your well-being.
  • Benefits Include:

Company info

  • This is a full-time, on-site position in Santa Monica.
  • EEOC Statement
  • We are an equal opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability status, protected veteran status, or any other characteristic protected by law.

This listing is sourced directly from favorited's careers page and normalized into a canonical job model.