Vynca Inc
Site Reliability Engineer
Remote - United States
Sponsorship not specifiedDetected 1 day ago
PythonGoDistributed SystemsPostgreSQLMySQLSnowflakeRedshiftAWSCloud PlatformsKubernetesTerraformHelmCI/CDLinuxPrometheusGrafanaDatadogDevOpsSite Reliability EngineeringPlatform EngineeringData EngineeringCybersecurityIncident ResponseCompliance
> stay_score
odds of building a lasting career here
37Risky
Cap-exempt (no lottery)0
Sponsors this role80
Entry-level history0
PERM / green-card track0
Lottery odds40
Fits your clock70
Thin sponsorship signal and lottery-bound. A low-probability bet with your clock running. Prioritize cap-exempt roles and proven entry-level sponsors first.
Lottery odds assume a STEM candidate.
Personalize to your clock →> community_outcomes
No reports yet — be the first to help the next applicant.
About the role
- Join the dynamic journey at Vynca, where we're passionate about transforming care for individuals with complex needs.
- Our shared commitment to caring for each other and those we serve is what sets us apart.
- Guided by our unwavering core values: Excellence, Compassion, Curiosity, and Integrity, we forge paths of success together.
Responsibilities
- Design, provision, and manage AWS infrastructure using Terraform as the source of truth.
- Operate, maintain, and scale production workloads running on Kubernetes.
- Package, deploy, and manage applications using Helm and infrastructure automation tools.
- Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms.
- Define, monitor, and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets to balance reliability and engineering velocity.
- Develop automation for deployment, scaling, monitoring, incident response, and operational workflows to reduce manual effort and improve system resilience.
- Own platform observability by implementing and maintaining metrics, logging, tracing, monitoring, and alerting solutions.
- Lead incident response efforts, facilitate blameless postmortems, and drive long-term corrective actions that improve system reliability.
- Partner with Product and Engineering teams on capacity planning, performance optimization, and resilient system design.
- Implement and maintain security best practices to support HIPAA, SOC 2, and other compliance requirements.
Requirements
- Strong hands-on experience operating production workloads within AWS environments.
- Proven experience managing infrastructure as code using Terraform, including module development, state management, and deployment automation.
- Experience operating and supporting production Kubernetes environments.
- Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault tolerance.
- Experience establishing and managing observability practices including monitoring, logging, tracing, alerting, and incident response.
- Strong understanding of Linux systems administration, networking, cloud architecture, and distributed systems fundamentals.
- Experience designing, implementing, and maintaining CI/CD pipelines and deployment automation.
- Strong problem-solving skills with the ability to troubleshoot complex infrastructure and application issues.
Nice to have
- Strong programming or scripting experience with Python, Go, or similar languages.
- Experience with observability platforms such as Prometheus, Grafana, Datadog, CloudWatch, SigNoz, or OpenTelemetry.
- Experience with GitOps tools such as ArgoCD or Flux.
- Experience managing databases such as PostgreSQL, MySQL, Redshift, or ClickHouse.
- Experience implementing secrets management solutions such as AWS Secrets Manager or HashiCorp Vault.
- Familiarity with data infrastructure technologies including Snowflake, Redshift, and ETL/ELT pipelines.
- Experience with database performance tuning and optimization.
- The hiring process for this role may consist of applying, followed by a phone screen, online assessment(s), interview(s), an offer, and background/reference checks.
Benefits
- We're looking for a Site Reliability Engineer (E3) to help build and operate the infrastructure that powers Vynca's healthcare technology platform.
- You'll play a critical role in maintaining the health of our production environment while helping shape the future architecture of our systems.
- Education: Bachelor's degree in Computer Science, Information Systems, Software Engineering, or a related technical field
Company info
- At Vynca, our mission is to provide comprehensive care for more quality days at home.
- At this time we are only considering applicants in the following states: Arizona, California, Colorado, Florida, Georgia, Illinois, Nevada, North Carolina, Oregon, Texas, Utah and Washington.
Equal opportunity
- Equal Opportunity Employer: At Vynca Inc., we embrace diversity and are committed to fostering an inclusive workplace.
Visa & Work Authorization
- ender identity, gender expression, sexual orientation, marital status, veteran status, disability, genetic information, citizenship status, or membership in any other protected group under federal, state, or local law
Apply directly at Vynca Inc →Create a free account for alerts like thisView Vynca Inc immigration profile
This listing is sourced directly from Vynca Inc's careers page and normalized into a canonical job model.