Commure
Staff Software Engineer, Data Warehouse
Mountain View, CA · Staff+
Stay score
odds of building a lasting career here
Thin sponsorship signal and lottery-bound. A low-probability bet with your clock running. Prioritize cap-exempt roles and proven entry-level sponsors first.
Lottery odds assume a STEM candidate.
Personalize to your clock →Employer immigration record
from this employer's Department of Labor filings
Green-card intent detected
Files H-1B transfers
Sourced from Department of Labor LCA, PERM and prevailing-wage disclosure data. Employer matching is by name, so figures may be split across an employer's legal entities. Absence of a filing means none appears in our copy of the data, not that none exists.
Community outcomes
No reports yet — be the first to help the next applicant.
About the role
- The core stack today: Debezium for CDC, StarRocks as our MPP query and serving engine, and dbt for transformation and modeling.
- This is a hands-on IC role with broad scope.
Responsibilities
- Own the data warehouse platform end-to-end: CDC pipelines, data lake, query layer, transformation layer, and the analytics-facing tooling that sits on top.
- Design and operate CDC pipelines with Debezium (and Kafka, Redpanda, or an equivalent streaming backbone) that move data from operational databases into the warehouse with low latency and high fidelity.
- Run and scale StarRocks (or adjacent MPP/lakehouse engines) as the query and serving layer - schema design, materialized views, ingestion patterns, tuning, and cost/performance trade-offs.
- Build the transformation layer with dbt: modeling standards, tests, documentation, and a semantic layer that gives every team a single source of truth for metrics.
- Partner with Security and Compliance on PHI/PII handling, access controls, lineage, and auditability so the platform meets HIPAA and SOC 2 bar by default.
- Set patterns and conventions: schema contracts, ingestion patterns, and self-serve tooling - that let product and analytics teams build on the platform without needing you in the loop for every decision.
- 6+ years of software engineering experience, with significant time building or operating data platforms at scale.
- Fluent in SQL, schema design, query optimization, and reasoning about cost and latency trade-offs on large datasets.
- Experience building semantic layers (dbt Semantic Layer, Cube) or data catalogs / lineage (DataHub, OpenMetadata, Amundsen).
- Our team works directly alongside clinicians, not through layers of process, which means the gap between what you build and its impact on patient care is immediate.
Requirements
- Experience across multiple clouds (AWS, GCP, Azure), infrastructure-as-code (Terraform, Pulumi) and Kubernetes controllers.
Nice to have
- Direct experience with Debezium, StarRocks, and dbt in production.
- Experience with HIPAA-regulated data (PHI handling, de-identification, and access governance).
Compensation
- Today, 500,000+ clinicians across 500+ healthcare organizations nationwide trust Commure to handle $25B+ in annual claims and support over 200 million patient interactions.
Benefits
- At Commure, we're building the AI Operating System for healthcare, the foundation that defines how care is delivered, documented, and financed.
- Healthcare carries a $1 trillion administrative burden and we're at the center of transforming it.
- Today, 500,000+ clinicians across 500+ healthcare organizations nationwide trust Commure to handle $25B+ in annual claims and support over 200 million patient interactions.
- The future of healthcare is being built right now.
This listing is sourced directly from Commure's careers page and normalized into a canonical job model.