Persona Identities
Senior Software Engineer, Data Platform
San Francisco · Senior · Contract
Sponsorship not specifiedDetected 22 days ago
PythonGoDistributed SystemsMySQLMongoDBGCPKubernetesRESTKafkaSparkData EngineeringData ScienceCommunicationCollaborationMentoring
About the role
- You'll work the full depth of that stack, from raw ingestion through storage layout and compute all the way to the query layer - owning the performance, reliability, and cost of each.
- This is a high-ownership role on a small team, working closely with the rest of engineering, data science, and infrastructure to make the platform fast, reliable, and cost-efficient at hundreds of billions of records per day.
- Operate the low-latency serving layer for real-time, high-concurrency access to platform data.
Responsibilities
- Build and scale our datalake - the pipelines that feed it and the storage design that keeps it fast and efficient.
- Develop high-throughput data-movement and ingestion services that bring data into the platform reliably and at scale.
- Own streaming and batch ingestion and the change-data-capture pipelines that absorb billions of records and terabytes of data per day.
- Build new query and analytics capabilities directly on the datalake - for use cases that don't fit a traditional warehouse.
- Partner with infrastructure and product teams to provision, secure, and scale the platform.
- Experience building and operating scalable data pipelines across structured and unstructured data.
Requirements
- 5+ years of experience in software engineering, with deep expertise in data infrastructure or large-scale distributed systems.
- Proficiency in Python - our primary language across pipelines.
- Hands-on experience with big-data or streaming technologies such as Spark, Kafka, or Flink, and interest in lakehouse table formats like Iceberg.
Nice to have
- ClickHouse, DuckDB, Google Cloud (GCP), Kubernetes, MySQL, MongoDB, Snowflake.
- This step is used solely to confirm that candidates are who they say they are, and will have no impact on hiring decisions.
- Nice to have: ClickHouse, DuckDB, Google Cloud (GCP), Kubernetes, MySQL, MongoDB, Snowflake.
Skills
- Excellent communication and collaboration skills across engineering, data science, and infrastructure teams.
Company info
- As part of our interview process, all candidates will be asked to verify their identity with Persona.
Apply directly at Persona Identities →Create a free account for alerts like thisView Persona Identities immigration profile
This listing is sourced directly from Persona Identities's careers page and normalized into a canonical job model.