Profluent
Senior Software Engineer, Data Platform
Emeryville, California, United States; Hybrid (2-3 days on-site) · Senior
Work authorization required$170k-$220kDetected 30 days ago
PythonCode ReviewGitPostgreSQLBigQueryGCPCloud PlatformsCI/CDMachine LearningSparkData EngineeringData ScienceAI OrchestrationBioinformaticsCRISPRResearchCollaboration
About the role
- It enables rapid machine learning, biological discovery, and secure collaboration across internal and external programs.
- You will work closely with ML, bioinformatics, and program teams to ensure Profluent's data is organized, governed, accessible, and protected.
- Sponsorship will not be provided now or at any time in the future for this position.
Responsibilities
- Design, build, and maintain scalable data infrastructure for protein engineering campaigns, including ingestion, transformation, validation, storage, and retrieval of large scientific datasets
- Develop secure data pipelines for internal and partner-generated data, with strong attention to access control, data siloing, provenance, auditability, and compliance with data use restrictions
- Own core components of Profluent's data warehouse and data platform, using Python, GCP, PostgreSQL, BigQuery, and related cloud-native technologies
- Build systems that transform raw experimental, computational, and partner data into structured, reliable, analysis-ready and model-ready datasets
- Collaborate with ML engineers, computational biologists, data scientists, and program stakeholders to understand data requirements and translate them into scalable technical systems
- Improve engineering quality through thoughtful system design, code review, testing, CI/CD, observability, and maintainable development workflows
- Profluent is an AI-first protein design company.
- This platform houses data from protein engineering campaigns, including protein designs, experimental results, partner datasets, analytical outputs, and model-ready training data.
- High-growth opportunity with meaningful impact on the future of protein design
Requirements
- 5+ years of software engineering, data engineering, or data platform experience
- Strong proficiency in Python and modern software development practices, including git, testing, code review, CI/CD, and production deployment
- Experience designing and operating production data pipelines, data warehouses, and data models at scale
- Hands-on experience with cloud platforms, preferably GCP, and technologies such as BigQuery, PostgreSQL, object storage, workflow orchestration, and containerized services
- Strong understanding of data security, access control, data partitioning or siloing, audit logging, and managing sensitive or restricted datasets
- BS, MS, or PhD in Computer Science, Engineering, Data Science, Bioinformatics, or a related technical field, or equivalent practical experience
- Preferences (but not required)
- Experience with scientific, biological, clinical, genomic, laboratory, or high-throughput experimental data
- Familiarity with data governance, lineage, metadata systems, schema registries, or data catalogs
- Experience with research data systems, LIMS, ELNs, Benchling, or adjacent scientific platforms
- Experience working with complex, heterogeneous datasets and building systems that make them reliable, discoverable, and usable
- Ability to work independently, make sound technical decisions, and drive projects from ambiguous requirements to production systems
- Experience managing external partner, customer, or restricted-access datasets
- Background working with ML, data science, computational biology, or cross-disciplinary technical teams
- Interest in learning biology, gene editing, protein design, or machine learning concepts
Compensation
- Competitive compensation package with equity participation
Benefits
- Competitive compensation package with equity participation
- Comprehensive benefits including health/dental/vision insurance
- Generous PTO policy and commitment to work-life balance
Company info
- Founded in 2022, we develop deep generative models to design and validate novel, functional proteins to revolutionize biomedicine.
- Based in Emeryville, CA, we are backed by leading investors including Altimeter Capital, Bezos Expeditions, Spark Capital, Insight Partners, Air Street Capital, AIX Ventures, and Convergent Ventures, and have raised over $150M to date.
- We're looking for a Senior Software Engineer to help design, build, and scale Profluent's data platform.
Equal opportunity
- equal opportunity employer promoting diversity and inclusion in the workspace.
Visa & Work Authorization
- Applicants must have ongoing work authorization in the United States that does not require employer sponsorship.
- Sponsorship will not be provided now or at any time in the future for this position.
Apply directly at Profluent →Create a free account for alerts like thisView Profluent immigration profile
This listing is sourced directly from Profluent's careers page and normalized into a canonical job model.