Jobedly Post a Job

Senior Software Engineer, Data Platform

Profluent · Emeryville, CA
Full-timeBioinformaticsTechnology$136,000–$184,000/yr
Apply on Jobedly ⚡ One-click AI Apply

About the Senior Software Engineer Data Platform role

Senior Software Engineer Data Platform positions focus on delivering results in their domain. This page aggregates open Senior Software Engineer Data Platform roles and what employers typically expect.

Profluent is the frontier AI lab for biology. Profluent builds powerful foundation models for all of life's molecules, unlocking solutions that transform medicine, agriculture, and beyond. Founded in 2022 and headquartered in Emeryville, CA, Profluent is backed by leading investors including Altimeter Capital, Bezos Expeditions, Spark Capital, Insight Partners, Air Street Capital, AIX Ventures, and Convergent Ventures and has raised over $150M to date. We’re looking for a Senior Software Engineer to help design, build, and scale Profluent’s data platform. This platform houses data from protein engineering campaigns, including protein designs, experimental results, partner datasets, analytical outputs, and model-ready training data. It enables rapid machine learning, biological discovery, and secure collaboration across internal and external programs. This role is ideal for an engineer who enjoys building robust data systems: secure ingestion pipelines, well-structured warehouses, reliable data models, access controls, auditability, and infrastructure that makes complex scientific data usable at scale. You will work closely with ML, bioinformatics, and program teams to ensure Profluent’s data is organized, governed, accessible, and protected. Responsibilities Design, build, and maintain scalable data infrastructure for protein engineering campaigns, including ingestion, transformation, validation, storage, and retrieval of large scientific datasets Develop secure data pipelines for internal and partner-generated data, with strong attention to access control, data siloing, provenance, auditability, and compliance with data use restrictions Own core components of Profluent’s data warehouse and data platform, using Python, GCP, PostgreSQL, BigQuery, and related cloud-native technologies Build systems that transform raw experimental, computational, and partner data into structured, reliable, analysis-ready and model-ready datasets Establish best practices for data modeling, metadata management, data quality checks, schema evolution, versioning, and documentation Collaborate with ML engineers, computational biologists, data scientists, and program stakeholders to understand data requirements and translate them into scalable technical systems Improve engineering quality through thoughtful system design, code review, testing, CI/CD, observability, and maintainable development workflows Contribute to architectural decisions for how Profluent stores, secures, organizes, and uses data across programs and partnerships Qualifications 5+ years of software engineering, data engineering, or data platform experience Strong proficiency in Python and modern software development practices, including git, testing, code review, CI/CD, and production deployment Experience designing and operating production data pipelines, data warehouses, and data models at scale Hands-on experience with cloud platforms, preferably GCP, and technologies such as BigQuery, PostgreSQL, object storage, workflow orchestration, and containerized services Strong understanding of data security, access control, data partitioning or siloing, audit logging, and managing sensitive or restricted datasets Experience working with complex, heterogeneous datasets and building systems that make them reliable, discoverable, and usable Ability to work independently, make sound technical decisions, and drive projects from ambiguous requirements to production systems BS, MS, or PhD in Computer Science, Engineering, Data Science, Bioinformatics, or a related technical field, or equivalent practical experience Preferences Experience with scientific, biological, clinical, genomic, laboratory, or high-throughput experimental data Experience managing external partner, customer, or restricted-access datasets Familiarity with data governance, lineage, metadata systems, schema registries, or data catalogs Experience with research data systems, LIMS, ELNs, Benchling, or adjacent scientific platf…

Salary estimate

$136,000 – $184,000/yr
Provided by the employer.

Skills for this role

PythonPostgresqlGCPCi/CdGITSparkMachine LearningData ScienceSecurity

Resume tips for Senior Software Engineer Data Platform applicants

Interview preparation

Prepare concrete STAR-format stories that show Senior Software Engineer Data Platform outcomes you drove.

Research the employer's product and recent news before the interview.

Be ready to explain how you'd approach a typical Senior Software Engineer Data Platform problem end to end.

Have thoughtful questions ready about the team, tools and success metrics.

About Profluent

Profluent is actively hiring on Jobedly. Explore their open roles and what it's like to work there.

Apply on Jobedly ⚡ One-click AI Apply

Similar jobs

Companies hiring for similar roles