JobLarper
JobLarperCompaniesProfluent Bio › Senior Software Engineer, Data Platform

Senior Software Engineer, Data Platform

Profluent Bio — Tracked from its greenhouse job board

Emeryville, California, United States; Hybrid (2-3 days on-site) Senior+ Posted May 5, 2026
PythonSQLREST/APIsCloudML

About the role

Profluent is the frontier AI lab for biology. Profluent builds powerful foundation models for all of life's molecules, unlocking solutions that transform medicine, agriculture, and beyond. Founded in 2022 and headquartered in Emeryville, CA, Profluent is backed by leading investors including Altimeter Capital, Bezos Expeditions, Spark Capital, Insight Partners, Air Street Capital, AIX Ventures, and Convergent Ventures and has raised over $150M to date.

We’re looking for a Senior Software Engineer to help design, build, and scale Profluent’s data platform. This platform houses data from protein engineering campaigns, including protein designs, experimental results, partner datasets, analytical outputs, and model-ready training data. It enables rapid machine learning, biological discovery, and secure collaboration across internal and external programs.

This role is ideal for an engineer who enjoys building robust data systems: secure ingestion pipelines, well-structured warehouses, reliable data models, access controls, auditability, and infrastructure that makes complex scientific data usable at scale. You will work closely with ML, bioinformatics, and program teams to ensure Profluent’s data is organized, governed, accessible, and protected.

Responsibilities

Design, build, and maintain scalable data infrastructure for protein engineering campaigns, including ingestion, transformation, validation, storage, and retrieval of large scientific datasets

Develop secure data pipelines for internal and partner-generated data, with strong attention to access control, data siloing, provenance, auditability, and compliance with data use restrictions

Own core components of Profluent’s data warehouse and data platform, using Python, GCP, PostgreSQL, BigQuery, and related cloud-native technologies

Build systems that transform raw experimental, computational, and partner data into structured, reliable, analysis-ready and model-ready datasets

Establish best practices for data modeling, metadata management, data quality checks, schema evolution, versioning, and documentation

Collaborate with ML engineers, computational biologists, data scientists, and program stakeholders to understand data requirements and translate them into scalable technical systems

Improve engineering quality through thoughtful system design, code review, testing, CI/CD, observability, and maintainable development workflows

Contribute to architectural decisions for how Profluent stores, secures, organizes, and uses data across programs and partnerships

Qualifications

5+ years of software engineering, data engineering, or data platform experience

Strong proficiency in Python and modern software development practices, including git, testing, code review, CI/CD, and production deployment

Experience designing and operating production data pipelines, data warehouses, and data models at scale

Hands-on experience with cloud platforms, preferably GCP, and technologies such as BigQuery, PostgreSQL, object storage, workflow orchestration, and containerized services

Strong understanding of data security, access control, data partitioning or siloing, audit logging, and managing sensitive or restricted datasets

Experience working with complex, heterogeneous datasets and building systems that make them reliable, discoverable, and usable

Ability to work independently, make sound technical decisions, and drive projects from ambiguous requirements to production systems

BS, MS, or PhD in Computer Science, Engineering, Data Science, Bioinformatics, or a related technical field, or equivalent practical experience

Preferences

Experience with scientific, biological, clinical, genomic, laboratory, or high-throughput experimental data

Experience managing external partner, customer, or restricted-access datasets

Familiarity with data governance, lineage, metadata systems, schema registries, or data catalogs

Experience with research data systems, LIMS, ELNs, Benchling, or adjacent scientific platforms

Background working with ML, data science, computational biology, or cross-disciplinary technical teams

Interest in learning biology, gene editing, protein design, or machine learning concepts

What We Offer

High-growth opportunity with meaningful impact on the future of protein design

Competitive compensation package with equity participation

401(k) with a strong employer match

Comprehensive benefits including health/dental/vision insurance

Generous PTO policy and commitment to work-life balance

Professional development opportunities in a cutting-edge field at the intersection of AI and biology

Profluent Bio, Inc is an equal opportunity employer promoting diversity and inclusion in the workspace. We do not discriminate on the basis of race, color, religion, marital status, age, national origin, ancestry, physical or mental disability, medical conditions, veteran status, sexual orientation, gender (including gender identity and gender expression), sex (which includes pregnancy, childbirth, and breastfeeding), genetic information, taking or requesting statutorily protected leave, or any other basis protected by law.

Work Authorization Requirement

Applicants must have ongoing work authorization in the United States that does not require employer sponsorship. Sponsorship will not be provided now or at any time in the future for this position.

Employment Eligibility Verification

Legal authorization to work in the United States is required. In compliance with federal law, all persons hired must verify their identity and work eligibility and complete the required employment verification form upon hire.
Hiring Salary Range
$170,000 $220,000 USD

What the index says about this role

  • First seen by JobLarper — Aug 2, 2026, 5 days ago. Older postings collect hundreds of applicants — a tailored résumé matters more the longer a role has been live.
  • No pay range in our index for this listing. Across 1,685 indexed Software Engineer roles in US that do publish one, the middle half sits between $185k and $285k, median $225k — JobLarper's read of the market, not a figure from Profluent Bio.
  • What Software Engineer roles ask for — across 10,230 indexed openings: REST/APIs (46%), Backend (42%), Cloud (38%), Python (38%), Go (25%). This posting names REST/APIs, Cloud, Python, ML.
  • Profluent Bio is hiring actively — 7 open roles indexed.

Derived from the 27,000 roles JobLarper indexes daily from official company boards — not from the job description above.

⚡ JobLarper watched this role appear on Profluent Bio's official board on May 5, 2026. Sign up free to get alerted minutes after roles like this go live, and tailor your real résumé to the exact description — nothing invented.

More open roles at Profluent Bio

All 7 open roles at Profluent Bio →

Similar Software Engineer roles at other companies