1

Biology Data Engineer Jobs (NOW HIRING)

Senior Staff Data Engineer

New York, NY · On-site +1

$270K - $338K/yr

The data that trains biological frontier models comes in dozens of modalities (sequences, images ... As a Senior Staff Data Engineer at Biohub, you'll be designing systems that ingest data from public ...

Sr. Data Engineer

Raleigh, NC · On-site

$111K - $133K/yr

Experience with scientific informatics solutions OR background in biological science, chemistry ... Data engineering related certification from cloud platform like Azure/AWS/Cloudera, and * Customer ...

Sr. Data Engineer

Raleigh, NC · On-site

$111K - $133K/yr

Experience with scientific informatics solutions OR background in biological science, chemistry ... Data engineering related certification from cloud platform like Azure/AWS/Cloudera, and * Customer ...

Sr. Data Engineer

Raleigh, NC · On-site

$111K - $133K/yr

Experience with scientific informatics solutions OR background in biological science, chemistry ... Data engineering related certification from cloud platform like Azure/AWS/Cloudera, and * Customer ...

Sr. Data Engineer

Milford, MA · On-site

$125K - $150K/yr

As a Sr Data Engineer, you will be responsible for implementing AI, process automation, data ... and biology. We collaborate with customers around the world to advance the release of effective ...

Sr. Data Engineer

Milford, MA · On-site

$125K - $150K/yr

As a Sr Data Engineer, you will be responsible for implementing AI, process automation, data ... and biology. We collaborate with customers around the world to advance the release of effective ...

Sr. Data Engineer

Milford, MA · On-site

$125K - $150K/yr

As a Sr Data Engineer, you will be responsible for implementing AI, process automation, data ... and biology. We collaborate with customers around the world to advance the release of effective ...

Biology with a computational emphasis * Or a related quantitative life sciences discipline ... Proficiency in Python, R, or another scientific programming language for biological data analysis.

Showing results 21-40

Biology Data Engineer information

See salary details

$44.5K

$129.7K

$177.5K

How much do biology data engineer jobs pay per year?

As of Sep 10, 2026, the average yearly pay for biology data engineer in the United States is $129,716.00, according to ZipRecruiter salary data. Most workers in this role earn between $114,500.00 and $137,500.00 per year, depending on experience, location, and employer.

What is a biology data engineer?

A Biology Data Engineer is a professional who designs, builds, and maintains data systems specifically for biological and life sciences research. They work with large and complex biological datasets, ensuring that data is efficiently collected, stored, and accessible for analysis. Their responsibilities often include creating data pipelines, integrating data from various sources like genomics, proteomics, or clinical studies, and ensuring data quality and security. Biology Data Engineers collaborate closely with bioinformaticians, researchers, and software developers to support scientific discovery. Their work is essential for enabling advanced analytics, such as machine learning, in biological research.

How do biology data engineers typically collaborate with biologists and other researchers on data projects?

Biology Data Engineers frequently work closely with biologists, bioinformaticians, and research scientists to understand the specific data requirements and biological context of projects. This often involves translating experimental needs into data pipelines, helping researchers manage large datasets, and ensuring data integrity and accessibility. Regular meetings, joint problem-solving sessions, and iterative feedback are common, enabling seamless integration of computational solutions with biological research. Strong communication skills and a willingness to learn domain-specific concepts are essential for success in this collaborative environment.

What are the key skills and qualifications needed to thrive as a biology data engineer, and why are they important?

To thrive as a Biology Data Engineer, you need a strong background in biology and computational data analysis, often supported by a degree in bioinformatics, computational biology, or computer science. Familiarity with programming languages (such as Python or R), biological databases, and data management platforms is typically required, as well as experience with cloud computing and big data tools. Strong problem-solving, collaboration, and communication skills are essential for translating complex biological data into actionable insights. These skills ensure the effective integration, analysis, and interpretation of large-scale biological datasets critical for research and innovation.

What is the difference between Biology Data Engineer vs Bioinformatics Data Scientist?

AspectBiology Data EngineerBioinformatics Data Scientist
Required CredentialsBachelor's or Master's in Biology, Data Science, or related fields; experience with data engineering toolsBachelor's or Master's in Bioinformatics, Computer Science, or related fields; strong programming skills
Work EnvironmentData pipelines, database management, cloud platforms in research or biotech companiesData analysis, algorithm development, research in healthcare or biotech sectors
Employer & Industry UsageBiotech firms, research institutions, pharmaceutical companiesResearch labs, healthcare organizations, biotech firms

The main difference between a Biology Data Engineer and a Bioinformatics Data Scientist lies in their focus areas. Biology Data Engineers primarily build and maintain data infrastructure for biological data, while Bioinformatics Data Scientists analyze and interpret biological data to derive insights. Both roles require strong technical skills and are vital in biotech and research industries, but they serve different functions within data management and analysis workflows.

What are popular job titles related to Biology Data Engineer jobs?

For Biology Data Engineer jobs, the most frequently searched job titles are:

Infographic showing various Biology Data Engineer job openings in the United States as of September 2026, with employment types broken down into 1% Internship, 1% As Needed, 83% Full Time, 12% Part Time, and 3% Contract. Highlights an 85% Physical, 3% Hybrid, and 12% Remote job distribution, with an average salary of $129,716 per year, or $62.4 per hour.

Senior Staff Data Engineer

New York, NY • On-site, Remote

$270K - $338K/yr

Full-time

Retirement, PTO

Re-posted 15 days ago


Key responsibilities

  • Design and build data pipelines that process genomic and imaging data at petabyte scale

  • Build agent-based systems for automated dataset curation, quality control, and workflow generation

  • Collaborate with AI Research teams and scientists to translate model requirements into data specifications and integrate data into large-scale AI-ready datasets


Job description

Biohub is the first large-scale initiative bringing frontier AI models, massive compute, and frontier experimental capabilities under one roof. We're building a general-purpose system to accelerate scientific discovery, integrating frontier AI models, biological foundation models, and lab capabilities, with the ultimate goal of curing disease. Our technology powers scientists around the world, translating AI capabilities into tools that accelerate research everywhere.
The Team
Biohub is a 501(c)(3) biomedical research organization building the first large-scale scientific initiative combining frontier AI with frontier biology to solve disease. We build the technology to help scientists around the world use AI-powered biology to study how cells operate, organize, and work as part of systems to understand why disease happens and how to correct it. With our compute capacity, AI research and engineering, and state-of-the-art technology for measuring, imaging, and programming biology, we are enabling scientists worldwide to use AI-powered biology to advance our understanding of human health.
The Opportunity
The role is part of the Data Engineering team, which focuses on owning the strategy, sourcing and implementation for data supporting AI research and development. Our goal is to maximize the speed, agility, and capability of biological AI research by connecting public data resources and Biohub's experimental platforms to AI systems. The data that trains biological frontier models comes in dozens of modalities (sequences, images, spatial coordinates, time series, molecular structures, metadata, publication artifacts) each with its own noise characteristics, biases, and information content. The question of how to represent this data for learning is one of the most important open problems in biological AI.
As a Senior Staff Data Engineer at Biohub, you'll be designing systems that ingest data from public repositories, transform heterogeneous biological formats into AI-ready datasets, combine that with proprietary datasets, and deliver training datasets to researchers pushing the boundaries of what's possible in biological AI. The infrastructure you build will directly shape what our models can learn.
We're a small team with significant resources and long time horizons. We use AI tools aggressively in our own work-Claude Code, agents for workflow automation, LLMs for metadata extraction. We care about code quality, operational reliability, and building systems that scale. And we care about the biology: we want engineers who can recognize when a pipeline output is technically correct but scientifically wrong.
If you want to work at the intersection of large-scale infrastructure and frontier science, with real autonomy and the chance to build something genuinely new, we'd like to talk.
What You'll Do
  • Design and build data pipelines that process genomic and imaging data at petabyte scale
  • Solve performance and bandwidth challenges with creative engineering
  • Build agent-based systems for automated dataset curation, quality control, and workflow generation
  • Create tooling for data cataloging and registration that makes datasets discoverable and accessible
  • Collaborate with AI Research teams to translate model requirements into data specifications, and with our scientists to integrate public and internal data into large-scale ai-ready datasets
  • Improve pipeline reliability and observability, working toward 99%+ success rates without manual intervention
What You'll Bring
  • 8+ years of experience building reliable, operable data systems at petabyte scale for Staff-level candidates; 12+ years for Senior Staff-level candidates.
  • Strong software engineering fundamentals
  • Experience deploying distributed computing frameworks like Databricks, Spark, or Ray for large-scale data processing
  • Experience building and deploying large scale data platform solutions using infrastructure as code like Terraform or CDK
  • Experience with cloud infrastructure (AWS preferred) and on prem infrastructure
  • Comfort with ambiguity; ability to make progress when requirements are evolving
  • Interest in AI-native development practices and tooling
  • Nice to have: Background in computational biology, bioinformatics, or life sciences and experience with genomics datasets and formats (FASTQ, BAM, VCF) or imaging formats (OME-Zarr, HDF5)
Compensation
The anticipated base pay ranges for this role in Redwood City, CA and New York City, NY are $241,000-$301,000 annually for the Staff level and $270,000-$338,000 annually for the Senior Staff level. Final compensation and placement within the applicable range are based on the level at which you are hired, as well as job-related skills, experience, and knowledge evaluated throughout the interview process.
Better Together
As we grow, we're excited to strengthen in-person connections and cultivate a collaborative, team-oriented environment. This role is a hybrid position requiring you to be onsite for at least 60% of the working month, approximately 3 days a week, with specific in-office days determined by the team's manager. The exact schedule will be at the hiring manager's discretion and communicated during the interview process.
Benefits for the Whole You
We're thankful to have an incredible team behind our work. To honor their commitment, we offer a wide range of benefits to support the people who make all we do possible.
  • Provides a generous employer match on employee 401(k) contributions to support planning for the future.
  • Paid time off to volunteer at an organization of your choice.
  • Funding for select family-forming benefits.
  • Relocation support for employees who need assistance moving

If you're interested in a role but your previous experience doesn't perfectly align with each qualification in the job description, we still encourage you to apply as you may be the perfect fit for this or another role.
#LI-Hybrid #LI-Onsite