1

Biology Data Engineer Jobs (NOW HIRING)

Scientific Data Engineer

Bodega Bay, CA · On-site

$135K - $163K/yr

Support machine learning and AI use of biological data by making it well-structured, documented, and efficiently accessible * Work closely with the community of developers of the Neurodata Without ...

Data Engineer

Bethesda, MD · On-site

$122K - $146K/yr

Black Canyon Consulting LLC is seeking Data Engineers to support their work for the National Center ... GKE, Google Store, Cloud functions • 5+ years of working with genetic and biological data • ...

Data Engineer

Bethesda, MD · On-site

$122K - $146K/yr

Black Canyon Consulting LLC is seeking Data Engineers to support their work for the National Center ... GKE, Google Store, Cloud functions • 5+ years of working with genetic and biological data • ...

Candidates with a strong proficiency in computer programming and big data management are preferred ... Assistant/Associate Professor (Computational Biology/Data Science) Posting Date:February 25, 2025 ...

Candidates with a strong proficiency in computer programming and big data management are preferred ... Assistant/Associate Professor (Computational Biology/Data Science) Posting Date:February 25, 2025 ...

Candidates with a strong proficiency in computer programming and big data management are preferred ... Assistant/Associate Professor (Computational Biology/Data Science) Posting Date:February 25, 2025 ...

Data Engineer

Durham, NC · On-site

$103K - $124K/yr

... complex biological and chemical interactions, predicts precision outcomes, and enables ... Position Summary The Data Engineer sits within the Data & Analytics organization and supports the ...

Data Engineer

Durham, NC · On-site

$110K - $132K/yr

... complex biological and chemical interactions, predicts precision outcomes, and enables next ... Position Summary The Data Engineer sits within the Data & Analytics organization and supports the ...

Data Engineer

Somerville, MA · On-site

$125K - $150K/yr

We call this biology's dark matter. It's signal-rich and mechanism-defining, yet almost entirely ... Position Overview Matterworks is seeking a Data Engineer to build and run the pipelines behind our ...

Data Engineer

Somerville, MA · On-site

$125K - $150K/yr

We call this biology's dark matter. It's signal-rich and mechanism-defining, yet almost entirely ... Position Overview Matterworks is seeking a Data Engineer to build and run the pipelines behind our ...

You will design and manage our ETL pipeline of diverse biological data, with an eye for both ... You will work closely with experimental scientists, device engineers, and software and operations ...

Data Engineer- Columbus OH

Columbus, OH · On-site

$110K - $132K/yr

... biology/chemistry/physics to develop sophisticated informatics solutions that drive efficiencies in content curation and workflow process. • Applies data transformation and other data-engineering ...

next page

Showing results 1-20

Biology Data Engineer information

See salary details

$44.5K

$129.7K

$177.5K

How much do biology data engineer jobs pay per year?

As of Sep 10, 2026, the average yearly pay for biology data engineer in the United States is $129,716.00, according to ZipRecruiter salary data. Most workers in this role earn between $114,500.00 and $137,500.00 per year, depending on experience, location, and employer.

What is a biology data engineer?

A Biology Data Engineer is a professional who designs, builds, and maintains data systems specifically for biological and life sciences research. They work with large and complex biological datasets, ensuring that data is efficiently collected, stored, and accessible for analysis. Their responsibilities often include creating data pipelines, integrating data from various sources like genomics, proteomics, or clinical studies, and ensuring data quality and security. Biology Data Engineers collaborate closely with bioinformaticians, researchers, and software developers to support scientific discovery. Their work is essential for enabling advanced analytics, such as machine learning, in biological research.

How do biology data engineers typically collaborate with biologists and other researchers on data projects?

Biology Data Engineers frequently work closely with biologists, bioinformaticians, and research scientists to understand the specific data requirements and biological context of projects. This often involves translating experimental needs into data pipelines, helping researchers manage large datasets, and ensuring data integrity and accessibility. Regular meetings, joint problem-solving sessions, and iterative feedback are common, enabling seamless integration of computational solutions with biological research. Strong communication skills and a willingness to learn domain-specific concepts are essential for success in this collaborative environment.

What are the key skills and qualifications needed to thrive as a biology data engineer, and why are they important?

To thrive as a Biology Data Engineer, you need a strong background in biology and computational data analysis, often supported by a degree in bioinformatics, computational biology, or computer science. Familiarity with programming languages (such as Python or R), biological databases, and data management platforms is typically required, as well as experience with cloud computing and big data tools. Strong problem-solving, collaboration, and communication skills are essential for translating complex biological data into actionable insights. These skills ensure the effective integration, analysis, and interpretation of large-scale biological datasets critical for research and innovation.

What is the difference between Biology Data Engineer vs Bioinformatics Data Scientist?

AspectBiology Data EngineerBioinformatics Data Scientist
Required CredentialsBachelor's or Master's in Biology, Data Science, or related fields; experience with data engineering toolsBachelor's or Master's in Bioinformatics, Computer Science, or related fields; strong programming skills
Work EnvironmentData pipelines, database management, cloud platforms in research or biotech companiesData analysis, algorithm development, research in healthcare or biotech sectors
Employer & Industry UsageBiotech firms, research institutions, pharmaceutical companiesResearch labs, healthcare organizations, biotech firms

The main difference between a Biology Data Engineer and a Bioinformatics Data Scientist lies in their focus areas. Biology Data Engineers primarily build and maintain data infrastructure for biological data, while Bioinformatics Data Scientists analyze and interpret biological data to derive insights. Both roles require strong technical skills and are vital in biotech and research industries, but they serve different functions within data management and analysis workflows.

What are popular job titles related to Biology Data Engineer jobs?

For Biology Data Engineer jobs, the most frequently searched job titles are:

Infographic showing various Biology Data Engineer job openings in the United States as of September 2026, with employment types broken down into 1% Internship, 1% As Needed, 83% Full Time, 12% Part Time, and 3% Contract. Highlights an 85% Physical, 3% Hybrid, and 12% Remote job distribution, with an average salary of $129,716 per year, or $62.4 per hour.

Scientific Data Engineer

Bodega Bay, CA • On-site

$135K - $163K/yr

Full-time

Medical, Retirement, PTO

Posted 19 days ago


Key responsibilities

  • Design and develop user-friendly software packages for scientific data management and analysis

  • Support machine learning and AI use of biological data by making it well-structured, documented, and efficiently accessible

  • Maintain and manage open source software products, including managing development priorities, software releases, continuous integration, and testing


Job description

Lawrence Berkeley National Laboratory is hiring a Scientific Data Engineer within the Scientific Data Division. 

The Computational Biosciences Group has an immediate opening for a software and data engineer in the area of multi-modal data modeling and analysis with applications to bioscience research. You will develop new methods and software tools that enable scientific knowledge discovery using modern data management and machine learning technologies and advance the state-of-the-art in data-intensive analysis. Your projects will focus on the domains of omics/structural biology data and neurophysiology data. Under limited instruction, you will be part of an experienced team conducting R&D in the areas of FAIR data science, AI, and modern methods for data understanding. You will be working as part of a multi-disciplinary team composed of computer scientists, data scientists, and bioinformaticians. Please note this is a scientific software/data engineering position- it is not a pure machine learning or AI research position, and it is not a pure data science or analytics position. 

You will:

  • Design and develop user-friendly software packages for scientific data management and analysis 

  • Work with domain experts to develop FAIR data models (i.e., models of the structure organization of the data) and management solutions for bioscience applications

  • Support machine learning and AI use of biological data by making it well-structured, documented, and efficiently accessible

  • Work closely with the community of developers of the Neurodata Without Borders and LinkML open source data ecosystems, as well as the Joint Genome Institute.

  • Maintain and manage open source software products, including managing development priorities, software releases, continuous integration, and testing

  • Design, implement and maintain high performance computing and cloud solutions for visualization and analysis of complex biological data

  • Develop machine learning and AI solutions for analysis of biological data in close collaboration with diverse teams of scientists 

  • Train scientists and research software engineers in the use of the developed software products at workshops and conferences

  • Demonstrate good judgment in selecting methods and techniques for obtaining solutions. 

  • Network with senior internal and external personnel in their own area of expertise.   

 

We are looking for:

  • Typically requires a minimum of 5 years of related experience with a Bachelor's degree in computer science, data science, machine learning, bioinformatics, or equivalent; or 3 years and a Master's degree; or equivalent work experience designing and developing software for data modeling or analysis; or a PhD in a relevant STEM field

  • Demonstrated experience developing software in a scientific or research context, such as in a research group, a scientific user facility, or on a scientific software project

  • Demonstrated hands-on experience in a production environment, developing scientific software, scientific data models, or scientific data pipelines

  • Strong programming experience in Python. Working proficiency in C++ or Javascript is a plus.

  • Experience testing large code bases

  • Experience contributing to community-driven open source software

  • Demonstrated experience in one or more of the following areas: data management, scientific data analysis, machine learning 

  • Works well in a collaborative team environment

  • Demonstrated capability with the Git version control and continuous integration systems, such as GitHub or GitLab

  • Ability to work effectively with domain scientists whose expertise is outside computing, and to translate their requirements into technical designs.

  • Excellent oral and written communication skills.

  • Demonstrated ability to work effectively as part of a cross-disciplinary team.

 

Desired skills/knowledge:

  • Master's or PhD in Computer Science or related field, with 5 or more years of professional experience designing and developing scientific data modeling or analysis software

  • Experience working with modern scientific data formats and database systems, such as HDF5, Zarr, MongoDB, PostgreSQL, MySQL, and Redis

  • Experience with Neurodata Without Borders, LinkML, or similar software ecosystems

  • Experience working with large biological data, such as in the areas of neurophysiology, microbiology, genomics, or protein design

  • Experience designing or working with structured data models, schemas, ontologies, or data standards

  • Familiarity with FAIR data principles, persistent identifiers, provenance, and controlled vocabularies and ontologies

  • Experience preparing scientific datasets for use by machine learning pipelines or LLM-based agents

  • Experience working with cloud object storage, cloud computing, High-Performance Computing, data lakehouse architecture, or containerization.

  • Experience developing web-based graphical user interfaces (GUIs) or application programming interfaces (APIs) for scientific data analysis and management

 

How to apply:

In addition to your resume, please include a brief cover letter (max 300 words) that describes one scientific software project you have contributed to, or one scientific data model, schema, standard, or pipeline you have helped design or implement. Describe your specific role and what the technical challenge was. Thesis, internship, research lab, and personal projects count. If the code is public, include a link.

We're here for the same mission, to bring science solutions to the world. Join our team and YOU will play a supporting role in our goal to address global challenges! Have a high level of impact and work for an organization associated with 17 Nobel Prizes!

Why join Berkeley Lab?

We invest in our employees by offering a total rewards package you can count on:

  • Exceptional health and retirement benefits, including pension or 401K-style plans

  • A culture where you'll belong - we are invested in our teams! 

  • In addition to accruing vacation and sick time, we also have a Winter Holiday Shutdown every year.

  • Parental bonding leave (for both mothers and fathers)

  • Pet insurance

Additional information:

  • Appointment type: This is a full-time, 2 years, term appointment with the possibility of extension or conversion to Career appointment based upon satisfactory job performance, continuing availability of funds and ongoing operational needs.

  • Salary range: The expected salary for this position is $131,760 - $161,064, which fits into the full salary of $117,132 - $197,676 depending upon the candidate's skills, knowledge, and abilities. This includes education, certifications, and years of experience.

  • Background check: This position is subject to a background check. Any convictions will be evaluated to determine if they directly relate to the responsibilities and requirements of the position. Having a conviction history will not automatically disqualify an applicant from being considered for employment.

  • Work modality: Work may be performed on-site, or hybrid. The primary location for this role is Lawrence Berkeley National Lab, 1 Cyclotron Road, Berkeley, CA. Work must be performed within the United States. A REAL ID or other acceptable form of identification is required to access Berkeley Lab sites (for more information click here).

Want to learn more about working at Berkeley Lab? Please visit: careers.lbl.gov

Equal Employment Opportunity Employer: The foundation of Berkeley Lab is our Stewardship Values: Team Science, Service, Trust, Innovation, and Respect; and we strive to build community with these shared values and commitments. Berkeley Lab is an Equal Opportunity Employer. We heartily welcome applications from all who could contribute to the Lab's mission of leading scientific discovery, excellence, and professionalism. In support of our rich global community, all qualified applicants will be considered for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, age, protected veteran status, or other protected categories under State and Federal law.

Misconduct Disclosure Requirement: As a condition of employment, the final candidate who accepts an offer of employment will be required to disclose if they have been subject to any final administrative or judicial decisions within the last seven years determining that they committed any misconduct; or have filed an appeal of a finding of substantiated misconduct with a previous employer. For additional information, click here.