1

Assistant Data Curation Jobs (NOW HIRING)

Data Scientist 2

Annapolis, MD · On-site

$122K - $168K/yr

Data Processing: (Data management and curation, data description and visualization, workflow and ... conversely, assist others with drawing appropriate conclusions from the analysis of such data.

Data Scientist 3

Annapolis, MD · On-site

$161K - $211K/yr

Data Processing: (Data management and curation, data description and visualization, workflow and ... conversely, assist others with drawing appropriate conclusions from the analysis of such data.

next page

Showing results 1-20

Assistant Data Curation information

What is the difference between Assistant Data Curation vs Data Analyst?

AspectAssistant Data CurationData Analyst
Primary FocusOrganizing, cleaning, and maintaining data setsAnalyzing data to extract insights and support decision-making
Skills & CertificationsData management, basic SQL, attention to detailStatistical analysis, data visualization, SQL, Excel
Work EnvironmentData repositories, database systems, data teamsBusiness units, analytics teams, reporting platforms
Typical EmployerTech companies, research institutions, data-driven organizationsFinance, marketing, consulting firms, tech companies

Assistant Data Curation primarily involves preparing and maintaining data for analysis, focusing on data quality and organization. Data Analysts, on the other hand, interpret data to generate insights and support strategic decisions. While both roles require data management skills, Data Analysts typically have stronger analytical and statistical expertise. Understanding these differences helps job seekers identify the right role based on their skills and career goals.

More about Assistant Data Curation jobs
What cities are hiring for Assistant Data Curation jobs? Cities with the most Assistant Data Curation job openings:
What are the most commonly searched types of Data Curation jobs? The most popular types of Data Curation jobs are:
What states have the most Assistant Data Curation jobs? States with the most job openings for Assistant Data Curation jobs include:
Infographic showing various Assistant Data Curation job openings in the United States as of July 2026, with employment types broken down into 80% Full Time, 10% Part Time, and 10% Contract. Highlights an 70% In-person, and 30% Remote job distribution.
Sr. Data Expert, Data Engineer (Clinical & Multimodal Data Integration)

Sr. Data Expert, Data Engineer (Clinical & Multimodal Data Integration)

Genentech, Inc.

South San Francisco, CA • On-site

$137K - $165K/yr

Full-time

This job post has expired 1 day ago. Applications are no longer accepted.


Genentech rating

8.8

Company rating: 8.8 out of 10

Based on 22 frontline employees who took The Breakroom Quiz

8th of 74 rated pharmaceutical


Job description

A healthier future. It's what drives us to innovate. To continuously advance science and ensure everyone has access to the healthcare they need today and for generations to come. Creating a world where we all have more time with the people we love. That's what makes us Roche.
Advances in AI, data and computational sciences are transforming drug discovery and development. Roche's Research and Early Development organizations at Genentech (gRED) and Pharma (pRED) have demonstrated how these technologies accelerate R&D, leveraging data and novel computational models to drive impact. Seamless data sharing and access to models across gRED and pRED are essential to maximising these opportunities. The Computational Sciences Center of Excellence (CS CoE) is a strategic, unified group whose goal is to harness the transformative power of data and Artificial Intelligence (AI) to assist our scientists in both pRED and gRED to deliver more innovative and life-changing medicines for patients worldwide.
The Computational Sciences Center of Excellence (CS CoE) brings together data, AI, and computational expertise to accelerate innovation across gRED and pRED. Within CS CoE, the Data and Digital Catalyst (DDC) organization leads the modernization of our data ecosystem, enabling scalable, data-driven science.
The Data Capability organization within DDC is responsible for establishing foundational data capabilities, including data connectivity, data compliance, scientific content management and data ingestion, curation, integration, and delivery. The team ensures that high-quality, well-structured datasets are available to power analytics, AI/ML, and scientific discovery across Research and Early Development.
The Opportunity:
We are seeking a Sr. Data Expert, Data Engineer to lead the integration and delivery of clinically anchored, multimodal scientific datasets spanning clinical, sequencing, imaging, proteomics, and other emerging data modalities.
In this role, you will:
  • Lead the integration and harmonization of clinical and multimodal scientific datasets, applying industry data standards and metadata frameworks to improve interoperability and scientific usability.
  • Own end-to-end data delivery by designing, validating, documenting, and delivering high-quality, analysis-ready datasets that support research, AI/ML, and computational biology initiatives.
  • Develop scalable data workflows that automate data ingestion, quality control, transformation, and metadata management across diverse scientific data sources.
  • Partner with computational scientists, bioinformaticians, and data engineers to understand scientific requirements and translate them into scalable, reusable data solutions.
  • Drive data quality and continuous improvement by implementing validation frameworks, metadata standards, and AI-assisted data curation practices that improve data discoverability and reuse.
  • Support emerging AI and foundation model initiatives by preparing interoperable, metadata-rich datasets optimized for downstream analytics and machine learning applications.

Who You Are:
  • You have a PhD with 2+ years, a Master's degree with 3-5 years, or a Bachelor's degree with 5+ years of experience in Bioinformatics, Data Science, Biomedical Engineering, Computer Science, Clinical Sciences, or a related discipline, with experience working with clinical, biomedical, or scientific datasets.
  • You have hands-on experience integrating clinical data with one or more scientific modalities, including sequencing, imaging, proteomics, or other omics datasets, and understand clinical data models and longitudinal patient data.
  • You are proficient in Python (Pandas), SQL, and scientific data processing, with experience working with scientific data formats such as FASTQ, BAM/CRAM, VCF, DICOM, AnnData, or Parquet, and familiarity with cloud data platforms (AWS or GCP).
  • You have experience developing or supporting automated data pipelines using workflow orchestration tools such as Airflow, Nextflow, Snakemake, or Prefect, and are comfortable using Git for collaborative software development.
  • You are a collaborative problem solver with a strong focus on data quality, metadata management, and scientific reproducibility, and enjoy partnering with multidisciplinary teams to deliver scalable data solutions.

Preferred Qualifications:
  • Experience with clinical data standards such as CDISC (SDTM/ADaM), OMOP, or FHIR.
  • Experience integrating multimodal datasets (e.g., clinical + genomics, imaging + transcriptomics, or multi-omics).
  • Familiarity with biomedical ontologies, controlled vocabularies, FAIR data principles, and metadata standards.
  • Experience preparing scientific datasets for AI/ML workflows or foundation model development.
  • Experience supporting translational research, biomarker discovery, or drug discovery programs.

Onsite presence, on our South San Francisco campus, is expected for at least 3 days a week.
Relocation benefits are not available for this job posting.
The expected salary range for this position based on the primary location of California is $119,800 - $222,400. Actual pay will be determined based on experience, qualifications, geographic location, and other job-related factors permitted by law. A discretionary annual bonus may be available based on individual and Company performance. This position also qualifies for the benefits detailed at the link provided below.
Benefits
#LI-JD1
#ComputationCoE
Genentech is an equal opportunity employer. It is our policy and practice to employ, promote, and otherwise treat any and all employees and applicants on the basis of merit, qualifications, and competence. The company's policy prohibits unlawful discrimination, including but not limited to, discrimination on the basis of Protected Veteran status, individuals with disabilities status, and consistent with all federal, state, or local laws.
If you have a disability and need an accommodation in relation to the online application process, please contact us by completing this form Accommodations for Applicants.

What Genentech employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Genentech logo

About Genentech

Sourced by ZipRecruiter

A member of the Roche Group, Genentech has been at the forefront of the biotechnology industry for more than 40 years, using human genetic information to develop novel medicines for serious and life-threatening diseases. Genentech has multiple therapies on the market for cancer & other serious illnesses. Please take this opportunity to learn about Genentech where we believe that our employees are our most important asset & are dedicated to remaining a great place to work.

Industry

Scientific research and development services

Company size

10,000+ Employees

Headquarters location

South San Francisco, CA, US

Year founded

1976

Social media