The RWDData Engineering is atechnicalrole thatsupportsthe end-to-end engineering vision for Lilly ... Collaborate closely with multi-functional partners - data scientists, statisticians, analytics ...
The RWDData Engineering is atechnicalrole thatsupportsthe end-to-end engineering vision for Lilly ... Collaborate closely with multi-functional partners - data scientists, statisticians, analytics ...
Managed Services - Senior Data Engineer - Manager
Indianapolis, IN · On-site
$99K - $232K/yr
... Research, Statistics, Engineering - Demonstrating proficiency in Business Intelligence and Reporting Tools (BIRT) - Utilizing Python and Java for data engineering tasks - Excelling in data ...
Managed Services - Senior Data Engineer - Manager
Indianapolis, IN · On-site
$99K - $232K/yr
... Research, Statistics, Engineering - Demonstrating proficiency in Business Intelligence and Reporting Tools (BIRT) - Utilizing Python and Java for data engineering tasks - Excelling in data ...
Engineering Services also provides technical consulting and support for other divisions within ... statistics for all Indiana University campuses, is available online . You may also request a ...
Engineering Services also provides technical consulting and support for other divisions within ... statistics for all Indiana University campuses, is available online . You may also request a ...
Proficiency in programming languages like Python and R. * Experience with BI tools such as Tableau ... Statistical package expertise in IBM SPSS/PASW, R, STATA or MS Excel statistics. Compensation ...
Proficiency in programming languages like Python and R. * Experience with BI tools such as Tableau ... Statistical package expertise in IBM SPSS/PASW, R, STATA or MS Excel statistics. Compensation ...
Completes Kaizen activity that applies statistical and other improvement tools. * Meets with ... Bachelor's degree in Manufacturing Engineering, Engineering Technology, Building Construction ...
Completes Kaizen activity that applies statistical and other improvement tools. * Meets with ... Bachelor's degree in Manufacturing Engineering, Engineering Technology, Building Construction ...
Associate Director of RWD Engineering
Indianapolis, IN · On-site
$241K/yr
Collaborate closely with multi-functional partners - data scientists, statisticians, analytics ... Core Engineering Skills * High proficiency in SQL optimization across cloud platforms - complex ...
Associate Director of RWD Engineering
Indianapolis, IN · On-site
$241K/yr
Collaborate closely with multi-functional partners - data scientists, statisticians, analytics ... Core Engineering Skills * High proficiency in SQL optimization across cloud platforms - complex ...
Completes Kaizen activity that applies statistical and other improvement tools. * Meets with ... Bachelor's degree in Manufacturing Engineering, Engineering Technology, Building Construction ...
Completes Kaizen activity that applies statistical and other improvement tools. * Meets with ... Bachelor's degree in Manufacturing Engineering, Engineering Technology, Building Construction ...
Prepare detailed reports and compile statistics, including the comparison and research, with ... * Assist engineers in performing work such as; preparing detailed drawings, memorandum of ...
Prepare detailed reports and compile statistics, including the comparison and research, with ... * Assist engineers in performing work such as; preparing detailed drawings, memorandum of ...
Engineering Technician
Madison, IN · On-site
$1.2K - $1.9K/wk
Engineering Technician Department: Maintenance Employment Type: Full Time Location: Clifty Creek ... Prepare detailed reports and compile statistics, including the comparison and research, with ...
Engineering Technician
Madison, IN · On-site
$1.2K - $1.9K/wk
Engineering Technician Department: Maintenance Employment Type: Full Time Location: Clifty Creek ... Prepare detailed reports and compile statistics, including the comparison and research, with ...
Senior Quality Engineer
$80K - $109K/yr
Strong background in quality control and quality engineering functions * Understanding GD&T * Experience with Statistical Software * Experience with Gage Repeatability and Reproducibility * Excellent ...
Senior Quality Engineer
$80K - $109K/yr
Strong background in quality control and quality engineering functions * Understanding GD&T * Experience with Statistical Software * Experience with Gage Repeatability and Reproducibility * Excellent ...
Managed Services - Senior Data Engineer - Senior Manager
Indianapolis, IN · On-site
$124K - $280K/yr
... Statistics, Engineering - Demonstrating advanced skills in data architecture development and data pipeline management - Utilizing Python and Java for complex data analysis and visualization ...
Managed Services - Senior Data Engineer - Senior Manager
Indianapolis, IN · On-site
$124K - $280K/yr
... Statistics, Engineering - Demonstrating advanced skills in data architecture development and data pipeline management - Utilizing Python and Java for complex data analysis and visualization ...
Manufacturing Process Improvement Engineer
Indianapolis, IN · On-site
$35 - $38/hr
This role focuses on applying data-driven manufacturing engineering methods to identify root causes ... Utilize statistical analysis and process control tools such as SPC, capability studies, regression ...
Quick apply
Manufacturing Process Improvement Engineer
Indianapolis, IN · On-site
$35 - $38/hr
This role focuses on applying data-driven manufacturing engineering methods to identify root causes ... Utilize statistical analysis and process control tools such as SPC, capability studies, regression ...
Quality Engineering Manager
Jasper, IN · On-site
The Quality Engineering Manager position will include direct communications with the customer(s ... Utilization of Statistical Analysis & Tools * Chi Square, ANOVA, DOE, SPC * ESD knowledge, GD&T ...
Quality Engineering Manager
Jasper, IN · On-site
The Quality Engineering Manager position will include direct communications with the customer(s ... Utilization of Statistical Analysis & Tools * Chi Square, ANOVA, DOE, SPC * ESD knowledge, GD&T ...
Proficiency in Microsoft Office tools and statistical software such as Minitab. * Ability to create ... Working knowledge of engineering documentation, document control, and revision processes. * Ability ...
Proficiency in Microsoft Office tools and statistical software such as Minitab. * Ability to create ... Working knowledge of engineering documentation, document control, and revision processes. * Ability ...
Collects production data and applies engineering principles, scientific and statistical methods to analyze, document, and diagram production processes. A manufacturing engineer is responsible for ...
Collects production data and applies engineering principles, scientific and statistical methods to analyze, document, and diagram production processes. A manufacturing engineer is responsible for ...
Facility Engineer - Manufacturing Engineering, Body
Lafayette, IN · On-site
$70K - $90K/yr
Collects production data and applies engineering principles, scientific and statistical methods to analyze, document, and diagram production processes. A manufacturing engineer is responsible for ...
Facility Engineer - Manufacturing Engineering, Body
Lafayette, IN · On-site
$70K - $90K/yr
Collects production data and applies engineering principles, scientific and statistical methods to analyze, document, and diagram production processes. A manufacturing engineer is responsible for ...
Senior Quality Engineer
Plymouth, IN · On-site
$80K - $109K/yr
Strong background in quality control and quality engineering functions * Understanding GD&T * Experience with Statistical Software * Experience with Gage Repeatability and Reproducibility * Excellent ...
Senior Quality Engineer
Plymouth, IN · On-site
$80K - $109K/yr
Strong background in quality control and quality engineering functions * Understanding GD&T * Experience with Statistical Software * Experience with Gage Repeatability and Reproducibility * Excellent ...
Experience with statistical process control (SPC) and quality management systems (QMS). * Certification in quality assurance or relevant engineering disciplines. * 3D Printing Experience, e.g., hands ...
Experience with statistical process control (SPC) and quality management systems (QMS). * Certification in quality assurance or relevant engineering disciplines. * 3D Printing Experience, e.g., hands ...
Manufacturing Engineer
Warsaw, IN · On-site
$70K - $90K/yr
... and statistical tools BOM & Router creation and update Ability to read and understand engineering drawings Knowledge of basic quality tools, risk analysis (PFMEA), statistics (SPC), Critical-to ...
Quick apply
Manufacturing Engineer
Warsaw, IN · On-site
$70K - $90K/yr
... and statistical tools BOM & Router creation and update Ability to read and understand engineering drawings Knowledge of basic quality tools, risk analysis (PFMEA), statistics (SPC), Critical-to ...
CTIO AI Engineering Manager
Indianapolis, IN · On-site
$73K - $244K/yr
Those in data science and machine learning engineering at PwC will focus on leveraging advanced ... You will work on developing predictive models, conducting statistical analysis, and creating data ...
CTIO AI Engineering Manager
Indianapolis, IN · On-site
$73K - $244K/yr
Those in data science and machine learning engineering at PwC will focus on leveraging advanced ... You will work on developing predictive models, conducting statistical analysis, and creating data ...
Statistical Engineering information
See Indiana salary details
$58.7K - $60.3K
6% of jobs
$60.3K - $62K
8% of jobs
$63.6K is the 25th percentile. Wages below this are outliers.
$62K - $63.6K
10% of jobs
$63.6K - $65.3K
8% of jobs
$65.3K - $66.9K
10% of jobs
The median wage is $68.2K / yr.
$66.9K - $68.6K
8% of jobs
$68.6K - $70.2K
10% of jobs
$70.2K - $71.9K
8% of jobs
$72.5K is the 75th percentile. Wages above this are outliers.
$71.9K - $73.5K
10% of jobs
$73.5K - $75.2K
8% of jobs
$75.2K - $76.8K
10% of jobs
$58.7K
$68.8K
$76.8K
How much do statistical engineering jobs pay per year?
What are the key skills and qualifications needed to thrive as a statistical engineer, and why are they important?
What is the difference between Statistical Engineering vs Data Scientist?
| Aspect | Statistical Engineering | Data Scientist |
|---|---|---|
| Required credentials | Statistics, Data Analysis, Engineering | Statistics, Computer Science, Data Analysis |
| Work environment | Manufacturing, R&D, Engineering teams | Business, Tech, Research sectors |
| Employer usage | Optimizing processes, designing experiments | Building models, insights, predictive analytics |
Statistical Engineering focuses on applying statistical methods to improve engineering processes and product development, often within manufacturing or R&D settings. Data Scientists analyze large datasets to extract insights, build predictive models, and support business decisions. While both roles require strong statistical skills, Statistical Engineering emphasizes process optimization and experimental design, whereas Data Scientists focus on data-driven insights across diverse industries.
How does a statistical engineer typically collaborate with cross-functional teams to implement data-driven solutions?
What is statistical engineering?
Full-time
Posted 21 days ago
Eli Lilly and Company rating
8.8
Based on 63 frontline employees who took The Breakroom Quiz
11th of 86 rated pharmaceutical
Job description
At Lilly, the work is demanding because patients are waiting. We unite caring with discovery to help make life better for people around the world, knowing that every decision, every detail, and every day matters. Headquartered in Indianapolis, Indiana, our over 50,000 employees around the globe take on complex challenges to discover and deliver life-changing medicines, strengthen how health is understood and managed, and support the communities we serve. This is hard, urgent, selfless work-but it's work worth doing. If you're driven by purpose and ready to bring your best to work that truly matters for patients, we invite you to join us.
Purpose:
The RWDData Engineering is atechnicalrole thatsupportsthe end-to-end engineering vision for Lilly's real-world data (RWD) infrastructure.These individual supportsdesignandexecutesthe scalable, cloud-native pipelines and data products that allow HEOR, SDIA, Statisticians, Medical, and Clinical teams to generate evidence faster, more reproducibly, and at greater scientific depth than is possible through traditional vendor engagements.
This job involves a depth of understanding ofthe multi-modal RWD ecosystemacross Lilly's therapeutic areasto contextualize and drive RWD to the necessary end-user data productsbysupportingthe creation of sophisticated data engineering products, creating processes for automation of data ingestion and product creation, and leading special projects forGlobal Medical Affairs - Health Economics and Outcomes functions and the broader enterprise.Further, this position willbe responsible foridentifyingand advocating standard processes across the data asset lifecycle, working closely with other datadomainand analytics leaders. Collaborating closely with multi-functional teams, you willleadthe technical implementation of data products, ensuring scalability, reliability, and performance. The ideal candidatepossessesdeepexpertisein data engineering, strong problem-solving skills, and a passion forleveragingdata to drive business outcomes. This position works with the Sr. Director -RWD Architecture - Engineering.
This positionreports to HEOR Central andis embedded within the BIA organization and works in close partnership with HEOR, SDIA, Statisticians, Medical, and Clinical teams. It reports to the Director -RWData Engineering.
Responsibilities:This job description is intended to provide a general overview of the job requirements at the time it was prepared. The job requirements of any role/position can change over time and can includeadditionalresponsibilities not specifically described in the job description. Consult with your supervisorregardingyour actual job responsibilities and any related duties that might berequiredfor the role/position.
- Lead the design, development, and implementation of cloud-native data products and high-throughput data pipelines that transform raw real-world data into scalable, reliable, analysis-ready assets supporting analytics, reporting, and evidence generation.
- Lead the Analytic Data Products Strategy to deliver key data assets that enable streamlined, compliant execution and analytics.
- Own the end-to-end lifecycle of RWD data products, from requirements gathering and prototyping through production deployment and optimization, ensuring scalability, reliability, performance, and reproducibility across cloud environments (e.g., Databricks, AWS S3, Azure Data Lake).
- Build,optimize, and maintain ETL/ELT ingestion and transformation pipelines for large-scale, multi-modal RWD - including claims, complex EHR data, and other linked healthcare datasets - handling data volumes ranging from tens of millions to billions of records.
- Implement and managelakehouse-style data architectures (e.g., medallion bronze/silver/gold patterns) using Databricks and cloud object storage (AWS S3, ADLS) to produce versioned, partitioned, and audit-ready data assets.
- Write andmaintainreusable, version-controlled transformation logic incorporating healthcare coding and terminology standards (e.g., ICD-10/ICD-9, NDC,RxNorm, SNOMED, CPT/HCPCS,LOINC) to produce domain-level datasets such as demographics, diagnoses, treatments, procedures, encounters, and labs.
- OptimizeSQL and distributed processing workloads (e.g., Spark-based jobs) for performance acrossvery largedatasets, applying partitioning, indexing, predicate pushdown, denormalization, and other optimization strategiesappropriate toanalytical workloads.
- Translate analytic, business, and research requirements into reproducible data extraction and transformation logic, supporting cohort construction, temporal logic, and consistent reuse of RWD across teams.
- Apply deep understanding of healthcare data structures and standards when engineering data products, ensuring datasets are fit for purpose for downstream analytics and compliant with scientific, regulatory, and audit expectations.
- Establish and implement standard engineering practices andmethodologyacross the data asset lifecycle, including automated data ingestion, data quality checks, integrity testing, validation, monitoring, alerting, and documentation from source table to analysis-ready output.
- Contribute to CI/CD pipeline setup, code review, and testing standards, ensuring all transformation code is version-controlled, tested, and deployable in a reproducible manner.
- Collaborate closely with multi-functional partners - data scientists, statisticians, analytics leaders, and other technical teams - to understand business and technical requirements and develop documentation of RWD engineering standards, transformation templates, code list repositories, and pipeline performance guidelines.
- Provide technical consultation to collaborators on appropriate use of data products and underlying RWD assets, including structural limitations of specific data sources, join strategies, and performance considerations; collaborate with the RWD Operations Lead to develop source-specific training materials for HEOR scientists, SDIA, and statisticians.
- Createan inclusive culturewhereproducing andmaintaininghigh-quality data is a core discipline.
Technical Skills:
Core Engineering Skills
- Highproficiencyin SQL optimization across cloud platforms - complex joins, window functions, query tuning, workload management - on AWS Redshift, Databricks SQL, Snowflake, orBigQuery.
- Python fluency: pandas,PySpark, Polars for large-scale data manipulation; workflow orchestration with Apache Airflow, Prefect, orDagsterfor production pipeline scheduling and monitoring - including Databricks Workflows for orchestrating multi-task jobs within the Lakehouse.
- Distributed computing: Apache Spark (PySpark),Dask, or Ray - ability to write, tune, and debug distributed jobs processing multi-terabyte datasets across partitioned cloud storage, including Databricks clusters with auto-scaling and spot instance optimization.
- Cloud data engineering: hands-on pipeline development on AWS (S3, Glue, Redshift, EMR), Azure (ADLS, Synapse, ADF), or GCP (BigQuery, Dataflow) - as well as Databricks on any major cloud (AWS, Azure, or GCP) using Unity Catalog for cross-workspace governance - not just configuration.
- Delta Lake / Apache Iceberg: time travel, schema evolution,upsert/merge operations, partition optimization - building versioned, ACID-compliant data assets at scale; Delta Lake experience ideally hands-on within the Databricks Lakehouse platform using Delta Live Tables (DLT) for declarative pipeline authoring.
- ETL/ELT tooling: AWS Glue,dbt,dbt-databricks, or equivalent for building tested, documented transformation pipelines with automated data quality checks; familiarity with Databricks Asset Bundles (DABs) for packaging and deploying notebooks, jobs, and DLT pipelines as code.
- DevOps fluency: Git, CI/CD pipelines (GitHub Actions, Azure DevOps), Docker -maintainingall pipeline code as version-controlled, testable, and deployable artifacts; experience deploying to Databricks via CLI, REST API, or Terraformprovidera plus.
- Strong fluency with data modeling: designing star/snowflake schemas, OMOP-compliant structures, and flat domain long filesoptimizedfor analytical workloads in a research context - including materializing these structures as managed Delta tables within Databricks Unity Catalog withappropriate accesscontrols and lineage tracking.
- R fluencya plusfor collaboration with statistical and HEOR teams on dataset validation and specification.
Healthcare RWD -reviews and EHR
- Medical and pharmacy claims: deep working knowledge of CCAE, OptumClinformatics, IQVIAPharMetrics, Truveta, and Komodo data structures - enrollment/eligibility tables, revenue codes, place-of-service codes, inpatient vs. outpatient claim splitting, drug identification at NDC and GPI level, and known structural quirks of each source.
- Complex EHR data: HL7 FHIR resources, Epic/Cerner/Truveta data models, clinical note schemas, problem list hierarchies, medication order and administration tables, lab result normalization, and vital sign time series - including semi-structured and nested JSON/XML from EHR exports and FHIR APIs.
- Healthcare terminology and ontology mapping: ICD-10-CM/PCS, ICD-9-CM, NDC,RxNorm, SNOMED CT, CPT-4, HCPCS, LOINC, ATC - building, versioning, and governing code list repositories joined to raw tables to produce concept-labeled analysis-ready datasets.
- OMOP CDM: transforming source claims and EHR data to OMOP v5.x including vocabulary loading (Athena), ETL specification documentation, and Achilles/DQD data quality checks.
- Phenotyping and cohort construction: translating clinical study protocols into reproducible extraction logic - index date derivation, washout periods, time-varying covariates, censoring - suitable for pharmacoepidemiology and HEOR studies.
- Regulatory and privacy frameworks: HIPAA-compliant data handling, de-identification standards (Safe Harbor, Expert Determination), DUA compliance, and audit trail requirements for FDA-grade RWE submissions.
AI Fluency
- Deploy NLP pipelines for structured extraction from unstructured clinical notes - operationalizing pre-trained biomedical language models (BioBERT,ClinicalBERT) for outcome ascertainment and phenotyping within the data pipeline.
- Implement AI-powered data profiling and anomaly detection to automate quality checks across large ingestion runs and surface issues before they reach analytical teams.
- Use LLM-assisted SQL generation and code review tools to accelerate pipeline development and reduce query errors at scale.
- Apply intelligent caching, query result reuse, and automated feature engineering tooptimizecompute costs and prepare multi-modal variables for downstream ML inputs.
- Familiarity withMLOpstooling (MLflow, SageMaker, Azure ML) for versioning and monitoring AI models integrated with data pipelines.
Minimum Qualification Requirements:
- Bachelor's degree in Computer Science, Engineering, Statistics, Information Technology, Bioinformatics, or Technical Field.
- Minimum of 3 years of hands-on data engineering experience with a demonstrated focus on healthcare or life sciences RWD.
- Minimum of 3 of applied expertise across Python,SQLJava, Spring, Spring Boot, Prefect, and/or other business intelligence tools, ,ETL/ELT pipelines, and cloud platforms (AWS Glue/EMR, Snowflake, or Databricks),- applied directly to real-world healthcare data at scale.
Additional Preferences:
- Master's degree in Computer Science, Engineering, Statistics, Information Technology, Bioinformatics.
- Experience with distributed computing frameworks (Spark,Dask) for large-scale RWD processing.
- Deep understanding of healthcare coding standards (ICD-10, NDC,RxNorm, SNOMED CT, CPT, LOINC), real-world data structures, and major RWD vendors and platforms (Truveta, Optum, IQVIA, Komodo,HealthVerity).
- Familiarity with DevOps and CI/CD practices relevant to data pipeline development and deployment.
- Strong problem-solving skills, attention to detail, and ability to work effectively in matrixed, cross-functional teams with both technical and scientific partners.
- Demonstrated ability to build production-grade ingestion pipelines for multi-terabyte, multi-modal datasets: claims, complexher..
- Advanced SQL optimization skills for AWS Redshift and/or S3/Databricks-based architectures, including query tuning and workload management.
- Experience with OMOP CDM transformation, vocabulary management (Athena), and federated analytics networks (OHDSI,PCORnet, Sentinel).
- Familiarity with bioinformatics workflow managers (Snakemake,Nextflow, WDL) and genomics cloud platforms (DNAnexus, Terra, AWS Genomics CLI).
- Experience with data pipeline observability and quality tooling (dbt, Great Expectations, Monte Carlo).
- Familiarity with FDA RWE Framework guidance, EMA RWD guidance, and audit-readiness requirements for observational studies.
- Knowledge of Agile/Scrum methodologies and project management tools such as JIRA
- Knowledge and experience with -omics RWD:Genomics(e.g., ingesting and processing VCF/GVCF files, PLINK binary formats (bed/bim/fam),Transcriptomics(e.g.,bulk RNA-seq count matrices (featureCounts, STAR/RSEM), single-cell and single-nucleus RNA-seq (AnnData/h5ad, 10x GenomicsCellRanger) - ingestion, normalization, and linkage to clinical phenotype data;Proteomics(e.g.,DIA/DDA mass spectrometry output (MaxQuant, DIA-NN) and OLINK).
- Recruitment open across sectors: pharma/biotech data teams, RWD vendors (Truveta, Komodo, Optum, IQVIA,HealthVerity), cloud providers (AWS, Google, Microsoft), genomics platforms (Illumina,DNAnexus, Broad Institute), or academic health systems with large-scale RWD programs.
Other Information:
- Location: Indianapolis, IN is strongly preferred. Relocation assistance is available for qualified candidates. Remote work may be considered based on business needs and candidate qualifications.
Lilly is dedicated to helping individuals with disabilities to actively engage in the workforce, ensuring equal opportunities when vying for positions. If you require accommodation to submit a resume for a position at Lilly, please complete the accommodation request form (https://careers.lilly.com/us/en/workplace-accommodation) for further assistance. Please note this is for individuals to request an accommodation as part of the application process and any other correspondence will not receiv...
What Eli Lilly and Company employees say
Pay
Benefits
Hours and flexibility
Workplace
Get the full story on Breakroom
About Eli Lilly
Sourced by ZipRecruiter
Eli Lilly, based in Indianapolis, IN, US, is one of the pioneers in the pharmaceutical industry with a rich history dating back to 1876. This global pharmaceutical company focuses on discovering, developing, manufacturing and selling pharmaceutical products in approximately 120 countries. The company's product categories include endocrinology, oncology, cardiovascular, neuroscience, and immunology. Having invested over $9 billion in research and development in the past decade, Eli Lilly is also committed to creating high-quality medicines that meet real needs. As a recipient of several awards and recognitions, Eli Lilly is known for its focus on life-saving research and drug development. Their mission is to make medicines that help people live longer, healthier, and more active lives.
Industry
Pharmaceutical product wholesalers
Company size
10,000+ Employees
Headquarters location
Indianapolis, IN, US
Year founded
1876