1

Azure Data Fabric Jobs in Indiana (NOW HIRING)

Data Engineer

Indianapolis, IN

$109K - $131K/yr

Hands-on experience with cloud data platforms and pipeline tools; experience with Azure (Data ... Exposure to Microsoft Fabric, Power BI, or similar modern data and analytics platforms preferred.

Data Engineer

Indianapolis, IN · On-site

$109K - $131K/yr

Hands-on experience with cloud data platforms and pipeline tools; experience with Azure (Data ... Exposure to Microsoft Fabric, Power BI, or similar modern data and analytics platforms preferred.

Data Engineer

Carmel, IN

$108K - $130K/yr

Utilize CI/CD pipelines (Azure DevOps, Git-based workflows, Fabric deployment pipelines, etc.) to ... Develop scalable data models for investment reporting, portfolio analytics, and capital markets ...

Data Engineer

Carmel, IN · On-site

$108K - $130K/yr

Utilize CI/CD pipelines (Azure DevOps, Git-based workflows, Fabric deployment pipelines, etc.) to ... Develop scalable data models for investment reporting, portfolio analytics, and capital markets ...

Senior Data Engineer

Indianapolis, IN · Hybrid

$101K - $137K/yr

Relevant Microsoft Fabric, Azure Data, Databricks, cloud architecture, or data-engineering certifications are preferred. Security certifications such as CISSP, ISSAP, ISSEP, or equivalent ...

Senior Data Engineer

Indianapolis, IN · On-site

$101K - $137K/yr

Relevant Microsoft Fabric, Azure Data, Databricks, cloud architecture, or data-engineering certifications are preferred. Security certifications such as CISSP, ISSAP, ISSEP, or equivalent ...

Senior Data Engineer

Indianapolis, IN · Hybrid

$101K - $137K/yr

Relevant Microsoft Fabric, Azure Data, Databricks, cloud architecture, or data-engineering certifications are preferred. Security certifications such as CISSP, ISSAP, ISSEP, or equivalent ...

Data Architect

Indianapolis, IN · Remote

$61 - $78.50/hr

Microsoft Fabric Ecosystem - Practical experience with Fabric Lakehouse architecture, Delta Lake ... cloud environments (Azure preferred). * Data Governance & Metadata Management - Experience ...

Data Architect

Indianapolis, IN · Remote

$61 - $78.50/hr

Microsoft Fabric Ecosystem - Practical experience with Fabric Lakehouse architecture, Delta Lake ... cloud environments (Azure preferred). * Data Governance & Metadata Management - Experience ...

Data Architect

Indianapolis, IN · Remote

$61 - $78.50/hr

Microsoft Fabric Ecosystem - Practical experience with Fabric Lakehouse architecture, Delta Lake ... cloud environments (Azure preferred). * Data Governance & Metadata Management - Experience ...

Data Engineer (in person)

Westfield, IN · On-site

$109K - $131K/yr

... platforms (Azure preferred, AWS and GCP experience also valued) • Familiarity with modern data platforms such as Databricks, Snowflake, Microsoft Fabric, or Redshift • Familiarity with ...

Showing results 21-40

Azure Data Fabric information

What are some common challenges faced by professionals working with Azure Data Fabric, and how can they be addressed?

Professionals working with Azure Data Fabric often encounter challenges related to integrating diverse data sources, ensuring data security, and optimizing data pipeline performance. Managing data consistency and governance across distributed environments can also be complex. These challenges can be addressed by leveraging Azure's built-in monitoring and security tools, implementing robust data management policies, and staying up-to-date with best practices through continuous learning. Collaboration with cross-functional teams, such as data engineers, architects, and security specialists, is essential for successful project delivery.

What is Azure Data Fabric?

Azure Data Fabric is a set of cloud-based tools and services provided by Microsoft Azure to enable seamless data integration, management, and analytics across various data sources and environments. It helps organizations unify their data infrastructure, making it easier to collect, store, process, and analyze data from on-premises, cloud, and hybrid sources. Azure Data Fabric typically leverages services like Azure Data Factory, Azure Synapse Analytics, and Azure Data Lake to create a comprehensive data solution. This approach helps streamline data workflows, improve data governance, and accelerate business insights.

What is Azure Data Fabric used for?

Azure Data Fabric is a data management platform that enables organizations to integrate, govern, and analyze large volumes of data across hybrid and multi-cloud environments. It provides tools for data virtualization, security, and real-time processing, supporting data engineers and analysts in building scalable data solutions. Familiarity with cloud services and data architecture is beneficial for roles involving Azure Data Fabric.

What are the key skills and qualifications needed to thrive as an Azure Data Fabric engineer?

To thrive as an Azure Data Fabric Engineer, you need expertise in data architecture, data integration, and cloud computing, typically with a background in computer science or a related field. Proficiency with Microsoft Azure services (such as Azure Data Factory, Synapse Analytics, and Data Lake), along with relevant certifications like Azure Data Engineer Associate, is crucial. Strong analytical thinking, problem-solving skills, and effective communication help you collaborate across teams and address complex data challenges. These skills ensure secure, scalable, and efficient data solutions that empower organizations to make data-driven decisions.

What is the difference between Azure Data Fabric vs Data Engineer?

AspectAzure Data FabricData Engineer
Primary FocusData integration, management, and platform services on AzureDesigning, building, and maintaining data pipelines and infrastructure
Required SkillsAzure services, data architecture, cloud computingSQL, ETL, programming, data modeling
CertificationsAzure Data Engineer, Azure Solutions ArchitectMicrosoft Certified: Data Engineer Associate
Work EnvironmentCloud platforms, enterprise data solutionsData warehouses, cloud/on-premises environments

Azure Data Fabric focuses on providing a comprehensive data platform on Azure, enabling data integration and management. Data Engineers build and maintain data pipelines and infrastructure within such platforms. While Azure Data Fabric is a cloud service, Data Engineers work across various environments, often utilizing Azure services. Both roles require knowledge of data architecture and cloud tools, but their core responsibilities differ: platform management versus pipeline development.

What cities in Indiana are hiring for Azure Data Fabric jobs? Cities in Indiana with the most Azure Data Fabric job openings:

Agentic AI Data Engineer - CMC Data Integration

Initial Therapeutics, Inc.

Indianapolis, IN • On-site

$65.25 - $169.40/hr

Other

Medical, Dental, Vision, Life, Retirement, PTO

Posted 3 days ago

New


Job description

At Lilly, the work is demanding because patients are waiting. We unite caring with discovery to help make life better for people around the world, knowing that every decision, every detail, and every day matters. Headquartered in Indianapolis, Indiana, our over 50,000 employees around the globe take on complex challenges to discover and deliver life‑changing medicines, strengthen how health is understood and managed, and support the communities we serve. This is hard, urgent, selfless work—but it’s work worth doing. If you’re driven by purpose and ready to bring your best to work that truly matters for patients, we invite you to join us.

Overview

The Bioproduct Research and Development organization strives to deliver creative medicines to patients by developing and commercializing insulins, monoclonal antibodies, novel therapeutic proteins, peptides, oligonucleotide therapies, and gene therapy systems. This multidisciplinary group works collaboratively with our discovery and manufacturing colleagues.

We are seeking an AI Data Engineer to build the data ingestion infrastructure and a unified data model that underpins the modernized CMC Data Backbone. This is a hands‑on engineering role with design influence — you will write production‑quality pipelines, define CMC data schemas, and work directly with scientists and digital architects to ensure data from internal LIMS/ELN systems and external CDMO partners flows reliably into a single data backbone.

You will work with a team of engineers and data scientists. You will have the autonomy to own your components end‑to‑end. If you want hands‑on experience at the intersection of pharmaceutical science and modern agentic AI data engineering — agentic pipelines, document AI, GxP‑compliant data infrastructure — this is the role to build that foundation.

Key Responsibilities

Agentic Pipeline Components:

  • Implement individual agent components (e.g., document extraction agent, schema mapping agent, validation agent) within the established orchestration framework (LangGraph, LlamaIndex, or equivalent).
  • Write tool‑calling logic, handle failure modes, and ensure each agent component is testable and observable with instrumented logging of inputs, outputs, and intermediate decisions.
  • Iterate on agent behavior based on real data performance; work with the senior engineer to identify and resolve failure patterns.
  • Participate in validation and qualification activities for AI‑assisted workflows, supporting documentation that demonstrates computational tools reflect scientific intent.

Human‑in‑the‑Loop (HITL) Workflow Implementation:

  • Build review queues and flagging logic that surface low‑confidence or out‑of‑specification extractions to scientific reviewers for approval before data is loaded.
  • Implement routing logic that captures reviewer decisions, logs outcomes with full audit trail, and reintegrates approved data into the pipeline per 21 CFR Part 11 electronic records requirements.
  • Tune flagging thresholds based on feedback from scientific owners; maintain and improve HITL logic as new data sources are onboarded.

Data Ingestion & Pipeline Engineering:

  • Design and build AI‑assisted ingestion pipelines that extract and structure the data from unstructured CDMO/CRO data sources: PDFs (Certificates of Analysis, batch records), Excel files, and vendor portal exports.
  • Implement validation, reconciliation, and exception‑handling logic to ensure data completeness and integrity before loading.
  • Build monitoring and alerting for pipeline health, data quality, and ingestion failures.
  • Design a data quality framework with automated checks, rejection handling, and audit trail logging.
  • Develop reusable pipeline templates and schema documentation that reduce onboarding time for new CDMO partners.
Required Qualifications
  • MS or PhD in Computer Science, Computer Engineering, Data Engineering, or related technical field with 1–2 years of relevant experience; OR
  • BS in Computer Science or Computer Engineering with 3–5 years of hands‑on data engineering experience.
  • Proficiency in Python and SQL; ability to write, review, and own production‑quality code.
  • Demonstrated experience building ETL/ELT pipelines from unstructured or semi‑structured sources (PDFs, Excel, JSON, XML).
  • Hands‑on experience building LLM‑powered applications: retrieval‑augmented generation, tool‑calling, multi‑step orchestration, or equivalent agentic patterns.
  • Hands‑on experience with cloud data platforms: Azure (Data Factory, Databricks, Fabric) or AWS (S3, Glue, Lambda, Redshift).
  • Solid understanding of relational data modeling, schema design, and data normalization principles.
  • Familiarity with data orchestration tools (Airflow, Azure Data Factory, Prefect, or similar).
Additional Preferences
  • Working knowledge of 21 CFR Part 11, ALCOA+, and GxP data integrity principles, or clear demonstrated ability to apply similar audit/compliance frameworks.
  • Experience integrating data from LIMS, ELN, SDMS, or CDS systems (Benchling, LabVantage, OpenLABS, or equivalent).
  • Familiarity with pharmaceutical CMC data types: analytical results, batch records, stability studies, specifications.
  • Experience with data mesh architecture or data product ownership models.
  • Knowledge of MLOps practices and preparing data for AI/ML model training in regulated environments.
  • Exposure to regulatory submission data formats (eCTD, CTD, CDISC SEND/SDTM).
  • Experience with CI/CD pipelines (GitHub Actions, Azure DevOps) applied to data engineering workloads.

Lilly is dedicated to helping individuals with disabilities to actively engage in the workforce, ensuring equal opportunities when vying for positions. If you require accommodation to submit a resume for a position at Lilly, please complete the accommodation request form https://careers.lilly.com/us/en/workplace-accommodation for further assistance. Please note this is for individuals to request an accommodation as part of the application process and any other correspondence will not receive a response.

Lilly is proud to be an EEO Employer and does not discriminate on the basis of age, race, color, religion, gender identity, sex, gender expression, sexual orientation, genetic information, ancestry, national origin, protected veteran status, disability, or any other legally protected status.

Actual compensation will depend on a candidate’s education, experience, skills, and geographic location. The anticipated wage for this position is $65,250 - $169,400

Full‑time equivalent employees also will be eligible for a company bonus (depending, in part, on company and individual performance). In addition, Lilly offers a comprehensive benefit program to eligible employees, including eligibility to participate in a company‑sponsored 401(k); pension; vacation benefits; eligibility for medical, dental, vision and prescription drug benefits; flexible benefits (e.g., healthcare and/or dependent day care flexible spending accounts); life insurance and death benefits; certain time off and leave of absence benefits; and well‑being benefits (e.g., employee assistance program, fitness benefits, and employee clubs and activities). Lilly reserves the right to amend, modify, or terminate its compensation and benefit programs in its sole discretion and Lilly’s compensation practices and guidelines will apply regarding the details of any promotion or transfer of Lilly employees.

#J-18808-Ljbffr