1

Data Engineer Jobs in York, PA (NOW HIRING)

Data Engineer

Lititz, PA ยท On-site

$106K - $127K/yr

The Data Engineer will work closely with the Data Architect and partner across Enterprise Technology, Architecture/AI, Integration, Infrastructure and Cloud, Information Security, and business teams ...

Data Engineer

Lititz, PA ยท On-site

$106K - $127K/yr

The Data Engineer will work closely with the Data Architect and partner across Enterprise Technology, Architecture/AI, Integration, Infrastructure and Cloud, Information Security, and business teams ...

Data Engineer - onsite

Hanover, PA ยท On-site

$110K - $132K/yr

Data Engineer Location: Hanover, MD Type: Contract Compensation: 55-70/hr W-2, depending on experience Work Model: Onsite US Citizenship required by federal client Responsibilities * Assign a ...

Data Engineer - onsite

Hanover, PA ยท On-site

$110K - $132K/yr

Data Engineer Location: Hanover, MD Type: Contract Compensation: 55-70/hr W-2, depending on experience Work Model: Onsite US Citizenship required by federal client Responsibilities * Assign a ...

Data Engineer - onsite

Hanover, PA ยท On-site

$110K - $132K/yr

Data Engineer Location: Hanover, MD Type: Contract Compensation: 55-70/hr W-2, depending on experience Work Model: Onsite US Citizenship required by federal client Responsibilities * Assign a ...

Data Engineer - onsite

Hanover, PA ยท On-site

$110K - $132K/yr

Data Engineer Location: Hanover, MD Type: Contract Compensation: 55-70/hr W-2, depending on experience Work Model: Onsite US Citizenship required by federal client Responsibilities * Assign a ...

Data Engineer - onsite

Hanover, PA ยท On-site

$110K - $132K/yr

Data Engineer Location: Hanover, MD Type: Contract Compensation: 55-70/hr W-2, depending on experience Work Model: Onsite US Citizenship required by federal client Responsibilities * Assign a ...

Data Engineer - onsite

Hanover, PA ยท On-site

$110K - $132K/yr

Data Engineer Location: Hanover, MD Type: Contract Compensation: 55-70/hr W-2, depending on experience Work Model: Onsite US Citizenship required by federal client Responsibilities * Assign a ...

Data Engineer Job

Lancaster, PA ยท On-site +1

$125K - $145K/yr

Proficiency in SQL and at least one programming language (Python, Scala, or Java). * Experience with data lakehouse technologies (Databricks, Snowflake). * Familiarity with data pipeline ...

Data Engineer Job

Lancaster, PA ยท On-site +1

$125K - $145K/yr

Armstrong is seeking a Data Engineer to design, build, and maintain scalable data pipelines and infrastructure that enable advanced analytics, AI, and data-driven decision-making across the ...

Data Engineer Job

Lancaster, PA ยท On-site +1

$125K - $145K/yr

Armstrong is seeking a Data Engineer to design, build, and maintain scalable data pipelines and infrastructure that enable advanced analytics, AI, and data-driven decision-making across the ...

Data Engineer Job

Lancaster, PA ยท On-site

$125K - $145K/yr

Armstrong is seeking a Data Engineer to design, build, and maintain scalable data pipelines and infrastructure that enable advanced analytics, AI, and data-driven decision-making across the ...

Data Engineer[Hybrid-remote]

Harrisburg, PA ยท On-site

$113K - $135K/yr

(Data Engineer) Location: Harrisburg, PA Visa: Any visa except CPT, TN and GC (PP number is mandatory to fetch I94) Interview Type: webcam interview Local candidates only DCED seeks a mid-level Data ...

OT Data Engineer Summary: The OT Data Engineer - Unified Namespace is responsible for designing and implementing scalable industrial data solutions that enable real-time visibility and integration ...

next page

Showing results 1-20

Data Engineer information

See York, PA salary details

$43.8K

$127.7K

$174.7K

How much do data engineer jobs pay per year?

As of Aug 28, 2026, the average yearly pay for data engineer in York, PA is $127,663.00, according to ZipRecruiter salary data. Most workers in this role earn between $112,700.00 and $135,300.00 per year, depending on experience, location, and employer.

What is a data engineer?

Data Engineers are IT professionals who design, construct, install, and maintain large-scale processing systems and other infrastructure for collecting, storing, and analyzing data. They build and optimize data pipelines and architectures that allow organizations to efficiently access and use data for business insights. Data Engineers work closely with data scientists, analysts, and other stakeholders to ensure that data is reliable, accessible, and secure. Their responsibilities often include working with databases, cloud platforms, and big data tools.

What are the key skills and qualifications needed to thrive as a data engineer, and why are they important?

To thrive as a Data Engineer, you need a strong background in computer science, data modeling, and programming languages such as Python or Java, often coupled with a relevant degree. Familiarity with ETL tools, big data frameworks (like Hadoop or Spark), and cloud platforms (such as AWS or Azure) is typically required, along with certifications like AWS Certified Data Analytics. Strong problem-solving skills, attention to detail, and effective communication set exceptional data engineers apart. These skills and qualities are essential for building robust data pipelines, ensuring data quality, and supporting data-driven decision-making across organizations.

What does a data engineer do?

The job duties of a data engineer involve helping with the development of systems, software, and infrastructure used to process, store and analyze data. Your responsibilities in this career include working to install data management software. Your employer may expect you to perform maintenance and install updates to all software and systems that they use for data acquisition, management, and analysis. Data engineers also analyze existing data systems to find ways to improve efficiency and accessibility. You then suggest upgrades or changes based on your assessment.

How do data engineers typically collaborate with data scientists and analysts within an organization?

Data Engineers play a crucial role in ensuring that Data Scientists and Analysts have reliable, well-structured data for their projects. This collaboration often involves building and maintaining data pipelines, optimizing data storage solutions, and troubleshooting data quality issues. Regular communication and agile teamwork are common, with Data Engineers frequently participating in meetings to understand analytical requirements and adjust data processes accordingly. By working closely together, these teams can quickly iterate on data models and deliver actionable insights to drive business decisions.

What is the difference between Data Engineer vs Data Scientist?

AspectData EngineerData Scientist
Primary FocusBuilding and maintaining data pipelines and infrastructureAnalyzing data to extract insights and create models
SkillsSQL, ETL, programming (Python, Java), database managementStatistics, machine learning, data analysis, programming (Python, R)
Work EnvironmentData warehouses, cloud platforms, backend systemsData analysis environments, research labs, visualization tools
Common ToolsApache Spark, Hadoop, Airflow, SQLJupyter, RStudio, Tableau, scikit-learn

Data Engineers focus on creating and maintaining the infrastructure that allows data to be collected, stored, and processed efficiently. Data Scientists analyze this data to generate insights, build predictive models, and support decision-making. While their skills overlap, Data Engineers are more involved in data pipeline development, whereas Data Scientists focus on data analysis and modeling.

Is a data engineer entry level?

Data engineering is typically an intermediate to senior-level role that requires experience with programming, databases, and data pipeline tools. Entry-level positions may be available for those with relevant internships or strong foundational skills, but most data engineering roles demand several years of experience or advanced knowledge of tools like SQL, Python, and cloud platforms.

What is the role of a data engineer?

A data engineer designs, builds, and maintains data pipelines and infrastructure to collect, process, and store large volumes of data. They work with tools like SQL, Python, and cloud platforms to ensure data is accessible and reliable for analysis and decision-making.

What job categories do people searching Data Engineer jobs in York, PA look for?

The top searched job categories for Data Engineer jobs in York, PA are:

What cities near York, PA are hiring for Data Engineer jobs?

Cities near York, PA with the most Data Engineer job openings:

Infographic showing various Data Engineer job openings in York, PA as of August 2026, with employment types broken down into 83% Full Time, 7% Temporary, and 10% Contract. Highlights an 79% In-person, 7% Hybrid, and 14% Remote job distribution, with an average salary of $127,663 per year, or $61.4 per hour.

Data Engineer

Lititz, PA โ€ข On-site

TAIT
Arts, Entertainment, and Recreationย โ€ขย 501 - 1,000 employees

$106K - $127K/yr

Full-time

Re-posted 6 days ago


Job description

TAIT partners with artists, brands, IP holders and place makers to bring culture-defining, never-before-seen experiences to life. With a legacy of innovation spanning over 45 years, TAIT has grown from pioneering in rock 'n' roll concert staging to setting the global standard for extraordinary live events and experiences through cutting-edge technology, precision engineering, and creative design. TAIT's 20 global offices have developed iconic productions and experiences in over 30 countries, all seven continents, and even outer space for renowned performers, theme parks, exhibits, and venues across the globe, including partnerships with Taylor Swift, Cirque Du Soleil, Royal Opera House, Nike, NASA, Bloomberg, Google, Beyoncรฉ, and The Olympics
TAIT is looking for a hands-on Data Engineer to build, scale, and operate the governed data platform that will underpin the next phase of the company's technology transformation. Reporting to the Data Architect, this role will turn the target data architecture into reliable pipelines, reusable data products, trusted analytical models, and production platform capabilities.
This role is ideal for an experienced engineer who can move between architecture and execution and who has a proven track record delivering production-grade Databricks solutions in AWS. The successful candidate will do more than move data from one system to another: they will have delivered projects that combine complex enterprise and operational data to enable sophisticated analysis, forecasting, optimization, AI/ML use cases, and confident executive decision-making.
The Data Engineer will work closely with the Data Architect and partner across Enterprise Technology, Architecture/AI, Integration, Infrastructure and Cloud, Information Security, and business teams including Finance, Sales, Project Delivery, Operations, Procurement, Manufacturing, and People/Resourcing. The role will help establish clear sources of truth and enable analytics across systems such as Salesforce CRM and CPQ, Kantata PSA, Epicor and the next-generation ERP, MES, Autodesk/PLM, HCM, service management platforms, and other internal and external data sources.
Responsibilities:
Databricks and AWS Data Platform Engineering
  • Design, build, deploy, and operate production data pipelines and lakehouse capabilities using Databricks on AWS.
  • Implement scalable ingestion, transformation, orchestration, and serving patterns for batch, near-real-time, streaming, API, file, and change-data-capture workloads.
  • Use AWS services and controls - including S3, IAM, KMS, VPC networking, Secrets Manager, CloudWatch, and appropriate catalog or integration services - to create secure and supportable data solutions.
  • Develop reliable data layers, optimizing workloads, storage design, cluster or serverless configuration, performance, and platform cost.
  • Establish automated deployment practices using Git, CI/CD, infrastructure as code, automated testing, and environment promotion across development, test, and production.
  • Implement monitoring, observability, alerting, recovery, service-level objectives, and operational runbooks for critical data products and pipelines.

Analytics Enablement and Data Products
  • Translate high-value business questions into well-defined data products, curated datasets, and semantic-ready models for Power BI, advanced analytics, and AI-enabled solutions.
  • Integrate data from multiple enterprise and operational sources to support strategic analytics, dashboards, KPIs, project and customer profitability, resource and capacity analysis, financial planning, forecasting, scenario modeling, operational optimization, and other sophisticated analytical use cases.
  • Partner with analysts, product owners, and business subject-matter experts to ensure data products are analytically fit for purpose, clearly defined, and adopted by their intended users.
  • Create reusable datasets and feature-ready data that accelerate experimentation and responsible delivery of machine learning, generative AI, and other advanced analytical capabilities.
  • Define and track technical and business measures for data-product adoption, quality, timeliness, reliability, and realized value.

Roadmap Delivery
  • Partner with the Integration team and Data Architect to define data contracts, APIs, event patterns, and reusable integration standards, including appropriate alignment with MuleSoft and other enterprise integration capabilities.
  • Support major platform implementations and migrations by developing conversion pipelines, reconciliation routines, archival strategies, historical-data models, and post-go-live analytical capabilities.
  • Create repeatable onboarding patterns that accelerate M&A data discovery, integration, harmonization, and reporting across acquired businesses.
  • Reduce data and technology debt by replacing brittle point-to-point extracts, undocumented transformations, and duplicate datasets with governed, reusable platform services.

Data Quality, Governance, and Security
  • Work with the Data Architect to implement enterprise data standards, naming conventions, reference architectures, reusable patterns, and engineering guardrails.
  • Build automated data-quality controls, source-to-target reconciliation, exception handling, freshness checks, and issue-management processes for critical data domains.
  • Implement metadata management, cataloging, lineage, ownership, and access controls using Databricks governance capabilities and supporting enterprise tools.
  • Apply security by design, including least-privilege access, encryption, secrets management, environment separation, auditability, retention, and appropriate handling of confidential or regulated data.
  • Help define and maintain trusted sources of truth, master and reference-data practices, data definitions, and documentation that make strategic metrics traceable and explainable.

Delivery, Collaboration, and Continuous Improvement
  • Own engineering work from discovery through production support, including estimation, design, development, testing, documentation, deployment, adoption, and continuous improvement.
  • Communicate technical tradeoffs, dependencies, risks, and recommendations clearly to both technical and non-technical audiences.
  • Participate in design reviews, code reviews, incident reviews, architecture forums, and roadmap planning; provide practical feedback that improves quality without slowing delivery unnecessarily.
  • Work effectively with global internal teams, nearshore and offshore partners, software vendors, and implementation partners while maintaining clear engineering accountability and standards.
  • Mentor less-experienced engineers and analysts, share reusable practices, and help build a strong data-engineering discipline within GTS.

Required Qualifications:
  • Bachelor's degree in computer science, engineering, information systems, data science, or a related discipline, or equivalent practical experience.
  • 7+ years of professional experience in data engineering, analytics engineering, data platform engineering, or a closely related role.
  • A proven track record designing, delivering, and operating production-grade solutions in Databricks on AWS, including responsibility for architecture implementation, pipeline development, deployment, support, performance, and cost.
  • Evidence of delivering multiple end-to-end projects where engineered data products enabled sophisticated analysis - such as cross-system KPI reporting, forecasting, optimization, profitability analysis, anomaly detection, predictive modeling, scenario analysis, or AI/ML - with clear business outcomes.
  • Advanced SQL and strong Python/PySpark skills, including experience developing maintainable, tested, production-quality code for large and complex datasets.
  • Strong knowledge of Spark, Delta Lake or comparable lakehouse patterns, dimensional and analytical data modeling, partitioning, schema evolution, data contracts, and performance tuning.
  • Hands-on experience with core AWS data-platform capabilities such as S3, IAM, KMS, VPC networking, CloudWatch, Secrets Manager, and related services used to secure, integrate, and operate Databricks workloads.
  • Experience building batch and incremental pipelines using APIs, files, databases, SaaS connectors, change data capture, and at least one orchestration framework or managed workflow capability.
  • Experience with Git-based development, CI/CD, automated data testing, code review, environment management, and infrastructure-as-code practices.
  • Practical experience implementing data quality, observability, cataloging, lineage, access control, and operational support for business-critical data products.
  • Strong problem-solving, documentation, and communication skills, including the ability to translate ambiguous business needs into pragmatic technical solutions and explain data issues in business terms.
  • Ability to work independently in a changing environment, prioritize based on business value and risk, and collaborate effectively across architecture, integration, infrastructure, security, applications, analytics, and business teams.

Preferred Qualifications:
  • Databricks certification and/or AWS certification relevant to data engineering, analytics, or cloud architecture.
  • Experience with Unity Catalog or comparable governance capabilities, including fine-grained access, lineage, cataloging, and multi-environment administration.
  • Experience preparing governed data models for Power BI, including semantic-model design, incremental refresh patterns, performance optimization, and row-level security.
  • Experience integrating enterprise applications such as Salesforce, Epicor or another ERP, Kantata or another PSA, CPQ, MES, PLM, HCM, EPM, and service management platforms.
  • Experience with MuleSoft or another enterprise integration platform and with designing API- or event-driven data exchange patterns.
  • Experience enabling machine learning or AI solutions through feature engineering, model-ready datasets, ML lifecycle tooling, vector or unstructured-data pipelines, or responsible AI data controls.
  • Experience in a global, project-based, manufacturing, engineering, live-entertainment, or operationally complex business environment.
  • Experience supporting enterprise transformation, ERP replacement, legacy-platform retirement, cloud migration, and/or M&A integration.

#LI-ML1
TAIT is an equal opportunity employer fully committed to diversity and inclusion in the workplace. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran or any other protected characteristic as outlined by international, national, state, or local laws.