1

Intern Streaming Data Engineer Jobs in Pennsylvania

Data Engineer - onsite

Hanover, PA ยท On-site

$110K - $132K/yr

Data Engineer Location: Hanover, MD Type: Contract Compensation: 55-70/hr W-2, depending on ... Hands-on experience with Azure Synapse, Azure Data Factory, Azure Databricks, and Azure Stream ...

$130K - $144K/yr

We are seeking a Senior Data Engineer to join our team and help us build the next generation of ... Spark, Databricks, Structured Streaming * Databases: Relational (Postgres) and NoSQL (DynamoDB ...

Sr Databricks Data Engineer

Philadelphia, PA

$115K - $138K/yr

... Streaming, Databricks Workflows, Apache Airflow, Unity Catalog, automated continuous integration and continuous deployment (CI/CD) pipelines, and performance optimization of data engineering ...

$150 - $200/hr

Strong understanding of batch and real-time data architectures, including streaming and change data ... software engineering practices for data systems. * Demonstrated ability to independently solve ...

Posted today

Senior Data Engineer

Plymouth Meeting, PA ยท On-site

$180K - $210K/yr

Strong understanding of batch and real-time data architectures, including streaming and change data ... software engineering practices for data systems. * Demonstrated ability to independently solve ...

New

Sr Databricks Data Engineer

Pittsburgh, PA

$111K - $133K/yr

... Streaming, Databricks Workflows, Apache Airflow, Unity Catalog, automated continuous integration and continuous deployment (CI/CD) pipelines, and performance optimization of data engineering ...

Lead Data Engineer

Coopersburg, PA ยท On-site

$97K - $127K/yr

Apply your experience in streaming pipelines, cloud platforms, and data modeling to deliver high ... Define engineering standards, and best practices for scalable data platforms across the data team.

Principal Data Engineer - AI

Philadelphia, PA ยท On-site

$136K - $182K/yr

Engineer feature-rich context pipelines that process large-scale enterprise data, balancing batch and streaming patterns seamlessly. * Optimize and scale large distributed queries and data ...

We are seeking a Junior Data Engineer who will work hands-on and collaboratively with our ... Some experience or coursework involving Apache Spark, real-time streaming technologies, or large ...

We are seeking a Junior Data Engineer who will work hands-on and collaboratively with our ... Some experience or coursework involving Apache Spark, real-time streaming technologies, or large ...

We are seeking a Junior Data Engineer who will work hands-on and collaboratively with our ... Some experience or coursework involving Apache Spark, real-time streaming technologies, or large ...

Data Engineering Job Category: Scientific/Technology All Job Posting Locations: Mooresville ... Develop blueprints for batch, streaming, and event-driven pipelines that support analytics, BI, AI ...

Lead Data Engineer

Coopersburg, PA ยท On-site

$97K - $127K/yr

Apply your experience in streaming pipelines, cloud platforms, and data modeling to deliver high ... Mentor Data Engineers through design reviews, code reviews, and technical coaching. * Design ...

A Senior ETL Developer specializing in Google Cloud Platform (GCP) and BigQuery designs, builds ... Create robust ingestion patterns for batch and streaming data. * Troubleshoot data integration ...

Showing results 41-60

Intern Streaming Data Engineer information

What does an intern streaming data engineer do?

An Intern Streaming Data Engineer assists in designing, developing, and maintaining systems that process real-time data streams. They typically work with technologies like Apache Kafka, Apache Flink, or Spark Streaming to collect, process, and analyze data as it arrives. Their responsibilities may include writing code, troubleshooting data pipelines, and collaborating with senior engineers to ensure data flows efficiently. The role is ideal for students or recent graduates looking to gain hands-on experience with big data and real-time analytics.

What types of projects or tasks can an intern streaming data engineer expect to work on during their internship?

As an Intern Streaming Data Engineer, you can expect to work on projects involving the development, testing, and optimization of real-time data pipelines. Typical tasks may include assisting with the integration of streaming platforms like Apache Kafka or AWS Kinesis, writing and debugging code to process large volumes of incoming data, and collaborating with senior engineers to ensure data quality and reliability. You'll often work within a team of data engineers and analysts, gaining hands-on experience with the latest big data tools and contributing to solutions that support real-time analytics and business decision-making.

What are the key skills and qualifications needed to thrive as an intern streaming data engineer, and why are they important?

To thrive as an Intern Streaming Data Engineer, you typically need foundational knowledge in computer science, data engineering concepts, and familiarity with real-time data processing. Experience with tools like Apache Kafka, Apache Flink, or Spark Streaming, and programming languages such as Python or Java, is often preferred. Strong problem-solving skills, attention to detail, and effective teamwork and communication abilities help set candidates apart. These skills and qualifications are crucial for efficiently building, maintaining, and troubleshooting streaming data pipelines in dynamic data-driven environments.

What is the difference between Intern Streaming Data Engineer vs Intern Data Analyst?

AspectIntern Streaming Data EngineerIntern Data Analyst
Required SkillsKnowledge of streaming platforms (e.g., Kafka, Spark Streaming), programming (Python, Java), data pipeline developmentData analysis, SQL, Excel, basic statistical skills
Work EnvironmentDeveloping real-time data pipelines, working with big data toolsAnalyzing stored data, generating reports and insights
Industry UsageTech, finance, e-commerce companies focusing on real-time data processingMarketing, business intelligence, research departments

The Intern Streaming Data Engineer focuses on building and maintaining real-time data pipelines using streaming technologies, requiring programming and big data skills. In contrast, the Intern Data Analyst primarily analyzes stored data to generate insights, emphasizing statistical and reporting skills. Both roles are common in data-driven industries but serve different functions within data management and analysis.

What job categories do people searching Intern Streaming Data Engineer jobs in Pennsylvania look for?

The top searched job categories for Intern Streaming Data Engineer jobs in Pennsylvania are:

Infographic showing various Intern Streaming Data Engineer job openings in Pennsylvania as of July 2026, with employment types broken down into 1% As Needed, 81% Full Time, 14% Part Time, 1% Temporary, and 3% Contract. Highlights an 88% Physical, 3% Hybrid, and 9% Remote job distribution.

Senior Data Engineer, Malvern, PA (Onsite)

Prohires

Malvern, PA โ€ข On-site

$112K - $134K/yr

Other

Posted 4 days ago


Job description

Senior Data Engineer

Location: Malvern, PA (Onsite)
Duration: Long-Term Contract
Experience Required: 8+ Years

Job Summary

We are seeking a highly skilled Senior Data Engineer with strong expertise in Python, SQL, and AWS cloud technologies to design, develop, and optimize scalable data solutions. The ideal candidate will have extensive experience building modern data pipelines, ETL/ELT frameworks, cloud-native architectures, and large-scale data processing systems. This role will be responsible for developing robust data platforms that support analytics, reporting, machine learning, and business intelligence initiatives.

Key Responsibilities
  • Design, develop, and maintain scalable data pipelines using Python and AWS services.
  • Build and optimize ETL/ELT workflows for ingesting, transforming, and loading large datasets.
  • Develop cloud-native data solutions leveraging AWS services such as Glue, S3, Lambda, Redshift, Athena, CloudWatch, and IAM.
  • Write complex and optimized SQL queries for data extraction, transformation, validation, and reporting.
  • Design and implement data models for analytics and business intelligence platforms.
  • Monitor, troubleshoot, and optimize data workflows to ensure reliability and performance.
  • Collaborate with Data Scientists, Analysts, Architects, and Business stakeholders to understand data requirements and deliver solutions.
  • Implement data quality, governance, security, and compliance standards across data platforms.
  • Build automation frameworks for data processing and operational support.
  • Participate in code reviews, architecture discussions, and technical design sessions.
  • Create technical documentation and maintain best practices for data engineering processes.
Required Qualifications
  • Bachelor's degree in Computer Science, Information Technology, Engineering, or related field.
  • 8+ years of experience in Data Engineering and Data Warehousing.
  • Strong hands-on experience with Advanced Python programming for data engineering applications.
  • Extensive experience with SQL and relational databases.
  • Strong AWS experience with:
    • AWS Glue
    • Amazon S3
    • AWS Lambda
    • Amazon Redshift
    • AWS Athena
    • CloudWatch
    • IAM
  • Experience building and maintaining large-scale ETL/ELT pipelines.
  • Strong understanding of data modeling, data warehousing, and dimensional modeling concepts.
  • Experience with performance tuning and optimization of SQL queries and data pipelines.
  • Knowledge of CI/CD practices and version control tools such as Git.
  • Strong troubleshooting and problem-solving skills.
Preferred Qualifications
  • Experience with Apache Spark, PySpark, or distributed data processing frameworks.
  • Experience with Airflow or other workflow orchestration tools.
  • Knowledge of streaming technologies such as Kafka or Kinesis.
  • Experience working in Agile/Scrum environments.
  • Exposure to Snowflake, Databricks, or other modern data platforms.
  • AWS certifications are a plus.
Technical Skills

Programming:

  • Python (Advanced)
  • SQL

Cloud & Data Services:

  • AWS Glue
  • Amazon S3
  • AWS Lambda
  • Amazon Redshift
  • AWS Athena
  • CloudWatch
  • IAM

Tools & Technologies:

  • Git
  • CI/CD Pipelines
  • ETL/ELT Frameworks
  • Data Warehousing
  • Data Modeling
Nice to Have
  • PySpark / Spark
  • Airflow
  • Kafka
  • Snowflake
  • Databricks
  • Machine Learning Data Pipelines

Keywords: Senior Data Engineer, Python Data Engineer, AWS Data Engineer, ETL Developer, Data Warehouse Engineer, AWS Glue, Redshift, Lambda, S3, SQL, PySpark, Airflow, Data Pipelines, Cloud Data Engineering.