2

Data Engineer Internship Remote Jobs in California

Senior Data Engineer

Los Angeles, CA · On-site +1

$160K - $190K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

... remote work days. To learn more about the work we do at EDO, please visit EDO Press . The Role As a ... What We Are Looking For * 6+ years of data engineering experience and three years of hands-on ...

Senior Data Engineer

San Francisco, CA · On-site +1

$124K - $169K/yr

  • PTO

Learn more in our CEO's funding announcement: We're a small, remote-first team. We take ownership ... As a Senior Data Engineer, you will be tasked with designing and implementing robust data solutions ...

Senior Data Engineer

San Francisco, CA · On-site +1

$124K - $169K/yr

  • Medical

  • Life

  • Retirement

  • PTO

About this role: Senior Data Engineer The Data, Analytics and Reporting Technology team is ... Hybrid schedule (3 days in office, 2 days remote) * Work Transparently: You always deal in an ...

New

Senior Data Engineer, Risk

Bodega Bay, CA · On-site +1

$125K - $170K/yr

Today, Cash App has thousands of employees working globally across office and remote locations ... The Role As a Data Engineer you will handle everything from data architecture and modeling to data ...

The role We're looking for a Senior Data Engineer to help the team deliver data science services ... We are a remote-first company for most positions so you may work from anywhere you like in the U.S ...

Senior AWS Data Engineer

Mountain View, CA · On-site +1

$125K - $169K/yr

Remote If you meet these qualifications and are pursuing new challenges, start your application on ... data integration patterns. * Familiarity with GitLab, Terraform, CI/CD, and AWS developer tools.

Data Ops Engineer

San Diego, CA · On-site +1

$200K - $240K/yr

None Potential for Remote Work: ORA_ON_SITE Description We are seeking a Data Ops Engineer to design, build, and maintain real-time data ingestion pipelines. In this role, you will be responsible for ...

Remote We have below longterm job opening. If you are interested , Please send your updated resume ... Mentor and guide data engineers while ensuring adherence to best practices. Optimize ETL/ELT ...

Data Engineering Manager

Pasadena, CA · On-site +1

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

Founded in 2006, Spokeo has built a dedicated, remote-first team with an average tenure of 6.9 ... About this Opportunity Spokeo is seeking a Data Engineering Manager to help build a ...

Data Engineering Manager

Pasadena, CA · On-site +1

$172K - $206K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

Founded in 2006, Spokeo has built a dedicated, remote-first team with an average tenure of 6.9 ... About this Opportunity Spokeo is seeking a Data Engineering Manager to help build a ...

Showing results 21-40

Data Engineer Internship Remote information

What is a data engineer internship remote?

A Data Engineer Internship Remote job is a temporary, online position where interns assist with building, maintaining, and optimizing data pipelines and infrastructure. Interns work with large datasets, databases, and ETL processes to support business intelligence and analytics teams. They gain experience in cloud platforms, SQL, Python, and data warehousing while collaborating with engineers and analysts. This remote role allows flexibility while providing hands-on experience in data engineering best practices.

What are the typical daily responsibilities of a remote data engineer intern?

As a remote Data Engineer Intern, your daily responsibilities often include assisting with data extraction, transformation, and loading (ETL) processes, cleaning and organizing datasets, and helping to build or maintain data pipelines. You may also work on tasks such as writing scripts in SQL or Python, contributing to database schema design, and documenting your work for team collaboration. Regular communication with your supervisor and team, participating in virtual meetings, and collaborating on version control platforms like Git are also common aspects of the role. These experiences offer valuable insight into real-world data engineering workflows and prepare you for more advanced responsibilities in the field.

What are the key skills and qualifications needed to thrive in the data engineer internship remote position, and why are they important?

To excel as a Data Engineer Intern in a remote setting, you need a solid grounding in computer science, data structures, and programming concepts, often supported by coursework or experience in related fields. Familiarity with data modeling, SQL, Python, cloud platforms (like AWS or Azure), and tools such as Apache Spark or Airflow is highly valuable, along with any project-based experience or relevant certifications. Strong communication, self-motivation, and time management skills are crucial for working efficiently and collaboratively in a remote environment. These abilities enable you to handle data workflows, solve technical problems, and contribute effectively to distributed teams in real-world projects.

What job categories do people searching Data Engineer Internship Remote jobs in California look for?

The top searched job categories for Data Engineer Internship Remote jobs in California are:

What cities in California are hiring for Data Engineer Internship Remote jobs?

Cities in California with the most Data Engineer Internship Remote job openings:

Infographic showing various Data Engineer Internship Remote job openings in California as of August 2026, with employment types broken down into 30% Internship, 50% Full Time, and 20% Contract. Highlights an 100% Remote job distribution.

AI Data Engineer - Scientific Data Platforms (Remote)

Astrix Inc

South San Francisco, CA • On-site, Remote

$35 - $38/hr

Full-time

Re-posted 6 days ago


Job description

Pay Rate Low: 35 | Pay Rate High: 40
Our client is a leading global biotechnology and pharmaceutical organization driven by a mission to innovate, continuously advance science, and ensure everyone has access to the healthcare they need.
Title: AI Data Engineer - Scientific Data Platforms
Location: Remote, Must work PST
Pay rate: $35-38/hr (Depends on experience level)
Schedule: Full-time (40 hours/week)
Duration: 1-year contract, (Plus benefits)
Position Overview
This role addresses a critical need in scaling our AI models for drug discovery by building largely automated, scalable, agent-driven data ingestion and curation pipelines for genomics data. This includes metadata inference, constructing performant query architectures, and transforming high-dimensional datasets (e.g., single-cell omics, clinical trials) into AI-ready training formats.
Key Responsibilities
  • Build an agentic data ingestion pipeline and move beyond bespoke steps toward agents that teams can reliably use as a shared, deployed service.
  • Triage and prioritize incoming requests to ingest specific datasets. Clean and organize data, building the first-pass cleaning and organization steps into the agentic flow.
  • Validate cross-modal linkage. Add automated checks that catch when ingested data does not connect correctly and flag low-quality or mismatched records.
  • Version every dataset, retaining and making prior versions addressable. Preserve raw data and provenance, ensuring agent workflows log validation and transformation steps so lineage is fully traceable.
  • Partner with AI, software engineering, and computational biology groups to co-define data standards and conventions.

Qualifications & Requirements
  • Demonstrated experience building multi-agent workflows or LLM workflows using tools/frameworks such as LangGraph or LlamaIndex, including tool/function calling and asynchronous task execution.
  • Strong Python skills for data manipulation, working with APIs and databases, and handling heterogeneous data formats.
  • Familiarity with dataset versioning approaches (e.g., DVC, lakeFS, or equivalent).
  • Comfortable with or showing a strong willingness to learn common omics data formats like AnnData, H5AD, and TileDB.
  • No deep bioinformatics expertise required; just a basic conceptual understanding of different modalities (e.g., RNA-seq vs. scRNA-seq vs. WES; genomics vs. transcriptomics vs. proteomics vs. metabolomics).
  • Comfortable writing unit and functional tests to ensure data processing workflows are reliable and reproducible.
  • Degree in a technical field or equivalent practical experience.
  • Must be Authorized to work in the United States without Sponsorship.
Nice to Have
  • Experience deploying agent workflows as a shared service (e.g., FastAPI or MCP endpoints).
  • Exposure to cloud platforms (AWS, GCP) and containerization (Docker).
  • Familiarity with scientific workflow managers such as Nextflow or Snakemake.

INDBH
#LI-MG1