1

Data Infrastructure Internship Jobs in California

AI Data Engineer

Los Angeles, CA

$123K - $148K/yr

Prior production engineering experience with data pipelines, integrations, or analytics infrastructure - internship, contract, or full-time all qualify * Daily fluency with a standard engineering ...

AI Data Engineer

Los Angeles, CA ยท On-site

$123K - $148K/yr

Prior production engineering experience with data pipelines, integrations, or analytics infrastructure - internship, contract, or full-time all qualify * Daily fluency with a standard engineering ...

Design, build, and maintain the data pipelines and infrastructure that power both our product and ... A combination of internships, research, and substantial project experience that clearly ...

Product Analyst

San Francisco, CA ยท On-site

$100K - $125K/yr

Construction data infrastructure The Role The client is hiring a sharp, semi-technical generalist ... A strong startup-internship signal (YC company or comparable early-stage) * Demonstrated systems ...

Data Engineer II

San Francisco, CA ยท On-site

$134K - $162K/yr

Design, build, and maintain the data pipelines and infrastructure that power both our product and ... A combination of internships, research, and substantial project experience that clearly ...

This internship provides hands-on experience in building production-ready data systems within a ... infrastructure and workflows Other Duties: Perform other related duties and ad hoc projects as ...

This internship provides hands-on experience in building production-ready data systems within a ... infrastructure and workflows Other Duties: Perform other related duties and ad hoc projects as ...

next page

Showing results 1-20

Data Infrastructure Internship information

What is the difference between Data Infrastructure Internship vs Data Engineer?

AspectData Infrastructure InternshipData Engineer
Required CredentialsTypically pursuing or recent graduate in Computer Science, Data Science, or related fieldsBachelor's or Master's in Computer Science, Software Engineering, or related fields; often requires professional experience
Work EnvironmentInternship setting, often in tech companies or data-driven organizations, with mentorshipFull-time professional role, involved in designing, building, and maintaining data systems
Employer & Industry UsageUsed by companies hiring interns to support data infrastructure projectsUsed by organizations to develop scalable data pipelines and infrastructure

The Data Infrastructure Internship is an entry-level position aimed at students or recent graduates gaining hands-on experience. In contrast, a Data Engineer is a full-time professional responsible for developing and maintaining data systems. Internships provide foundational exposure, while Data Engineers handle complex, ongoing data infrastructure tasks.

What are the most commonly searched types of Data Infrastructure jobs in California? The most popular types of Data Infrastructure jobs in California are:
What cities in California are hiring for Data Infrastructure Internship jobs? Cities in California with the most Data Infrastructure Internship job openings:
Infographic showing various Data Infrastructure Internship job openings in California as of July 2026, with employment types broken down into 1% As Needed, 83% Full Time, 12% Part Time, 1% Temporary, and 3% Contract. Highlights an 88% Physical, 3% Hybrid, and 9% Remote job distribution.

Staff+ Software Engineer, Data Infrastructure

Anthropic

San Francisco, CA โ€ข On-site

$134K - $162K/yr

Other

Posted 15 days ago


Job description

About the role

Data Infrastructure designs, operates, and scales secure, privacy-respecting systems that power data-driven decisions across Anthropic. Our mission is to provide data processing, storage, and access that are trusted, fast, and easy to use.

We're looking for infrastructure engineers who thrive working at the intersection of data systems, security, and scalability. You'll tackle diverse challenges ranging from building financial reporting pipelines to architecting access control systems to ensuring cloud storage reliability. This role offers the opportunity to work directly with data scientists, analysts, and business stakeholders while diving deep into cloud infrastructure primitives.

Responsibilities:

Within Data Infra, you may be matched to critical business areas including:ย 

  • Data Governance & Access Control: Design and implement robust access control systems ensuring only authorized users can access sensitive data. Build infrastructure for permission management, audit logging, and compliance requirements. Work on IAM policies, ACLs, and security controls that scale across thousands of users and systems.

  • Financial Data Infrastructure: Build and maintain data pipelines and warehouses powering business-critical reporting. Ensure data integrity, accuracy, and availability for complex financial systems, including third party revenue ingestion pipelines; manage the external relationships as needed to drive upstream dependencies. Own the reliability of systems processing revenue, usage, and business metrics.

  • Cloud Storage & Reliability: Architect disaster recovery, backup, and replication systems for petabyte-scale data. Ensure high availability and durability of data stored in cloud object storage (GCS, S3). Build systems that protect against data loss and enable rapid recovery.

  • Data Platform & Tooling: Scale data processing infrastructure using technologies like BigQuery, BigTable, Airflow, dbt, and Spark. Optimize query performance, manage costs, and enable self-service analytics across the organization.

You might be a good fit if you:
  • Have 10+ years (not including internships or co-ops) of experience in a Software Engineer role, building data infrastructure, storage systems, or related distributed systems
  • Have 3+ years (not including internships or co-ops) of experience leading large scale, complex projects or teams as an engineer or tech lead
  • Can set technical direction for a team, not just execute within it
  • Have deep experience with at least one of:
  • Strong proficiency in programming languages like Python, Go, Java, or similar
  • Experience with infrastructure-as-code (Terraform, Pulumi) and cloud platforms (GCP, AWS)
  • Can navigate complex technical tradeoffs between performance, cost, security, and maintainability
  • Have excellent collaboration skills - you work well with both technical and non-technical stakeholders
Strong candidates may also have:
  • Experience with security and compliance requirements (ITGC, GDPR, financial controls)
  • Background in data warehousing, ETL/ELT pipelines, or analytics infrastructure
  • Experience with Kubernetes, containerization, and cloud-native architectures
  • Track record of improving data reliability, availability, or cost efficiency at scale
  • Knowledge of column-oriented databases, OLAP systems, or big data processing frameworks
  • Experience working in fintech, financial services, or highly regulated environments
  • Security engineering background with focus on data protection and access controls
Technologies We Use:
  • Data: BigQuery, BigTable, Airflow, Cloud Composer, dbt, Spark, Segment, Fivetran
  • Storage: GCS, S3
  • Infrastructure: Terraform, Kubernetes, GCP, AWS
  • Languages: Python, Go, SQL

Deadline to apply: None. Applications will be reviewed on a rolling basis.