1

Big Data Engineer Intern Jobs in California (NOW HIRING)

Data Engineer Intern

Los Angeles, CA ยท On-site

$123K - $148K/yr

Reporting to Data Team Lead, the Data Engineer intern will participate in the acquisition and ... Experience building and optimizing 'big data' data pipelines, architectures and data sets. * Solid ...

Data Engineer Intern

Los Angeles, CA

$123K - $148K/yr

Reporting to Data Team Lead, the Data Engineer intern will participate in the acquisition and ... Experience building and optimizing 'big data' data pipelines, architectures and data sets. * Solid ...

Big Data Engineer

Los Angeles, CA ยท On-site

$60 - $79.50/hr

Company Description Intelliswift Software, Inc As a Big Data Engineer, you will be an integral member of our threat intelligence service, i.e. auto focus, team responsible for architecture, design ...

Big Data Engineer

Mountain View, CA ยท On-site

$65.75 - $87/hr

... Big Data & Java etc. * Work with business partners directly to seek clarity on requirements ... Electrical Engineering, Information Systems or other technical discipline; advanced degree ...

Big Data Engineer

San Diego, CA ยท On-site

$59.25 - $78.25/hr

Company Description Jobsbridge Key Qualifications: 4+ years of experience in Big Data, Business Intelligence and Data Warehousing environment 1+ years of experience in Hadoop/HBase systems 2+ years ...

Big Data Engineer

Palo Alto, CA ยท On-site

$65.50 - $86.75/hr

Role: Big Data Engineer w/Spark Location: Palo Alto, CA Duration: Potential Contract to Hire Mode of Interview: Skype After Phone !! Need Green Card OR US Citizen OR EAD GC Candidates Only

Java Big Data Engineer

Cupertino, CA

$68.75 - $91/hr

Be it core Java, full-stack Java, Web/UI designers, Big Data or Cloud or Mobility developers/architects, we have them all. Proficiency with Big Data processing technologies (Hadoop, Hive, Spark ...

Java Big Data Engineer

Cupertino, CA

$68.75 - $91/hr

Be it core Java, full-stack Java, Web/UI designers, Big Data or Cloud or Mobility developers/architects, we have them all. Proficiency with Big Data processing technologies (Hadoop, Hive, Spark ...

We are building the next-generation, big data talent platform that aims to revolutionize the hiring ... As a QA Engineer Intern, you will work alongside our small team of engineers to develop new ...

Big Data Developer

Glendale, CA

$56.25 - $72.75/hr

Skill Big Data Developer Java Hadoop DaaS Java Hadoop / Hive / MapReduce / Pig / Sqoop / Flume SQL XML JSO Location Glendale, CA Total Experience 3 yrs. Max Salary Not Mentioned Employment Type ...

We are building the next-generation, big data talent platform that aims to revolutionize the hiring ... As a software engineering intern, you will work alongside our small team of engineers to create new ...

Big Data

Sunnyvale, CA ยท On-site

$62.25 - $80.75/hr

Must have strong programming knowledge of Core Java or Scala - Objects & Classes, Data Types ... Must have experience working on Big Data Processing Frameworks and Tools - MapReduce, YARN, Hive ...

Big Data Developer

Los Angeles, CA ยท On-site

$57 - $74/hr

Company Description Intelliswift Software, Inc Big Data & Java experience is mandatory Experience in Big Data Tech stack - Mapr, Hive, PIG is mandatory Must have Data Management experience in large ...

next page

Showing results 1-20

Big Data Engineer Intern information

What does a Big Data Engineer Intern do?

A Big Data Engineer Intern assists with designing, building, and maintaining large-scale data processing systems. They often work with technologies such as Hadoop, Spark, and SQL to manage and analyze large datasets. Interns may help develop data pipelines, clean and organize data, and support senior engineers in optimizing data workflows. This role provides hands-on experience in handling big data tools and working on real-world data engineering projects.

What are the key skills and qualifications needed to thrive as a Big Data Engineer Intern?

To thrive as a Big Data Engineer Intern, you need a solid understanding of programming (especially Python, Java, or Scala), data structures, and basic knowledge of distributed computing concepts, typically supported by coursework or relevant projects. Familiarity with big data tools like Hadoop, Spark, and data querying languages such as SQL, as well as exposure to cloud platforms like AWS or Azure, is highly valuable. Strong analytical thinking, problem-solving abilities, and communication skills help interns collaborate effectively and learn quickly in dynamic environments. These skills and qualities are crucial for handling large-scale datasets and supporting the development of data-driven solutions within engineering teams.

What are some common challenges a Big Data Engineer Intern might face when working with large datasets?

As a Big Data Engineer Intern, you'll often encounter challenges related to managing the scale and complexity of massive datasets. These can include optimizing data ingestion pipelines for speed and efficiency, troubleshooting data quality issues, and ensuring data privacy and security. Additionally, you may need to quickly learn new tools or frameworks, such as Hadoop or Spark, and collaborate closely with data scientists and engineers to ensure data is structured and accessible for analysis. Developing problem-solving skills and being proactive in seeking help from your team can help you overcome these hurdles.

What is the difference between Big Data Engineer Intern vs Data Engineer?

AspectBig Data Engineer InternData Engineer
CredentialsRelevant coursework, some internshipsBachelor's or master's in CS, experience preferred
Work EnvironmentInternship, learning-focused, entry-level projectsFull-time, professional projects, team collaboration
Industry UsageTech, finance, healthcare, startupsSame industries, more responsibility
Search & Comparison IntentEntry-level, internship opportunitiesCareer advancement, full-time roles

The main difference between a Big Data Engineer Intern and a Data Engineer is experience level and responsibility. Interns are typically students or early learners gaining exposure, while Data Engineers are full-time professionals managing complex data systems. Internships serve as stepping stones toward full-time data engineering careers.

What are the most commonly searched types of Big Data Engineer jobs in California?

The most popular types of Big Data Engineer jobs in California are:

What cities in California are hiring for Big Data Engineer Intern jobs?

Cities in California with the most Big Data Engineer Intern job openings:

Data Engineer Intern

Voxelcloud

Los Angeles, CA โ€ข On-site

$123K - $148K/yr

Full-time

Medical, Dental, Vision, Retirement, PTO

Re-posted 6 days ago


Job description

Company Description
Founded in 2016, VoxelCloud is a Los Angeles-based leader worldwide in artificial intelligence (AI) analysis of medical images. Backed by Sequoia and Tencent. We help healthcare providers make better/earlier diagnoses and other clinical decisions. http://www.voxelcloud.ai
Job Description
The Data team at VoxelCloud (Westwood, Los Angeles, CA) manages and maintains large-scale medical and healthcare data at the core of all our R&D activities. Reporting to Data Team Lead, the Data Engineer intern will participate in the acquisition and manipulation of massive datasets in multi-modal formats (medical images, text(EMR), etc.) on cloud storage. The ideal candidate is an experienced data pipeline builder and data wrangler who enjoys optimizing data systems and building them from the ground up. The Data Engineer intern will support our software developers and machine learning engineers on product/research initiatives and will create an optimal data delivery pipeline that is consistent across ongoing projects. They must be self-directed and comfortable supporting the data needs of multiple teams, systems, and products. The right candidate will be excited by the prospect of optimizing or even re-designing our company's data architecture to support our next generation of products and data initiatives.
Responsibilities:
  • Create and maintain optimal data pipelines to support machine learning research and development
  • Identify, design, and implement internal process improvements: automating data QA, optimizing data delivery, re-designing infrastructure for greater scalability, etc.
  • Build the infrastructure required for optimal extraction, transformation, and loading of data from a wide variety of data sources using SQL and AWS/AliCloud big data technologies.
  • Build analytics tools that utilize the data pipeline to provide actionable insights into product utilization and operational efficiency.
  • Keep our data separated and secure across national boundaries both locally and on cloud storage.

Qualifications
  • Proficient with at least one object-oriented/object function scripting languages: Python, Java, C++, Scala, etc
  • Working SQL knowledge and experience working with relational databases, query authoring (SQL) as well as working familiarity with a variety of databases (Postgres).
  • Experience building and optimizing 'big data' data pipelines, architectures and data sets.
  • Solid understanding of information retrieval, statistics and machine learning. Experience with Computer Vision and NLP is a plus.
  • Prefer 1+ years in big data and related technology (e.g. DFS); experience with high-performance and scalable distributed system.
  • Prefer experience with AWS cloud services: EC2, EMR, RDS, Redshift
  • Skillful with automation tasks, but willing to get hands dirty for quality control.
  • Detail-oriented, well organized and self-motivated with a continuous drive to learn, explore and challenge; good communication skills and team player.
  • Experience supporting and working with cross-functional teams in a dynamic environment.
  • MS, BA/BS degree in computer science, statistics or related field.

Additional Information
We Offer...
  • An outstanding start-up culture;
  • Transparent, collaborative work environment;
  • Competitive compensation
  • Excellent Medical, Dental, and Vision coverage
  • 401k, paid Vacation and Holiday

All your information will be kept confidential according to EEO guidelines.