2

Data Science Remote Internship Jobs in Berkeley, CA

... science, data engineering, or applied data role, ideally with exposure to messy, real-world or ... moving, remote environment with ambiguous, evolving priorities Salary Range: Salary ranges are ...

Data Engineer

San Francisco, CA ยท On-site +1

$145K/yr

This is a remote position. Duties * Support production systems and help triage issues during live ... Knowledge of data science and machine learning concepts * Professional working experience with MLB ...

Data Engineer

San Francisco, CA ยท On-site +1

$160K/yr

This is a remote position. Duties * Support production systems and help triage issues during live ... Knowledge of data science and machine learning concepts * A strong interest in sports and sports ...

Sr. Data Platform Engineer

San Francisco, CA ยท On-site +1

$134K - $161K/yr

You will work closely with data engineers, data scientists, ML/AI engineers, BI developers, and ... Employee divides their time between in-office and remote work. Access to an office location is ...

Support data science workflows by building infrastructure that enables model training, feature ... We are a remote-first company for most positions so you may work from anywhere you like in the U.S ...

Data Engineer

San Francisco, CA ยท Remote

$145K/yr

This is a remote position. Duties * Support production systems and help triage issues during live ... Knowledge of data science and machine learning concepts * Professional working experience with MLB ...

Data Engineer

San Francisco, CA ยท Remote

$160K/yr

This is a remote position. Duties * Support production systems and help triage issues during live ... Knowledge of data science and machine learning concepts * A strong interest in sports and sports ...

Software Engineer, Data Engineering

San Francisco, CA ยท On-site +1

$134K - $162K/yr

United States (remote) What is Verse? The race to AI has become the race to power. Every ... Partnering closely with our Data Science and data-heavy internal teams to support various ...

Showing results 41-60

Data Science Remote Internship information

See Berkeley, CA salary details

$14

$27

$51

How much do data science remote internship jobs pay per hour?

As of Aug 23, 2026, the average hourly pay for data science remote internship in Berkeley, CA is $27.56, according to ZipRecruiter salary data. Most workers in this role earn between $21.20 and $30.00 per hour, depending on experience, location, and employer.

What is a data science remote internship?

A Data Science Remote Internship is a temporary, practical work experience opportunity in the field of data science that is completed remotely, usually from your own home or any location with internet access. Interns work on real-world projects involving data analysis, machine learning, and statistical modeling, often collaborating with teams through online communication tools. This type of internship is ideal for gaining hands-on experience, building a portfolio, and developing skills relevant to data science careers, all while offering flexibility and eliminating the need to relocate.

What are the key skills and qualifications needed to thrive as a data science remote intern?

To thrive as a Data Science Remote Intern, you need a solid understanding of statistics, data analysis, and programming languages like Python or R, often supported by coursework or projects in data science or related fields. Familiarity with tools such as Jupyter Notebook, SQL, and machine learning libraries (e.g., scikit-learn, TensorFlow) is typically expected. Strong problem-solving abilities, self-motivation, and effective communication are essential soft skills for collaborating remotely and conveying analytical insights. These competencies ensure you can independently contribute to projects, adapt to remote workflows, and deliver actionable data-driven solutions.

What types of projects can I expect to work on during a remote data science internship, and how is project collaboration typically managed?

During a remote data science internship, you can expect to work on projects such as data cleaning, exploratory data analysis, model development, and visualization tasks that support ongoing business needs. Collaboration is commonly managed through virtual tools like Slack, Zoom, and project management platforms (e.g., Jira or Trello), with regular check-ins and code reviews from your mentor or team. Interns often participate in team meetings, contribute to group presentations, and use version control systems like Git to share code and receive feedback. This structure ensures you gain practical experience while staying connected with your team, even in a remote setting.

What is the difference between Data Science Remote Internship vs Data Analyst Remote Internship?

AspectData Science Remote InternshipData Analyst Remote Internship
Required CredentialsTypically pursuing or recent graduate in Data Science, Statistics, or related fieldsOften pursuing or recent graduate in Data Analysis, Business, or related fields
Work EnvironmentRemote, collaborative with data science teams, using programming languages like Python or RRemote, focusing on data interpretation, visualization, and reporting tools like Excel, SQL, Tableau
Employer & Industry UsageTech companies, finance, healthcare, startupsBusiness, marketing, finance, consulting firms

While both roles involve working with data remotely, Data Science Remote Internships focus on building predictive models and programming skills, whereas Data Analyst Remote Internships emphasize data interpretation, visualization, and reporting. The choice depends on your career goals and skill set.

What are popular job titles related to Data Science Remote Internship jobs in Berkeley, CA?

For Data Science Remote Internship jobs in Berkeley, CA, the most frequently searched job titles are:

What cities near Berkeley, CA are hiring for Data Science Remote Internship jobs?

Cities near Berkeley, CA with the most Data Science Remote Internship job openings:

Chemical Data Scientist

Valdera

San Francisco, CA โ€ข On-site, Remote

Full-time

Re-posted 7 days ago


Job description

About Valdera:
At Valdera, we empower innovators to turn ideas into reality by transforming how manufacturers source materials. We make it effortless for companies to find the best materials and suppliers for their needs, enabling them to build high-quality products at scale and deliver them to millions of consumers worldwide.
We are a team of ambitious, results-driven individuals with a proven track record of working with Fortune 500 industrial manufacturers, beauty brands, and chemical companies. We are a fast-growing company that hires talented, hardworking people who excel in high-performance environments and want to grow their careers quickly.
Our culture is built for exceptional individuals to take on meaningful challenges, collaborate with the top minds in our industry, and see the direct impact of their work. If you're looking for a fast-paced environment where your ideas will drive real change, Valdera is the place for you.
Join us, and let's shape the future of manufacturing together.
Role Description:
We are hiring a Chemical Data Scientist to build and maintain the pipelines that keep Valdera's supplier and chemical product data accurate, current, and structured - the foundation every buyer and supplier relies on across Valdera's procurement platform.
Data quality plays a critical role at Valdera. When a buyer launches a request, they expect to be matched with the right suppliers and accurate specs on the first try. Delivering that depends on clean, current data - CAS numbers, specifications, certifications, and regulatory documents pulled from thousands of inconsistent, often messy sources. This requires strong data engineering fundamentals and a working knowledge of chemical industry data. For example, you might take dozens of differently structured chemical supplier catalogs and turn them into one clean, standardized product database.
You will take ownership of the full data pipeline - from scrapers and ETL workflows to data cleaning, matching, and classification models that connect suppliers to buyer requirements. You're energized by messy, real-world data and confident partnering with Supplier Management and Engineering to close coverage gaps. As a data-obsessed professional, you're dedicated to the accuracy our buyers and suppliers depend on.
Role Responsibilities:
  • Design and build pipelines to collect supplier data and chemical product information (specifications, CAS numbers, certifications, SDS/regulatory documents, NAICS classification of manufacturing plants) from supplier sites, distributor catalogs, trade databases, and other public and semi-structured sources
  • Develop and maintain web scrapers and automated ETL workflows to keep supplier and product data current at scale
  • Clean, normalize, and reconcile inconsistent supplier data into structured, standardized formats suitable for internal tools and analytics
  • Apply chemical domain knowledge to validate and enrich data - resolving product names, CAS numbers, synonyms, and specifications across suppliers
  • Evaluate and improve matching and classification models to map suppliers and products to buyer requirements, and to identify overlapping or equivalent chemical offerings
  • Partner with Supplier Management and Engineering to define data quality standards, identify gaps in supplier coverage, and prioritize new data sources.
  • Own pipeline health and data quality, and drive the KPIs that measure overall data coverage

Experience & Qualifications:
  • 5+ years of experience in a data science, data engineering, or applied data role, ideally with exposure to messy, real-world or industrial datasets.
  • Working knowledge of chemistry or chemical industry data - comfort with CAS numbers, chemical properties, SDS documents, NAICS classification, and supplier certifications
  • Strong Python skills, with experience building web scrapers and data pipelines
  • Experience with data cleaning and normalization at scale, and a good eye for spotting inconsistencies in unstructured data
  • Familiarity with building or applying matching, deduplication, or classification models (traditional ML or LLM-based approaches)
  • Hands-on experience using AI tools and LLMs to accelerate data extraction, enrichment, or engineering workflows
  • Startup mindset with a strong sense of ownership - comfortable working independently in a fast-moving, remote environment with ambiguous, evolving priorities

Salary Range:
Salary ranges are determined by multiple factors, including the labor market, market compensation bands, internal parity, and budget considerations. The final offer will be based on the candidate's individual skills, qualifications, location, and experience relative to the requirements of the role.
Benefits:
Valdera offers generous benefits to employees. You will be provided a more detailed breakdown of your options prior to joining Valdera.
Equal Opportunity Employer Statement:
Valdera is an equal-opportunity employer committed to building a diverse and inclusive team. We welcome applicants of all backgrounds and celebrate a culture that values varied perspectives, skills, and experiences. We are dedicated to maintaining a workplace free from discrimination, where everyone feels valued, respected, and empowered to contribute.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.