1

Trainee Databricks Data Engineer Jobs in Pasadena, CA

Senior Databricks Data Engineer Location-Type: Onsite, Santa Monica, CA (office relocation to Century City anticipated within approximately one year) Start Date: ASAP Duration: Permanent Compensation ...

Data Engineer

Pasadena, CA

$124K - $150K/yr

Azure or Databricks certifications (e.g., Azure Data Engineer Associate, Azure Solutions Architect Expert, Databricks Data Engineer Professional) are a plus. We are GEI. Some of the world's most ...

Data Engineer

Pasadena, CA · On-site

$124K - $150K/yr

Azure or Databricks certifications (e.g., Azure Data Engineer Associate, Azure Solutions Architect Expert, Databricks Data Engineer Professional) are a plus. We are GEI. Some of the world's most ...

Principle Data Engineer

Alhambra, CA · On-site

$121K - $145K/yr

Responsibilities : • Five (5) + years of professional work experience specializing in Databricks with significant expertise in data engineering, data integration, data warehousing, and cloud ...

... Architect, Databricks Data Engineer Associate] is a plus - Designing and implementing thorough data architecture strategies - Developing and documenting data models and data flow diagrams ...

AWS Data Engineer

Los Angeles, CA · On-site

$123K - $148K/yr

The ideal candidate will have strong expertise in AWS data services, Databricks, Python, and CI/CD ... Mentor junior engineers, perform code reviews, and promote engineering best practices. Participate ...

AWS Data Engineer

Los Angeles, CA · On-site

$123K - $148K/yr

The ideal candidate will have strong expertise in AWS data services, Databricks, Python, and CI/CD ... Mentor junior engineers, perform code reviews, and promote engineering best practices.

Senior Data Engineer

Glendale, CA · On-site

$112K - $152K/yr

Databricks and Python expertise. Qualifications: * 5+ years of data engineering experience developing large data pipelines * Proficiency in at least one major programming language. * Expertise on ...

Data Engineer

Glendale, CA · On-site

$121K - $145K/yr

... Databricks, Delta Lake, Kubernetes and AWS ● Collaborate with product managers, architects, and other engineers to drive the success of the Core Data platform ● Contribute to developing and ...

Data Engineer

Los Angeles, CA · On-site

$123K - $148K/yr

Together. Summary The Data Engineer, Solutions & Data role designs, builds, and operates data ... Data pipeline tooling and cloud data services experience (Azure Data Factory, Azure Databricks ...

Build and deliver compelling proofs-of-concept and live demos on the Databricks Platform that drive technical wins * Own frontline technical relationships with customer engineers, data teams, and ...

Job Title Data Engineer Client Confidential Location Los Angeles, CA (5 days - Onsite) Type of Hire ... Experience with cloud-based environments (AWS, GCP, Databricks) and MLOps practices * Strong ...

next page

Showing results 1-20

Trainee Databricks Data Engineer information

See Pasadena, CA salary details

$48.5K

$141.5K

$193.6K

How much do trainee databricks data engineer jobs pay per year?

As of Aug 28, 2026, the average yearly pay for trainee databricks data engineer in Pasadena, CA is $141,495.00, according to ZipRecruiter salary data. Most workers in this role earn between $124,900.00 and $150,000.00 per year, depending on experience, location, and employer.

What is the difference between Trainee Databricks Data Engineer vs Junior Data Engineer?

AspectTrainee Databricks Data EngineerJunior Data Engineer
Required CredentialsBasic knowledge of Databricks, SQL, and data fundamentalsDegree in Computer Science or related field, some experience with data tools
Work EnvironmentTraining programs, mentorship, entry-level projects on Databricks platformEntry-level to mid-level data teams, real-world data projects
Employer & Industry UsageTech companies, data consulting firms, startups focusing on cloud data platformsVariety of industries including finance, healthcare, retail, with data teams

The Trainee Databricks Data Engineer is an entry-level role focused on learning Databricks and data engineering fundamentals, often within training programs. In contrast, a Junior Data Engineer typically has some hands-on experience and works on real data projects. Both roles are common in tech-driven industries, but the trainee position emphasizes skill development, while the junior role involves more independent work.

What are the most commonly searched types of Databricks Data Engineer jobs in Pasadena, CA?

The most popular types of Databricks Data Engineer jobs in Pasadena, CA are:

What are popular job titles related to Trainee Databricks Data Engineer jobs in Pasadena, CA?

For Trainee Databricks Data Engineer jobs in Pasadena, CA, the most frequently searched job titles are:

What job categories do people searching Trainee Databricks Data Engineer jobs in Pasadena, CA look for?

The top searched job categories for Trainee Databricks Data Engineer jobs in Pasadena, CA are:

What cities near Pasadena, CA are hiring for Trainee Databricks Data Engineer jobs?

Cities near Pasadena, CA with the most Trainee Databricks Data Engineer job openings:

AWS Databricks Data Engineer

Tror AI for everyone

Los Angeles, CA • On-site

$123K - $148K/yr

Contractor

Re-posted 17 days ago


Job description

Job Title: AWS Databricks Data Engineer

Job Location: Los Angeles CA (Hybrid)

Hire type: FTE / CTH

Note: Only Locals to California

 

Job Description –

We are seeking a highly skilled AWS Data Engineer with strong expertise in SQL, Python, PySpark, Data Warehousing, and Cloud-based ETL to join our data engineering team. The ideal candidate will design, implement, and optimize large-scale data pipelines, ensuring scalability, reliability, and high performance. This role requires close collaboration with cross-functional teams and business stakeholders to deliver modern, efficient data solutions.

Key Responsibilities

1. Data Pipeline Development

  • Build and maintain scalable ETL/ELT pipelines using Databricks on AWS.
  • Leverage PySpark/Spark and SQL to transform and process large, complex datasets.
  • Integrate data from multiple sources including S3, relational/non-relational databases, and AWS-native services.

2. Collaboration & Analysis

  • Partner with downstream teams to prepare data for dashboards, analytics, and BI tools.
  • Work closely with business stakeholders to understand requirements and deliver tailored, high‑quality data solutions.

3. Performance & Optimization

  • Optimize Databricks workloads for cost, performance, and efficient compute utilization.
  • Monitor and troubleshoot pipelines to ensure reliability, accuracy, and SLA adherence.
  • Apply query optimization, Spark tuning, and shuffle minimization best practices when handling tens of millions of rows.

4. Governance & Security

  • Implement and manage data governance, access control, and security policies using Unity Catalog.
  • Ensure compliance with organizational and regulatory data‑handling standards.

5. Deployment & DevOps

  • Use Databricks Asset Bundles for deployment of jobs, notebooks, and configuration across environments.
  • Maintain effective version control of Databricks artifacts using GitLab or similar tools.
  • Use CI/CD pipelines to support automated deployments and environment setups.

Technical Skills (Required)

  • Strong expertise in Databricks (Delta Lake, Unity Catalog, Lakehouse Architecture, Table Triggers, Workflows, Delta Live Pipelines, Databricks Runtime, etc.).
  • Proven ability to implement robust PySpark solutions.
  • Hands‑on experience with Databricks Workflows & orchestration.
  • Solid knowledge of Medallion Architecture (Bronze/Silver/Gold).
  • Significant experience designing or rebuilding batch‑heavy data pipelines.
  • Strong background in query optimization, performance tuning, and Spark shuffle optimization.
  • Ability to handle and process tens of millions of records efficiently.
  • Familiarity with Genie enablement concepts (understanding required; deep experience optional).
  • Experience with CI/CD, environment setup, and Git-based development workflows.
  • Solid understanding of AWS cloud, including:
  • IAM
  • Networking fundamentals
  • Storage integration (S3, Glue Catalog, etc.)

Preferred Experience

  • Experience with Databricks Runtime configurations and advanced features.
  • Knowledge of streaming frameworks such as Spark Structured Streaming.
  • Experience developing real-time or near real-time data solutions.
  • Exposure to GitLab pipelines or similar CI/CD systems.

Certifications (Optional)

  • Databricks Certified Data Engineer Associate / Professional
  • AWS Data Engineer or AWS Solutions Architect certification

Thanks & Regards

Akhil

akhil@tror.ai