2

Entry Level Databricks Data Engineer Jobs in Houston, TX

Data Engineer

Houston, TX Β· On-site

$109K - $131K/yr

Design, build, and operate scalable batch and real‑time/streaming data pipelines on Databricks ... Apply software engineering discipline to data: version control, code review, automated testing, and ...

Data Engineer

Houston, TX Β· On-site

$109K - $131K/yr

Design, build, and operate scalable batch and real‑time/streaming data pipelines on Databricks ... Apply software engineering discipline to data: version control, code review, automated testing, and ...

Data Engineer

Houston, TX Β· On-site

$109K - $131K/yr

Design, build, and operate scalable batch and real-time/streaming data pipelines on Databricks and ... Apply software engineering discipline to data: version control, code review, automated testing, and ...

Data Engineer

Houston, TX Β· On-site

$109K - $131K/yr

Design, build, and operate scalable batch and real-time/streaming data pipelines on Databricks and ... Apply software engineering discipline to data: version control, code review, automated testing, and ...

Data Engineer

Houston, TX Β· On-site

$53K - $88K/yr

Guidehouse seeks a Data Engineer I to support the development, maintenance, and enhancement of data ... Exposure to Databricks, Spark, or large-scale data processing technologies. * Familiarity with ...

Data Engineer

Houston, TX Β· Hybrid

$109K - $131K/yr

Together. Summary The Data Engineer, Solutions & Data role designs, builds, and operates data ... Data pipeline tooling and cloud data services experience (Azure Data Factory, Azure Databricks ...

This is a greenfield project, and experience with Databricks or Snowflake would be highly beneficial. Requirements * Strong hands-on experience as an AWS Data Engineer * Advanced Python and SQL

New

Public Health Data Engineer

Houston, TX Β· On-site

$109K - $131K/yr

Guidehouse seeks a Data Engineer to support the development, maintenance, and enhancement of data ... Exposure to Databricks, Spark, or large-scale data processing technologies. * Familiarity with AWS ...

Experience with Databricks, Azure data services, Azure storage, or similar modern cloud data platforms * Experience with Azure DevOps or comparable tools for repositories, pipelines, release ...

Advisor II, Data Engineer

Houston, TX Β· On-site

$108K - $133K/yr

Experience with Databricks, Azure data services, Azure storage, or similar modern cloud data platforms * Experience with Azure DevOps or comparable tools for repositories, pipelines, release ...

Azure Data Engineer

Houston, TX Β· On-site

$105K - $126K/yr

Azure Data Engineer Houston TX - Onsite from day 1 Project Overview: New project- credit card ... Datalake, Databricks logic app, tableau β€’ Experience communicating with Architects, Senior ...

Data & AI Engineer

Houston, TX Β· On-site

$109K - $131K/yr

Build and optimize solutions within Databricks, including notebooks, workflows/jobs, Delta Lake ... Data Engineering . Software Development . Databricks . Snowflake . Data Integration/ETL/ELT

Azure DataLake Engineer [remote]

Houston, TX Β· On-site

$52.50 - $65.25/hr

Requirements Β· Strong Expertise in DATA Azure data lake Azure Databricks, Azure SQL, Power BI Β· Expertise to Implementing Data warehousing Solutions experience as Data Engineer in Azure Big Data ...

AI Engineer

Houston, TX Β· On-site

$55K - $187K/yr

... Microsoft Azure, Databricks, Snowflake, or related data and AI credentials - Utilizing AI ... PwC does not intend to hire experienced or entry level job seekers who will need, now or in the ...

In this role at PwC, you will apply data, algorithms, and software engineering to build and deploy ... PwC does not intend to hire experienced or entry level job seekers who will need, now or in the ...

next page

Showing results 1-20

Entry Level Databricks Data Engineer information

See Houston, TX salary details

$28.6K

$66.2K

$112.7K

How much do entry level databricks data engineer jobs pay per year?

As of Sep 14, 2026, the average yearly pay for entry level databricks data engineer in Houston, TX is $66,239.00, according to ZipRecruiter salary data. Most workers in this role earn between $49,200.00 and $75,000.00 per year, depending on experience, location, and employer.

What is an entry level Databricks data engineer?

An Entry Level Databricks Data Engineer is a professional who uses Databricks, a cloud-based data analytics platform, to design, build, and maintain data pipelines. They are responsible for preparing and processing large datasets, ensuring data quality, and enabling analytics and machine learning workflows. Typically, they work with tools such as Apache Spark, SQL, and Python, and collaborate with data analysts and data scientists to deliver data-driven solutions. As entry-level engineers, they are expected to have foundational knowledge of data engineering concepts and be eager to learn more advanced techniques on the job.

What are the key skills and qualifications needed to thrive as an entry level Databricks data engineer?

To thrive as an Entry Level Databricks Data Engineer, you need a foundational understanding of data engineering concepts, SQL, and Python or Scala, typically supported by a relevant degree in computer science or a related field. Familiarity with Databricks, Apache Spark, cloud platforms (like AWS or Azure), and optional certifications such as Databricks Data Engineer Associate are highly valuable. Strong analytical thinking, attention to detail, and effective communication skills help you collaborate with teams and solve complex data challenges. These skills and qualities are essential for building reliable data pipelines, ensuring data quality, and delivering actionable insights in a fast-paced environment.

What are some common challenges faced by entry level Databricks data engineers, and how can they effectively overcome them?

Entry-level Databricks Data Engineers often face challenges such as learning to optimize Apache Spark jobs, managing complex data pipelines, and understanding cloud-based workflows. To overcome these, it's important to dedicate time to hands-on practice with Databricks notebooks, collaborate closely with more experienced engineers, and actively participate in code reviews and team discussions. Leveraging Databricks' extensive documentation and community forums can also help troubleshoot issues and stay updated on best practices.

What are the most commonly searched types of Databricks Data Engineer jobs in Houston, TX?

The most popular types of Databricks Data Engineer jobs in Houston, TX are:

What are popular job titles related to Entry Level Databricks Data Engineer jobs in Houston, TX?

For Entry Level Databricks Data Engineer jobs in Houston, TX, the most frequently searched job titles are:

What job categories do people searching Entry Level Databricks Data Engineer jobs in Houston, TX look for?

The top searched job categories for Entry Level Databricks Data Engineer jobs in Houston, TX are:

What cities near Houston, TX are hiring for Entry Level Databricks Data Engineer jobs?

Cities near Houston, TX with the most Entry Level Databricks Data Engineer job openings:

Data Engineer

Houston, TX β€’ On-site

$109K - $131K/yr

Other

Medical, Dental, Vision, Life, Retirement, PTO

Posted 12 days ago


Job description

Fervo is building the most cost-effective, repeatable geothermal power plants in the world. Scaling that mission depends on a trustworthy, well-governed data foundation that turns raw sensor signals, drilling and completions records, and power plant telemetry into reliable, decision-ready information. The Data Engineer, within the Data & AI team, designs, builds, and operates the pipelines, models, and platforms that move data from the field to the people and systems that act on it β€” engineers, operators, data scientists, and the analytical and agentic applications built on top.

The Data Engineer owns data products end to end β€” from ingestion and modeling, through quality, governance, and serving, to monitoring in production. Working across Data Science, AI Engineering, IT Infrastructure, domain SMEs, and business stakeholders, this role establishes reusable patterns for real-time and batch processing, IoT/historian integration, data quality and entity linkage, and self-service analytics on our Azure, Databricks, and Snowflake stack. Success requires strong hands‑on engineering depth in distributed data processing, sound data modeling and architecture judgment, and pragmatism about what to ship versus what to defer.

Responsibilities Data Pipeline & Platform Engineering
  • Design, build, and operate scalable batch and real‑time/streaming data pipelines on Databricks and Azure Data Factory, landing data in Azure Data Lake Storage (ADLS) and Snowflake
  • Implement the medallion (bronze/silver/gold) architecture using Delta Lake and Delta Live Tables, with reliable incremental processing, schema evolution, and change data capture
  • Build and tune Apache Spark jobs (PySpark/Spark SQL) for large-scale, parallel data processing β€” partitioning, shuffles, caching, broadcast joins, and cost/performance optimization
  • Ingest and process high-volume IoT and historian data (sensor, SCADA, time‑series) via streaming frameworks (Structured Streaming, Event Hubs/Kafka) and micro‑batch patterns
Data Modeling, Quality & Governance
  • Model curated, analytics-ready datasets and serving layers that are well-documented, performant, and easy for downstream consumers to use
  • Implement automated data quality frameworks β€” validation, profiling, anomaly detection, freshness and completeness checks β€” with clear alerting and remediation paths
  • Build entity resolution and record linkage logic to unify wells, pads, assets, equipment, and events across heterogeneous source systems
  • Establish and enforce data governance using Unity Catalog β€” access controls, lineage, data classification, and a shared semantic/metadata layer that makes business concepts queryable and trustworthy
Reliability, CI/CD & Production Operations
  • Apply software engineering discipline to data: version control, code review, automated testing, and CI/CD pipelines (Azure DevOps or GitHub Actions) for data and infrastructure
  • Implement monitoring, logging, and observability across pipelines to support debugging, SLA tracking, cost monitoring, and continuous improvement
  • Support production incidents and platform-level issues impacting data pipelines and downstream consumers; develop runbooks and reduce toil through automation
Analytics Enablement & Collaboration
  • Partner with analysts and stakeholders to deliver datasets and semantic models that power dashboards in Power BI and Spotfire
  • Collaborate with Data Science and AI Engineering to provision clean, governed, feature-ready data for ML and agentic workflows
  • Translate domain problems from drilling, completions, production, geophysics, and power plant operations into well-scoped, reliable data products with clear ownership and success metrics
Required Qualifications
  • Bachelor’s or Master’s degree in Computer Science, Data Engineering, Software Engineering, Information Systems, Applied Mathematics, Physics, or a related technical field β€” or equivalent practical experience demonstrated through a portfolio of shipped data systems. Master’s preferred.
  • 2+ years of hands‑on experience building and operating production data pipelines, not just prototypes or notebooks
  • Deep understanding of the Apache Spark framework and distributed, parallel data processing β€” partitioning, shuffles, joins, caching, and performance tuning at scale
  • Strong programming skills in Python (PySpark) and SQL, including writing testable, maintainable production code
  • Hands‑on experience with Databricks, including Delta Lake, Delta Live Tables, and Unity Catalog
  • Experience with Azure Data Factory and Azure Data Lake Storage (ADLS), or equivalent cloud data services with willingness to work in our Azure-first environment
  • Experience with cloud data warehousing on Snowflake (or equivalent: BigQuery, Redshift, Databricks SQL)
  • Experience building both real‑time/streaming and batch pipelines (Structured Streaming, Event Hubs/Kafka, or similar)
  • Solid data modeling skills (dimensional, medallion/lakehouse, or normalized) and a track record of building well-documented, consumable datasets
  • Experience implementing data quality, validation, and observability for pipelines
  • Strong Git and CI/CD experience (Azure DevOps or GitHub Actions), including version control discipline, code review, and automated testing
  • Experience delivering data to BI/analytics tools such as Power BI and/or Spotfire
Preferred Qualifications
  • Experience with data governance and cataloging β€” Unity Catalog, Microsoft Purview, lineage tooling, and semantic/metric layers (Snowflake Semantic Views, Databricks Metric Views, dbt Semantic Layer)
  • Experience with entity resolution / record linkage and master data management across messy, multi-source data
  • Background in time-series data, signal processing, or industrial IoT (MQTT, OPC UA, SparkplugB) and historian systems (Canary, Ignition)
  • Experience with infrastructure-as-code (Terraform preferred) and containerization (Docker)
  • Experience with orchestration tooling (Databricks Workflows, Airflow, ADF pipelines) and dbt for transformation
  • Familiarity with data security and access control patterns β€” Entra ID, row/column-level security, masking, Key Vault-managed secrets
  • Oil and gas or energy industry experience, including familiarity with drilling, completions, production, geophysics, or industrial historian data
  • Exposure to ML/AI data needs β€” feature engineering, feature stores, or provisioning data for agentic/LLM applications
  • Experience optimizing cloud data cost and performance at scale
Location

Fervo Energy is headquartered in Houston, TX, with growing offices in Golden, CO, Reno, NV, and Oakland, CA, and Salt Lake City, UT. This position will be eligible for some hybrid work flexibility, but regular in‑office presence at the Oakland or Houston office will be required.

Compensation & Benefits

Fervo provides a comprehensive suite of benefits including medical, dental, vision, life, short‑term and long‑term disability, flexible paid time off, and paid parental leave. Additionally, Fervo offers an incentive stock options program, a bonus incentive program, and a 401(k) plan with an employer match.

Fervo Energy is providing the compensation range and general description of other compensation and benefits that the company in good faith believes it might pay and/or offer for this position based on the successful applicant’s education, experience, knowledge, skills, and abilities in addition to internal equity and geographic location. Expected Salary: $x – $x based on location and experience.

Fervo Energy reserves the right to ultimately pay more or less than the posted range and offer other compensation, depending on circumstances not related to an applicant’s sex or other status protected by local, state, or federal law.

Fervo Energy is an Equal Opportunity Employer and does not discriminate on the basis of race, color, creed, gender, religion, marital status, registered domestic partner status, age, national origin, ancestry, physical or mental disability, medical condition, sex, genetic information, sexual orientation, military and veteran status or any other consideration made unlawful by federal, state, or local laws. It also prohibits unlawful discrimination based on the perception that anyone has any of those characteristics or is associated with a person who has or is perceived as having any of those characteristics.

#J-18808-Ljbffr