1

Distinguished Data Engineer Jobs (NOW HIRING)

As a Distinguished Engineer on the Data Engineering team, you'll own some of the hardest infrastructure problems at CloudZero: shaping the next-generation streaming data platform, the dimensional ...

AVP, Lead Data Engineer

Philadelphia, PA

$109K - $131K/yr

The company is distinguished by its extensive product and service offerings, broad distribution ... data engineering experience, including hands-on ETL/ELT development, data warehouse design, and ...

COSMOS - Data Engineer III

Little Rock, AR ยท On-site

$109K - $131K/yr

The Data Engineer III will collect, manage, and convert raw data into usable information for ... Nitin Agarwal (nxagarwal@ualr.edu), Maulden-Entergy Endowed Chair and Distinguished Professor and ...

COSMOS - Data Engineer III

Little Rock, AR ยท On-site

$109K - $131K/yr

Nitin Agarwal (nxagarwal@ualr.edu), Maulden-Entergy Endowed Chair, Distinguished Professor, and ... Lead a team of data engineers; * Collecting and analyzing raw data from various sources, including ...

COSMOS - Data Engineer IV

Little Rock, AR ยท On-site

$109K - $131K/yr

The Data Engineer IV will collect, manage, and convert raw data into usable information for ... Nitin Agarwal (nxagarwal@ualr.edu), Maulden-Entergy Endowed Chair and Distinguished Professor and ...

COSMOS - Data Engineer IV

Little Rock, AR

$109K - $131K/yr

Nitin Agarwal (nxagarwal@ualr.edu), Maulden-Entergy Endowed Chair, Distinguished Professor, and ... Lead a team of data engineers; * Collecting and analyzing raw data from various sources, including ...

AI - Data Engineer

Nashville, TN ยท On-site

$110K - $132K/yr

Distinguished by its culture of collaboration and exceptional client service, CAA's diverse ... Strong programming skills in Python and experience building data and AI pipelines * Experience with ...

The Project/Program Specialist (Data Engineer II) position is a full-time provisional with the ... Nitin Agarwal (nxagarwal@ualr.edu), Maulden-Entergy Endowed Chair and Distinguished Professor and ...

Showing results 41-60

Distinguished Data Engineer information

See salary details

$44.5K

$129.7K

$177.5K

How much do distinguished data engineer jobs pay per year?

As of Aug 6, 2026, the average yearly pay for distinguished data engineer in the United States is $129,716.00, according to ZipRecruiter salary data. Most workers in this role earn between $114,500.00 and $137,500.00 per year, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive as a distinguished data engineer?

To thrive as a Distinguished Data Engineer, you need advanced expertise in data architecture, distributed systems, and programming languages such as Python, Java, or Scala, often supported by a degree in computer science or a related field. Familiarity with big data platforms (e.g., Hadoop, Spark), cloud data services (AWS, GCP, Azure), and relevant certifications like Google Professional Data Engineer or AWS Certified Data Analytics are typical. Exceptional problem-solving, leadership, and communication skills help drive cross-functional initiatives and mentor teams. These skills and qualities are crucial for designing scalable data solutions, ensuring data integrity, and delivering strategic business value.

What is the difference between Distinguished Data Engineer vs Data Engineer?

AspectDistinguished Data EngineerData Engineer
Required CredentialsBachelor's or Master's in CS, certifications like Google Cloud Professional Data EngineerBachelor's in CS or related field, certifications optional
Work EnvironmentLarge tech companies, research institutions, specialized projectsVaried industries, startups, enterprises
Employer & Industry UsageUsed in organizations emphasizing innovation and advanced data solutionsCommon across industries for building data pipelines

The main difference is that a Distinguished Data Engineer typically holds a higher level of expertise, often with advanced certifications and experience working on complex projects in leading organizations. Data Engineers focus on designing and maintaining data pipelines across various industries. The Distinguished Data Engineer role is more specialized and recognized for leadership in data engineering innovation.

What are some common challenges faced by distinguished data engineers in cross-functional teams?

Distinguished Data Engineers often collaborate with product managers, data scientists, and software engineers to architect large-scale data solutions. A common challenge is bridging the gap between technical complexity and business needs, ensuring solutions are both robust and aligned with organizational goals. Additionally, they may need to mentor junior engineers while balancing project deadlines and evolving technology stacks. Effective communication and strong leadership skills are essential to address these challenges and drive impactful data initiatives.

What is a distinguished data engineer?

Distinguished Data Engineers are highly experienced technical experts recognized for their deep knowledge and leadership in designing, building, and optimizing large-scale data systems. They often set technical direction, mentor teams, and drive innovation within organizations. Unlike typical data engineers, distinguished data engineers influence company-wide strategies, establish best practices, and contribute to the broader data engineering community. Their role combines advanced technical skills with strategic vision and thought leadership.
More about Distinguished Data Engineer jobs
Infographic showing various Distinguished Data Engineer job openings in the United States as of August 2026, with employment types broken down into 1% As Needed, 83% Full Time, 12% Part Time, and 4% Contract. Highlights an 87% Physical, 3% Hybrid, and 10% Remote job distribution, with an average salary of $129,716 per year, or $62.4 per hour.

Distinguished Engineer, Data Platform

CloudZero

San Francisco, CA โ€ข Remote

$275K - $330K/yr

Full-time

Re-posted 14 days ago


Job description

About the Role

CloudZero is growing fast. Our customer base is expanding, the data challenges we're solving are getting more complex, and the platform is scaling to match. As a Distinguished Engineer on the Data Engineering team, you'll own some of the hardest infrastructure problems at CloudZero: shaping the next-generation streaming data platform, the dimensional cost model underlying every attribution decision, the hot/cold storage architecture serving both real-time and historical queries, and the query engine that powers our entire product.

This is real platform architecture work at real scale, not a consulting role or a review-and-advise job. You'll define the roadmap, drive the foundational decisions, and be a force multiplier for a talented engineering team — evolving CloudZero from batch-oriented pipelines toward a streaming-first architecture where cost attribution reaches engineers within seconds of a resource being used, not the next morning.

This role is ideal for an architect who has built systems like this before, has the scars to prove it, and wants to see their decisions matter in direct and measurable ways for customers and for the business.

What You'll Do

Define the Data Platform Architecture

  • Lead end-to-end technical design for CloudZero's next-generation data platform, from event ingestion and stream processing through hot/cold storage and the query layer to the API surface

  • Document architectural decisions, tradeoffs, and migration strategies with the rigor of an RFC-driven process

  • Shape and drive every layer of the new architecture: event ingestion, stream processing and enrichment, real-time serving, analytical storage, query layer, and API

Drive Streaming Infrastructure to Production

  • Design and deliver CloudZero's real-time data pipeline from ingestion through enrichment to serving

  • Establish SLOs for throughput, latency, and correctness, and build the operational playbooks that make this system trustworthy enough to replace the batch pipelines our entire product currently depends on

  • Tackle real-time streaming at scale across thousands of customers simultaneously, with fault tolerance, backpressure awareness, and correctness as non-negotiables

Tackle the Dimension Cardinality Problem

  • Redesign CloudZero's dimensional cost model to support high-cardinality, multi-dimensional cost attribution without runaway materialization costs

  • Drive incremental, delta-based materialization strategies using modern open table formats, dramatically reducing expensive full-rebuild jobs and unlocking millions in annual infrastructure savings

Evolve the Query Layer

  • Assess CloudZero's current query infrastructure, drive in-flight migrations to completion, and lead the evolution of the query engine layer going forward

  • Own performance optimization across partition pruning, predicate pushdown, and query planning, and set the vision for how the query layer grows as data volumes scale 10x

Extend Cost Attribution to Real-Time

  • Evolve CloudZero's proprietary cost attribution engine from a batch-oriented model to one that assigns complex cost dimensions by team, feature, and customer within seconds of resource usage

  • Rethink enrichment, data lineage, and correctness guarantees in a streaming context

Shape the Data Engineering Roadmap

  • Partner with product, infrastructure, and analytics engineering to define a multi-year data platform roadmap

  • Build consensus across engineering leadership on foundational investments including table formats, streaming frameworks, query engines, and schema management

Elevate the Engineering Team

  • Participate in architecture reviews, contribute to design patterns and best practices, and mentor senior and staff engineers through code review, pairing, and structured feedback

  • Make everyone around you better, not by directing, but by raising the collective craft

What You Bring

Data Platform & Architecture

  • 10+ years in data engineering with a clear trajectory toward principal or staff-level architecture

  • Built and operated large-scale data platforms serving tens of millions of events per day in production

  • Deep experience with streaming systems such as Kafka, Kinesis, Flink, or Spark Streaming at real production throughput

  • Strong hands-on fluency with modern open table formats including Apache Iceberg, Delta Lake, and Hudi, including compaction, partitioning strategy, and time-travel queries

  • Designed hot/cold storage architectures with explicit latency SLOs per tier

  • Proven ability to drive a data platform end to end, not just a single layer

Data Modeling & Dimensional Design

  • Expert in dimensional data modeling including fact/dimension schema design, slowly changing dimensions, and cardinality management

  • Deep understanding of the materialization tradeoff space: full vs. incremental, push vs. pull, pre-aggregate vs. query-time

  • Experience with cost attribution, showback/chargeback, or multi-tenant data partitioning patterns

  • Strong SQL and query optimization background across predicate pushdown, partition pruning, and cost-based query planning

Query Engines & Compute

  • Hands-on with distributed query engines such as Trino, Presto, Spark SQL, or DuckDB including configuration, optimization, and production operations

  • Understands catalog and metadata management and how it couples to query engines

  • Comfortable with cloud data warehouses such as Snowflake, BigQuery, and Redshift and how they integrate with open table formats

  • Experience driving query engine migrations while maintaining production SLAs

Engineering Leadership

  • Track record as a technical anchor for a data platform or data engineering team

  • Writes clear ADRs, RFCs, and technical design docs that bring engineers along

  • Can drive multi-month, multi-team technical initiatives from inception to production without heavy process overhead

  • Communicates complex tradeoffs to non-technical stakeholders including product and business leadership

  • Comfortable in a high-autonomy environment: builds consensus, influences through expertise, and helps teams move forward

Bonus If You Have...
  • FinOps or cloud cost domain experience

  • Multi-cloud data ingestion across AWS, Azure, and GCP

  • Apache Flink at production scale

  • Lakehouse architecture patterns

  • Real-time feature engineering for ML

  • Data mesh or domain-oriented design patterns

  • Prior startup or high-growth SaaS experience

  • Open source contributions to the data ecosystem

About CloudZero
CloudZero is the AI ROI Company. We built the financial control plane for AI: the system finance, IT, and engineering use to connect every AI dollar to the outcome it produced. Across every provider. In real time.
AI spend is the fastest-growing line on enterprise P&Ls and the least understood. Only 14% of CFOs can prove AI ROI today. CloudZero answers the question no one else can: what did it cost to produce this outcome, for this customer, on this model.
The largest cloud spenders on the planet already run on CloudZero, including Coinbase, Duolingo, DoorDash, and Shutterstock. We processed 14 trillion billing events in the last twelve months. We're the first listed partner on Anthropic's cost and usage API. We've raised over $119 million, including a $56 million Series C backed by leading venture capital firms.

Why Join Our Team?
At CloudZero, you’ll find a collaborative, fast-moving environment where your work makes a direct impact. We’re a team that values ownership, creativity, and curiosity — and we’re tackling some of the most complex challenges in the cloud space. If you’re excited by working with cutting-edge technology, driving meaningful outcomes, and growing with a company that’s scaling fast, we’d love to hear from you!

Compensation Range: $275K - $330K