1

Duckdb Jobs in Atlanta, GA (NOW HIRING)

Lead Data Engineer

Alpharetta, GA · On-site

$120 - $160/hr

Trino / Presto / DuckDB * CDC & streaming: Debezium (SQL Server CDC, Postgres logical replication), Kafka / Redpanda * Orchestration: Dagster (or Airflow) * Storage: S3 / MinIO * SQL Server and ...

Lead Data Engineer

Alpharetta, GA · On-site

$111K - $134K/yr

Trino / Presto / DuckDB CDC & streaming: Debezium (SQL Server CDC, Postgres logical replication), Kafka / Redpanda Orchestration: Dagster (or Airflow) Storage: S3 / MinIO SQL Server and PostgreSQL ...

Lead Data Engineer

Alpharetta, GA · On-site

$111K - $134K/yr

Trino / Presto / DuckDB CDC & streaming: Debezium (SQL Server CDC, Postgres logical replication), Kafka / Redpanda Orchestration: Dagster (or Airflow) Storage: S3 / MinIO SQL Server and PostgreSQL ...

Lead Data Engineer

Alpharetta, GA

$111K - $134K/yr

Trino / Presto / DuckDB · CDC & streaming: Debezium (SQL Server CDC, Postgres logical replication), Kafka / Redpanda · Orchestration: Dagster (or Airflow) · Storage: S3 / MinIO · SQL Server and ...

Lead Data Engineer

Alpharetta, GA · On-site

$111K - $134K/yr

Trino / Presto / DuckDB • CDC & streaming: Debezium (SQL Server CDC, Postgres logical replication), Kafka / Redpanda • Orchestration: Dagster (or Airflow) • Storage: S3 / MinIO • SQL Server ...

Duckdb information

What is DuckDB?

DuckDB is an in-process SQL OLAP (Online Analytical Processing) database management system designed for fast analytical data processing. Unlike traditional client-server databases, DuckDB runs directly within your application and works well with data science tools and workflows. It is lightweight, easy to install, and supports SQL queries on data stored in CSV, Parquet, and other formats, making it popular for data analysis and research use cases.

What are some common challenges faced by professionals working with DuckDB in data engineering roles?

Professionals using DuckDB in data engineering often encounter challenges such as optimizing query performance for large-scale datasets, integrating DuckDB with existing data pipelines, and ensuring compatibility with various data formats and sources. Additionally, since DuckDB is relatively new compared to other database systems, users may find limited community support or resources for advanced use cases. Collaborating with data scientists and analysts is key, as DuckDB is often used for interactive analytics, requiring close communication to tailor solutions that meet both performance and usability needs.

What are the key skills and qualifications needed to thrive as a DuckDB developer, and why are they important?

To thrive as a DuckDB Developer, you need strong SQL proficiency, a background in database management, and experience with analytical data processing, ideally supported by a degree in computer science or a related field. Familiarity with DuckDB, data integration tools, and programming languages such as Python or R is essential. Attention to detail, problem-solving abilities, and effective communication are valuable soft skills in this role. These skills ensure efficient data analysis, seamless database integration, and the ability to convey insights to technical and non-technical stakeholders.

What is the difference between Duckdb vs Data Analyst?

AspectDuckdbData Analyst
Primary RoleEmbedded analytical database engine for data processingInterprets data, creates reports, and provides insights
Required SkillsSQL, data management, database integrationData visualization, statistical analysis, SQL
Work EnvironmentDevelopers, data engineers, embedded systemsBusiness environments, analytics teams, offices
CertificationsNone specific, SQL knowledge preferredCertified Data Analyst, SQL certifications

Duckdb is a database engine used mainly by developers and data engineers for embedded data processing, while Data Analysts focus on interpreting data and generating insights. Both roles require SQL skills, but their work environments and objectives differ significantly.

What can I do with DuckDB?

As a data analyst or developer, you can use DuckDB to perform efficient in-process SQL queries on large datasets, often within Python or R environments. It supports complex analytical queries, data transformation, and integration with data science workflows, making it suitable for data analysis, machine learning, and reporting tasks.
Infographic showing various Duckdb job openings in Atlanta, GA as of August 2026, with employment types broken down into 96% Full Time, and 4% Contract. Highlights an 84% Physical, 5% Hybrid, and 11% Remote job distribution.

Lead Data Engineer

Navanta, LLC

Alpharetta, GA • On-site

$120 - $160/hr

Other

Posted 6 days ago


Job description

Overview

Summary/Objective

The Lead Data Engineer owns the Navanta data backbone — public Call Report data in the early build, and secure ingestion from bank cores into lakehouses as each client’s on-premises environment is stood up. Working under the SVP of Technology and Commercial AI and in close partnership with the AI/ML, security, and platform teams, this role builds the architecturally clean, well-modeled, reconcilable data foundation that makes it possible for the Navanta AI platforms to give numbers a banker will act on.

Responsibilities Essential Functions
  • Design the lakehouse: Apache Iceberg (or similar technology) on object storage, a catalog for table management and per-bank isolation, dbt models, and a query engine
  • Build secure, least-privilege ingestion from bank systems — log-based CDC where permitted, with query-based and batch/SFTP fallbacks, plus an in-bank collector pattern
  • Own data modeling for the semantic and metric layer (deposits, concentration, uninsured exposure, asset quality, and peer groups)
  • Handle schema drift, data quality, and reconciliation; make ingestion observable and recoverable
  • Partner with the AI/ML team on the structured-query path and with Security on PII classification at landing, in alignment with regulatory data-handling requirements
  • Document data lineage, transformation logic, and access controls to support audit and exam readiness
  • Define and enforce data contracts, quality thresholds, and alerting for pipeline failures
Core Competencies
  • End-to-end ownership of ingestion-through-serving pipelines, with a bias toward reliability and observability
  • Rigorous data modeling for analytics — semantic layers, metric definitions, and reconcilable outputs
  • Security and compliance mindset: PII handling, least-privilege access, and data governance aligned to regulatory guidance
  • Cross-functional partnership with AI/ML and platform engineering to deliver governed, queryable data products
KPIs
  • Data freshness and pipeline reliability — SLAs met for data ingestion and bank-core feeds
  • Data quality score across key metrics versus source reconciliation
  • Time to onboard a new bank’s data environment, from kickoff to queryable lakehouse
  • PII classification coverage at landing and zero unauthorized data-access incidents
  • Semantic layer adoption — percentage of assistant queries resolved via governed metrics versus ad hoc SQL
Qualifications Education and/or Experience
  • Bachelor’s degree in computer science, mathematics, information systems, or a related field, or equivalent hands‑on experience
  • Experience in the financial services industry or a regulated data environment strongly preferred
Work Structure & Expectations
  • Full-time role combining ongoing pipeline operations with initiative-based lakehouse build-out and new bank onboarding
  • Close collaboration with AI/ML, platform engineering, and security teams; on-call rotation covering data pipeline reliability
Physical Demands

The physical demands described here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions.

While performing the duties of this job, the employee is regularly required to sit and use hands to finger, handle, or touch objects, tools, or controls. The employee frequently is required to talk or hear. The employee is occasionally required to stand; walk; and stoop, kneel, crouch, or crawl. The employee must occasionally lift and/or move up to 10 pounds, usually waist high, up to 50 feet away. Specific vision abilities required by this job include close vision and the ability to adjust focus.

Work Environment

The work environment characteristics described here are representative of those an employee encounters while performing the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions.

  • Typical office environment
  • Up to 20% travel time may be required
Core Technologies
  • Languages: Python, SQL (deep)
  • Lakehouse & catalog: Apache Iceberg; Polaris / Nessie / Lakekeeper
  • Transform & query: dbt; Trino / Presto / DuckDB
  • CDC & streaming: Debezium (SQL Server CDC, Postgres logical replication), Kafka / Redpanda
  • Orchestration: Dagster (or Airflow)
  • Storage: S3 / MinIO
  • SQL Server and PostgreSQL data modeling, pgvector (or equivalent)
Nice to Have
  • Experience with financial or core-banking data, or FFIEC / Call Report data specifically
  • Strong SQL Server familiarity
  • Data contracts, lineage, and governance practices
Who is Navanta?

Navanta is the trusted technology and services partner for community financial institutions, unifying critical systems, security, cloud infrastructure, and support into one seamless, purpose built experience. With more than 35 years of banking expertise — from Managed IT to Core Banking, CRM, and Advisory Services — Navanta helps institutions simplify complexity, reduce risk, and strengthen daily operations. Navanta empowers community bankers and their people to thrive together. Go Bankers, Go.

#J-18808-Ljbffr