2

Remote Databricks Developer Jobs in Sterling, VA

Data Engineer

Washington, DC ยท On-site +1

$129K - $155K/yr

Location: 100% Remote Years' Experience: 5+ years Professional Experience Education: Bachelor ... Experience with Databricks, Structured Streaming, Delta Lake concepts, and Delta Live Tables ...

Data Engineer

Chantilly, VA ยท On-site +1

$77K - $176K/yr

  • Medical

  • Life

  • Retirement

  • PTO

Remote Work: No Job Number: R0245235 Location: Chantilly,VA,US Share job via: Share Data Engineer ... Experience with Databricks * Experience with data warehousing using AWS Redshift, MySQL, or ...

Data Engineer

Arlington, VA ยท On-site +1

$62K - $141K/yr

  • Medical

  • Life

  • Retirement

  • PTO

Remote Work: Hybrid Job Number: R0242942 Location: Arlington,VA,US Share job via: Share Data ... Experience with Databricks, including Delta Lake, Spark, notebooks, and workflows or jobs, and ...

Data Engineer

Chantilly, VA ยท On-site +1

$77K - $176K/yr

  • Medical

  • Life

  • Retirement

  • PTO

Remote Work: No Job Number: R0245362 Location: Chantilly,VA,US Share job via: Share Data Engineer ... Experience with Databricks * Experience with data warehousing using AWS Redshift, MySQL, or ...

Data Engineer

Arlington, VA ยท On-site +1

$62K - $141K/yr

  • Medical

  • Life

  • Retirement

  • PTO

Remote Work: Hybrid Job Number: R0242536 Location: Arlington,VA,US Share job via: Share Data ... Experience with Databricks, including Delta Lake, Spark, notebooks, and workflows or jobs

AI/ML Operationalization Specialist

Mclean, VA ยท On-site +1

$98K - $163K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

You will work across both on site and remote environments to ensure the reliable deployment of ML ... Collaborate with data scientists, engineers, and mission stakeholders to ensure models meet ...

New

AI/ML Operationalization Specialist

Washington, DC ยท On-site +1

$98K - $163K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

You will work across both on site and remote environments to ensure the reliable deployment of ML ... Collaborate with data scientists, engineers, and mission stakeholders to ensure models meet ...

New

Senior Data Engineer - Remote (USA)

Reston, VA ยท Remote

$110K - $149K/yr

Work with DevOps engineers on CI, CD, and IaC * Read specs and translate them into code and design ... Experience building job workflows with the Databricks platform * Strong understanding of AWS ...

This position is currently remote; however, in accordance with federal contract requirements and ... Working knowledge of a data/lakehouse platform (Databricks, Snowflake) and SQL beyond CRUD.

Senior AI Platform & Data Engineer

Mclean, VA ยท Remote

$107K - $145K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

Senior AI Platform & Data Engineer Job Number: 868 This is a remote position. Ad Hoc is a ... Python (primary language), SQL, Spark/Databricks (for large-scale data processing), Kubernetes/ECS ...

Agentic Solutions Engineer

Mclean, VA ยท On-site +1

$69K - $158K/yr

  • Medical

  • Life

  • Retirement

  • PTO

Remote Work: Hybrid Job Number: R0244800 Location: McLean,VA,US Share job via: Share Agentic ... data lakes powered by Databricks, writing Python across notebooks, scripts, and full web ...

Showing results 41-60

Remote Databricks Developer information

What is a remote Databricks developer?

A Remote Databricks Developer is a software professional who specializes in building, managing, and optimizing data pipelines and analytics workflows on the Databricks platform, while working from a remote location. They use Databricks, which is based on Apache Spark, to process large datasets, develop ETL processes, implement machine learning models, and collaborate with data teams. Their responsibilities often include writing code in languages like Python, Scala, or SQL, integrating with cloud services, and ensuring data quality and security. Working remotely, they communicate with teams online and use cloud-based tools to complete their tasks efficiently.

How does a remote Databricks developer typically collaborate with cross-functional teams while working from different locations?

Remote Databricks Developers often work closely with data engineers, data scientists, and business analysts through virtual collaboration tools like Slack, Jira, and Zoom. Since team members may be distributed across various time zones, clear communication, regular stand-up meetings, and thorough documentation are essential for ensuring alignment on project goals and deadlines. Developers are also expected to participate in code reviews and shared knowledge sessions to maintain coding standards and support a collaborative environment. This structure helps ensure that complex data solutions are delivered efficiently and meet business requirements.

What are the key skills and qualifications needed to thrive as a remote Databricks developer, and why are they important?

To thrive as a Remote Databricks Developer, you need strong expertise in data engineering, programming languages like Python or Scala, and experience with big data frameworks, typically supported by a degree in computer science or a related field. Proficiency with Databricks platform, Apache Spark, cloud services (such as AWS or Azure), and relevant certifications like Databricks Certified Associate Developer are commonly required. Strong problem-solving, communication, and self-motivation are crucial soft skills for remote collaboration and project delivery. These skills and qualities ensure efficient development, scalable data solutions, and effective teamwork in distributed environments.

What is the difference between Remote Databricks Developer vs Data Engineer?

AspectRemote Databricks DeveloperData Engineer
Required SkillsProficiency in Databricks, Spark, Python, SQLProficiency in data pipelines, ETL, cloud platforms, SQL
Work EnvironmentCollaborates on data projects using Databricks platformBuilds and maintains data infrastructure across cloud environments
CertificationsDatabricks certifications often preferredCloud certifications (AWS, Azure), data engineering certifications

While both roles involve working with data and cloud platforms, a Remote Databricks Developer specializes in developing solutions within the Databricks environment, focusing on Spark and data analytics. A Data Engineer has a broader scope, designing and maintaining data pipelines and infrastructure across various platforms. The roles overlap in skills like SQL and cloud knowledge, but their primary focus and tools differ.

What are popular job titles related to Remote Databricks Developer jobs in Sterling, VA?

For Remote Databricks Developer jobs in Sterling, VA, the most frequently searched job titles are:

What job categories do people searching Remote Databricks Developer jobs in Sterling, VA look for?

The top searched job categories for Remote Databricks Developer jobs in Sterling, VA are:

What cities near Sterling, VA are hiring for Remote Databricks Developer jobs?

Cities near Sterling, VA with the most Remote Databricks Developer job openings:

Infographic showing various Remote Databricks Developer job openings in Sterling, VA as of August 2026, with employment types broken down into 68% Full Time, 26% Contract, and 6% Nights. Highlights an 100% Remote job distribution.

Data Engineer

Sparibis

Washington, DC โ€ข On-site, Remote

$129K - $155K/yr

Full-time

Re-posted 28 days ago


Job description

Location: 100% Remote
Years' Experience: 5+ years Professional Experience
Education: Bachelor's Degree in IT related field
Clearance: Applicants must be able to obtain and maintain a secret security clearance. United States Citizenship is required as part of the eligibility criteria to be able to obtain this type of security clearance.
Required Certifications:
  • CompTIA Security +

Key Skills:
  • 5+ years of IT experience focusing on enterprise data architecture and management to include data flow charts, diagrams, and other technical documentation.
  • Experience with Databricks, Structured Streaming, Delta Lake concepts, and Delta Live Tables required.
  • Python development experience required.
  • Experience with ETL and ELT tools such as SSIS, Pentaho, and/or Data Migration Services, and the ability to incorporate Python as required.
  • Advanced level SQL experience (Joins, Aggregation, Windowing functions, Common Table Expressions, RDBMS schema design, Postgres performance optimization).
  • Proficiency using Git for version control, including repository management, branching, merging, and pull requests.
  • Active CompTIA Security+ certification preferred. If selected, must be able to obtain a CompTIA Security+ certification prior to beginning supporting the program.

Responsibilities
  • Plan, create, and maintain data architectures, ensuring alignment with business requirements.
  • Obtain data, formulate dataset processes, and store optimized data.
  • Identify problems and inefficiencies and apply solutions.
  • Determine tasks where manual participation can be eliminated with automation.
  • Identify and optimize data bottlenecks, leveraging automation where possible.
  • Create and manage data lifecycle policies (retention, backups/restore, etc).
  • In-depth knowledge for creating, maintaining, and managing ETL/ELT pipelines.
  • Create, maintain, and manage data transformations.
  • Maintain/update documentation.
  • Create, maintain, and manage data pipeline schedules.
  • Monitor data pipelines.
  • Create, maintain, and manage data quality gates (Great Expectations) to ensure high data quality.
  • Support AI/ML teams with optimizing feature engineering code.
  • Expertise in Spark/Python/Databricks, Data Lake and SQL.
  • Create, maintain, and manage Spark Structured Steaming jobs, including using the newer Delta Live Tables and/or DBT.
  • Research existing data in the data lake to determine best sources for data.
  • Create, manage, and maintain ksqlDB and Kafka Streams queries/code
  • Data driven testing for data quality.
  • Maintain and update Python-based data processing scripts executed on AWS Lambdas.
  • Unit tests for all the Spark, Python data processing and Lambda codes.
  • Maintain PCIS Reporting Database data lake with optimizations and maintenance (performance tuning, etc).
  • Streamlining data processing experience including formalizing concepts of how to handle lake data, defining windows, and how window definitions impact data freshness.

Qualifications
  • 5+ years of IT experience focusing on enterprise data architecture and management.
  • Must have an active Secret security clearance.
  • Bachelor degree required.
  • CompTIA Security+ certification preferred. If selected, must be able to obtain a CompTIA Security+ certification prior to begin supporting the program.
  • Experience in Conceptual/Logical/Physical Data Modeling & expertise in Relational and Dimensional Data Modeling.
  • Experience with Databricks and Python Development, Structured Streaming, Delta Lake concepts, and Delta Live Tables required.
    • Additional experience with Spark, Spark SQL, Spark DataFrames and DataSets, and PySpark.
    • Data Lake concepts such as time travel and schema evolution and optimization.
    • Structured Streaming and Delta Live Tables with Databricks a bonus.
  • Knowledge of Python (Python 3.X) for CI/CD pipelines required.
    • Familiarity with Pytest and Unittest a bonus.
  • Experience leading and architecting enterprise-wide initiatives specifically system integration, data migration, transformation, data warehouse build, data mart build, and data lakes implementation / support.
    • Advanced level understanding of streaming data pipelines and how they differ from batch systems.
    • Formalize concepts of how to handle late data, defining windows, and data freshness.
    • Advanced understanding of ETL and ELT and ETL/ELT tools such as SSIS, Pentaho, Data Migration Service etc.
    • Understanding of concepts and implementation strategies for different incremental data loads such as tumbling window, sliding window, high watermark, etc.
    • Familiarity and/or expertise with Great Expectations or other data quality/data validation frameworks a bonus.
    • Understanding of streaming data pipelines and batch systems.
    • Familiarity with concepts such as late data, defining windows, and how window definitions impact data freshness.
  • Advanced level SQL experience (Joins, Aggregation, Windowing functions, Common Table Expressions, RDBMS schema design, Postgres performance optimization).
  • Indexing and partitioning strategy experience.
  • Debug, troubleshoot, design and implement solutions to complex technical issues.
  • Experience with large-scale, high-performance enterprise big data application deployment and solution.
  • Understanding how to create DAGs to define workflows.
  • Familiarity with CI/CD pipelines, containerization, and pipeline orchestration tools such as Airflow, Prefect, etc a bonus but not required.
  • Architecture experience in AWS environment a bonus.
    • Familiarity working with Kinesis and/or Lambda specifically with how to push and pull data, how to use AWS tools to view data in Kinesis streams, and for processing massive data at scale a bonus.
    • Experience with Docker, Jenkins, and CloudWatch.
    • Ability to write and maintain Jenkinsfiles for supporting CI/CD pipelines.
    • Experience working with AWS Lambdas for configuration and optimization.
    • Experience working with DynamoDB to query and write data.
    • Experience with S3.
  • Experience working with JSON and defining JSON Schemas a bonus.
  • Experience setting up and management Confluent/Kafka topics and ensuring performance using Kafka a bonus.
    • Familiarity with Schema Registry, message formats such as Avro, ORC, etc.
    • Understanding how to manage ksqlDB SQL files and migrations and Kafka Streams.
  • Ability to thrive in a team-based environment.
  • Experience briefing the benefits and constraints of technology solutions to technology partners, stakeholders, team members, and senior level of management.
  • Proficiency using Git for version control, including repository management, branching, merging, and pull requests.
    • Repository setup and management.
    • Branching strategies (feature, develop, main).
    • Merging and resolving conflicts.
    • Creating and reviewing pull requests.
    • Commit best practices (clear messages, atomic commits).
    • Tagging and release management.

About Sparibis
Sparibis LLC is a professional solution firm that Clients rely on to access the best talent to drive their business success.
Sparibis is an equal opportunity employer that values diversity at all levels. All individuals, regardless of personal characteristics, are encouraged to apply.