2

Remote Databricks Data Engineer Jobs in Toronto, ON

25-199 - Data Engineer

Oshawa, ON ยท Remote

$85 - $95/hr

MP4, $80/hr - $95/hr INC Duration: 11 Months Hours of work: 35 hours Location: 1908 Colonel Sam Drive, Oshawa (Hybrid - 3 days remote) Job Overview As an Azure and Databricks Data Engineer, you will ...

... remote) Job Overview As a Senior Data Developer, you will be responsible for building and ... Databricks, Collibra, and Power Bl. Work within the agile SCRUM work management framework in ...

This can be a remote role, however for those who would like to come into the office our offices are ... Implement modern data and analytics practices using tools such as Databricks, Snowflake, Salesforce ...

Lead, Data Engineer

Mississauga, ON ยท On-site +1

CA$122K - CA$162K/yr

Summary The Lead Data Engineer is a senior individual contributor within McKesson's Decision ... Databricks, Snowflake, Azure Data Factory * Confluent Kafka / Azure Event Hub * PySpark and ...

Lead, Data Engineer

Mississauga, ON ยท On-site +1

CA$122K - CA$162K/yr

Summary The Lead Data Engineer is a senior individual contributor within McKesson's Decision ... Databricks, Snowflake, Azure Data Factory * Confluent Kafka / Azure Event Hub * PySpark and ...

The Senior / Lead Data Engineer will bepart of McKesson Decision Intelligence team, and ... Specific experience with Snowflake, Databricks, Azure data factory,PySpark,Analytical SQL,Splunk ...

The Senior / Lead Data Engineer will bepart of McKesson Decision Intelligence team, and ... Specific experience with Snowflake, Databricks, Azure data factory,PySpark,Analytical SQL,Splunk ...

Data Engineer

Toronto, ON ยท Remote

CA$140K - CA$240K/yr

This is a fully remote position that offers a competitive salary range of $140,000 to $240,000 USD ... Databricks, or equivalent technologies * Experience with streaming and real-time data platforms ...

Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate complex data engineering tasks. * Review model-generated implementations involving ETL pipelines , data ...

Lead Data Engineer

Toronto, ON ยท Remote

CA$220K - CA$300K/yr

This is a fully remote position that offers a competitive salary range of $220,000 to $300,000 USD ... Databricks, or equivalent technologies * Experience with streaming systems such as Kafka, Kinesis ...

Data Engineer Location: Fully Remote (Canada, EST time zone) Compensation: Salary + Bonus + Health Benefits About Veem Veem is transforming global money movement. Traditional cross-border payments ...

25-026 DevOps Engineer

Toronto, ON ยท Remote

CA$80 - CA$100/hr

... Remote) Job Overview We are seeking a skilled DevOps Engineer to support our data analytics ... Ability to work collaboratively Experience with Databricks workflow automation, Delta Lake, Azure

We are seeking a Senior Data Engineer to help design and build the next generation of our Data ... Our team is 100% distributed and remote. Responsibilities: * Design, build, and evolve the core ...

next page

Showing results 1-20

Remote Databricks Data Engineer information

What are the key skills and qualifications needed to thrive as a Remote Databricks Data Engineer, and why are they important?

To thrive as a Remote Databricks Data Engineer, you need a solid background in data engineering, strong programming skills in Python or Scala, and experience with big data frameworks, often supported by a degree in computer science or a related field. Proficiency with Databricks, Apache Spark, cloud platforms (such as AWS or Azure), and relevant certifications like Databricks Certified Data Engineer are highly valuable. Strong problem-solving abilities, effective remote communication, and collaboration skills set top performers apart in distributed teams. These skills and qualities ensure efficient data pipeline development, seamless integration, and successful project delivery in remote environments.

What is a Remote Databricks Data Engineer?

A Remote Databricks Data Engineer is a professional who designs, develops, and manages large-scale data processing systems using the Databricks platform, often working from a remote location. They focus on building data pipelines, integrating data sources, and optimizing workflows for analytics and machine learning, leveraging tools like Apache Spark within Databricks. These engineers collaborate with data scientists, analysts, and other stakeholders to ensure data is accessible, reliable, and scalable for business needs. Remote roles offer flexibility in work location while still requiring strong communication and technical skills.

What are some common challenges faced by remote Databricks Data Engineers and how can they be addressed?

Remote Databricks Data Engineers often encounter challenges such as coordinating efficiently with distributed teams, managing access to secure data environments, and ensuring smooth pipeline deployments across different cloud platforms. To overcome these, it's important to leverage communication tools for regular check-ins, follow strict data governance protocols, and utilize collaborative features in Databricks such as shared notebooks and version control. Proactively documenting your work and staying updated with platform updates can also help streamline remote collaboration and problem-solving.
What are the most commonly searched types of Databricks Data Engineer jobs in Toronto, ON? The most popular types of Databricks Data Engineer jobs in Toronto, ON are:
What are popular job titles related to Remote Databricks Data Engineer jobs in Toronto, ON? For Remote Databricks Data Engineer jobs in Toronto, ON, the most frequently searched job titles are:
What job categories do people searching Remote Databricks Data Engineer jobs in Toronto, ON look for? The top searched job categories for Remote Databricks Data Engineer jobs in Toronto, ON are:
Infographic showing various Remote Databricks Data Engineer job openings in Toronto, ON as of July 2026, with employment types broken down into 1% As Needed, 80% Full Time, 10% Part Time, and 9% Contract. Highlights an 81% Physical, 3% Hybrid, and 16% Remote job distribution.

Senior Data Engineer (Databricks/AWS)

Fusemachines

Toronto, ON โ€ข Remote

Contractor

Posted 7 days ago


Job description

About Fusemachines

Fusemachines is a 12+ year old AI company, dedicated to delivering state-of-the-art AI products and solutions to a diverse range of industries. Founded by Sameer Maskey, Ph.D., an Adjunct Associate Professor at Columbia University, our company is on a steadfast mission to democratize AI and harness the power of global AI talent from underserved communities. With a robust presence in four countries and a dedicated team of over 400 full-time employees, we are committed to fostering AI transformation journeys for businesses worldwide. At Fusemachines, we not only bridge the gap between AI advancement and its global impact but also strive to deliver the most advanced technology solutions to the world.
Type: Full-time, Remote
ย 

About the role

This is a full-time, high-impact position for a Senior Data Engineer with expertise in Databricks, dbt, and Apache Airflow to support a critical CRM data architecture migration for a key client in the Life Sciences industry.

In this role, you will join an urgent initiative to backfill key engineering capabilities and maintain momentum during an ongoing CRM system transition. The project involves migrating enterprise customer data from Veeva CRM to Salesforce Life Sciences Cloud, integrated with an underlying AWS S3 cloud environment and Databricks data warehouse. Your main focus will be building out, configuring, and redirecting data ingestion pipelines out of Life Sciences Cloud into the data warehouse, while implementing dbt models and Airflow orchestrations to ensure complete data accuracy.

Candidates must be able to operate strictly on US East Coast business hours (location is flexible across North America, LATAM, or remote with full Eastern Time overlap).

Qualification / Skill Set Requirement:

  • Core Technical Expertise:

    • 5+ years of hands-on data engineering experience with deep expertise in AWS, Databricks, dbt, and Apache Airflow.

    • Strong programming proficiency in Python / PySpark and Advanced SQL (complex joins, analytical window functions).

    • Hands-on expertise in Databricks platform architecture, Lakehouse implementation, Delta Lake, Unity Catalog, and cluster performance tuning.

  • Architecture & Migration:

    • Proven track record of architecting and executing migrations.

    • Demonstrated experience scaling platform performance.

  • Pipeline Orchestration & Modeling:

    • Proven experience building scalable transformations pipelines using dbt for data transformation, testing, and documentation.

    • Solid background orchestrating complex workflow DAGs with Apache Airflow.

    • Experience working with AWS cloud infrastructure, specifically AWS S3 as an underlying data lake storage layer.

  • CRM Integration & Domain Knowledge:

    • Hands-on experience developing integrations and data ingestion pipelines for CRM platforms, specifically Salesforce, Salesforce Life Sciences Cloud, and/or Veeva CRM.

    • Understanding of data structures, customer master data, and analytics workflows within the Life Sciences.

  • DevOps & Governance:

    • Deep understanding of SDLC/Agile and DevOps for CI/CD and artifact management.

    • Knowledge of AWS and Databricks security best practices and compliance standards.

  • Certifications Preferred: Databricks Certified Data Engineer Associate/Professional, Databricks Spark Developer, and major cloud certifications in AWS.

  • Logistics & Shift Overlap:

    • Ability to maintain 100% full working time overlap with US East Coast business hours (ET). Flexible location (US, Canada, LATAM, or remote ET).

Responsibilities

  • Pipeline Development & Integration: Architect, build, and deploy data integration pipelines connecting Salesforce Life Sciences Cloud to the clientโ€™s Databricks warehouse environment.

  • CRM Migration Support: Execute pipeline modifications to transition legacy data feeds from Veeva CRM to Salesforce Life Sciences Cloud, updating warehouse models accordingly.

  • Transformation & Workflow Orchestration: Write clean, modular dbt transformation models and organize end-to-end DAG execution using Apache Airflow.

  • Data Warehouse & Storage Optimization: Manage Delta tables and optimize Databricks clusters and AWS S3 storage for high performance and cost efficiency.

  • Data Validation & Quality Assurance: Implement data quality testing, schemas, and verification rules in dbt and Python to guarantee accurate data delivery.

  • Monitoring & Alerting: Build and enforce proactive monitoring frameworks.

  • Agile Collaboration: Work closely with project leads, solution architects, and technical stakeholders during US East Coast hours to ensure rapid iteration and goal completion.

Equal Opportunity Employer: Race, Color, Religion, Sex, Sexual Orientation, Gender Identity, National Origin, Age, Genetic Information, Disability, Protected Veteran Status, or any other legally protected group status.

Powered by JazzHR

vQfwmy5HZ5