1

Hadoop Python Jobs in Ontario (NOW HIRING)

Write and optimize advanced queries using SQL and Python to improve data processing, performance ... Familiarity with big data and distributed technologies such as Hadoop, Kafka, and distributed file ...

Design, develop, and maintain largescale Spark applications using Python, Scala, and/or Java ... Experience migrating Spark jobs from onprem Hadoop/Cloudera to cloud platforms an asset.

Write and optimize advanced queries using SQL and Python to improve data processing, performance ... Familiarity with big data and distributed technologies such as Hadoop, Kafka, and distributed file ...

Logstash, Python, SQL Server, Kafka, Hadoop, Spark, Trino * API: Streamlit, Flask, Node.JS, Django, and Microservices technologies * Automation/DevOps: Github Actions, Airflow, UCD, Selenium and ...

In-depth experience with Python, SQL Experience using advanced tools such as Hadoop, HIVE etc. and working knowledge with relational and non-relational modeling, database applications and ...

In-depth experience with Python, SQL Experience using advanced tools such as Hadoop, HIVE etc. and working knowledge with relational and non-relational modeling, database applications and ...

Strong programming skills in Python and Scala required. Experience in other programming languages ... Hadoop, Spark, Kafka) required. * Experience with building and deploying API's with Docker and ...

Spark, Hadoop, Kafka) * Data Warehousing & ETL (Ex: SQL, OLTP/OLAP/DSS) * Data Science and Machine ... Python, Scala, Java, or R. * [Desired] Built solutions with public cloud providers such as AWS ...

... Hadoop, Spark, Hive) * Experience training and deploying machine learning models using common Python opensource frameworks (i.e., scikit-learn) and/or Python DevOps * Microsoft Office (Excel, Word ...

New

Develop Python programs and perform advanced data analytics * Contribute to the build-out of ... Familiarity with Spark, Hadoop, or other big data technologies WHAT'S IN IT FOR YOU? We thrive on ...

Familiarity with Hadoop, Spark, Databricks or other distributed computing systems. * Understanding ... Experience writing code/scripts in Python. * Experience with Spring Boot. * Nice to have React ...

New

High proficiency using Python for data and AI engineering, experience building end to end ML ... Experience using AWS Sagemaker, S3, Snowflake, Databrick, Airflow, Hadoop, PySpark * Familiar with ...

Showing results 21-40

Hadoop Python information

What are the key skills and qualifications needed to thrive as a Hadoop Python developer?

To thrive as a Hadoop Python Developer, you need a strong understanding of distributed computing, Hadoop ecosystem components (like HDFS, MapReduce, Hive, or Pig), and advanced Python programming skills, often supported by a degree in computer science or related field. Familiarity with tools such as Apache Spark, Sqoop, and workflow schedulers (like Oozie or Airflow), along with experience in handling big data platforms, is typically required. Problem-solving abilities, attention to detail, and effective communication help developers collaborate with teams and translate business requirements into scalable data solutions. These skills and qualifications are essential for efficiently processing and analyzing large datasets, ensuring data reliability, and driving business insights.

What is the difference between Hadoop Python vs Hadoop Java Developer?

AspectHadoop PythonHadoop Java Developer
Required CredentialsPython programming skills, Hadoop certificationsJava programming skills, Hadoop certifications
Work EnvironmentData analysis, scripting, data pipeline developmentCore development, system integration, big data application coding
Industry UsageData science, analytics, machine learning projectsData infrastructure, platform development, system optimization

Hadoop Python and Hadoop Java Developer roles both involve working with Hadoop ecosystems, but Python focuses more on data analysis and scripting, while Java is geared towards core development and system integration. The choice depends on your programming expertise and career goals within big data environments.

What is a Hadoop Python developer?

A Hadoop Python developer is a software professional who specializes in using Python programming language to develop, implement, and maintain applications that process and analyze large datasets within the Hadoop ecosystem. They leverage Python libraries like PySpark to write scalable data processing scripts, interact with Hadoop components such as HDFS, and optimize big data workflows. These developers play a critical role in building data pipelines, performing data transformation, and supporting analytics projects in organizations that handle vast amounts of data.

How do Hadoop Python developers typically collaborate with data engineers and analysts on large-scale data projects?

Hadoop Python developers frequently work alongside data engineers and analysts to design, implement, and optimize data pipelines for handling vast datasets. They are responsible for writing Python scripts that interface with Hadoop components, ensuring data is processed efficiently and meets project requirements. Regular communication with data engineers helps align on infrastructure and architectural decisions, while close collaboration with analysts ensures data outputs are accurate and actionable. Agile methodologies and daily stand-ups are common, fostering teamwork and quick problem-solving.
What are popular job titles related to Hadoop Python jobs in Ontario? For Hadoop Python jobs in Ontario, the most frequently searched job titles are:
What job categories do people searching Hadoop Python jobs in Ontario look for? The top searched job categories for Hadoop Python jobs in Ontario are:
Infographic showing various Hadoop Python job openings in Ontario as of July 2026, with employment types broken down into 88% Full Time, 4% Part Time, and 8% Contract. Highlights an 85% Physical, 6% Hybrid, and 9% Remote job distribution.

Sr Data Specialist

McKesson

Mississauga, ON • On-site, Remote

Full-time

Re-posted 26 days ago


McKesson rating

7.9

Company rating: 7.9 out of 10

Based on 209 frontline employees who took The Breakroom Quiz

48th of 86 rated pharmaceutical


Job description

McKesson is an impact-driven, Fortune 10 company that touches virtually every aspect of healthcare. We are known for delivering insights, products, and services that make quality care more accessible and affordable. Here, we focus on the health, happiness, and well-being of you and those we serve - we care.

What you do at McKesson matters. We foster a culture where you can grow, make an impact, and are empowered to bring new ideas. Together, we thrive as we shape the future of health for patients, our communities, and our people. If you want to be part of tomorrow's health today, we want to hear from you.

Position Summary :

The Sr Data Specialist provides technical leadership in designing, implementing, and optimizing McKesson's enterprise data infrastructure to enable advanced analytics and decision-making. This position is responsible for building scalable data pipelines, developing ETL programs, and ensuring data integrity, reliability, and compliance within a regulated environment.

The role involves creating custom software components, maintaining metadata repositories, and implementing processes for data standardization and quality improvements. With an automation-first and enterprise-first mindset, the Sr Data Specialist collaborates with architects, analysts, data scientists, and governance teams to deliver scalable, reusable, and production-ready data solutions, including support for advanced analytics and AI-driven use cases that drive strategic and operational outcomes.

Key Responsibilities:

  • Design and implement scalable data pipelines and ETL/ELT programs to integrate complex data sources across internal and external systems, supporting both batch and real-time processing.
  • Write and optimize advanced queries using SQL and Python to improve data processing, performance, and analytical outcomes.
  • Lead data exploration, requirements analysis, and data source identification for analytical and operational use cases.
  • Develop and maintain metadata repositories, including data definitions, lineage, and business rules, ensuring data integrity and usability across enterprise systems.
  • Create and implement processes for data standardization, reliability, quality improvement, and governance compliance.
  • Troubleshoot and resolve complex data analytic issues across production and development environments, including database and pipeline performance tuning.
  • Develop and maintain reusable query libraries and custom software components to support analytics, reporting, and AI/ML solutions.
  • Recommend and implement modern tools, technologies, and best practices to enhance data engineering capabilities and platform performance.
  • Ensure adherence to governance, security, and regulatory compliance through rigorous data quality checks, validation processes, and documentation.
  • Support testing, monitoring, and validation of data pipelines, including development of test cases and quality checks.
  • Collaborate with cross-functional teams to align data solutions with enterprise architecture, governance standards, and business priorities.
  • Provide technical leadership, influence design decisions, and mentor junior engineering teams to promote best practices and continuous improvement.

Minimum Job Qualifications (Knowledge, Skills, & Abilities):

  • Expertise in designing and maintaining scalable data pipelines, ETL/ELT processes, and data integration across complex enterprise systems.
  • Advanced proficiency in SQL and Python for data processing, query optimization, and analytics enablement (R is a plus).
  • Strong knowledge of data modeling, data architecture, metadata management, and data governance practices.
  • Hands-on experience with modern data platforms and tools such as Databricks, Snowflake, Azure Data Factory, and PySpark.
  • Familiarity with big data and distributed technologies such as Hadoop, Kafka, and distributed file systems.
  • Ability to implement processes for data standardization, reliability, quality improvement, and stewardship.
  • Competence in troubleshooting complex data issues, resolving data model conflicts, and optimizing performance.
  • Experience developing custom components, analytics applications, and enabling advanced analytics/AI use cases.
  • Familiarity with cloud platforms and architectures (Azure preferred) and concepts including SaaS, PaaS, and IaaS.
  • Effective communication and leadership skills for guiding technical decisions, influencing stakeholders, and mentoring team members.

Business Experience:

  • Bachelor's degree or equivalent combination of education and experience required (Master's preferred).
  • Typically requires 7+ years of relevant professional experience in data engineering, analytics, or related fields.
  • Hands-on experience with advanced ETL/ELT development, SQL/Python scripting, and enterprise data integration across multiple platforms.
  • Experience with modern data ecosystems and large-scale data processing frameworks preferred.
  • Experience in healthcare or other regulated industries is strongly preferred.

Working Conditions:

  • In office requirement, we are Flex and Connect with 2 days a week in office

We are proud to offer a competitive compensation package at McKesson as part of our Total Rewards. This is determined by several factors, including performance, experience and skills, equity, regular job market evaluations, and geographical markets. The pay range shown below is aligned with McKesson's pay philosophy, and pay will always be compliant with any applicable regulations. In addition to base pay, other compensation, such as an annual bonus or long-term incentive opportunities may be offered. For more information regarding benefits at McKesson, pleaseclick here.

Our Base Pay Range for this position

$99,100 - $132,100

McKesson has become aware of online recruiting-related scams in which individuals who are not affiliated with or authorized by McKesson are using McKesson's (or affiliated entities, like CoverMyMeds or RxCrossroads) name in fraudulent emails, job postings or social media messages. In light of these scams, please bear the following in mind:
McKesson Talent Advisors will never solicit money or credit card information in connection with a McKesson job application.


McKesson Talent Advisors do not communicate with candidates via online chatrooms or using email accounts such as Gmail or Hotmail. Note that McKesson does rely on a virtual assistant (Gia) for certain recruiting-related communications with candidates.

McKesson job postings are posted on our career site: careers.mckesson.com.

McKesson is an Equal Opportunity Employer

McKesson provides equal employment opportunities to applicants and employees, without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status, disability, age, genetic information, or any other legally protected category. For additional information on McKesson's full Equal Employment Opportunity policies, visit our Equal Employment Opportunity page.

McKesson is committed to being an Equal Employment Opportunity Employer and offers opportunities to all job seekers including job seekers with disabilities. If you need a reasonable accommodation to assist with your job search or application for employment, please contact us by sending an email to (United States) Disability_Accommodation@McKesson.com or (Canada) Accessibility@mckesson.ca. Resumes or CVs submitted to this email box will not be accepted.

Join us at McKesson!


What McKesson employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom