Sabio Infotech

5 jobs near Columbus, OH

AWS Data Engineer

Charlotte, NC · On-site

$111K - $134K/yr

Job Title: Sr.Data Engineer- AWS & Streaming Location: Fort Mill SC (OR) Charlotte, NC (OR) New York, NY (2-3 Days Hybrid) Experience level - 12+ Years Key Responsibilities: * Develop and maintain ...

Product Owner Pittsburgh PA Full-Time Position Summary: The main function of a product owner is to maximize the value of products created by the scrum development team. A typical product owner will ...

New

Sr. Machine Learning Engineer

Cincinnati, OH · On-site

$100K - $137K/yr

Sr. Machince Learning Engineer Location: Cincinnati OH (Hybrid - 3 Days Onsite) Duration: 1+ Year Key Responsibilities · Design, develop, deploy, and maintain scalable machine learning and ...

Sr. Databricks Architect

Dallas, TX · On-site

$64.25 - $84.50/hr

Job Title: Senior Databricks Architect Domain : Financial Services Location : Dallas, TX/Florham Park, NJ Key Responsibilities: · Architect and implement end-to-end data solutions using Databricks ...

AWS Data Engineer(12+ Years)

Charlotte, NC · On-site

$111K - $134K/yr

Job Title: Sr.Data Engineer AWS & Streaming Location: Fort Mill SC or New York, NY (2-3 Days Hybrid) Experience level 10- 15 Years Role Summary: We are seeking a Mid Senior Data Engineer with strong ...

AWS Data Engineer

Sabio infotech

Charlotte, NC • On-site

$111K - $134K/yr

Other

Re-posted 11 days ago


Job description

Job Title: Sr.Data Engineer– AWS & Streaming

Location: Fort Mill SC (OR) Charlotte, NC (OR) New York, NY (2-3 Days Hybrid)

Experience level – 12+ Years

Key Responsibilities:

  • Develop and maintain scalable ETL/ELT pipelines using AWS Glue, PySpark, and Python

  • Build event-driven workflows using AWS Lambda

  • Design and manage real-time streaming solutions using Kafka, KSQL, and Apache Flink

  • Implement and enforce comprehensive data quality frameworks, including validation, profiling, monitoring, and reconciliation

  • Optimize data processing performance, scalability, reliability, and cost in cloud environments

  • Collaborate with cross-functional teams to deliver reliable, production-grade data platforms and ensure data integrity across the pipeline

 Required Skills:

  • Strong hands-on experience with Python and PySpark

  • Proven expertise in AWS Glue, Lambda, and other cloud-native data services

  • Solid experience with the Kafka ecosystem (topics, partitions, consumer groups, streaming patterns)

  • Demonstrated experience building and supporting data quality frameworks (validation rules, reconciliation checks, profiling, anomaly detection)

  • Strong understanding of distributed data processing and scalable architecture patterns