Job Title:ย AWS Data Engineer
Location:ย Remote
Employment Type:ย Contract / Full-Time
Job Summary
We are seeking a highly skilledย AWS Data Engineerย with strong experience in building scalable cloud-native data platforms and ETL pipelines on AWS. The ideal candidate should have hands-on expertise withย AWS Glue, Spark ETL, Step Functions, Amazon EKS, Kubernetes, Lambda, S3 Data Lake architectures, SQS, KEDA, Glue Data Catalog, Glue Data Quality, and CloudWatch.
Key Responsibilities
ยทDesign, develop, and maintain scalable ETL pipelines usingย AWS Glueย andย Apache Spark (PySpark).
ยทBuild and orchestrate data workflows usingย AWS Step Functions.
ยทDesign and implementย S3 Data Lakeย architectures following AWS best practices.
ยทDevelop and deploy containerized applications onย Amazon EKSย usingย Kubernetes.
ยทBuild event-driven data processing solutions usingย Amazon SQS,ย AWS Lambda, andย KEDAย for auto-scaling.
ยทManage metadata usingย AWS Glue Data Catalog.
ยทImplement data validation and governance usingย AWS Glue Data Quality.
ยทMonitor applications and data pipelines usingย Amazon CloudWatch.
ยทOptimize ETL jobs, Spark workloads, and Kubernetes deployments for performance and scalability.
ยทCollaborate with architects, developers, DevOps, and business stakeholders to deliver robust cloud data solutions.
ยทParticipate in Agile ceremonies including sprint planning, code reviews, and production deployments.
Required Skills
ยท5+ years of experience in Data Engineering or Cloud Engineering.
ยทStrong hands-on experience withย AWS Glue.
ยทExperience buildingย Spark ETLย pipelines usingย PySpark.
ยทExperience withย AWS Step Functions.
ยทStrong knowledge ofย Amazon S3 Data Lakeย architecture and storage patterns.
ยทHands-on experience withย Amazon EKSย andย Kubernetes.
ยทExperience implementing event-driven architectures usingย Amazon SQS,ย AWS Lambda, andย KEDA.
ยทExperience withย AWS Glue Data Catalogย andย Glue Data Quality.
ยทStrong experience monitoring AWS workloads usingย Amazon CloudWatch.
ยทGood understanding of distributed data processing and cloud-native application design.
ยทExperience working in Agile/Scrum environments.
Preferred Skills
ยทPython / PySpark
ยทSQL
ยทTerraform or CloudFormation
ยทDocker
ยทGit
ยทCI/CD (Jenkins, GitHub Actions, CodePipeline)
ยทKafka
ยทApache Airflow
ยทDelta Lake / Iceberg / Lake Formation
ยทIAM, VPC, CloudTrail
Education
ยทBachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
Mandatory Skills
ยทAWS Glue
ยทSpark ETL / PySpark
ยทAWS Step Functions
ยทAmazon S3 Data Lake
ยทAmazon EKS
ยทKubernetes
ยทAmazon SQS
ยทKEDA
ยทAWS Lambda
ยทGlue Data Catalog
ยทGlue Data Quality
ยทAmazon CloudWatch
Nice to Have
ยทTerraform
ยทDocker
ยทPython
ยทSQL
ยทCI/CD
ยทKafka
ยทAWS Certifications (Solutions Architect, Developer, Data Engineer, or DevOps Engineer)