AWS Data Engineer (Pyspark)_Indore,Pune,Hyderabad,,_6 to 9 yrs
Job Description
Key Responsibilities
\nETL & Pipeline Development: Design and optimize scalable ETL batch pipelines in AWS for high performance and reliability.
\nOrchestration: Manage data workflows using tools like AWS Glue, MWAA, or Step Functions.
\nData Processing: Develop large-scale processing jobs using PySpark while ensuring data quality and integrity.
\nInfrastructure as Code (IaC): Automate and manage AWS infrastructure using Terraform.
\nContainerization: Deploy and manage applications using Amazon EKS and Docker.
\nRequired Skills & Qualifications
\nAWS Services: Hands-on expertise with Glue, S3, IAM, KMS, SNS, Athena, Lambda, SQS, CloudWatch, and EC2.
\nProgramming: Proficiency in PySpark for complex data transformations.
\nMigration: Proven experience moving data from on-premises systems to AWS.
\nDevOps Tools: Skilled in Terraform for IaC and Docker/Containers for application packaging.