Request ID: 106204-1
Title: AWS Data Engineer (Python / PySpark / Redshift)
Locations: Malvern, PA Duration: 06+ Months with Possible Extension.
We are seeking an experienced AWS Data Engineer with strong expertise in SQL, Python, PySpark, AWS Glue, Amazon Redshift, and Amazon S3 to design, develop, and support scalable data engineering solutions. The ideal candidate will have hands-on experience building ETL pipelines, optimizing data processing workflows, and implementing cloud-based data warehouse solutions.
The candidate will be responsible for developing reliable data pipelines, performing data transformations, optimizing queries, and ensuring data quality across enterprise AWS data platforms.
Key Responsibilities:
Design, develop, and implement scalable data engineering solutions using AWS cloud technologies.
Develop and maintain ETL pipelines using AWS Glue for data extraction, transformation, and loading.
Build and optimize PySpark-based data processing jobs within AWS Glue.
Design and manage data solutions using:
Amazon Redshift
Redshift Spectrum
Amazon S3
Load, organize, and query large datasets stored in Amazon S3.
Create and manage AWS Glue Data Catalog configurations.
Develop repeatable and reliable Glue ETL jobs while ensuring data quality and performance.
Write optimized SQL queries for Redshift data processing and analytics.
Build data validation frameworks to compare source and target outputs.
Perform code reviews, testing, debugging, and performance tuning.
Troubleshoot technical issues and production defects.
Collaborate with cross-functional teams to deliver high-quality data solutions aligned with business objectives.
Required Skills & Experience:
6 8 years of experience in Data Engineering.
Strong hands-on experience with:
SQL
Python
PySpark
AWS Cloud Services
Strong experience with:
AWS Glue
Amazon Redshift
Amazon S3
Redshift Spectrum
Glue Data Catalog
Experience designing and developing ETL/ELT pipelines.
Experience with data transformation, ingestion, validation, and optimization.
Strong understanding of data warehousing concepts and cloud-based data platforms.
Ability to work independently and collaborate with technical and business teams.