Data Engineer with 4+ years of experience in designing, developing, and optimizing scalable data pipelines and cloud-based data platforms. Proficient in Python, PySpark, SQL, AWS Glue, Amazon S3, Amazon Redshift, Athena, Pandas, and Apache Spark. Experienced in building robust ETL pipelines, processing large-scale datasets, and developing data lake and data warehouse solutions on AWS. Skilled in data modeling, star schema design, query optimization, partitioning, and performance tuning to improve processing efficiency and reporting. Hands-on experience with AWS Glue Workflows, Apache Airflow, AWS Step Functions, CloudWatch, and CloudTrail for orchestration and monitoring. Strong understanding of data quality, validation, and distributed data processing. Adept at collaborating with cross-functional teams to deliver reliable, scalable, and business-driven data solutions.