Associate Analyst
Capgemini Technology Services
Nov 2019 - Oct 2021 (2 years)
• Developed and implemented data processing pipeline, automating data collection, cleaning, and transformation processes. resulted in increasing data processing efficiency by 40% and improved data accuracy by 30%. • Utilized SQL to execute complex database queries, data extraction, and manipulation across large datasets of over 2 TB, deriving critical insights that influenced key business decisions and strategies. • Employed PySpark for big data processing, efficiently handling and analyzing datasets in a distributed computing environment, which streamlined data processing times by 50%. • Created dynamic dashboards using Power BI, enhancing data visualization and reporting capabilities, which provided insights into business metrics for