Damodarrao Thakkalapelli

Damodarrao Thakkalapelli

About

Detail

Senior Data Engineer / Architect at Kroger
Herndon, Virginia, United States

Contact Damodarrao regarding: 
work
Full-time jobs

Timeline


work
Job

Résumé


Jobs verified_user 0% verified
  • Kroger
    Lead – Data Engineer/Architect
    Kroger
    May 2024 - Feb 2025 (10 months)
    • Demonstrated the ability to validate join conditions through comparative analysis of results from various join methods, enhancing data integrity and accuracy. • Developed and implemented a complex Python script to merge two large datasets, enhancing processing efficiency and creating a combined table tailored for management analysis. • Developed and maintained data pipelines in Databricks to handle large-scale data processing and ETL workflows, ensuring data quality and system reliability. • Managed Databricks clusters, including configuration, resource allocation, and scaling, to balance cost efficiency with performance. • Set up and maintained Azure-based ETL pipelines, ensuring secure and reliable data integration from various sou
  • B
    – Data Engineer/ Architect
    Bank of America (Data Mortgage Hub)
    Mar 2018 - Sep 2024 (6 years 7 months)
    • Analyze, design, and build modern data solutions using Azure PaaS service to support visualization of data. • Implemented solutions for ingesting data from various sources and processing the Data-at-Rest utilizing Big Data technologies such as Hadoop, Map Reduce Frameworks, HBase, and Hive. • Involved in Extracting Transforming and Loading data from Sources Systems to Azure Data Storage services using a combination of Azure Data Factory, T-SQL, Spark SQL, and U-SQL Azure Data Lake Analytics Data Ingestion to Azure Services - (Azure Data Lake, Azure Storage, Azure SQL, Azure DW) and processing the data in Azure Databricks. • Worked on the building, deploying of the Service Azure fabric applications as microservices. • Experience with
  • F
    Big Data Engineer/Data Architect
    Fisher Investment
    Feb 2016 - Mar 2018 (2 years 2 months)
    • Evaluate deep learning algorithms for text summarization using Python, Keras and TensorFlow on Cloudera Hadoop System and Use Spark API for Machine learning and translate a predictive model from SAS code to Spark and used Spark API over Cloudera Hadoop YARN to perform analytics on data in Hive. • Recreating existing application logic and functionality in the Azure Data Lake, Data Factory, SQL Database and SQL data warehouse environment • Exploring with Spark for improving the performance and optimization of the existing algorithms in Hadoop using Spark Context, Spark-SQL, Data Frame, Pair RDD's, Spark YARN. • Utilized the Azure polybase utility to run t-sql queries on external data in Hadoop and as well as to import and export data fr
  • GEICO
    Data Engineer/ETL Developer
    GEICO
    Feb 2014 - Jan 2016 (2 years)
    • Gathered business requirements, working closely with business users, project leaders and developers and analyzed the business requirements and designed data models and analyzed existing systems and propose improvements in processes and systems for usage of modern scheduling tools like Airflow and migrating the legacy systems into an Enterprise data lake. • Design and Develop ETL Processes in AWS Glue to migrate Campaign data from external sources like S3, ORC/Parquet/Text Files into Snowflake and involve with Data Extraction, aggregations, and consolidation of Adobe data within AWS Glue using PySpark. • Build ETL pipeline end to end from AWS S3 to Key, Value store DynamoDB, and Snowflake Datawarehouse for analytical queries and specifi
  • Bank of America
    ETL Developer
    Bank of America
    Mar 2012 - Jan 2014 (1 year 11 months)
    • Analyzed the source system and involved in designing the ETL data load. • Responsible for developing, support and maintenance for the ETL (Extract, Transform and Load) • processes using Informatica PowerCenter • Developed/designed Informatica mappings by translating the business requirements. • Worked in various transformations like Lookup, Joiner, Sorter, Aggregator, Router, Rank and SourceQualifier to create complex mapping. • Involved in performance tuning of the Informatica mappings using various components like Parameter Files, round robin and Key range partitioning to ensure source and target bottlenecks were removed. • Data Analysis & Profiling. Extracted business rule and implemented business logic to extract and load SQL s
  • Banner Health
    SQL Developer
    Banner Health
    Jan 2010 - Feb 2012 (2 years 2 months)
    • Collaborated in analyzing business requirements and understanding the functional workflow of information from source systems to destination systems. • Expertise in Always-On Availability Groups (AAG) and HA/DR Setup and Configuration • Migrated SQL server 2005 to SQL Server 2008 in Microsoft Windows Server 2008 Enterprise Edition. • Installed and administered Eight-Node cluster Active/Passive on Microsoft Windows Server 208 Enterprise Edition on SAN environment. • Involved in fine tuning queries and stored procedures using SQL Profiler and Execution plan in SQL Server. • Created and Modified tables, views, Indexes, Functions, Stored Procedures, cursors in SQL Server. • Written the complex queries using joints, Sub-queries, common t
  • N
    Python Developer
    Northern Mass Telephone Workers Community Credit Union
    Aug 2009 - Jan 2010 (6 months)
    • Developed scripts in python for configuring and tracking progress in one Siebel and B2B CRM. • Developed web-service for client-handler tier and connected with Database tier in Oracle Flex cube. • Worked on design and development of Unix shell scripting as a part of the ETL process to automate the process of loading. • Worked on ETL tasks like pulling, pushing data from and too various servers. • Wrote Python and batch scripts to automate the ETL scripts runs every hour. Developed Wrote scripts to database. Used Python to storage deletion content. Rewrite existing to deliver data.
Education verified_user 0% verified
  • G
    MS
    Gannon University PA -US
This is a community-created genome.