Damodarrao Thakkalapelli
Damodarrao Thakkalapelli
About
Detail
Senior Data Engineer / Architect at Kroger
Herndon, Virginia, United States
More than 15 years of experience in the IT industry including Data Architecture, Data Engineering, Hadoop Ecosystem. Expert on migrating SQL database to Azure data Lake, Azure data lake Analytics, Azure SQL Database, Logic Apps, Data Bricks and Azure SQL data warehouse and controlling and granting database access and migrating on premise and real time databases to Azure Data Lake store using Azure Data Factory (ADF) and Azure Fabric Services. Experience in leading multiple data migration and data transformation implementations in Banking, Investment, Insurance, Retail and Healthcare industries. Hands on experience in Hadoop ecosystem including Spark, Kafka, HBase, Scala, Pig, Impala, Sqoop, Oozie, Flume, Storm, big data technologies and experienced in developing custom UDFs for Pig and Hive to incorporate methods and functionality of Python/Scala, Java into Pig Latin and HQL (HiveQL). Experience in implementing large Lambda architectures using Azure Data platform capabilities like Azure Data Lake, Azure Data Factory, HD Insight, Azure SQL Server, Azure ML and Power BI. Hands on experience in Azure Cloud Services (PaaS & IaaS), Azure Synapse Analytics, SQLmAzure, Data Factory, Azure Analysis services, Application Insights, Azure Monitoring, Key Vault, Azure Data Lake, and Azure Data Factory Logic Apps. Designed and developed ETL Processes in AWS Glue to migrate data from external sources like S3, ORC/Parquet/Text Files into Snowflake and involve with Data Extraction, aggregations, and consolidation of Adobe data within AWS Glue using PySpark. Built ETL pipeline end to end from AWS S3 to Key, Value store DynamoDB, and Snowflake Datawarehouse for analytical queries and specifically for cloud data. Design and construct of AWS Data pipelines using various resources in AWS including AWS API Gateway to receives response from AWS lambda and retrieve data from snowflake using lambda function and convert the response into Json format using Database as Snowflake, DynamoDB, AWS Lambda function and AWS S3. Expert in developing Spark applications using Spark - SQL in Databricks for data extraction, transformation, and aggregation from multiple file formats for analyzing & transforming the data to uncover in-sights into the customer usage patterns. Good experience in Tableau for Data Visualization and analysis on large data sets, drawing various conclusions and leveraged and integrated Google Cloud Storage and Big Query applications, which connected to Tableau for end user web-based dashboards and reports. Experience in database design and development with Business Intelligence using SQL Server 2014/2016, Integration Services (SSIS), DTS Packages, SQL Server Analysis Services (SSAS), DAX, OLAP Cubes, Star Schema and Snowflake Schema. Exploring on Azure Cognitive Services (LUIS) and Machine Learning and IOT technologies. Experienced in development of Big Data projects using Hadoop, Hive, HDP, Pig, Flume, Storm and Map Reduce open-source tools and experience in installation, configuration, supporting and managing Hadoop clusters. Experience in Data Modeling and Data Analysis using Dimensional Data Modeling and Relational Data Modeling, Star Schema/Snowflake Modeling, FACT & Dimensions tables, Physical & Logical Data Modeling. Expert in understanding Big Data infrastructure, distributed file systems -HDFS, parallel processing - Map Reduce framework. Experience in installation, configuration, supporting and managing - Cloudera Hadoop platform along with CDH4 & CDH5 clusters and Hortonwors. Extensive experience in migrating on premise ETLs to Google Cloud Platform (GCP) using cloud native tools such as BIG query, Cloud Data Proc, Google Cloud Storage, Composer. Very keen on knowing newer techno stack that Google Cloud platform (GCP) adds. Expert in working parallelly in both GCP and Azure Clouds coherently. Experience in building efficient pipelines for moving data between GCP and Azure using Azure Data Factory. Experience in building power bi reports on Azure Analysis services for better performance when comparing that to direct query using GCP BigQuery. Extensive use of cloud shell SDK in GCP to configure/deploy the services Data Proc, Storage, and BigQuery. Experience in using stackdriver service/ dataproc clusters in GCP for accessing logs for debugging.
Contact Damodarrao regarding:
work
Full-time jobs