S

Sai Kiran

About

Detail

Analytics | Data Engineer | Bigdata | Azure | AWS I Hadoop | ETL | Talend | SQL | Snowflake | BI | Looking for C2C/C2H roles, Data Engineer around 8 years of...
Denton, Texas, United States

Contact Sai regarding: 
work
Full-time jobs

Timeline


work
Job
school
Education

Résumé


Jobs verified_user 0% verified
  • Gap
    Senior Data Engineer
    Gap
    Oct 2022 - Current (3 years 11 months)
    • Participated in the analysis, design, and development phase of the Software Development Lifecycle (SDLC). Experience in the agile environment, used to have sprint planning meeting, scrum calls and retro meetings for every sprint, Used JIRA for the project management and GitHub for the version control. • Designed and implemented real-time data processing pipelines using Apache Kafka, Apache Spark, and Spark Streaming to process and analyze large volumes of data. Developed and maintained HBase clusters on Hortonworks and Cloudera distributed platforms for storing large scale unstructured data. • Worked on various Azure services such as Azure Data Lake, Data Factory, Azure Databricks, and Azure SQL Database to store and process data. Utilize
  • National Grid
    Senior Data Engineer
    National Grid
    Jul 2021 - Oct 2022 (1 year 4 months)
    • Responsible for sessions with business, project manager, Business Analyst, and other key people to understand the business needs and propose a solution from a Warehouse standpoint. • Experience in Developing Spark applications using Spark/PySpark - SQL in Databricks for data extraction, transformation, and aggregation from multiple file formats for analyzing & transforming the data to uncover insights into the customer usage, consumption patterns, and behavior. • Skilled dimensional modeling, forecasting using large-scale datasets (Star schema, Snowflake schema), transactional modeling, and SCD (Slowly changing dimension). • Developed scripts to transfer data from FTP server to the ingestion layer using Azure CLI commands. • Created Azure
  • Ford Motor Company
    Senior Data Engineer
    Ford Motor Company
    Aug 2020 - Jun 2021 (11 months)
    • Experienced with Cloud Service Providers such as GCP. • Used Python programming and Django for the backend development, Bootstrap and Angular for frontend connectivity and MongoDB for database. • Extensive experience in building ETL jobs using Jupyter notebooks with Apache Spark. • Heavily used Jupyter Notebooks to analyze and connect the data from multiple sources. • Developed data integration strategies for data flow between disparate source systems and Big Data enabled, Enterprise Data Lake. • By Cloud Composer we managed workflow orchestration service, enabling to create, Monitor, schedule and manage workflow pipelines that span across clouds and on-premises data centers. • Cloud Composer is built on the popular Apache Airflow open so
  • Ferguson
    Data Engineer
    Ferguson
    Mar 2018 - Jul 2020 (2 years 5 months)
    • Experienced in building and architecting multiple Data pipelines, end to end ETL and ELT process for Data Ingestion and transformation in AWS and Spark. • Leveraged cloud and GPU computing technologies for automated machine learning and analytics pipelines, such as AWS • Participated in all phases of data mining; data collection, data cleaning, developing models, validation, visualization, performed Gap analysis provide feedback to the business team to improve the software delivery. • Data Mining with large datasets of Structured and Unstructured data, Data Acquisition, Data Validation, Predictive modeling, Data Visualization on provider, member, claims, and service fund data. • Involved in Developing RESTful API's (Microservices) using P
  • Technosoft
    Data Engineer
    Technosoft
    Apr 2015 - Nov 2017 (2 years 8 months)
    • Installed and configured Hadoop Map Reduce, HDFS, developed multiple Map Reduce jobs in java and Scala for data cleaning and preprocessing. • Experienced in installing, configuring and using Hadoop Ecosystem components. • Experienced in Importing and exporting data into HDFS and Hive using Sqoop. • Participated in development/implementation of Cloudera Hadoop environment. • Experienced in running query-using Impala and used BI tools to run ad-hoc queries directly on Hadoop. • Integrated Cassandra as a distributed persistent metadata store to provide metadata resolution for network entities on the network • Involved in various NOSQL databases like Hbase, Cassandra in implementing and integration. • Installed and configured Hive and written
  • B
    Data Modeler
    Brio Technologies Inc.
    Jun 2014 - Mar 2015 (10 months)
    • Interacted with business users to analyze the business process and requirements and transformed requirements into Conceptual, logical and Physical Data Models, designing database, documenting and rolling out the deliverables. • Analyze various data sources such as SQL server, Oracle, Flat files etc. and identify potential source systems for projects. • Created data models for AWS Redshift and Athena from dimensional data models. • Coordinated data profiling/data mapping with business subject matter experts, data stewards, data architects, ETL developers, and data modelers. • Developed logical/physical data models using Erwin tool across the subject areas based on the specifications and established referential integrity of the system. • Re
Education verified_user 0% verified
  • Jawaharlal Nehru Technological University
    bachelor, Bachelor’s computer science and engineering, Computer Science
    Jawaharlal Nehru Technological University
    Jan 2010 - Jan 2014 (4 years 1 month)
This is a community-created genome.