s

suraj kumar

About

Detail

.
United States

Contact suraj regarding: 
work
Full-time jobs

Timeline


work
Job

Résumé


Jobs verified_user 0% verified
  • TIAA
    Senior Data Engineer – Azure
    TIAA
    Jan 2022 - Current (4 years 7 months)
    • Designed and implemented end-to-end data solutions on Azure, levev raging services such as Azure Data Factory, Azure Databricks, Azure Data Lake Storage, and Azure SQL Database. • Developed data pipelines using Azure Data Factory to orchestrate data movement and transformations across diverse data sources and destinations. • Built and optimized scalable data processing workflows using Azure Databricks, leveraging Spark for data ingestion, transformation, and analysis. • Proficient in writing Spark jobs using Scala and Python in Azure Databricks notebooks for data cleansing, feature engineering, and advanced analytics. • Implemented real-time data processing solutions using Azure Event Hubs, Azure Stream Analytics, and Databricks Str
  • Simplify Healthcare
    Data Engineer
    Simplify Healthcare
    Jul 2020 - Dec 2021 (1 year 6 months)
    • Extracted, converted, and loaded data from various source to Azure Data Storage Services utilizing Azure data factory and T-SQL for data lake analytics. • Performed data transformations for ML OPs, including adding calculated columns, maintaining relationships, establishing various metrics, merging & appending queries, changing values, splitting columns, and grouping by Date & Time Column. • Data Ingestion to Azure Services, including Azure Data Lake, Azure Storage, Azure SQL, and Azure DW, as well as data processing in Azure Databricks. • Using Server Manager, created batches and sessions to transport data at preset intervals and on-demand. • Combined several data connections and produced multiple joins across different data sourc
  • Edifecs
    Data Engineer
    Edifecs
    Feb 2019 - Jul 2020 (1 year 6 months)
    • As a Big Data Developer implemented solutions for ingesting data from various sources and processing the Data-at Rest utilizing Big Data technologies such as Hadoop, MapReduce Frameworks, MongoDB, Hive, Oozie, Flume, Sqoop and Talend etc. • Migrated an in-house database to AWS Cloud and designed, built, and deployed a multitude of applications utilizing the AWS stack (Including EC2, RDS) by focusing on high-availability and auto-scaling. • Worked on analyzing Hadoop clusters using different big data analytic tools including Flume, Pig, Hive, HBase, Oozie, Zookeeper, Sqoop, Spark and Kafka. • Developed Spark code using Python and Spark-SQL/Streaming for faster testing and processing of data and • Data Extraction, aggregations, and co
Education verified_user 0% verified
  • Wichita State University
    Masters in
    Wichita State University