R

RUPENDRA LANKE

About

Detail

Data Engineer
United States

Contact RUPENDRA regarding: 
work
Full-time jobs

Timeline


work
Job

Résumé


Jobs verified_user 0% verified
  • Allstate
    Data Engineer
    Allstate
    Jan 2024 - Current (2 years 8 months)
    • Developed reusable Spark libraries for common data processing tasks, leading to a decrease in development time for future pipelines. • Established data lineage tracking within Databricks to trace data origin and transformations, improving data governance and auditability. • Improved data decoupling by decoupling data producers and consumers using Kafka, allowing for independent development and scalability of data pipelines. • Automated data ingestion and transformation pipelines using AWS Glue and AWS Pipeline, reducing manual effort and improving data processing speed. • Implement Airflow pipelines that orchestrate batch processing of financial data at specific intervals for regulatory reporting, ensuring adherence to reporting dead
  • G
    Data Engineer
    Genius SoftTech,
    Jan 2019 - Jul 2022 (3 years 7 months)
    • Designed and implemented highly scalable data pipelines on Databricks using PySpark, ingesting terabytes of data per day from various sources (relational databases, log files, APIs). • Developed and maintained reusable Airflow DAGs (Directed Acyclic Graphs) for various data processing tasks, promoting code maintainability and scalability. • Created a Power BI dashboard to visualize key business KPIs, resulting in a weekly time savings of 2 hours on manual reporting. • Employed a robust data streaming architecture using Kafka, efficiently handling the ingestion and distribution of real -time data across diverse applications. • Established serverless functions using Lambda to trigger data processing tasks upon specific events in S3, ac
Education verified_user 0% verified
  • University of North Texas
    Master of Science
    University of North Texas