Raju Gangula

Raju Gangula

About

Detail

Senior Data Engineer in Change Healthcare| AWS Redshift | Azure Data Factory | Hadoop | Python | Pyspark | | C2C/C2H | Open to work |
Lombard, Illinois, United States

Contact Raju regarding: 
work
Full-time jobs

Timeline


work
Job
school
Education

Résumé


Jobs verified_user 0% verified
  • Change Healthcare
    Senior Data Engineer
    Change Healthcare
    Jun 2021 - Current (5 years 3 months)
    • Designed and implemented scalable data architectures on AWS, utilizing services like Amazon S3, Redshift, and RDS, to support efficient data storage and processing. • Expertise in using SSIS for designing and implementing ETL processes, data integration, and transformation. • Deployed and managed Docker containers for efficient application packaging and deployment, ensuring scalability and consistency across multiple environments. • Created end-to-end Python data pipelines, handling data extraction, transformation, and loading from various sources to data warehouses, enabling data-driven decision-making • Cleaned and preprocess raw data to ensure it is suitable for machine learning, including handling missing values, outlier detection, an
  • Molina Healthcare
    Azure Data Engineer
    Molina Healthcare
    Nov 2019 - May 2021 (1 year 7 months)
    • Leveraged Azure Databricks for big data analytics, developing and optimizing data processing jobs using Apache Spark for large-scale data sets. • Design and implement database solutions in Azure SQL Data Warehouse, Azure SQL. • Design and implement end-to-end data solutions (storage, integration, processing, visualization) in Azure. • Implement Copy activity and Custom Azure Data Factory Pipeline Activities. • Developed ELT processes from the files from abinitio, google sheets in GCP, with compute being data prep, data proc (pyspark), and Bigquery. • Involved in identifying production bugs in the data using stack driver logs in GCP. • Build data pipelines in airflow in GCP for ETL-related jobs using different airflow operators. • Deployed
  • First Republic
    Data Engineer
    First Republic
    Jun 2018 - Oct 2019 (1 year 5 months)
    Developed Spark programs and created the data frames and worked on transformations. Involved in loading data from Linux file systems, servers, Python web services using Kafka producers and partitions. Deployed Spark application and java web services in pivotal cloud foundry. Applied Kafka custom encoders for custom input format to load data into Kafka Partitions.  Implement POC with Hadoop. Extract data with Spark into HDFS. Used Spark SQL with Scala for creating data frames and performed transformations on data frames. Developed code to read data stream from Kafka and send it to respective bolts through respective stream. Creating Databricks notebooks using SQL, Python and automated notebooks using jobs. Creating Spark clusters an
  • Nationwide
    Senior Data Engineer
    Nationwide
    Nov 2017 - May 2018 (7 months)
    Created several types of data visualizations using Python and Tableau. Creating reusable User defined functions in java/python Developed and analyzed the SQL scripts and designed the solution to implement using spark Developed a data pipeline using Kafka, Spark and Hive to ingest, transform and analysing data. Supported MapReduce Programs and distributed applications running on the Hadoop cluster and scripting Hadoop package installation and configuration to support fully automated deployments. Migrated existing on-premises application to AWS and used AWS services like EC2 and S3 for large data sets processing and storage and worked with ELASTIC MAPREDUCE and setup Hadoop environment in AWS EC2 Instances. Worked with systems engineeri
  • Dhruvsoft Services Private Limited
    Big Data Developer
    Dhruvsoft Services Private Limited
    Mar 2016 - Aug 2017 (1 year 6 months)
    • Developed multiple MapReduce jobs in Python for data cleaning and preprocessing and assisted with data capacity planning and node forecasting. • Involved in design and ongoing operation of several Hadoop clusters and Configured and deployed Hive Meta store using MySQL and thrift server • Conducted performance tuning and troubleshooting of Tableau reports to optimize resource utilization and reduce report load times. • Uploaded and processed more than 30 terabytes of data from various structured and unstructured sources into HDFS (AWS cloud) using Sqoop and Flume. • Prepared complete description documentation as per the Knowledge Transferred about the Phase-II Talend Job Design and goal and prepared documentation about the Support and M
  • Jabong
    Data Analyst
    Jabong
    Jul 2014 - Feb 2016 (1 year 8 months)
    Wrote SQL scripts to meet the business requirement. Analyzed views and produced reports. Tested cleansed data for integrity and uniqueness. Automated the existing system to achieve faster and accurate data loading. Generated weekly, bi-weekly reports and sent to client business team using business objects. Data Cleaning, merging and exporting the dataset was done in Tableau Prep. Learned to create Business Process Models. Ability to manage multiple projects simultaneously tracking them towards varying timelines effectively through a combination of business and technical skills. Good Understanding of clinical practice management, medical and laboratory billing and insurance claim with processing with process flow diagrams. Assisted
Education verified_user 0% verified
  • JNTUH College of Engineering Hyderabad
    Bachelor of Technology - BTech, Computer Science
    JNTUH College of Engineering Hyderabad
    Aug 2011 - Aug 2015 (4 years 1 month)