S

Sohail Anjum

About

Detail

Islamabad, Pakistan

Timeline


work
Job
school
Education

Résumé


Jobs verified_user 0% verified
  • Fusemachines
    Data Engineer
    Fusemachines
    Jun 2024 - Current (2 years 2 months)
    • Developed and optimized PySpark applications for processing large-scale datasets in a distributed environment. • Defined AWS Redshift Schema, Identified Dimension and Fact tables. SCD type implementation. • Designed and implemented AWS-based data pipelines using S3, ECR and Lambda for scalable and efficient data processing. • Containerized applications using Docker to ensure seamless deployment and environment consistency. • Optimized performance for big data solutions in a cloud-based ecosystem. • Data Ingestion in various target destination such AWS Redshift, PostgreSQL, S3 and Glue Catalog. • Automated workflows and DAGs deployments using AWS Managed Workflow for Apache Airflow. • Used AWS Docker Image for in-house development
  • DPL
    Senior Data Engineer
    DPL
    Jan 2022 - Jun 2024 (2 years 6 months)
    • Structured and programmed scalable data pipelines using AWS Glue and Lambda functions to process large volumes of data from various sources. • Utilized AWS Glue allowing seamless data integration, transformation, and loading into data lakes, databases and warehouses. • Designed and implemented multiple databases using Amazon RDS, to support the company's web, mobile applications and IoT data • Implemented data processing pipelines using PySpark, a Python library for Apache Spark, to handle large scale data transformations efficiently. • Utilized AWS Glue for ETL processes, allowing seamless data integration, transformation, and loading into data lakes and warehouses. • Designed and optimized data workflows to ensure scalability and
  • S&P Global Market Intelligence
    Senior Data Engineer
    S&P Global Market Intelligence
    Feb 2016 - Dec 2021 (5 years 11 months)
    • Experienced in designing and deploying data pipelines and analytics solutions using Databricks on AWS. • Developed ETL workflows using Databricks and Delta Lake. • Implementing complex business transformation using PySpark language and storing transformations into delta/parquet format, performed Insert, Updates and UpSert using merge statement. • Performed aggregates and analytics using PySpark/Databricks and produced quality results. • Persistence of data using JDBC in target databases such as Aurora PostgreSQL, MS SQL Server etc. • Configured changeable parameters and credentials in Databricks Secrets. • Proficient in Agile Scrum methodology, with experience in sprint planning, backlog grooming, daily standups, sprint reviews, an
  • F
    Software Engineer
    F3 TechnologiesHealth Care Technologies
    Mar 2014 - Jan 2016 (1 year 11 months)
    • Analyzed and implemented best coding practices into the project code. • Full knowledge of how to design, program, implement, and maintain. • Identified and developed areas for revisions in current projects. • Executing and implementing software tests. • Developed quality assurance procedures for software projects. • Coordinating the efforts and cooperating with other developers, designers, system and business analysts, etc. • Break down bigger tasks into smaller, easily managed ones • Documented every part of the development process for further work and maintenance.
  • T
    Software Engineer
    Trees Technologies
    Apr 2013 - Jan 2014 (10 months)
    • Coordinated with the Technical Lead on current programming tasks. • Collaborated with other programmers to design and implement features. • Quickly produce well-organized, optimized, and documented source code. • Create and document software tools required by artists or other developers. • Debug existing source code and polish feature sets. • Contributed to technical design documentation. • Work independently when required.
Education verified_user 0% verified
  • U
    BS (CS)
    University of Arid Agriculture
    Jan 2008 - Jan 2012 (4 years 1 month)