C

Chaitanya O

About

Detail

Sr Data Engineer | Cloud Platforms | Data Processing & Analytics
United States

Contact Chaitanya regarding: 
work
Full-time jobs

Timeline


work
Job

Résumé


Jobs verified_user 0% verified
  • City National Bank
    Sr Data Engineer
    City National Bank
    Sep 2025 - Current (1 year)
    Senior Data Engineer leading enterprise data governance and cloud modernization. Implemented Informatica Intelligent Data Management Cloud (IDMC) capabilities including Data Catalog, Data Quality, Data Lineage, and Data Marketplace concepts for banking datasets. Built metadata-driven ETL pipelines using AWS Glue, PySpark, and Spark SQL with full lineage tracking across AWS and Azure systems. Enabled governed data products for analytics and AI-ready datasets. Improved data discoverability, reduced manual validation through automation, enabled self-service analytics, and strengthened audit compliance through end-to-end lineage and governance controls.
  • Ford Motor Company
    Sr Data Engineer
    Ford Motor Company
    May 2025 - Sep 2025 (5 months)
    Designed and deployed scalable data pipelines, Data Lake architectures, and real-time processing systems using tools like BigQuery, Azure Synapse, Dataflow, and Databricks. Expertise in ETL/ELT development, data warehousing, and data governance, supporting analytics and reporting across complex enterprise environments. Automated infrastructure provisioning with Terraform and optimized workflows for performance and cost efficiency. Adept at working with cross-functional teams to deliver secure, reliable, and high-performance data solutions that enable smarter decision-making and drive digital transformation.
  • UPS
    Sr Data Engineer
    UPS
    Mar 2025 - Aug 2025 (6 months)
    Implemented data governance frameworks and lineage models for GCP pipelines using BigQuery, Dataflow, and Dataproc, along with AWS-based data platforms using S3, Glue, and Redshift for enterprise analytics workloads. Built governed data products using PySpark and integrated CI/CD pipelines using Terraform. Applied Informatica Intelligent Data Management Cloud (IDMC) governance concepts for metadata-driven data management and classification. Improved data traceability, reduced inconsistencies via validation frameworks, and increased dataset reuse through standardized data product structures across both AWS and GCP environments.
  • BMO
    Sr Data Engineer
    BMO
    Sep 2024 - Feb 2025 (6 months)
    Built end-to-end ETL/ELT pipelines, developed robust Data Lake architectures, and implemented real-time data processing using Dataflow, BigQuery, Azure Data Factory, and Databricks. Applied Spark, SQL, and Python to transform and analyze large datasets. Optimized system performance and cost efficiency, enforced data governance, and delivered actionable insights through Looker, Power BI, and BigQuery. Collaborated with cross-functional teams to deliver reliable data platforms that supported strategic decisions and improved operational outcomes.
  • Caterpillar Inc
    Sr Data Engineer
    Caterpillar Inc
    May 2024 - Feb 2025 (10 months)
    Designed AWS data lake architecture with governance, metadata, and lineage tracking. Built streaming pipelines using Kafka and Kinesis for real-time IoT analytics. Applied Informatica Intelligent Data Management Cloud (IDMC) concepts for metadata enrichment, classification, and governed data structures. Improved real-time processing efficiency, reduced data quality issues, and enhanced data discovery through metadata-driven governance.
  • Siemens
    Data Engineer
    Siemens
    May 2023 - May 2024 (1 year 1 month)
    Built AWS data lake with governance and metadata frameworks aligned to modern data governance practices. Configured Glue Catalog for metadata management and implemented certified dataset publishing aligned with Informatica Intelligent Data Management Cloud (IDMC) governance model. Improved data reliability through automated validation, accelerated reporting through curated datasets, and strengthened compliance via standardized metadata and lineage tracking.
  • Siemens
    Sr Data Engineer
    Siemens
    May 2023 - May 2024 (1 year 1 month)
    As a Cloud Data Engineer, I successfully led the migration of on-premises data solutions to Google Cloud Platform (GCP), managing ETL processes from data extraction to transformation and loading. I developed scalable and secure RESTful APIs using Flask for real-time data processing with Cloud SQL and Cloud Spanner. My experience also includes optimizing query performance in BigQuery through partitioning and clustering, and using Kafka and Spark for real-time data ingestion. I implemented DevOps practices with CI/CD pipelines using GitLab, Jenkins, Docker, and Kubernetes, enhancing development workflows and automating infrastructure with Terraform. I worked extensively with NoSQL databases like HBase, integrating them with PySpark for real-t
  • Merck
    Data Engineer
    Merck
    Aug 2022 - Apr 2023 (9 months)
    Built governed GCP pipelines (BigQuery, Dataflow, Dataproc) with a strong focus on metadata, lineage, and compliance for regulated pharma datasets. Introduced Informatica Intelligent Data Management Cloud (IDMC) governance concepts for metadata management, data quality, and data lineage alignment in enterprise pipelines. Improved regulatory compliance, ensured consistency across batch and streaming systems, reduced manual reconciliation, and enabled trusted analytics datasets.
  • Edward Jones
    Data Engineer
    Edward Jones
    Feb 2022 - Jul 2022 (6 months)
    Developed AWS ETL pipelines using Glue, S3, and Redshift with data quality validation and streaming ingestion using Kafka and Kinesis. Applied early Informatica Intelligent Data Management Cloud (IDMC) Data Quality concepts for financial datasets, including validation and reconciliation frameworks. Improved data accuracy, reduced pipeline failures through schema enforcement, strengthened audit readiness through lineage tracking, and improved real-time analytics availability.
  • Capgemini
    Data Engineer
    Capgemini
    Sep 2021 - Jul 2022 (11 months)
    Experienced Data Engineer with expertise in designing and developing scalable ETL pipelines using Azure Data Factory, Azure Databricks, and Snowflake. Skilled in cloud-based data storage and processing solutions, including Azure Data Lake and SQL Database, along with AWS S3 and Redshift. Proficient in building and optimizing data transformations, utilizing PySpark, SQL, and Python. Extensive experience in implementing CI/CD pipelines with Azure DevOps, Jenkins, and GitHub. Adept at working in Agile teams, creating data visualizations with Power BI, Tableau, and DOMO, and ensuring performance optimization and automation of data workflows. Strong background in cloud platforms, BI tools, and data management.
  • Warner Bros Entertainment
    Cloud Data Engineer
    Warner Bros Entertainment
    Sep 2021 - Feb 2022 (6 months)
    Built AWS ETL pipelines using Glue and Spark with Informatica PowerCenter integration for structured data movement into Redshift. Improved pipeline stability, reduced transformation errors, and ensured consistent data delivery for analytics systems.
  • Walmart
    Big Data Engineer
    Walmart
    Feb 2021 - Sep 2021 (8 months)
    Designed GCP data lake architecture with governance, classification, and access control frameworks using Dataflow and Composer. Improved data security through IAM controls, enhanced data discoverability, and reduced manual intervention in pipeline operations.
  • Genpact
    Big Data Engineer
    Genpact
    Feb 2021 - Sep 2021 (8 months)
    Successfully ingested data from relational databases into HDFS using Sqoop, including creating and evaluating Sqoop jobs and incremental loading. I optimized Hive data modeling with partitions, bucketing, and indexing. I installed and configured essential tools such as Hive, Pig, Sqoop, Flume, and Oozie on Hadoop clusters and developed reusable Hive UDF libraries to meet business needs. I transformed and analyzed raw data using Hive queries and Pig scripts and converted SQL queries into Spark transformations. I built a recommendation system using Spark MLlib, integrated Kafka for real-time data processing, and optimized workflows in Oozie. Additionally, I have strong experience with MapReduce, HBase, Spark Streaming, and Tableau reporting w
  • Clover Infotech
    ETL Data Engineer
    Clover Infotech
    Jan 2019 - Jan 2021 (2 years 1 month)
    Gathered business requirements and prepared technical design documents, target-to-source mapping, and mapping specification documents. Extensively worked on Informatica Power Center and Ab-initio to develop ETL mappings and data transformations. Optimized query performance using Oracle hints and indexes. Developed and managed Sqoop jobs for Hive tables. Led research for technical improvements, enhancing application reliability. Skilled in UNIX Shell Scripting for automation and file management. Performed unit, integration, and system testing and assisted in UAT Testing. Experienced with AWS S3, Redshift, Data Pipeline, and SAP HANA environments.
  • Morgan Stanley
    ETL Data Engineer
    Morgan Stanley
    Jan 2019 - Jan 2021 (2 years 1 month)
    Developed Informatica PowerCenter and SSIS ETL workflows with metadata-driven lineage and governance support. Improved data integration efficiency, strengthened traceability, and reduced reporting delays through optimized ETL frameworks.
  • S
    SQL Data Engineer
    Sagacious Infosystems
    Mar 2017 - Dec 2018 (1 year 10 months)
    Experienced Data Analyst with a strong background in SQL Server, database design, ETL processes, and data conversion. Proficient in creating complex T-SQL queries, optimizing SQL Server performance, and developing SSIS packages for data transformation. Demonstrated expertise in designing high-level ETL architecture and working with diverse data sources. Skilled in data cleansing, mapping, and mining to ensure data integrity and efficiency. Possess strong analytical and problem-solving skills with a proven track record of enhancing data processing and reporting capabilities. Utilized MS SQL Server, T-SQL, PL/SQL, SSRS, SSAS, SSIS, and Visual Studio effectively.
  • State Farm
    SQL Data Engineer
    State Farm
    Mar 2017 - Dec 2018 (1 year 10 months)
    Insurance Data Warehouse & ETL Automation - Built SQL Server, SSIS, and Informatica-based ETL solutions for insurance data processing and reporting. Automated data workflows and optimized T-SQL procedures, reducing manual reporting effort by ~30% and improving data consistency across reporting systems.
  • Creditsafe
    SQL Developer
    Creditsafe
    Jun 2015 - Mar 2017 (1 year 10 months)
    Experienced SQL Developer with a strong background in database design, development, and performance optimization, specializing in Oracle environments (11g/9i) and UNIX systems. Designed and implemented complex relational database structures, including tables, views, indexes, triggers, and stored procedures to support high-volume enterprise applications. Developed and tuned advanced SQL and PL/SQL scripts for data processing, ensuring performance efficiency and accuracy. Automated batch processing and backup tasks using UNIX Shell Scripts and configured ActiveBatch for job scheduling across ETL pipelines and Python/SQL jobs. Led ETL development using SSIS, crafted detailed SSRS reports, and implemented robust data migration strategies with m
Education verified_user 0% verified
  • Indiana Wesleyan University
    Masters, Data Analytics
    Indiana Wesleyan University
  • PBR Visvodaya Institute of Technology  Science
    Bachelor of Technology - BTech, Electrical and Electronics Engineering
    PBR Visvodaya Institute of Technology Science
This is a community-created genome.