Ashwini Solanki
Ashwini Solanki
About
Detail
Passionate Data Engineer Seeking Exciting Opportunities
Houston, Texas, United States
I bring four years of professional software development expertise, specializing for the past three years in the dynamic realms of Big Data, Hadoop Ecosystem, Cloud Engineering, and Data Warehousing. My journey encompasses harnessing the power of AWS services like EC2, S3, EMR, RDS, VPC, Elastic Load Balancing, IAM, Auto Scaling, CloudFront, CloudWatch, and Lambda to orchestrate and trigger resources.
In the realm of Azure, I've crafted data pipelines with finesse using Azure Data Factory and Azure Databricks, seamlessly loading data into Azure Data Lake, Azure SQL Database, and Azure SQL Data Warehouse, ensuring precise control and secure database access. My expertise extends to HDInsight, Stream Analytics, Active Directory, Blob Storage, Cosmos DB, and Storage Explorer within the Azure ecosystem.
Mastering the rich tapestry of tools and services across major Hadoop Distributions—Cloudera, Amazon EMR, Azure HDInsight, and Hortonworks—has been a cornerstone of my professional journey. I've deftly managed and ingested massive volumes of streaming data, orchestrated automation through tools like Oozie and Airflow, and handled diverse data types, from Kafka and Spark streaming to batch data.
Spark applications have been a focal point of my development prowess, leveraging Spark SQL, Data Frames, Datasets, Spark-ML, and Spark Streaming to craft production-ready solutions. I've designed and implemented Kafka Producers and Consumers, storing stream data to HDFS and processing it efficiently using Spark.
Scripting with Python (PySpark), Scala, and Spark-SQL has been my forte, enabling agile development and aggregation from diverse file formats like XML, JSON, CSV, and Parquet. I've delved into the intricacies of HiveQL, Hive-ACID tables, Pig Latin queries, and custom MapReduce programs to extract meaningful insights from data.
My journey extends beyond development to encompass all facets of the data life cycle, from acquisition to warehousing, modeling, processing, and transformation. I've overseen the growth and storage estimation for large MongoDB clusters, executing ad-hoc queries, indexing, replication, load balancing, and aggregation with precision.
My journey is woven with a commitment to quality, utilizing bug tracking and ticketing systems such as Jira and Remedy, while ensuring version control through Git and SVN. Ready to embrace new challenges, I bring a wealth of experience and innovation to the intersection of software development and cutting-edge technologies.
Contact Ashwini regarding:
work
Full-time jobs