H

Het Shah

About

Detail

Bengaluru, Karnataka, India

Contact Het regarding: 
Flexible work

Timeline


work
Job
school
Education
flag
Award
auto_stories
Publication

Résumé


Jobs verified_user 0% verified
  • G
    Open Source ML Contributor
    Google Summer of Code (GSoC) 2026
    Jun 2026 - Current (4 months)
    • Develop and optimize semantic segmentation architectures (U-Net, SegNet) in PyTorch on multi-temporal satellite imagery, increasing coastline reconstruction accuracy by 16% across 1000+ km of Alaskan territory.
    • Build longitudinal time-series forecasting models on large-scale remote sensing data arrays, automating environmental tracking pipelines and reducing regional erosion predictive errors by 11%.
  • Wadhwani AI
    Machine Learning Engineer Intern
    Wadhwani AI
    Jan 2026 - Jun 2026 (6 months)
    • Data Pipeline Infrastructure (AWS/GCP): Architected an enterprise agricultural knowledge base utilizing the Mistral OCR API, orchestrating automated cloud-native ETL jobs via AWS S3 and GCP Cloud Functions to ingest 8,000+ complex multi-modal documents.
    • Production Deployment (FastAPI/vLLM): Containerized and deployed a production-grade agentic RAG pipeline using FastAPI, vLLM, and Docker to serve live agronomic advisory queries, achieving an end-to-end pipeline latency of ∼5 seconds.
    • MLOps & Experiment Tracking (WandB): Systematically benchmarked 5+ distinct RAG configurations and open-source LLMs across 1,500+ golden evaluation queries; monitored retrieval precision and generation alignment end-to-end via Weights & Biases (
  • I
    Undergraduate Researcher
    Indraprastha Institute of Information Technology, Delhi
    Oct 2025 - Dec 2025 (3 months)
    • Implemented bidirectional program slicing on C/C++ codebases, reducing vulnerability search space by 42% across 1.5M+ lines of production code.
    • Fine-tuned Large Language Models (LLMs) on sliced code representations to classify vulnerabilities, achieving a 25% improvement in F1-score detection performance over established baselines.
    • First-authored a research paper detailing the LLM-driven vulnerability detection pipeline, evaluating models on code vulnerabilities, currently under peer review at an A●-ranked conference.
  • I
    Teaching Assistant - Reinforcement Learning
    Indraprastha Institute of Information Technology, Delhi
    Aug 2025 - Dec 2025 (5 months)
    • Instructed 80+ undergraduate students in deep reinforcement learning theory (PPO, GRPO, Q-learning) and guided development of PyTorch-based agent training frameworks, yielding a 92% positive instructional rating.
  • I
    Research Intern
    IIT Patna
    May 2025 - Oct 2025 (6 months)
    • Pioneered a customized GRPO (Group Relative Policy Optimization) alignment architecture, optimizing a binary reward function (0/1) for exact ground-truth matching, which secured a 27% exact-match accuracy boost on a complex cultural QA dataset (Certificate).
  • A
    Research Intern
    AI Institute, University of South Carolina (AIISC)
    Jan 2025 - Oct 2025 (10 months)
    • Engineered an agentic multi-modal AI execution pipeline to parse complex lecture slides, autonomously generating difficulty-graded quiz questions and assessments with a 94% factual accuracy rate against the source material.
    • Programmed a personalized learning engine driven by hybrid reinforcement learning for real-time difficulty adaptation, integrating Llama-3.2 summary pipelines to boost active student retention by 19%.
    • Co-authored system-architecture paper ‘‘PAL: Personal Adaptive Learner’’, peer-reviewed and accepted for publication at AAAI 2026 (Certificate).
  • F
    Undergraduate Researcher
    FLaME.nlp Research Lab, IIIT Delhi
    Jan 2024 - Jun 2025 (1 year 6 months)
    • Designed a novel architecture based on knowledge distillation and Chain-of-Thought (CoT) reasoning for conversational AI models, successfully improving classification accuracy by 19% on the F1-score.
    • Curated a specialized dataset of 15,000 annotated dialogue records, constructing an automated quality-control validation pipeline using Pandas that programmatically cut label noise by 14%.
    • Co-authored ‘‘Measuring What Matters: Assessing Therapeutic Principles in Mental-Health Conversations’’, accepted for an Oral Presentation at the ACL 2026 Main Conference (Certificate).
Education verified_user 0% verified
  • C
    Machine Learning Specialization
    Coursera (Andrew Ng)
    Jan 2025
  • C
    Deep Learning Specialization
    Coursera (Andrew Ng)
    Jan 2025
  • I
    Bachelor of Technology in Computer Science and Engineering
    Indraprastha Institute of Information Technology, Delhi
    Aug 2022 - May 2026 (3 years 10 months)
    CGPA: 8.08 / 10.0
Projects (professional or personal) verified_user 0% verified
  • Independent
    Economic Recession Cycles Prediction
    Independent
    • Built a robust forecasting model on 40+ years of U.S. macroeconomic data to predict GDP trajectories 6 months in advance, achieving a 0.91 R-squared score.
    • Detected historical recession events with an 89% F1-score using ensemble feature engineering and statistical/regression analysis on yield curves, unemployment rates, and leading economic indicators.
  • Independent
    Harmful Meme Classification Pipeline
    Independent
    • Implemented a multi-modal PyTorch neural classifier utilizing a cross-attention transformer layer to blend visual features (ViT-hub) and text representations (LLM-based) over a curated dataset of 10,000+ high-noise targets.
    • Established a new state-of-the-art (SoTA) performance threshold on the benchmark task, eclipsing prior baseline models by 4.5% on macro F1-score evaluation metrics.
Awards verified_user 0% verified
  • B
    Best Undergraduate Researcher
    Jan 2025
  • A
    All India Rank (AIR) 152
    Jan 2025
Publications verified_user 0% verified
  • M
    Measuring What Matters: Assessing Therapeutic Principles in Mental-Health Conversations
    Jan 2026
  • P
    PAL: Personal Adaptive Learner
    Jan 2026
  • R
    Rethinking Reward Models! A Conceptual Framework for Enhancing LLM Reasoning through Intrinsic Traits
    Jan 2026