About
Computer Science undergraduate specializing in data engineering, ETL pipeline development, and data analytics. Experienced in Python, SQL, PySpark, Databricks, and Excel through academic projects and data analytics internships.
Experience
-
Data Analytics InternVK Packaging & Automations · Faridabad, HaryanaMay 2026 – Jun 2026
Analyzed multi-source operational and manufacturing datasets using Python (Pandas) and SQL to detect process bottlenecks and workflow inefficiencies. Constructed structured data visualizations and executive reporting formats in Excel/Power BI, aiding leadership in operational decision-making. Performed exploratory data analysis (EDA) and statistical evaluations to establish data quality standards across internal analytics workflows.
-
Data Analytics InternSamatrix Consulting Pvt. Ltd. · Jaipur, Rajasthan2025 – 2025
Applied structured data manipulation and statistical interpretation techniques to deliver consulting project deliverables. Conducted requirement gathering and data modeling routines to transform raw client datasets into structured analytical models.
Education
-
B.Tech in Computer Science EngineeringData Science & Analytics · 2023 – 2027
Skills
Tools / apps / platforms
Projects
-
Interactive Sales Performance DashboardMS Excel
Formulated an enterprise sales dashboard utilizing dynamic charts, Pivot Tables, and complex lookup logic to evaluate multi-regional performance KPIs. Streamlined data preparation workflows, reducing monthly executive report generation time by 35%.
-
Student Records Data Pipeline & AnalysisPython, SQL, Pandas
Designed SQL ETL transformation scripts and Pandas data-cleaning pipelines to standardize and process 5,000+ unstructured student academic records. Applied advanced SQL queries and Seaborn visualizations to identify key statistical variables driving student performance and retention.
-
Data Drift Analyzer & Pipeline MonitorPython, SciPy
Developed an automated statistical monitoring script using Python to execute Kolmogorov-Smirnov (KS) tests across incoming dataset distributions. Identified dataset schema and distribution shifts caused by evolving customer behavior and market trends, mitigating a potential 15% drop in downstream model accuracy. Documented automated anomaly detection thresholds to flag shifting historical vs. real-time production data patterns.
-
Automated Bus Scheduling & Route Management SystemPython, Smart India Hackathon 2024, DTC
Architected a real-time dynamic rerouting pipeline for Delhi Transport Corporation (DTC) to optimize bus scheduling and fleet management during urban transit congestion. Engineered data ingestion logic to process historical traffic volumes, vehicle speeds, and roadside sensor streams, improving simulated route planning efficiency by 18%. Integrated continuous data processing capabilities to dynamically re-route buses during traffic bottlenecks, delays, and road closures.