I am a Data Engineering Intern with a Bachelor of Technology background in Mechatronics and Automation Engineering. I work with structured data to build reliable pipelines for ingestion, transformation, validation, and analytics.
I have hands-on knowledge of Python, SQL, C++, ETL, data warehousing, data modeling, batch processing, and data quality practices. My work includes data cleaning, deduplication, integrity checks, and writing optimized SQL queries for reporting and analysis.
I have experience with big-data technologies including Apache Spark, Hadoop, HDFS, Hive, Sqoop, HBase, MapReduce, and YARN. I am familiar with relational and NoSQL database concepts and have worked with MySQL, PostgreSQL, and SQLite.
I have built end-to-end data engineering projects, including a PySpark financial transaction pipeline using a Bronze-Silver-Gold architecture and a real-time Spark Streaming application built with Scala. I am comfortable developing REST APIs with FastAPI and Flask and working in Linux-based development environments.
I have been recognized through competitive programming and technical challenges, including a Pre-Placement Interview through HackWithInfy 2025 and a semi-finalist position in Flipkart GRiD 7.0. I have solved more than 500 data structures and algorithms problems across coding platforms.
I am interested in data engineering roles where I can contribute to scalable data processing workflows, analytics systems, and cloud-based data solutions while continuing to deepen my expertise in modern data platforms.