I'm Pranjal Verma, a Data Engineer at Accenture in Pune, specializing in building scalable ETL/ELT pipelines, distributed data processing systems, and cloud-native data solutions.
I work primarily with Python, SQL, PySpark, Apache Spark, Airflow, and AWS, with a focus on data pipelines, automation, and reliable data architectures. I also enjoy building open-source projects and writing about data engineering at blog.pvcodes.in.

Experience
Accenture
- Built scalable ETL/ELT pipelines using AWS Glue, PySpark, and Apache Airflow for reliable data processing and analytics.
- Optimized distributed Spark workloads on Amazon EMR, increasing pipeline throughput by 3x through partitioning, cluster sizing, and Parquet optimization.
Walkover
- Developed backend services and REST APIs using Node.js, Express, PostgreSQL, and MongoDB for production applications.
- Implemented asynchronous workflows and optimized database operations to improve backend performance and reliability.
Projects
End-to-end cloud data pipeline for scraping and transforming VALORANT esports data using a Medallion Architecture, enabling analytics on player performance, team compositions, and map statistics.
A multi-model LLM chatbot platform supporting different large language models with a unified chat interface.
A predictive research model that identifies kidney stone risk by analyzing individual health factors such as high blood pressure and dietary saturated fatty acid intake.
A fine-tuned Qwen2.5-VL model that converts database ER diagrams into structured JSON schemas, achieving 89.2% table accuracy and 90% relationship accuracy—outperforming the base model.
Education
Devi Ahilya Vishwavidyalaya, Indore
Master of Computer Applications (MCA) in Computer Applications
Integral University, Lucknow
Bachelor of Computer Applications (BCA) in Computer Applications