Services
Priority DM . a day reply
Ask me anything
Having any question related to Data engineering?
Video meeting . 20 mins
Career guidance and roadmap for freshers/students
confused about how to make a career in data?
Video meeting . 30 mins
Career roadmap for experienced
how to pivot your career into data roles?
Video meeting . 15 mins
Video meeting . 20 mins
Data Engineering Roles and Responsibilities
what work you can expect and work on in your data role
Video meeting . 30 mins
Cracking the Data Engineering Interviews
How to be interview ready for data engineering jobs
About me
I am Harsh Prateek Singh, a Senior Data Engineer with over 10 years of experience designing and implementing large-scale data solutions. I specialize in building robust, scalable data pipelines and optimizing distributed data processing systems.
My expertise spans across:
Big Data technologies: Apache Spark, Hadoop, Hive, HDFS, Sqoop
Programming languages: Scala, Python (PySpark), SQL
Cloud platforms: AWS (S3, EMR, Glue, Redshift), with hands-on experience in deploying data workloads in production
Data orchestration tools: Apache Airflow, Oozie
Data modeling & ETL: End-to-end development and optimization of data pipelines for batch and near-real-time systems
I’ve worked on a wide range of data-driven projects — from data lake architecture to performance tuning of Spark jobs handling terabytes of data. My focus is always on writing clean, maintainable code, improving pipeline efficiency, and enabling teams with reliable, high-quality data.
I enjoy mentoring aspiring data engineers and helping others solve challenging problems in the data ecosystem.