Services

Priority DM . a day reply

Ask me anything

Having any question related to Data engineering?
60
Popular
Video meeting . 20 mins

Career guidance and roadmap for freshers/students

confused about how to make a career in data?
150
Popular
Video meeting . 30 mins

Career roadmap for experienced

how to pivot your career into data roles?
300
Video meeting . 15 mins
100
Video meeting . 20 mins

Data Engineering Roles and Responsibilities

what work you can expect and work on in your data role
200
Video meeting . 30 mins

Cracking the Data Engineering Interviews

How to be interview ready for data engineering jobs
500

About me

I am Harsh Prateek Singh, a Senior Data Engineer with over 10 years of experience designing and implementing large-scale data solutions. I specialize in building robust, scalable data pipelines and optimizing distributed data processing systems. My expertise spans across: Big Data technologies: Apache Spark, Hadoop, Hive, HDFS, Sqoop Programming languages: Scala, Python (PySpark), SQL Cloud platforms: AWS (S3, EMR, Glue, Redshift), with hands-on experience in deploying data workloads in production Data orchestration tools: Apache Airflow, Oozie Data modeling & ETL: End-to-end development and optimization of data pipelines for batch and near-real-time systems I’ve worked on a wide range of data-driven projects — from data lake architecture to performance tuning of Spark jobs handling terabytes of data. My focus is always on writing clean, maintainable code, improving pipeline efficiency, and enabling teams with reliable, high-quality data. I enjoy mentoring aspiring data engineers and helping others solve challenging problems in the data ecosystem.