Testimonials
Helpful
Thank you so much Sankalp. Very informative and helpful.
Explaining the whole feature around interview
Services
Video meeting . 30 mins
Video meeting . 30 mins
Priority DM . 2 days reply
Priority DM . 2 days reply
5
Video meeting . 60 mins
Video meeting . 30 mins
Video meeting . 30 mins
Priority DM . 2 days reply
Video meeting . 60 mins
5
About me
• Sankalp is a Full Stack data scientist with over 9 years of experience. He has agile software engineering experience with client facing consulting experience; translating business requirements into machine learning solutions.
• Experience in the field of Data Science extensively in the algorithms which are used across the ML landscape and also includes: unsupervised clustering (Kmeans, Hierarchical, SOM), Ensemble methods for classification and regression problems (SVM, XGBoost, NueralNet), time series forecasting, and more general optimization and simulation development.
• Experience in Big Data Analytics technologies using Hadoop, HDFS, Spark, Scala, Spark-SQL, Linux, MCS(MapR Control System), AMBARI, HUE, Hive, Pig, HBase, Sqoop, YARN, Oozie and Flume.
• Work experience with Hadoop distributions like MapR and Hortonworks.
• Work experience with Cloud Distribution like AWS and GCP.
• Experienced with the Spark application in creating algorithm, improving the performance and optimization of the algorithms in Hadoop using Scala/Python, Spark Context, Spark-SQL, Data Frame, and Pair RDD's.
• Worked on Spark Streaming to receive real time data from the flume and store the stream data to HDFS using Scala.
• Good working knowledge in using Sqoop to transfer bulk data between relational databases & HDFS and Flume for ingesting streaming data into HDFS.
• Wrote Hive queries for data analysis to meet the business requirements and involved in migrating Hive queries into Spark transformations using Data frames, Spark SQL, SQL Context, and Scala.
• Knowledge of developing Data Ingestion and data clearing scripts using Shell script.
• Experience in Oozie workflow for scheduling the jobs and creating dependencies between these jobs.
• In depth knowledge of Job Tracker, Task Tracker, Name Node, Data Node, Resource Manager, Node Manager, YARN concepts.
• Worked with agile, Scrum and Sprint software development framework for managing product development.