Services
Priority DM . 3 days reply
Video meeting . 60 mins
Video meeting . 30 mins
About me
1. Worked on Cloudera - HDP (Hortonworks Data Platform, Ambari) and have Detailed understanding of
Hadoop Architecture like Meta store, Hiveserver2, HS2 Interactive for spark, Name node, Data node,
Node Managers
etc.
2. Worked on Name node/ Data node issues, Handled Name node Failure, failure cause analysis, Major
production downtime issues effectively.
3. Written PySpark Code to effectively process Apache Ranger Audit Logs to derive meaningful data in
Tabular/ CSV format.
4. Written Sqoop Commands to import Transactional Records from MySQL to Hadoop and vice-versa.
5. Worked on Hive Compaction Issue, solved using insert overwrite, Major/ Minor Compactions.
6. Written Multiple SQL queries to derive meaningful report from Hive Meta store DB on MySQL.
7. Transfer of Inter-Cluster Data/ Hive Transactional/ Non-transactional, Managed/ External Tables with
Terabytes data using export, distcp commands.
8. Worked on Key tab generation of Service accounts, encryption of passwords, Kerberos principles.
9. Worked on different File Formats Like Columnar File Format (Orc, Parquet), and row-based file-format
such as Avro, text files and partitioned, bucketed hive tables.
10.Troubleshoot Daily BAU Jobs in case of failures, root cause analysis from application and HiveServer/
Meta store logs, solution for failures.