Skip to main content

Posts

Showing posts with the label interview questions hadoop

interview questions on big data group

  Difference between DWH and Datalake? 1) What is Edge Node? 2)What are the client and Cluster-Mode? 3)What is the best approach while running your spark application in Prod? 4)What is Dynamic Memory Allocation in spark and why do we need it? 5)What is shuffle Services and how to enable it? 6)DataFrame and Dataset? 7)Serialization and Deserialization, Encoders and Decoders? 8)Avro, Orc and Parquet differences? 7)Hive ACID properties? 8)What is narrow and Wide Transformation? 9)When stages will be  Created in Spark UI? 10)Question from GIT And GITHUB 11)Few Question from Maven 12) What is the process to deploy your code in production? 13)What is Kyro Serialization? 14)What are the optimization techniques used in Spark? How to connect data node, what is the password type means is that static or dynamic, How you will get the password to connect datanode and apart from that what security you are follow to protect datanode? 2.assume that the data is in JSON fo...