WorksheetsHadoop
Total questions: 15
Worksheet time: 10mins
.....................................is a Top level project, open-source implementation of frameworks for reliable, scalable, distributed computing and data storage.
Hadoop
Hive
MongoDB
Cassandra
....................................is the storage layer, .....................................is the resource management layers and ----------------------------------is data processing layer of hadoop
Map Reduce, YARN and HDFS
HDFS, MapReduce, YARN
YARN, MapReduce and HDFS
HDFS, YARN and MapReduce
NameNode is the ...................................... in the Apache Hadoop HDFS Architecture that maintains and manages the blocks present on the DataNodes .......................
Master node, and Slave node
Job Tracker and Node tracker
Slave node and Master Node
Task Tracker and Application manager
The ........................................................ works concurrently with the primary NameNode as a helper daemon.
Secondary NameNode
Master Name Node
Application Manager
Node Manager
Default size of each block in Apache Hadoop 2.x is
64MB
128KB
128GB
128 MB
I have a file “example.txt” of size 514 MB. If default configuration of block size, which is 128 MB. Then, how many blocks will be created?
4
5
2
6
if you are storing a file of 100 MB in HDFS using the default configuration, and if it perform replication management policy then how much of data is residing in different data node
100 MB
200 MB
600 MB
300 MB
...................................... is the utility that enables us to create or run MapReduce scripts in any language either, java or non-java, as mapper/reducer.
Hadoop streaming
Hadoop YARN
Hadoop HDFS
Hadoop Apache
Hadoop framework is written in which language ?
Python
C
Java
Go
Hadoop YARN Resource Manager (RM)
RM manages the global assignments of resources (CPU and memory) among all the applications. It arbitrates system resources between competing applications
is responsible for containers monitoring their resource usage and reporting
•tracks the health of the node on which it is running.
•It negotiates resources from the resource manager and works with the node manager.
Moving data in and out of Hadoop, is referred as data .........................
Input and output
Input format and output format
Job tracker and node tracker
Ingress and egress
What is the name of the Manager in YARN that controls Data Node?
Node Manager
Application Manager
Resource Manager
Server Manager
Apache Spark is a............................
Batch Processing
Stream Processing
Graph Processing
All of the above
Which tool in Hadoop ecosystem is responsible for Machine learning processing?
Sqoop
Mahout
ambari
hive
Which of these is not a type/flavor of Hadoop?
Cloudera
MapR
Hortonworks
AWS-EC2
