Wayground logo

Free Printable Worksheets

Font size

S
M
L
XL
Worksheets

Hadoop

Total questions: 15

Worksheet time: 10mins

Name
Class
Date
1.

.....................................is a Top level project, open-source implementation of frameworks for reliable, scalable, distributed computing and data storage.

a)

Hadoop

b)

Hive

c)

MongoDB

d)

Cassandra

2.

....................................is the storage layer, .....................................is the resource management layers and ----------------------------------is data processing layer of hadoop

a)

Map Reduce, YARN and HDFS

b)

HDFS, MapReduce, YARN

c)

YARN, MapReduce and HDFS

d)

HDFS, YARN and MapReduce

3.

NameNode is the ...................................... in the Apache Hadoop HDFS Architecture that maintains and manages the blocks present on the DataNodes .......................

a)

Master node, and Slave node

b)

Job Tracker and Node tracker

c)

Slave node and Master Node

d)

Task Tracker and Application manager

4.

The ........................................................ works concurrently with the primary NameNode as a helper daemon.

a)

Secondary NameNode

b)

Master Name Node

c)

Application Manager

d)

Node Manager

5.

Default size of each block in Apache Hadoop 2.x is

a)

64MB

b)

128KB

c)

128GB

d)

128 MB

6.

I have a file “example.txt” of size 514 MB. If default configuration of block size, which is 128 MB. Then, how many blocks will be created?

a)

4

b)

5

c)

2

d)

6

7.

if you are storing a file of 100 MB in HDFS using the default configuration, and if it perform replication management policy then how much of data is residing in different data node

a)

100 MB

b)

200 MB

c)

600 MB

d)

300 MB

8.

...................................... is the utility that enables us to create or run MapReduce scripts in any language either, java or non-java, as mapper/reducer.

a)

Hadoop streaming

b)

Hadoop YARN

c)

Hadoop HDFS

d)

Hadoop Apache

9.

Hadoop framework is written in which language ?

a)

Python

b)

C

c)

Java

d)

Go

10.

Hadoop YARN Resource Manager (RM)

a)

RM manages the global assignments of resources (CPU and memory) among all the applications. It arbitrates system resources between competing applications

b)

is responsible for containers monitoring their resource usage and reporting

c)

•tracks the health of the node on which it is running.

d)

•It negotiates resources from the resource manager and works with the node manager.

11.

Moving data in and out of Hadoop, is referred as data .........................

a)

Input and output

b)

Input format and output format

c)

Job tracker and node tracker

d)

Ingress and egress

12.

What is the name of the Manager in YARN that controls Data Node?

a)

Node Manager

b)

Application Manager

c)

Resource Manager

d)

Server Manager

13.

Apache Spark is a............................

a)

Batch Processing

b)

Stream Processing

c)

Graph Processing

d)

All of the above

14.

Which tool in Hadoop ecosystem is responsible for Machine learning processing?

a)

Sqoop

b)

Mahout

c)

ambari

d)

hive

15.

Which of these is not a type/flavor of Hadoop?

a)

Cloudera

b)

MapR

c)

Hortonworks

d)

AWS-EC2