NEW
Font size
WorksheetsQuiz on Big data technologies
Total questions: 15
Worksheet time: 15mins
All of the following accurately describes Apache Hadoop except
Distributed processing
Java based
Open source
Best for little amounts of data also
Hadoop is named after
Doug cuttings father
Doug cuttings company
Doug cuttings son toy elephant
Doug cutting
--------------------- is the component used to write sql like queries in Hadoop
Ambari
Hive
Mahooth
Oozie
Default replication factor of Hadoop is ----------
1
2
3
4
Expansion of HDFS
Hadoop data field system
Hive distributed file system
Hadoop data file system
Hadoop distributed file system
---------------------- takes the sorted output of Mapper as its input.
Mapper
Splitter
Sorter
Reducer
------------------ controls various components of Hadoop system
OOzie
Zoo keeper
Hive
Pig
This is the master node which saves the meta data of all data blocks.
Data node
Name node
Secondary node
Secondary name node
------------------ can be described as a programming model used to develop hadoop based applications for distributed processing
Hive
Hbase
Map reduce
HDFS
Transactional data of any bank is
Structured data
Unstructured data
Semi structured data
None of the above
The following are 3 V's of Big data
Velocity variety and Variance
Volume varasity and Variance
Volume variety and Variance
Velocity variety and Volume
Columnar storage of Hadoop is
Hive
Pig
Hbase
HDFS
The following are advantages of Big data processing
Cost efficency
Speed
Fault tolerance
All of the above
Analyzing the data and predicting the future is called as --------------------------- analytics
Descriptive analytics
Prescriptive analytics
Predictive analytics
None of the above
First word in a line is taken as key in the following input format
Text input format
n line input format
Sequence file input format
Key value text input format
