Font size
WorksheetsQuiz - Post Course - Big Data Ecosystem
Total questions: 30
Worksheet time: 16mins
Data in ___________ bytes size is called Big Data.
Tera
Giga
Peta
Meta
All of the following accurately describe Hadoop, EXCEPT ____________
Open-source
Distributed computing approach
Java-based
Real-time
What are the main components of big data?
YARN
Map Reduce
HDFS
All of these
Which of the following genres does Hadoop produce?
Distributed file system
Java message service
JAX-RS
Relational database management system
The data node and name node in HADOOP are
Both Worker Nodes
Worker Node and Master Node respectively
Master Node and Worker Node respectively
Both Master Nodes
________ is general-purpose computing model and runtime system for distributed data analytics.
MapReduce
Oozie
Drill
All of the above
Do Hadoop need specialized hardware to process the data?
Yes
No
May be
Can't say
Which of the following can be used to extract structured information from unstructured data?
Organize data into information
Identify patterns in data
Create frameworks for information
All of the above
For situations that may be too dangerous, costly or otherwise too difficult to test in the real world, what do computer scientists create in order to draw help discover new knowledge and create new hypothesis related to what situation they are studying?
Classifications
Infographics
Simulations
Regressions
What is the true definition of big data
huge amount of space
large, diverse sets of information
unlimited speed data internet connection
small, undiverse sets of information
Variety is a characteristic of big data. Which type of formats does Variety comes from?
Structured
Unstructured
Semi-structured
All of the above
What does Velocity refers to?
processing speed of a data
speed of an internet
the volume to the city
the speed of network in Servers Racks
Volume is one of the characteristics of big data. What does Volume refers to?
The amount of water in a glass bottle
The amount of money
The amount of love given
The amount of Youtube videos that existed
The amount of data
Big Data is significant because of all of the following except it___________________.
helps us in identifying trends in data
allows us to determine how much data a database can store
can be used to make connections in data
helps us solve questions we might have
What is not the advantage of the HDFS
Scalable
Cost Effective
Flexible
Fix physical location
A ________ serves as the master and there is only one NameNode per cluster.
Data Node
NameNode
Data block
Replication
HDFS works in a __________ fashion
master-worker
master-slave
worker/slave
worker/master
For YARN, the ___________ Manager UI provides host and port information.
Data Node
NameNode
Resource
Replication
....................................is the storage layer, .....................................is the resource management layers and ----------------------------------is data processing layer of hadoop
Map Reduce, YARN and HDFS
HDFS, MapReduce, YARN
YARN, MapReduce and HDFS
HDFS, YARN and MapReduce
Default size of each block in Apache Hadoop 2.x is
64MB
128KB
128GB
128 MB
I have a file “example.txt” of size 514 MB. If default configuration of block size, which is 128 MB. Then, how many blocks will be created?
4
5
2
6
if you are storing a file of 100 MB in HDFS using the default configuration, and if it perform replication management policy then how much of data is residing in different data node
100 MB
200 MB
600 MB
300 MB
5.Who created the hadoop?
dennis ritchie
james goasling
Doug Cutting
Carlo strozzi
8.What are the two components of Hadoop?
Hadoop and YARN
Hadoop and Sqoop
Hadoop and Shuffler
Hadoop and MapReduce
14. The Blocks of a file are replicated for ____________
Bilateral tolerances
Unilateral tolerances
Compound tolerances.
Fault tolerances.
Apache Hadoop YARN stands for:
Yet Another Reserve Negotiator
Yet Another Resource Network
Yet Another Resource Negotiator
Yet Another Resource Manager
What is NOT a characteristic of big data?
Volume
Variety
Vision
Velocity
Spreadsheets and databases are examples of software tools that allow us to process Big Data.
True
False
data stored in databases, in an ordered manner is :
Structured data
semi-Structured data
Unstructured data
used by companies to figure their customer behavior and make the appropriate business decisions and modifications :
machine-generated
Human-generated
all of above
