NEW
Font size
WorksheetsBIG DATA
Total questions: 25
Worksheet time: 16mins
Data in ___________ bytes size is called Big Data.
Tera
Giga
Peta
Meta
All of the following accurately describe Hadoop, EXCEPT ____________
Open-source
Distributed computing approach
Java-based
Real-time
__________ has the world’s largest Hadoop cluster.
Apple
Datamatics
None of the above
Facebook Tackles Big Data With _______ based on Hadoop.
‘Project Prism’
‘Prism’
'Project Big’
‘Project Data’
What are the main components of big data?
YARN
Map Reduce
HDFS
All of these
What was Handloop written in?
Java(programming language)
Java(software platform)
Perl
Lua(programming language)
Which of the following genres does Handloop produce?
Distributed file system
Java message service
JAX-RS
Relational database management system
Which of the following platform does Handloop run?
Debain
Bare metal
Cross-platform
Unix-like
The data node and name node in HADOOP are
Both Worker Nodes
Worker Node and Master Node respectively
Master Node and Worker Node respectively
Both Master Nodes
Point out the wrong statement
Non-Relational databases require that schemas be defined before you can add data.
NoSQL databases are built to allow the insertion of data without a predefined schema.
All of the options.
NewSQL databases are built to allow the insertion of data without a predefined schema.
Hadoop (a big data tool) works with number of related tools. Choose from the following, the common tools included into Hadoop:
MySQl, Google API and Map reduce
Map reduce, Scala and hummer
Map reduce, H base and Hive
Map reduce, hummer and Heron
Big Data is generally characterised by three Vs that stand for ______, _____ and ______.
Volume ; Viscosity ; Variety
Variety ; Velocity ; Vivid
Volume ; Variety ; Velocity
Viscosity ; Volume ; Velocity
Which company developed Apache Kafka?
Microsoft
Amazon
In which year Apache Kafka was developed ?
2010
2011
2012
2013
________ is general-purpose computing model and runtime system for distributed data analytics.
MapReduce
Oozie
Drill
All of the above
The unit of data that flows through a Flume agent is ________
Low
Event
Drill
All of the above
_______ splits the gap between structured and unstructured data, which, using the right datasets, can make it a huge asset.
Structured Data
Unstructured Data
Semi-structured Data
None of the above
Big Data Analysis performs the following except __________
Collects data
Analyzes data
Spreads data
None of the above
Which among the following is not a Big Data technology ____________.
Data Lakes
Hadoop Ecosystem
R Programming
None of the above
Do Hadoop need specialized hardware to process the data?
Yes
No
May be
Can't say
Application of Big Data is _______
Education
Banking and Securities
Insurance
All of the above
Is Qubole a Big data tool ?
No
Yes
May be
Can't say
Due to Big Data, new, innovative, and cost-effective technologies are constantly emerging and improving.
True
False
None of the above
Can't say
______ is a data processing framework in Big Data that can quickly perform processing tasks on very large data sets.
Apache Cassandra
Apache Lite
Apache Spark & Apache Cassandra
Apache Spark
Which of the following statement/s is/are true?
(i) Facebook has the world’s largest Hadoop cluster.
(ii) Hadoop 2.0 allows live stream processing of real time data
Both (i) and (ii)
Neither (i) nor (ii)
(i) only
(ii) only
