wayground logo

Free Printable Worksheets

Font size

S
M
L
XL
Worksheets

Q1 Big Data Engineer - M1 to M4

Total questions: 25

Worksheet time: 25mins

Name
Class
Date
1.

Logo of Hadoop

(a)  

2.

Examples of Big Data

a)

Metallurgy

b)

Biology

c)

Biochemical

d)

Biogeochemical

3.

What is 5th V

(a)  

4.

Web Crawling is also known as

(a)  

5.

Variety of Big Data includes

a)

Structured Data

b)

Unstructured Data

c)

Symmetric Data

d)

Structured and Unstructured Data

6.

Velocity means

a)

distance

b)

acceleration

c)

speed

d)

force

7.

Uncertainty of data

a)

Volume

b)

Velocity

c)

Veracity

d)

Variety

8.

What is defined as that data becomes so large that it cannot be processed using conventional methods

a)

Data Analytics

b)

Data Science

c)

Big Data

d)

Big Quality

9.

Which of the following is not the type of Big Data?

a)

Structured data

b)

Unstructured data

c)

Semi-structured data

d)

Semi-unstructured data

10.

which of the following is not a common use cases applied to Big Data?

a)

Graph creation and Analysis

b)

Predictive Models

c)

Risk Assessments

d)

Risk Management

11.

What are 4 E's ?

(a)  

12.

A growing number of businesses and industries are finding innovative ways to apply ________ to a variety of use-case scenarios because it affords a unique perspective on the analysis of networked entities and their relationships

a)

graph analytics

b)

tiger analytics

c)

big data analytics

d)

all the three

13.

What is self-configuring and adaptive system consisting of networks of sensors and smart objects whose purpose is to interconnect "all" things, including every day and industrial objects, in such a way as to, make them intelligent, programmable, and more capable of interacting with humans

a)

Digital Senses

b)

Internet of Things

c)

Big Data

d)

Data Analytics

14.

What is HDP?

(a)  

15.

Which type of database is optimized for read-heavy workloads and is suitable for analytical queries in Big Data applications?

a)

OLTP

b)

OLAP

c)

NoSQL

d)

RDBMS

16.

Which component of the Hadoop ecosystem is used for real-time data stream processing?

a)

HDFS

b)

HIVE

c)

SPARK

d)

PIG

17.

What is the primary role of Apache Kafka in the Big Data ecosystem?

a)

Real-time data streaming and message queuing

b)

Data storage and retrieval

c)

Data cleansing and transformation

d)

Data visualization and reporting

18.

In Big Data processing, what does "ETL" stand for?

a)

Extract, Transform, Load

b)

Extract, Test, Log

c)

Except, Test, Load

d)

Extract, Transform, Log

19.

What is the main function of the MapReduce framework in Hadoop?

a)

Data storage

b)

Data analysis

c)

Data cleaning

d)

Data visualization

20.

What is the primary purpose of Hadoop's HDFS (Hadoop Distributed File System)?

a)

Real-time data processing

b)

Data storage and management

c)

Data visualization

d)

Data cleaning

21.

Which programming language is commonly used for data processing in Big Data applications?

a)

Python

b)

Javascript

c)

Java

d)

Both Java and Javascript

22.

What is the term for the process of extracting valuable information from large datasets?

a)

Data Collection

b)

Data Cleaning

c)

Data Analytics

d)

Data Mining

23.

Which of the following is a key characteristic of Big Data?

a)

Low volume

b)

Structured data only

c)

High velocity

d)

Homogeneous data

24.

Which of the following is a key challenge in Big Data processing, related to ensuring data accuracy and consistency?

a)

Data replication

b)

Data governance

c)

Data Sampling

d)

Data Modelling

25.

What does the term "Batch Processing" refer to in the context of Big Data?

a)

Processing data in real-time as it arrives

b)

Processing data in small, incremental batches

c)

Processing data in large, non-real-time batches

d)

Processing data using a NoSQL database