WorksheetsModule 1 Overview
Total questions: 10
Worksheet time: 5mins
Identify the feature that distinguishes semi-structured data from traditional relational data models:
A strict separation between data and schema is enforced
It conforms to normalized relational schemas across all entities
Tags are used to segregate semantic elements and enforce hierarchical records
All entities in the same class must share an identical set of attributes
Refer to the XML example of a bookstore. Which element is the root of the document?
According to the Module 1 outline, which area contrasts legacy approaches with modern data paradigms?
Traditional Business Intelligence vs Big Data
Relational Databases vs Spreadsheet Tools
Batch Processing vs Stream Processing
Cloud vs On‑premises Networking
According to the XML example, what is the price of the book with
9.99
10.99
12.99
19.25
Which representation uses angle-bracket tags to define elements and attributes: the provided XML or the provided JSON?
The provided XML
The provided JSON
Both representations
Neither representation
You are building a system to find patterns in customer reviews and chat logs. Which technique list aligns with methods introduced for analyzing unstructured data?
Association rule mining, regression analysis, collaborative filtering; text analytics; NLP; noisy text analysis; manual tagging with metadata; part-of-speech tagging; UIMA platform
Sorting algorithms, B-tree indexing, ACID transaction logging; star schema design; OLAP cube generation
Fourier transforms only, Kalman filtering only, and PID control tuning
Primary key selection, foreign key constraints, and third normal form normalization
Which characteristic of data focuses on the structure, including source, granularity, types, and whether the data is static or real-time streaming?
Composition
Condition
Context
Correlation
In the provided timeline figure, which era is associated with being data driven and handling structured, unstructured, and multimedia data?
1970s and before
1980s and 1990s
2000s and beyond
Pre-1960s
Which sequence best reflects the progression outlined: from early storage to modern data landscapes?
Cloud-native systems → Mainframes → Relational databases
Mainframes with basic storage → Relational databases enabling data-intensive applications → Data-driven era with structured, unstructured, and multimedia data
Relational databases → Mainframes → WWW and IoT
IoT proliferation → Mainframes → Data centers
Based on the Hadoop environment figure, which targets are shown receiving outputs from Hadoop?
Only OLAP cubes and dashboards
HDFS, Operational systems, Data warehouse, Data marts, ODS
Legacy systems and third‑party apps only
Mobile apps and content delivery networks
