wayground logo

Free Printable Worksheets

NEW

Font size

S
M
L
XL
Worksheets

The Recap Quiz !

Total questions: 25

Worksheet time: 19mins

Name
Class
Date
1.

During a class discussion, Jerin and Chrisha are talking about where data comes from in their project. Which statement best describes a data source?

a)

Only internal databases

b)

Any system or location where data originates

c)

Only systems generating structured data

d)

Only APIs

2.

John and Chrisha are working on a data integration project. What is the primary difference between ETL and ELT?

a)

ELT requires no transformations

b)

ETL loads data only into cloud systems

c)

ETL transforms before load; ELT transforms after load

d)

ELT is only for batch processing

3.

Jeevan and Rashima are working on a project to move company data from various sources to a central database. Which of the following is NOT a purpose of data ingestion?

a)

Extracting data from sources

b)

Profiling business processes

c)

Delivering data to a target system

d)

Scheduling and orchestrating loads

4.

During a data profiling exercise in Jeevan and Jeshna's project, a high null percentage is detected in a mandatory column. What does this indicate?

a)

Valid scenario

b)

Data quality issue

c)

Metadata enrichment

d)

Business rule compliance

5.

While working on a data quality project, Rashima and Jerlin are using data profiling to assess their dataset. Data profiling primarily supports which DQ dimension?

a)

Accuracy

b)

Completeness

c)

Timeliness

d)

All of the above

6.

At a financial firm where Jerin and Jeshna work, what is a common consequence of poor data quality?

a)

Increased automation

b)

Regulatory fines

c)

Reduced storage costs

d)

Faster processing

7.

Adil and Jerin are working on a project where they need to store and exchange data between different systems. They are considering different file formats. Which is considered a semi-structured source?

a)

SQL table

b)

XML file

c)

PDF report

d)

MP4 file

8.

While working on a data analysis project, Jeshna and Jeevan use column profiling to help identify:

a)

Cross-system lineage

b)

Data precision issues

c)

User access violations

d)

Backup failures

9.

While working on a data project, Rashima and Jeevan noticed that ELT became more popular due to:

a)

Mainframes

b)

Cloud warehouses capable of large-scale transformation

c)

Lack of ETL tools

d)

Less need for governance

10.

In a school database, if Adil and Agnes are both assigned the same student ID as a primary key, which DQ dimension is violated?

a)

Usability

b)

Integrity

c)

Accuracy

d)

Lineage

11.

Jeevan and Chrisha are starting a new data analytics project. Which is a key driver for profiling early in their project?

a)

To speed up data load

b)

To understand actual data behavior before writing rules

c)

To avoid mapping

d)

To reduce metadata

12.

When Adil and Agnes work with data from external vendors for their project, which additional challenge might they face?

a)

You can control their schema fully

b)

Zero latency

c)

No ownership

d)

Volatile formats or unexpected changes

13.

In a school database, Jerin and Jeshna's gender values are recorded as 'M' and 'F'. A transformation is applied to standardize all gender values to 'M' or 'F'. This improves:

a)

Auditability

b)

Completeness

c)

Consistency

d)

Security

14.

Jerin and Agnes are building an AI system for their school project. Impact of poor data in their AI system commonly includes:

a)

Faster model training

b)

Model bias and poor predictions

c)

More accurate outcomes

d)

Unlimited scaling

15.

Rashima and John are managing a source system with rapidly changing schemas. What does their system require?

a)

No monitoring

b)

Strong metadata tracking

c)

Manual patching

d)

No documentation

16.

John and Chrisha are working on an ETL project for their class. Which step in ETL is responsible for handling schema alignment?

a)

Extract

b)

Transform

c)

Load

d)

Archive

17.

During a group analytics project, Jerlin and Agnes detected data quality defects late in their analysis. This situation most likely leads to:

a)

Faster dashboards

b)

Expensive rework

c)

Safer decisions

d)

Better latency

18.

Adil and Agnes are working on a data validation project and need to use profiling repeated patterns (regex-style). Which task are they most likely supporting?

a)

Security classification

b)

Pattern-based validation rule creation

c)

Pipeline scheduling

d)

Log backup

19.

Rashima and Chrisha are discussing their company's database systems. Which is true about source systems?

a)

Capture data for operational needs

b)

Designed for analytical consumption only

c)

Always have perfect data

d)

Cannot be modified

20.

While working on an ETL project, Agnes and John noticed several failures. These failures often occur because:

a)

Too much metadata

b)

Poor understanding of source data

c)

Lack of cloud systems

d)

No SQL usage

21.

In a university database, Jerin is entering student enrollment records. A missing foreign key link in the enrollment table reflects a failure in:

a)

Completeness

b)

Integrity

c)

Timeliness

d)

Security

22.

John and Chrisha are working on a new business analytics project. The first step they should take before designing business rules is:

a)

Load data into dashboards

b)

Profile data to understand behavior

c)

Create reports

d)

Define KPIs

23.

Chrisha and Jeevan are working on a project where they regularly analyze data to ensure its quality. Why is profiling a continuous activity?

a)

Data never changes

b)

Systems remain static

c)

Data evolves and new anomalies appear

d)

It replaces monitoring

24.

During a data migration project, Jeevan and Chrisha noticed that a poorly designed extract step may cause:

a)

Loss of lineage

b)

Transform errors

c)

Latency increases

d)

All of the above

25.

In a project managed by Rashima and Jeshna, which is the biggest hidden cost of poor data?

a)

Storage

b)

Manual workarounds

c)

ETL licensing

d)

Cloud compute