wayground logo

Free Printable Worksheets

NEW

Font size

S
M
L
XL
Worksheets

Baseline Test Data science TE B

Total questions: 25

Worksheet time: 13mins

Name
Class
Date
1.
What is Data Science primarily concerned with?
a)
Writing software
b)
Extracting insights from data
c)
Managing hardware
d)
Designing networks
2.
Which of the following best describes data?
a)
Final reports
b)
Raw facts and figures
c)
Cleaned information
d)
Visual summaries
3.
Data organized in rows and columns is called
a)
Unstructured data
b)
Semi-structured data
c)
Structured data
d)
Raw data
4.
Which is an example of unstructured data?
a)
CSV file
b)
Excel sheet
c)
Social media text
d)
Database table
5.
Which programming language is most popular in Data Science?
a)
C
b)
Java
c)
Python
d)
Assembly
6.
Which Python library is used for numerical computation?
a)
Pandas
b)
NumPy
c)
Seaborn
d)
TensorFlow
7.
Which library is mainly used for data manipulation?
a)
Matplotlib
b)
NumPy
c)
Pandas
d)
Scikit-learn
8.
What is data preprocessing?
a)
Data storage
b)
Data cleaning and preparation
c)
Model training
d)
Result visualization
9.
Which of the following is NOT a type of data?
a)
Structured
b)
Unstructured
c)
Semi-structured
d)
Compiled
10.
What does EDA stand for?
a)
Efficient Data Analysis
b)
Exploratory Data Analysis
c)
External Data Access
d)
Extracted Data Algorithm
11.
What is the purpose of Exploratory Data Analysis?
a)
Delete data
b)
Understand data patterns
c)
Train models
d)
Deploy models
12.
Which statistical measure represents average?
a)
Mode
b)
Median
c)
Mean
d)
Range
13.
Which visualization is best to show trends over time?
a)
Pie chart
b)
Bar chart
c)
Line chart
d)
Histogram
14.
What does CSV stand for?
a)
Common Separated Values
b)
Comma Separated Values
c)
Column Stored Variables
d)
Common Statistical Vector
15.
Which type of machine learning uses labeled data?
a)
Unsupervised learning
b)
Reinforcement learning
c)
Supervised learning
d)
Deep learning
16.
Which technique is used to predict continuous values?
a)
Classification
b)
Clustering
c)
Regression
d)
Association
17.
Which algorithm is used for clustering?
a)
Decision Tree
b)
Linear Regression
c)
K-Means
d)
Naive Bayes
18.
Big Data is typically defined by
a)
Accuracy and security
b)
Volume, Velocity, Variety
c)
Size and speed
d)
Tables and files
19.
Which Python library is commonly used for data visualization?
a)
NumPy
b)
Pandas
c)
Matplotlib
d)
Scikit-learn
20.
What is the main goal of data visualization?
a)
Store data
b)
Encrypt data
c)
Communicate insights
d)
Remove noise
21.
Which of the following is categorical data?
a)
Height
b)
Weight
c)
Age
d)
Gender
22.
A dataset is best described as
a)
Single data value
b)
Collection of related data
c)
Only numeric data
d)
Only textual data
23.
Which tool is commonly used for interactive data analysis?
a)
MS Word
b)
Notepad
c)
Jupyter Notebook
d)
File Explorer
24.
Which step follows data collection in Data Science?
a)
Deployment
b)
Data preprocessing
c)
Decision making
d)
Reporting
25.
Which of the following is an application of Data Science?
a)
Recommendation systems
b)
Fraud detection
c)
Weather forecasting
d)
All of the above