Wayground logo

Free Printable Worksheets

Font size

S
M
L
XL
Worksheets

Page 1

Total questions: 20

Worksheet time: 10mins

Name
Class
Date
1.

For sample standard deviation, why is N−1 used in the denominator instead of N?

a)

Ensures mean equals median

b)

Eliminates outliers automatically

c)

Gives closer real-world estimate

d)

Simplifies computation steps

2.

Why might N−1 be used instead of N when computing kurtosis numerators?

a)

To match Pearson’s skew index

b)

Bias correction in finite samples

c)

To increase numerical stability

d)

Because σ is unknown always

3.

Given deviations |x − μ| are summed and divided by N, what statistic is obtained?

a)

Mean absolute deviation

b)

Interquartile range

c)

Standard deviation

d)

Coefficient of variation

4.

What is the primary purpose of CV in data analysis?

a)

Transform skewed distributions to normal

b)

Detect outliers in a single dataset

c)

Compare datasets with different units

d)

Estimate the population mean accurately

5.

A Q–Q plot primarily compares what?

a)

Sample means to population means

b)

Dataset medians to interquartile ranges

c)

Sample quantiles to theoretical quantiles

d)

Observed deviations to absolute deviations

6.

When points deviate strongly from the Q–Q plot’s reference line, what is the most appropriate next step?

a)

Conduct careful analysis before interpretation

b)

Immediately assume data are uniformly distributed

c)

Discard outliers and proceed with conclusions

d)

Transform data without checking assumptions

7.

If a scatter plot were drawn from the table, which line type would most likely fit the trend?

a)

A downward-sloping linear fit

b)

A horizontal line through the mean

c)

A sinusoidal curve with multiple peaks

d)

An upward-sloping linear fit

8.

Which statement best describes the primary purpose of a scatter plot in data analysis?

a)

Display bivariate relationships with two variables

b)

Summarize categorical counts by one variable

c)

Show ranked ordering of discrete categories

d)

Visualize hierarchical groupings across levels

9.

Before calculating a correlation coefficient, which plot is useful for exploratory checks of anomalies?

a)

Box plot of grouped categories

b)

Pie chart of categorical segments

c)

Histogram of a single variable

d)

Scatter plot of the two variables

10.

In a system where A is an m×n matrix and y is an n-dimensional vector, how is the term ‘augmentation matrix’ formed during Gaussian elimination?

a)

By stacking A and y rowwise

b)

By multiplying A and y elementwise

c)

By appending y to A columnwise

d)

By appending x to y columnwise

11.

When a linear system has exactly one solution, how is the system classified?

a)

Consistent dependent system

b)

Consistent independent system

c)

Overdetermined inconsistent system

d)

Indeterminate contradictory system

12.

What is the goal of reducing a matrix to reduced echelon form in Gaussian elimination?

a)

To maximize sparsity of A

b)

To compute eigenvectors quickly

c)

To verify orthogonality of rows

d)

To isolate pivot positions for solving

13.

After reaching reduced echelon form, which method retrieves remaining unknowns starting from xn?

a)

LU factorization with partial pivoting

b)

Back-substitution from xn downward

c)

Iterative gradient descent on x

d)

Forward substitution from x1 upward

14.

What does the Probability Density Function (PDF) represent for a continuous variable?

a)

Cumulative probability up to any value

b)

Exact probability at a single point

c)

Frequency counts of discrete events

d)

Shape of distribution via density values

15.

Which function computes the probability of observing a value less than or equal to x?

a)

Likelihood Function (LF)

b)

Probability Mass Function (PMF)

c)

Probability Density Function (PDF)

d)

Cumulative Distribution Function (CDF)

16.

Which statement about machine learning datasets and distributions is accurate?

a)

They do not require sampling theory at all

b)

They may be generated by multiple distributions

c)

They always follow one perfect distribution

d)

They never involve random variables

17.

For the Poisson distribution with parameter λ, what is the standard deviation?

a)

√λ

b)

λ

c)

1/√λ

d)

√(λ/2)

18.

In the course registration table, what is the total number of students?

a)

90

b)

80

c)

100

d)

110

19.

Which conclusion is justified if the computed χ² is below the critical value for df = 1?

a)

Accept significant difference in registration rates

b)

Reject independence between gender and registration

c)

Conclude boys register more than girls definitively

d)

Fail to reject the null hypothesis of independence

20.

Why does PCA lead to a reduced dimension representation?

a)

It removes outliers using robust loss functions

b)

It binarizes continuous features for simplicity

c)

It exploits information redundancies across measurements

d)

It balances class distributions via resampling