wayground logo

Free Printable Worksheets

Font size

S
M
L
XL
Worksheets

MCSE Quiz 1

Total questions: 36

Worksheet time: 19mins

Name
Class
Date
1.

Which of the following is the primary goal of descriptive statistics?

a)

To test a hypothesis about a population parameter.

b)

To make predictions about future outcomes based on sample data.

c)

To organize, summarize, and present the main features of a dataset.

d)

To determine the causal relationship between two variables.

2.

What type of data is the number of active users on a website?

a)

Continuous

b)

Categorical

c)

Discrete

d)

Ordinal

3.

Which variable type best describes 'operating system type'?

a)

Quantitative

b)

Qualitative

c)

Continuous

d)

Interval

4.

Which of the following is an example of qualitative data?

a)

Height

b)

Age

c)

Gender

d)

Weight

5.

Which sampling method ensures representation from all strata of a population?

a)

Simple random sampling

b)

Cluster sampling

c)

Stratified sampling

d)

Systematic sampling

6.

Which of the following is NOT a source of sampling error?

a)

Random variation

b)

Non-response bias

c)

Sampling frame errors

d)

Increasing sample size

7.

Which of the following are types of statistics? (Select all that apply)

a)

Descriptive

b)

Inferential

c)

Analytical

d)

Computational

8.

An engineer calculates the average tensile strength of 50 steel rods from a production batch. This single calculated average value is an example of:

a)

Inferential statistics

b)

A population parameter

c)

A descriptive statistic

d)

Hypothesis testing

9.

A market analyst uses poll data from a sample of 1,200 voters to predict the outcome of a national election. This is an application of:

a)

Descriptive statistics

b)

Inferential statistics

c)

Data visualization

d)

Population enumeration

10.

A quality control engineer at a semiconductor manufacturing plant collects data on the defect rate of 1,000 microchips. Which of the following actions fall under the domain of descriptive statistics? (Select all that apply)

a)

Calculating the mean and standard deviation of the defect rate for the 1,000-chip sample.

b)

Constructing a histogram to visualize the distribution of defects in the sample.

c)

Using the sample defect rate to create a 95% confidence interval for the defect rate of the entire production run.

d)

Concluding that a change in the manufacturing process has caused a statistically significant reduction in defects based on the sample.

11.

A data scientist is analyzing user engagement on a new mobile application. The workflow involves collecting data from a sample of 5,000 users. Which of the following sequences represents a logically correct statistical workflow?

a)

First, use a t-test to infer if the average engagement time is greater than 10 minutes; then, calculate the sample mean and variance.

b)

First, create a bar chart of user activity types and calculate the median session duration; then, use this sample data to generalize about the behavior of all application users.

c)

First, make a prediction about the churn rate for the next quarter; then, collect the sample data to verify the prediction.

d)

First, use the sample to draw conclusions about the entire user population; then, present the sample data in tables and charts.

12.

A car manufacturer classifies its vehicles by color (e.g., Red, Blue, Black, White). What level of measurement does this data represent?

a)

Nominal

b)

Ordinal

c)

Interval

d)

Ratio

13.

A survey asks respondents to rate their satisfaction with a software product on a scale of 'Very Dissatisfied,' 'Dissatisfied,' 'Neutral,' 'Satisfied,' 'Very Satisfied.' This is an example of which type of data?

a)

Nominal

b)

Ordinal

c)

Interval

d)

Ratio

14.

The temperature in degrees Celsius is measured by a sensor. This data is an example of which measurement scale, and why?

a)

Nominal, because it categorizes temperature.

b)

Ordinal, because 20°C is warmer than 10°C.

c)

Interval, because the difference between 10°C and 20°C is the same as between 20°C and 30°C, but there is no true zero.

d)

Ratio, because 20°C is twice as hot as 10°C.

15.

An engineer measures the weight of a component in kilograms. This measurement is on a ratio scale because:

a)

It is a numerical value.

b)

The values can be ordered from lightest to heaviest.

c)

The difference between 10 kg and 20 kg is meaningful.

d)

It has a true zero point, where 0 kg means no weight, and ratios like '20 kg is twice as heavy as 10 kg' are meaningful.

16.

An engineering team is analyzing performance data from a set of electric motors. They have collected the following variables: (1) Motor model name, (2) Operating temperature in Kelvin, (3) Performance rating ('Low', 'Medium', 'High'), and (4) Year of manufacture. Which of the following statistical operations are valid for the given data types? (Select all that apply)

a)

Calculating the average (mean) of the motor model names.

b)

Calculating the ratio of the operating temperatures of two motors (e.g., Motor A is 1.5 times hotter than Motor B).

c)

Calculating the median of the performance ratings.

d)

Calculating the mean of the years of manufacture.

17.

In a(n) ______________, researchers manipulate one or more variables to observe the effect on other variables, while in a(n) ______________, researchers measure characteristics without attempting to influence them.

a)

observational study; controlled experiment

b)

controlled experiment; observational study

c)

survey; census

d)

case-control study; cohort study

18.

A research team wants to determine if a new fertilizer causes an increase in crop yield. They divide a field into 100 plots, randomly assigning 50 plots to receive the new fertilizer and 50 plots to receive a standard fertilizer. They then compare the yields. What type of study is this?

a)

An observational cohort study

b)

A randomized controlled experiment

c)

A case-control study

d)

A descriptive survey

19.

An epidemiologist studies the medical records of 1,000 individuals, 500 of whom have a rare disease (cases) and 500 who do not (controls), to identify past exposures that may be linked to the disease. This is an example of:

a)

A randomized controlled trial

b)

A prospective cohort study

c)

A case-control observational study

d)

A cross-sectional study

20.

A study finds a strong positive association between the number of hours engineers spend using a new CAD software and their project completion speed. Which of the following conclusions are validly supported by this information alone? (Select all that apply)

a)

Using the new CAD software causes engineers to complete projects faster.

b)

There is a correlation between the hours spent on the new CAD software and project completion speed.

c)

It is possible that faster engineers are more likely to adopt and use new software, and the software itself is not the cause of the speed increase.

d)

The study must have been a randomized controlled experiment.

21.

A pharmaceutical company wants to test the effectiveness of a new drug for reducing blood pressure. Which of the following study designs would provide the strongest evidence for a causal relationship between the drug and a reduction in blood pressure?

a)

An observational study tracking 5,000 patients who are already taking the drug and comparing their outcomes to the general population.

b)

A case-control study identifying patients with low blood pressure and looking at their past medication history.

c)

A randomized controlled trial where patients are randomly assigned to receive either the new drug or a placebo, with neither the patients nor the researchers knowing who received which.

d)

A survey asking doctors for their opinions on the effectiveness of the new drug.

22.

A researcher wants to survey 100 employees from a company of 1,000. They assign each employee a number from 1 to 1,000 and use a random number generator to select 100 unique numbers. This technique is known as:

a)

Stratified sampling

b)

Systematic sampling

c)

Simple random sampling

d)

Cluster sampling

23.

To ensure a sample of university students accurately reflects the gender distribution of the university (60% female, 40% male), a researcher first divides the student population by gender and then draws random samples from each group in the correct proportion. This method is called:

a)

Simple random sampling

b)

Cluster sampling

c)

Convenience sampling

d)

Stratified sampling

24.

A quality inspector needs to check every 50th item coming off a production line. The first item is chosen randomly from the first 50 items. This is an example of:

a)

Systematic sampling

b)

Quota sampling

c)

Cluster sampling

d)

Simple random sampling

25.

To study the performance of schools in a large state, a researcher randomly selects 20 school districts from the state and then collects data from every school within those selected districts. This sampling technique is:

a)

Stratified sampling

b)

Simple random sampling

c)

Cluster sampling

d)

Systematic sampling

26.

An engineer is tasked with assessing the overall quality of a large shipment of 100,000 resistors, which are packaged in 1,000 boxes of 100 resistors each. The engineer has limited time and budget. Which of the following sampling strategies represent a valid trade-off between logistical feasibility and statistical principles? (Select all that apply)

a)

Select the first 10 boxes from the top of the shipment and test all resistors inside them, as this is the quickest method.

b)

Randomly select 20 boxes (clusters) and then test every resistor within those 20 boxes.

c)

Assign a number to every single one of the 100,000 resistors and draw 2,000 unique random numbers to identify which resistors to test.

d)

Test all resistors from the 5 boxes located closest to the loading dock door.

27.

The difference between a sample statistic (e.g., sample mean) and the true population parameter, which occurs purely due to chance because a sample is not the entire population, is known as:

a)

Non-response error

b)

Measurement error

c)

Sampling error

d)

Coverage error

28.

A political pollster conducts a survey by calling landline telephones. This method may introduce which type of error because it excludes people who only use mobile phones or have no phone at all?

a)

Random sampling error

b)

Coverage error (or sample frame error)

c)

Non-response error

d)

Measurement error

29.

A web-based survey on internet usage habits is advertised on a tech news website. The results are likely to be skewed because only people who are interested in technology and visit that specific site will participate. This is a classic example of:

a)

Selection error (or self-selection bias)

b)

Simple random error

c)

Systematic sampling error

d)

Cluster sampling error

30.

A research firm conducts a large-scale survey with 50,000 respondents to predict consumer preference for a new product. The survey is administered only via an online platform, and the link is promoted on social media. Despite the very large sample size leading to a small margin of error, the prediction turns out to be highly inaccurate. Which of the following are plausible explanations for this failure? (Select all that apply)

a)

The random sampling error was too large due to the sample size.

b)

The study suffered from coverage error, as it excluded individuals who are not active on social media or lack internet access.

c)

The large sample size magnified the effect of a systematic bias in the sampling method.

d)

The study suffered from selection bias, as individuals with a strong interest in the product category were more likely to participate.

31.

Which of the following scenarios describes a non-sampling error? (Select all that apply)

a)

A researcher uses a well-calibrated instrument, but by chance, the 20 components selected for a sample have a slightly higher average weight than the entire population of components.

b)

A survey question is worded ambiguously, leading respondents to interpret it in different ways and provide inaccurate answers.

c)

A data entry clerk accidentally transposes digits while recording measurements from a lab notebook into a spreadsheet.

d)

In a telephone survey, a significant number of individuals selected for the sample refuse to participate.

32.

A random variable is:

a)

A constant value

b)

A function assigning numerical values to outcomes

c)

Always an integer

d)

A fixed percentage

33.

A dataset of house prices has a few multi-million dollar mansions, while the rest of the prices are clustered in a much lower range. Which measure of central tendency would be the most misleading or inflated representation of a "typical" house price?

a)

Median

b)

Mode

c)

Mean

d)

Interquartile Range (IQR)

34.

What does the standard deviation of a dataset measure?

a)

The total sum of all data points.

b)

The average distance of each data point from the median.

c)

The average spread or dispersion of data points around the mean.

d)

The difference between the maximum and minimum values.

35.

A dataset has a first quartile (Q1) of 50 and a third quartile (Q3) of 80. According to the common 1.5×IQR rule for outlier detection, which of the following data points would be considered an outlier?

a)

120

b)

95

c)

20

d)

130

36.

If a student's exam score is at the 85th percentile, what does this mean?

a)

The student scored better than or equal to 85% of the other test-takers.

b)

The student answered 85% of the questions correctly.

c)

Only 15% of students scored higher than this student.

d)

The student's score is 85% higher than the average score.