wayground logo

Free Printable Worksheets

Font size

S
M
L
XL
Worksheets

AP Stats Exam Review Units 1 – 4

Total questions: 87

Worksheet time: 44mins

Name
Class
Date
1.

A(n) ________________ is the entire group of individuals we want information about, while a(n) ________________ is a subset of the population examined.

a)

population, sample

b)

sample, population

c)

group, subset

d)

individual, group

2.

A(n) ______________ variable places an individual into one of several groups or categories.

a)

categorical

b)

continuous

c)

dependent

d)

quantitative

3.

A(n) ______________ variable takes numerical values for which it makes sense to find an average.

a)

quantitative

b)

categorical

c)

nominal

d)

ordinal

4.

Data from categorical variables are displayed with ______________ or ______________ charts.

a)

bar, pie

b)

line, scatter

c)

histogram, box

d)

area, bubble

5.

Data from quantitative variables are displayed with ________________, ________ plots, or ________ plots.

a)

histograms, dot, stem

b)

bar, pie, scatter

c)

box, line, area

d)

frequency, pie, bar

6.

When describing a quantitative distribution, you should always include ______________, ______________, ______________, ______________, and ______________.

a)

shape, center, spread, odd features, context

b)

mean, median, mode, range, variance

c)

center, spread, frequency, probability, context

d)

shape, median, mode, outliers, context

7.

Choose the correct comparison of the mean and median for each shape given.

a)

Skewed Left: Mean < Median;
Symmetric: Mean = Median;
Skewed right, Mean > Median

b)

Skewed Left: Mean > Median; Symmetric: Mean < Median;
Skewed right: Mean = Median

c)

Skewed Left: Mean = Median;
Symmetric: Mean > Median; Skewed right: Mean < Median

8.

A(n) ____________________ is an individual value that falls outside the overall pattern of the distribution.

a)

outlier

b)

median

c)

mode

d)

mean

9.

The ________________ is the average value, calculated by summing all values and dividing by the count. It is ________________ to outliers.

a)

mean, resistant

b)

median, resistant

c)

mean, nonresistant

d)

range, unaffected

10.

The ____________________ is the midpoint of the distribution, with half the observations smaller and half larger. It is _______________ to outliers.

a)

median; resistant

b)

mean; nonresistant

c)

median; nonresistant

d)

range; nonresistant

11.

The ____________________ describes the variation of the data set. It is the average distance from the mean. It is ________________ to outliers.

a)

standard deviation, resistant

b)

mean, nonresistant

c)

median, resistant

d)

standard deviation, nonresistant

12.

The Interquartile Range (________) is the distance between the first quartile (________) and the third quartile (________). It is ____________________ to outliers.

a)

IQR, Q1, Q3, resistant

b)

IQR, Q2, Q4, sensitive

c)

IQR, Q1, Q2, affected

d)

IQR, Q2, Q3, vulnerable

13.

The ____________________ summarizes a distribution using the Minimum, ________, Median, ________, and Maximum. These values create a boxplot.

a)

five-number summary, Q1, Q3

b)

mean summary, Q2, Q4

c)

quartile summary, Q1, Q2

d)

statistical summary, Q2, Q3

14.

A rule for identifying outliers is any value falling outside the interval ________________ to ________________.

a)

Q1 - 1.5(IQR), Q3 + 1.5(IQR)

b)

Q1 - 2(IQR), Q3 + 2(IQR)

c)

Q1 - 1(IQR), Q3 + 1(IQR)

d)

Q1 - 0.5(IQR), Q3 + 0.5(IQR)

15.

A ________________ measures how many standard deviations a value is from the mean.

a)

z-score

b)

mean

c)

variance

d)

median

16.

Adding a constant, c, to every observation __________ the measures of center (mean, median, quartiles) by c but ________________ the measures of spread (range, IQR, standard deviation).

a)

changes; does not change

b)

decreases; increases

c)

does not change; increases

d)

increases; decreases

17.

Multiplying every observation by a positive constant, b, ________________ both the measures of center and measures of spread by b.

a)

multiplies

b)

divides

c)

subtracts

d)

adds

18.

The empirical rule states that _______% of data is within 1 standard deviation of the mean in a normal distribution, _______% is within 2 standard deviations, and _______% is within 3 standard deviations.

a)

68; 95; 99.7

b)

70; 90; 99

c)

65; 92; 99.9

d)

60; 85; 98

19.

To find the percent of data lying within set boundaries in a normal curve, use ________________________ on the calculator.

a)

normalcdf

b)

normalpdf

c)

invNorm

d)

stdDev

20.

To find the value that lies at a certain percentage in a normal curve, use ________________________ on the calculator.

a)

invNorm

b)

normalcdf

c)

stdDev

d)

mean

21.

A ________________ displays the relationship between two ________________ variables.

a)

scatterplot; quantitative

b)

bar graph; categorical

c)

pie chart; qualitative

d)

histogram; discrete

22.

When describing a scatterplot, you should discuss ____________, ________________, ________________, and ________________.

a)

direction; shape; strength; outliers

b)

mean; median; mode; range

c)

slope; intercept; residuals; correlation

d)

variance; standard deviation; skewness; kurtosis

23.

The ________ variable is plotted on the horizontal (x) axis, and the ________ variable is plotted on the vertical (y) axis.

a)

explanatory; response

b)

response; explanatory

c)

dependent; independent

d)

independent; dependent

24.

The ________________, denoted by r, measures the ____________ and ____________ of the linear relationship between two quantitative variables.

a)

correlation coefficient; direction; strength

b)

regression; magnitude; frequency

c)

variance; spread; central tendency

d)

mean; association; variability

25.

The value of r must be between ________ and ________. A value close to 0 indicates a ________ linear relationship.

a)

-1; 1; weak

b)

0; 1; strong

c)

-1; 0; strong

d)

-1; 1; strong

26.

____________________ does not imply ____________________.

a)

Correlation; causation

b)

Causation; correlation

c)

Observation; inference

d)

Prediction; certainty

27.

The ________________ regression line is the line that minimizes the sum of the ________________ of the ________________ (vertical distances) from the data points to the line.

a)

least-squares; squares; residuals

b)

maximum-likelihood; cubes; errors

c)

mean; absolute values; deviations

d)

linear; products; differences

28.

The general form of the least-squares regression line is ________________, where ŷ is the ________________ value of the response variable.

a)

ŷ = a + bx; predicted

b)

ŷ = bx + a; observed

c)

ŷ = a - bx; calculated

d)

ŷ = bx - a; measured

29.

The ________ of the line represents the ____________ change in the ________ variable for every one-unit increase in the ____________ variable.

a)

slope; predicted; response; explanatory

b)

intercept; actual; explanatory; response

c)

slope; actual; explanatory; response

d)

intercept; predicted; response; explanatory

30.

The ________ is the ________________ value of the ________ variable when the ________ variable is zero.

a)

y-intercept; predicted; response; explanatory

b)

slope; observed; explanatory; response

c)

mean; actual; response; predictor

d)

coefficient; estimated; predictor; response

31.

In a computer output, the values needed for an LSRL are in the __________ column of numbers. The slope of the line is _______ and the y-intercept is _______.

a)

first; b; a

b)

first; a; b

c)

Regression; m; c

d)

Output; x; y

32.

______________ is the use of a regression line to make ________ for x values that are far outside the range of the ____________ data. These predictions are often _____________.

a)

Extrapolation; predictions; observed; unreliable

b)

Interpolation; estimations; measured; reliable

c)

Extrapolation; estimations; predicted; accurate

d)

Interpolation; predictions; observed; reliable

33.

A ________ is the difference between an observed value and the value predicted by the regression line.

a)

residual

b)

outlier

c)

mean

d)

variance

34.

A(n) _______ plot is a scatterplot of the _______ against the _______ variable.

a)

residual; residuals; explanatory

b)

scatter; explanatory; residuals

c)

line; residuals; response

d)

trend; response; explanatory

35.

A good fit for a linear model is indicated by a residual plot that shows ________ pattern.

a)

no obvious

b)

linear

c)

curved

d)

cyclical

36.

The ________________, denoted by r² is the percent of the variation in the ________ variable that is accounted for by the ____________ with the ________ variable.

a)

coefficient of determination; response; linear relationship; explanatory

b)

correlation coefficient; explanatory; nonlinear relationship; response

c)

regression coefficient; response; quadratic relationship; explanatory

d)

standard deviation; explanatory; linear relationship; response

37.

A(n) ________________ is an observation that lies outside the overall pattern of the other observations.

a)

outlier

b)

median

c)

mode

d)

range

38.

A(n) _____________ point is an observation that, if removed, would significantly change the ____________ or ____________ of the regression line.

a)

influential; slope; y-intercept

b)

outlier; mean; median

c)

critical; variance; standard deviation

d)

leverage; correlation; residual

39.

Points that are ____________ in the x direction relative to the rest of the data are ____________________ points.

a)

outliers; high leverage

b)

close; influential

c)

far; outlier

d)

close; outlier

40.

The standard deviation of the residuals is ______. It measures the typical _______________________________.

a)

s; residual

b)

r; correlation strength

c)

b; slope of the line

d)

a; intercept value

41.

A(n) ________________ attempts to collect data from every individual in the population.

a)

census

b)

survey

c)

sample

d)

estimate

42.

A(n) ________________ is a value from a sample to used to estimate a population ___________________.

a)

statistic; parameter

b)

parameter; statistic

c)

mean; median

d)

sample; population

43.

A(n) ________________ imposes a treatment on individuals to measure their responses.

a)

experiment

b)

survey

c)

observation

d)

census

44.

A(n) ____________________________ observes individuals and measures variables without attempting to influence the responses.

a)

observational study

b)

experimental study

c)

case-control study

d)

survey

45.

A(n) ________ is a method of choosing a sample where every individual and every possible group of size n has an equal chance of being selected.

a)

simple random sample (SRS)

b)

systematic sample

c)

convenience sample

d)

stratified sample

46.

In _____________ sampling, the population is first divided into similar groups called ______________, and then a SRS is chosen from each group.

a)

stratified; strata

b)

cluster; clusters

c)

systematic; systems

d)

random; groups

47.

In _____________ sampling, the population is first divided into representative groups called __________, and then all individuals from a randomly chosen subset of these groups are selected.

a)

cluster; clusters

b)

stratified; strata

c)

systematic; systems

d)

simple random; samples

48.

____________ sampling selects individuals who are easiest to reach. This method often leads to ________________ bias, because not all of the population was eligible to be chosen.

a)

Convenience; undercoverage

b)

Random; response

c)

Systematic; measurement

d)

Stratified; nonresponse

49.

______________ sampling allows individuals to choose to be in the sample by responding to a general invitation (e.g., an online poll). This often leads to ______________ bias.

a)

Voluntary response; voluntary response

b)

Random; selection

c)

Stratified; measurement

d)

Systematic; nonresponse

50.

__________________ is the difference between the results of multiple samples taken from a population.

a)

Sampling variability

b)

Sampling bias

c)

Measurement error

d)

Nonresponse error

51.

______________ is a systematic error in the design of the study that tends to favor certain outcomes.

a)

Bias

b)

Randomization

c)

Sampling

d)

Blinding

52.

__________________ bias occurs when an individual chosen for the sample cannot be contacted or refuses to cooperate.

a)

Nonresponse

b)

Selection

c)

Response

d)

Measurement

53.

__________________ bias occurs when respondents lie or give inaccurate answers, often due to the wording of the question or the identity of the interviewer.

a)

Response

b)

Selection

c)

Sampling

d)

Measurement

54.

The four key principles of experimental design are ________________, ________________, ________________, and ________________.

a)

Control, randomization, replication, comparison

b)

Control, observation, prediction, analysis

c)

Randomization, sampling, inference, estimation

d)

Replication, correlation, causation, measurement

55.

There are three steps needed to describe the random assignment of treatments in an experiment: ________________, ________________, ________________.

a)

Label, random selection, assign

b)

Observation, measurement, conclusion

c)

Selection, grouping, analysis

d)

Preparation, execution, evaluation

56.

The individuals to whom the treatments are applied are called ______________ units. When they are human, they are called ______________.

a)

Experimental; subjects

b)

Control; patients

c)

Sample; volunteers

d)

Test; participants

57.

A(n) ______________ is a condition applied to the experimental units. The treatments are formed by combining different levels of the ______________ variables (or factors).

a)

Treatment; explanatory

b)

Experiment; response

c)

Variable; dependent

d)

Factor; independent

58.

In a(n) ______________ experiment, neither the subjects nor the people who interact with them and measure the response variable know which treatment a subject received.

a)

Double-blind

b)

Single-blind

c)

Open-label

d)

Placebo-controlled

59.

In a(n) ______________ experiment, either the subjects or the people who interact with them know the treatment a subject received, but not both.

a)

Single-blind

b)

Double-blind

c)

Randomized

d)

Controlled

60.

The ______________ effect occurs when subjects not receiving an active treatment show a response simply because they believe they are receiving a treatment.

a)

Placebo

b)

Nocebo

c)

Hawthorne

d)

Observer

61.

In a(n) _____________________ experimental design, experimental units are randomly assigned to treatments.

a)

Completely randomized

b)

Matched pairs

c)

Block

d)

Factorial

62.

In a(n) __________________ experimental design, experimental units are assigned to groups based on a common characteristic, then treatments are assigned within each group.

a)

Randomized block

b)

Completely randomized

c)

Matched pairs

d)

Factorial

63.

In a(n) ____________ experimental design, experimental units are either matched up with one other unit or matched with themselves. Then treatments are randomly assigned.

a)

Matched pairs

b)

Completely randomized

c)

Block

d)

Factorial

64.

Random assignment of treatments is important to show __________________.

a)

Causation

b)

Correlation

c)

Generalization

d)

Bias

65.

Random sampling is important for _____________ results to a __________________.

a)

Generalizing; population

b)

Specifying; sample

c)

Analyzing; variable

d)

Comparing; group

66.

The set of all possible outcomes of a chance process is called the _____________.

a)

sample space

b)

event

c)

probability

d)

experiment

67.

A(n) ______________ is any collection of outcomes from some chance process.

a)

event

b)

sample

c)

experiment

d)

variable

68.

The ______________ of any event must be a number between 0 and 1.

a)

probability

b)

mean

c)

median

d)

mode

69.

The __________________________ states that if we observe more and more repetitions of a chance process, the proportion of times that a specific outcome occurs approaches a single value.

a)

law of large numbers

b)

central limit theorem

c)

random variable principle

d)

probability distribution rule

70.

If two events have no outcomes in common, they are ____________________ (or disjoint). This means P(A ∩ B) = __________.

a)

mutually exclusive, 0

b)

independent, 1

c)

complementary, A+B

d)

dependent, AB

71.

The probability that event B occurs given that event A has already occurred is called ________ ________ and is written as ___________.

a)

conditional probability, P(B|A)

b)

joint probability, P(A∩B)

c)

marginal probability, P(B)

d)

independent probability, P(A|B)

72.

Two events A and B are ________________ if the occurrence of one event does not affect the probability that the other event occurs. This means P(A ∩ B) = ____________

a)

independent, P(A) × P(B)

b)

dependent, P(A) + P(B)

c)

mutually exclusive, 0

d)

complementary, 1

73.

A ________________, or a ____________________ can be helpful tools for organizing and solving probability problems.

a)

Venn diagram, two-way table

b)

bar graph, pie chart,

c)

scatter plot, box plot

d)

dot plot, line graph

74.

A ____________________ is a variable whose value is a numerical outcome of a chance process.

a)

random variable

b)

dependent variable

c)

categorical variable

d)

constant

75.

A ____________ random variable X takes a fixed set of possible values with gaps between them.

a)

discrete

b)

continuous

c)

uniform

d)

normal

76.

A ____________ random variable Y takes all values in an interval of numbers.

a)

continuous

b)

discrete

c)

categorical

d)

binary

77.

The __________ value of a random variable is the mean of the outcomes, calculated as ______________

a)

expected, Σx*p(x)

b)

actual, simple addition

c)

possible, mean

d)

average, frequency count

78.

When adding or subtracting two random variables, X and Y, the new ________ is always the sum or difference of the individual means.

a)

mean

b)

variance

c)

standard deviation

d)

median

79.

When adding or subtracting two independent random variables, the new ________ is the square root of the sum of the______________.

a)

standard deviation, variances

b)

variance, standard deviation

c)

variance, multiply

d)

median, divide

80.

The conditions for a ________________ distribution are ____________________, ____________________, ____________________, and ______________________________.

a)

binomial, fixed number of trials, independent trials, two possible outcomes, constant probability of success

b)

normal, continuous data, symmetric distribution, mean equals median, bell-shaped curve

c)

poisson, rare events, independent events, constant average rate, discrete outcomes

d)

uniform, equal probability, fixed range, independent outcomes, constant distribution

81.

In a ____________ distribution, the variable of interest is the number of ______________ needed to get the ______________ success.

a)

geometric, trials, first

b)

binomial, successes, last

c)

poisson, events, next

d)

normal, samples, average

82.

The mean (expected value) of a Binomial random variable is __________. The standard deviation of a Binomial random variable is ____________.

a)

np, np(1p)np,\ \sqrt[]{np\left(1-p\right)}

b)

np, npnp,\ \sqrt[]{np}

c)

np, n(1p)np,\ \sqrt[]{n\left(1-p\right)}

d)

p, np(1p)p,\ \sqrt[]{np\left(1-p\right)}

83.

The mean (expected value) of a Geometric random variable is ____________. The standard deviation of a Geometric random variable is ____________.

a)

1p\frac{1}{p} , 1pp\frac{\sqrt[]{1-p}}{p}

b)

p, 1pp\frac{\sqrt[]{1-p}}{p}

c)

1p\frac{1}{p} , 1pp\sqrt[]{\frac{1-p}{p}}

d)

p, 1pp2\sqrt[]{\frac{1-p}{p^2}}

84.

When calculating a binomial probability of an exact value, use binom _______. When calculating a binomial geometric probability of a ≤, use binom_______.

a)

pdf, cdf

b)

cdf, pdf

c)

mean, variance

d)

mode, median

85.

What is the meaning of the term 'Categorical'?

a)

A variable that can be divided into groups or categories that do not have a numerical value.

b)

A variable that always represents a numerical value.

c)

A variable that can only take continuous values.

d)

A variable that is always used for mathematical calculations.

86.

What does 'Standard Deviation' measure?

a)

The typical distance from a mean.

b)

It measures the average value in a set of data.

c)

It measures the highest value in a set of data.

d)

It measures the total sum of all values in a set of data.

87.

Which term refers to the strength and direction of a linear relationship between two variables?

a)

Correlation Coefficient

b)

Mean

c)

Standard Deviation

d)

Coefficient of Determination