Font size
WorksheetsMock Test_6-2_Advance Statistics
Total questions: 165
Worksheet time: 8hrs 15mins
What research design aims to determine a cause from already existing effects?
Descriptive Research Design
Correlational Research Design
Quasi-Experimental Research
Ex Post Facto
What research design is often conducted in a controlled setting with corresponding research treatment?
Correlational
Ex post facto
Survey Research
Experimental
What is the suited research design for this research title, “The Effects of Twitter on the Communication Etiquette of Students”?
Ex post facto
Experimental
Descriptive
Correlational
Mr. James Rivera would like to know further the type of social media used between the male and female SHS students of Nicolas B. Barreras National High School. What is the appropriate research design to be used in his study?
Quasi-Experimental
Correlational
Experimental
Descriptive
What is the aim of Ex post facto research design?
determine a cause from already existing effe
establish cause and effect relationship
observe and describe a phenomenon
identify association among variables
“Effects of Type of Music to Aesthetic Performance of Ballet Dancers”, what is the appropriate research design for the given title?
Correlational
Descriptive
Survey Research
Experimental
Mr. Manzano would like to know further the type of social media used between the male and female SHS students of East Pagat National High School.
What appropriate statistical test should Mr. Manzano used to answer his research problem?
T-test for two dependent samples
Spearman’s rho
Chi-square
ANOVA
Which of the following statements is true about the conduct of experimental research?
There is no random assignment of individuals.
Individual subjects are randomly assigned.
Groups are exposed to presumed cause.
Intact groups are used.
What is the difference between quasi-experimental research and experimental research?
Only one dependent variable is used in quasi-experimental research, while multiple dependent variables can be used in quasi-experimental research.
Intact groups are used in experimental, while quasi-experimental randomly assigned individuals into groups.
Participants for groups are randomly selected in experimental, but not quasi-experimental research.
The researcher controls the intervention in the experimental group, but not quasi-experimental research.
Why would a researcher choose to use Simple Random Sampling as a sampling technique?
To consider giving equal chance to the member of accessible population being selected as part of the study.
To make sure that all subcategories of the population are represented in the selection of sample.
To group the entire population into clusters since the location of the samples are widely spread.
To systematically choose samples from a given list of individuals.
When can we consider a research sample as "best?"
representative of population
systematically chosen
conveniently represented
purposely selected
Given that your study will use stratified random sampling, wherein population of your scope is 250 with a computed sample size of 152, how many samples for each stratum will you have if group 1 has 92, group 2 has 86, and group 3 has 72 population?
Group 1 = 52, Group 2 = 54, Group 3 = 46
Group 1 = 56, Group 2 = 45, Group 3 = 51
Group 1 = 52, Group 2 = 52, Group 3 = 44
Group 1 = 54, Group 2 = 56, Group 3 = 41
What type of reliability is measured by administering two tests identical in all aspects except the actual wording of items?
Internal Consistency Reliability
Equivalent Forms Reliability
Test-retest reliability
Inter-rater Reliability
What type of validity is when an instrument produces results similar to those of another instrument that will be employed in the future?
Predictive Validity
Face Validity
Criterion Validity
Content Validity
The Ability Test has been proven to predict the mathematical skills of Senior High School students. What type of test validity is shown in the example?
Construct Validity
Criterion Validity
Content Validity
Face Validity
What indicator of a good research instrument when items are arranged from simple to complex?
Easily Tabulated
Sequential
Valid and Reliable
Concise
What is the purpose of Pearson’s r as a statistical technique? To test the
difference between sets of data from different groups
difference between two sets of data from one group.
degree of effect research intervention or treatment.
relationship between two continuous variables.
In this scale, a series of bipolar adjectives will be rated by the respondents. This scale seems to be more advantageous since it is more flexible and easier to construct.
Semantic Differential
Likert Scale
Spearman's rho
Face Validity
Data gathering is done through interview or questionnaire
observation
survey
experiment
none of these
(a) refers to the tools used in research for the purpose of gathering the data.
It is the science that deals with the collection, organization, presentation, analysis, and interpretation of data.
Descriptive
Inferential
Probability
Statistics
An area in Statistics that makes conjecture about the population based on data from sample
Descriptive
Inferential
Probability
Statistics
Statistical tests used to process data that are not normally distributed.
Parametric
Non-parametric
Multivariate Analysis
Regression Analysis
It is a parametric tests that determines differences between a parameter of more than 2 groups.
F-test
T-test
Z-Test
Pearson Product Moment Correlation
A parametric test that determines difference between sample mean of two independent groups
F-test
T-test
Z-test
Pearson Product Moment Correlation
It is a parametric test that determines the association of two variables.
F-test
t-test
z-test
Pearson Product Moment Correlation
Which among the following does not belong to the group of statistical tests
Chi-square
Kolmogorov-Smirnov
Mann-Whitney
Z-test
Which among the following does not belong to the group of statistical tests.
Chi-square
Pearson Product Moment Correlation
Mann-Whitney
Spearman Rank
It is a non-parametric test that determines difference between 3 or more groups.
Kolmogorov Smirnov Test
Mann-Whitney Test
Kruskall Wallis Test
Chi-Square Test
It examines the relationship of an independent variable to one or more dependent variables. In the end it will give a predictive model of the independent variable.
Chi-square
Correlation
Multivariate ANOVA
Regression
These are samples that are selected randomly so that its observation do not depend on the values on other observations.
independent samples
binomial test
x2 one sample test
one sample run test
It means having only two possible values such as yes or no, male or female, pass or fail, head or tail.
Fisher's Exact Test
Median Test
Dichotomous Variable
Contingency Table
The purpose of this test is to evaluate the differences between two discrete dichotomous variables, where responses of two independent groups fall exclusively into one category or the other.
Fisher's Exact Test for 2x2 Tables
Mood's Median Test
Kolmogorov-Smirnov Two-Sample Test
Wilcoxon-Mann-Whitney U Test
All, except one, is used to analyze two independent sample.
Wilcoxon-Mann-Whitney U Test
Median Test
Pearson Test
Hodges-Lehmann Test
This is a typical statistical theory which suggests that no statistical relationship and significance exists in a set of given single observed variable.
Null Hypothesis
Alternative Hypothesis
Independent Sample
Dependent Sample
SPSS is a major market occupier in terms of statistical packaging tolls which can efficiently be used as the derivative for the data manipulation or storage. SPSS means...
statistical performance for specific study
solution to problems in statistical study
statistical portfolio to solve statistics
statistical package for social science
One of its advantage is its insensitivity to departures from homogeneity of variance.
Median Test
Fisher's Exact Test
Wilcoxon-Mann Whitney U Test
Kolmogorov-Smirnov Test
This is also known as Mood's Test for two independent samples. It is useful for studies that investigate individual reactions to certain situations, such as medications, diets, therapies, and exercise programs.
Wilcoxon-Rank-Sum Test
Lehmann's Test
U Test
Median Test
This test is more powerful than the median test because the rank of each observation is considered instead of only the relation of a score to the median value in the distribution.
Wilcoxon-Mann-Whitney U test
Fisher's Exact Test
Mood's Median Test
Kolmogorov-Smirnov Test
If there is no difference between the average ranks for the two groups in Wilcoxon Test, then the average group ranks should be approximately equal. However, if the sum of the ranks is quite different in size, then we may suspect that the two groups were not from from the same distribution.
True
False
Half True
Half Flase
What is the measure of central tendency that represents the most frequently occurring value in a dataset?
Mode
Median
Standard Deviation
Mean
Which distribution is used to model the number of successful events in a fixed interval of time or space?
Poisson Distribution
Exponential Distribution
Normal Distribution
Binomial Distribution
What is the probability of an event A occurring given that event B has already occurred?
Marginal Probability
Bayes' Theorem
Conditional Probability
Joint Probability
Which statistical test is used to determine if there is a significant association between two variables?
ANOVA
Fisher's Exact Test
Chi-Square Test
T-Test
What is the measure of the strength and direction of the linear relationship between two variables?
Covariance
Correlation Coefficient
Coefficient of Determination
Slope of Regression Line
What is the probability distribution used to model the number of successful events in a fixed interval of time or space?
Poisson Distribution
Exponential Distribution
Normal Distribution
Binomial Distribution
Which distribution is commonly used to model the number of occurrences of an event in a fixed interval of time or space?
Poisson Distribution
Exponential Distribution
Normal Distribution
Binomial Distribution
What is the name of the distribution used to represent the number of events occurring in a fixed interval of time or space?
Poisson Distribution
Exponential Distribution
Normal Distribution
Binomial Distribution
What is the appropriate statistical test to determine if there is a significant difference between the means of two independent groups?
ANOVA
Fisher's Exact Test
Chi-Square Test
T-Test
Which statistical test is used to determine if there is a significant relationship between two categorical variables?
ANOVA
Fisher's Exact Test
Chi-Square Test
T-Test
Which measure of central tendency represents the most common value in a dataset?
Mode
Median
Standard Deviation
Mean
What is the statistical measure that represents the most frequently occurring value in a dataset?
Mode
Median
Standard Deviation
Mean
What is the measure of central tendency that represents the value that appears most frequently in a dataset?
Mode
Median
Standard Deviation
Mean
What is the probability of an event C occurring given that event D has already occurred?
Marginal Probability
Bayes' Theorem
Conditional Probability
Joint Probability
What is the probability of an event E occurring given that event F has already occurred?
Marginal Probability
Bayes' Theorem
Conditional Probability
Joint Probability
What is the probability of an event G occurring given that event H has already occurred?
Marginal Probability
Bayes' Theorem
Conditional Probability
Joint Probability
Which statistical measure represents the proportion of the variance for a dependent variable that's explained by an independent variable?
Covariance
Correlation Coefficient
Coefficient of Determination
Slope of Regression Line
What is the measure of the steepness of a line that represents the relationship between the independent and dependent variables?
Covariance
Correlation Coefficient
Coefficient of Determination
Slope of Regression Line
What is the statistical measure of the strength and direction of the linear relationship between two variables?
Covariance
Correlation Coefficient
Coefficient of Determination
Slope of Regression Line
In Bayesian statistics, the value equivalent to a p value in Null Hypothesis Significance testing is denoted as:
B
D
K
P
Which of the following statements is true about Bayesian statistical tests?
There is a solid cutoff value for significance/non-significance
H0 is defined as the null hypothesis
The ratio of probabilities is known as the Bayes Inferential (BI)
Measures of Credible Intervals are taken instead of Confidence Intervals
A correlation value of 0.57 between variables x and y suggests that:
As the value of x increases, the value of y increases
As the value of x increases, the value of y decreases
There is no relationship between variables x and y
Impossible to tell from this information
I want to investigate whether attendance, previous exam marks, and attitudes towards statistics will predict scores obtained in a statistics exam. What type of analysis should I conduct with the data I have collected?
Simple linear regression
Multiple linear regression
Pearson's r correlation
Spearman's rho correlation
After conducting a simple linear regression on a dataset, I am given an R squared value of 0.37. What percentage of variance in my outcome variable can be explained by the predictor variable in my model?
0.37%
3.7%
37%
0.037%
Which of the following is not an assumption that should be met when conducting a Pearson's correlation?
There should be an absence of significant outliers.
The relationship between x and y should form a curved line.
Variables x and y should be continuous.
Each participant or observation should have a related pair of values.
A company tries to advertise their latest product using Facebook adverts, and finds no increase in product sales. They decide not to do this again because they believe this method of advertising does not work. However, a second company uses the same advertising method and sees an increase in product sales. What kind of error did company one make?
Type I error
Type II error
Standard error
No error
The unexplained variance in a between groups ANOVA is:
Residual variance
Variance between groups
Variance within groups
Variance from the population
A table showing how often observations fall within a particular category is also known as a:
Contingency table
Frequency table
Cumulative frequency table
Expected values table
For a Chi-square analysis, we reject the null hypothesis when the observed value is:
The same as the critical Chi-square value
Less than the critical Chi-square value
Greater than the critical Chi-square value
Twice as high as the critical Chi-square value
When all the values in a distribution are the same, the variance is ___________.
one
zero
maximum
cannot be determined
Which of the following is CORRECT about the variance?
The smaller the variance, the more heterogenous are the data values.
The bigger the variance, the data tends to deviate less from the mean.
The lesser the variance, the data values tend to deviate more from the mean.
The bigger the variance, the data values tend to deviate more from the mean.
Which among the following is NOT true about the measures of variation?
It is also known as measures of spread.
The lower the variability the lesser the spread of scores.
The higher the variability the higher the spread of scores.
It does not tell us that the scores are compressed or spread out.
The average score in the test of section A is 36. Section B also gained the same average scores on the same test. However, section A has a standard deviation of 3.8 while section B has a standard deviation of 1.9. In which section are the scores less dispersed?
section A
section B
no enough information
could not be determined
A financial analyst sampled the top six companies' book value (in billions of pesos). They are:
7, 15, 18, 22, 25, 33
Approximately, what is the sample mean, as well as the sample standard deviation?
20 and 7.2 respectively
20 and 8.9 respectively
120 and 9.2 respectively
120 and 8.9 respectively
Which of the following statements is correct?
Group A is more varied than Group B because the sample size is larger.
Group A is less varied than Group B because Group A's standard deviation is bigger.
Group A is less varied than Group B because its standard deviation per animal is smaller.
Group A is relatively less varied than Group B because Group A's coefficient of variation is smaller.
Which of the following is a statistical method that measures the strength of the linear relationship between two variables?
z-value
scatterplot
testing hypothesis
Pearson correlation
In the Pearson r, what does n represent?
sum of x-values
sum of square x-values
number of paired values
sum of the products of paired values x and y
Which of the following values CANNOT represent a correlation coefficient r ?
-1
0
0.35
1.05
Which of the following is the range of the correlation coefficient (r)?
-1 < r < 1
-1 ≤ r ≤ 1
0 ≤ r ≤ 1
1 ≤ r ≤ -1
Which among the choices is the correct completed table?
A
B
C
D
What conclusion can the researcher draw based on the scatter plot?
The variables do not correlate.
The variables have a strong correlation.
The variables have a perfect correlation.
The variables have a moderate correlation.
Given the bivariate data table below what will be the value of r?
-0.5
-0.25
0.75
0.95
What is assumed to be its ρ?
-1
0
1
2
If you have two values tied for the 10th and the 11th place, then what rank do you put for both of them?
10
10.5
11
11.5
The spearman's rank correlation coefficient for the Grade 9 STATISTICS and CONSUMER CHEMISTRY scores is 0.91. Which of the following is true?
There is a weak positive correlation.
There is a weak negative correlation.
There is a strong positive correlation.
There is a strong negative correlation.
Spearman's correlation coefficient computation is unique in what way?
It can be used with more than two variables at a time.
It uses the differences in the data as part of the calculation.
It uses the ranked value of the input data as part of the calculation.
It is better at handling random data than other correlation techniques.
What assumption about the data is measured by Spearman's correlation coefficient in comparison to Pearson's correlation calculation?
That the data is random
That the data has a linear trend
That the data has a monotonic trend
That the data has an exponential trend
When will the prediction of an individual's score on the criterion variable based on our knowledge of this individual's value on the predictor variable be more accurate?
When there is no correlation between the two variables.
When the correlation between the two variables is weak.
When the correlation between the two variables is strong.
When the correlation between the two variables is moderate.
What is the possible interpretation of the
ρ = -0.38?
Low Positive Correlation, as the variable increases, the other variable partially increases.
Low Positive Correlation, as the variable increases, the other variable partially decreases.
Low Negative Correlation, as the variable increases, the other variable partially increases.
Low Negative Correlation, as the variable increases, the other variable partially decreases.
What is the Spearman rank correlation coefficient?
0.1636
0.794
0.8364
0.8606
What is the set of statistical methods used to describe the relationship between independent variables and a dependent variable?
regression analysis
regression equation
regression form
regression line
A researcher is investigating the relationship between the number of hours students study per week and their scores on a standardized test. What is the independent variable (x) in this study?
The students
The standardized test score
The number of hours studied per week
The relationship between study time and test scores
What is the estimated slope of the regression equation?
– 213.2
– 1.56
1.56
213.2
What is the y-intercept of the regression equation?
– 213.2
1.56
– 1.56
213.2
What is the regression equation based on the data given?
ŷ = – 213.2 – 1.56x
ŷ = – 213.2 + 1.56x
ŷ = 213.2 – 1.56x
ŷ = 213.2 + 1.56x
What is the average number of potato chips sold if the price of a bag of potato chips is ₱ 75.00?
56
65
76
96
The linear regression equation is y = 61.93x - 1.79. Use the equation to predict how far this person will travel after 5 hours of driving.
307.86 miles
0.19 miles
10 miles
500 miles
If the regression equation is ŷ = -4 + 5x, how would you predict the value of
the dependent variable if the value of the independent variable is 5?
Solve for x using the slope formula.
Solve for y using the slope formula.
Substitute ŷ = 5 into ŷ = -4 + 5x and solve.
Substitute x = 5 into ŷ = -4 + 5x and solve.
Colgate is one of the manufacturers of toothpaste which uses a machine to dispense liquid ingredients into bottles that move along a filling line. The owner claims that the machine can dispense at an average of 50 grams with a standard deviation of 0.7 grams. A sample of 35 bottles was selected and it was found that the average amount dispensed in the sample was 49.3 grams. Test the claims of the owner of the company at a 5% level of significance.
What test statistic will be used from the situation given?
t-test
z-test
population mean
standard deviation
Colgate is one of the manufacturers of toothpaste which uses a machine to dispense liquid ingredients into bottles that move along a filling line. The owner claims that the machine can dispense at an average of 50 grams with a standard deviation of 0.7 grams. A sample of 35 bottles was selected and it was found that the average amount dispensed in the sample was 49.3 grams. Test the claims of the owner of the company at a 5% level of significance.
Which of the following information is CORRECT?
𝛼 = 0.5
𝜎 = 0.7
𝑥̅ = 35
𝜇 = 49.3
Based on the situation given, what is the computed value?
-5.916
-4.950
4.950
5.916
A coffee shop needs to produce bottles that will hold 12 ounces of liquid. Periodically, the company gets complaints that their bottles are not holding enough liquid. To test this claim, the bottling company randomly sampled 36 bottles. Suppose the p-value of this test turned out to be 0.0455. State the proper conclusion.
At α = 0.05, reject the null hypothesis.
At α = 0.025, reject the null hypothesis.
At α = 0.035, accept the null hypothesis.
At α = 0.085, fail to reject the null hypothesis.
An investigator randomly assigns 30 college students into three equal size study groups (early- morning, afternoon, late-night) to determine if the period of the day at which people study affects their retention. The students live in a controlled environment for one week, on the third day of the experimental treatment is administered (study of predetermined material). On the seventh day the investigator tests for retention. In computing his ANOVA table, he sees that his MS within groups is larger than his MS between groups. What does this result indicate?
An error in the calculations was made.
There should have been additional controls in the experiment.
There was more than the expected amount of variability between groups.
There was more variability between subjects within the same group than there was between groups.
A botanist is studying the growth of a particular species of plant. A previous study indicated that the average height of the plant at maturity was 15 inches. The botanist suspects that a new fertilizer might increase the average height. They take a random sample of 25 plants grown with the new fertilizer and measure their heights. Let μ represent the average height of this plant species grown with the new fertilizer.
Which of the following is the appropriate set of hypotheses for the botanist's significance test?
H₀: μ = 15 inches Hₐ: μ < 15 inches
H₀: μ = 15 inches Hₐ: μ > 15 inches
H₀: μ = 15 inches Hₐ: μ ≠ 15 inches
H₀: μ > 15 inches Hₐ: μ = 15 inches
A light bulb manufacturer claims that its light bulbs have a mean life of 40,000 km. A random sample of 46 of these light bulbs is tested and the sample mean is 38,000 km. Assume that the population’s standard deviation is 2,000 km and the lives of the light bulbs are approximately normally distributed. Determine the computed value at 5% level of significance.
-6.782
-3.033
3.033
6.782
A researcher is studying the effects of three different types of background music (classical, pop, and no music) on test performance. Participants are randomly assigned to one of the three groups, and their test scores are recorded. Which statistical test is most appropriate for comparing the average test scores of the three groups?
t-test
chi-square test
ANOVA (analysis of variance)
correlation coefficient
The null hypothesis for an ANOVA states that __________.
All of the population means are different from each other.
There are no differences between any of the population means.
At least one of the population means is different from the others.
There are some differences between any of the population means.
What two pieces of information are needed to determine the critical value?
MSb, MSw
Sample size, number of groups
Mean, sample standard deviation.
Expected frequency, obtained frequency.
A researcher is investigating the effectiveness of four different learning methods (online modules, in-person lectures, group projects, and self-study) on student test scores. Students are randomly assigned to one of the four methods. What type of ANOVA test is most appropriate to analyze the data?
One-Way ANOVA
Two-Way ANOVA
Three-Way ANOVA
Four-Way ANOVA
A researcher is studying the effects of four different fertilizers on plant growth. Four groups of plants are grown, each receiving a different fertilizer. After a set period, the height of each plant is measured.
If the researcher conducts an ANOVA test on the plant height data, what will a significant F-statistic indicate?
The fertilizers have no effect on plant growth.
All four fertilizers have significantly different effects on plant growth.
There is no significant difference in plant growth between the four groups.
At least one fertilizer has a significantly different effect on plant growth compared to the others.
A medical researcher is testing a new diagnostic test for a rare disease.
Null Hypothesis (H0): The person does not have the disease.
Alternative Hypothesis (H1): The person has the disease.
The researcher tests a patient.
When is a Type I error committed?
The test is positive, and the patient has the disease.
The test is negative, and the patient does not have the disease.
The test is negative, but the patient has the disease.
The test is positive, but the patient does not have the disease.
What is the statistical method used in making decisions using experimental data?
Simple analysis
Analytical testing
Hypothesis testing
Experimental testing
If the coefficient determination is 0.95, what does it mean?
95% of the y values are positive.
95% of the y values are negative.
95% of the y values are predicted correctly by the model.
95% of the variation in y can be explained by the variation in x.
