wayground logo

Free Printable Worksheets

Font size

S
M
L
XL
Worksheets

EDU215_Item Analysis and Validation

Total questions: 45

Worksheet time: 18mins

Name
Class
Date
1.

It refers to the consistency of the scores obtained.

a)

Validity

b)

Reliability

c)

Practicality

d)

Authenticity

2.

It is a measure of internal consistency, that is, how closely related a set of items are as a group.

a)

Cronbach's Alpha

b)

Stability Test

c)

Internal Consistency

d)

Pearson Product Moment Correlation

3.

The researcher determines the validity by looking at the features of the instrument.

a)

Construct Validity

b)

Content Validity

c)

Face Validity

d)

Predictive Validity

4.

It measures by subjecting the instrument to an analysis by a group of experts who are knowledgeable about the subject both in theory and practices.

a)

Construct Validity

b)

Content Validity

c)

Face Validity

d)

Predictive Validity

5.

This refers whether the test corresponds to its theoretical construct.

a)

Construct Validity

b)

Content Validity

c)

Face Validity

d)

Divergent Validity

6.

It is determined by administering both the new test and the standardized test to a group of respondents, then finding the correlation between the two sets of the scores.

a)

Concurrent

b)

Predictive

c)

Consistency

d)

Convergent

7.

It refers to how well the test predicts some future behavior of the examinees.

a)

Consistency

b)

Concurrent

c)

Predictive

d)

Convergent

8.

The ability of the instrument to measure what it intends to measure.

a)

Validity

b)

Reliability

c)

Item Difficulty

d)

Discrimination Index

9.

The same test is given to a group of respondents twice.

a)

Internal Consistency

b)

Test-retest or Stability Test

c)

Equivalence Test

d)

Inter-Rater

10.

Items sought must be correlated with each other and the test should be internally consistent.

a)

Internal Consistency

b)

Test-retest or Stability Test

c)

Parallel Test

d)

Spli-Half

11.

Which of the following are ways of assessing reliability?

a)

Test-retest reliability

b)

Content Reliability

c)

Face reliability

d)

Concurrent reliability

12.

Which of the following are ways of assessing validity?

a)

Test-retest validity

b)

Inter-observer validity

c)

Face validity

d)

Concurrent validity

13.

When correlating results to assess reliability, what correlation coefficient must be obtained for the data to be considered reliable?

a)

0.8

b)

0.6

c)

0.9

d)

0.7

14.

Which of the following statements are true?

a)

Inter-observer reliability must involve the use of one observer.

b)

Inter-observer reliability must involve participants taking part in the research more than once.

c)

Test-retest reliability must involve participants taking part in the research more than once.

d)

Test-retest reliability must involve the use of more than one observer.

15.

When using a statistical test to assess test-retest or inter-observer reliability, which of the following criteria MUST that test meet?

a)

It must be a test of difference

b)

It must be a test of correlation

c)

It must be appropriate for nominal data

d)

It must be appropriate for a repeated measures design

16.

Which is a way in establishing test reliability?

a)

The test is examined if free from errors and properly administered.

b)

Scores in a test with different versions are correlated to test if they are parallel.

c)

The components or factors of the test contain items that are strongly uncorrelated.

d)

Two or more measures are correlated to show the same characteristics of the examinee.

17.

What is being established if items in the test are consistently answered by the students?

a)

Internal Consistency

b)

Inter-rater Reliability

c)

Test-retest

d)

Split-half

18.

Which type of validity was established if the components or factors of a test are hypothesized to have a negative correlation?

a)

Construct Validity

b)

Predictive Validity

c)

Content Validity

d)

Divergent Validity

19.

How do we determine if an item is easy or difficult?

a)

An item is easy if majority of students are not able to provide the correct answer. The item is difficult if majority of the students are able to answer correctly.

b)

An item is difficult if majority of the students are not able to provide the correct answer. The item is easy if majority of the students are able to answer correctly.

c)

An item can be determined difficult if the examinees who are high in the test can answer more the items correctly than the examinees who got lows scores. If not, the item is easy.

d)

An item can be determined easy if the examinees who are high in the test can answer more the items correctly than the examinees who got low scores. If not, the item is difficult.

20.

Which is used when the scores of the two variables measured by a test taken at two different times by the same participants are correlated?

a)

Pearson r correlation

b)

Linear Regression

c)

Significance of the Correlation

d)

Cronbach Alpha

21.

A school psychologist administers the same intelligence test to a student on two different occasions, three months apart. The scores are highly consistent. Which type of reliability was established?

a)

Test-retest reliability

b)

Split-half reliability

c)

Inter-rater reliability

d)

Internal Consistency

22.

Two different clinicians independently score a patient's behavioral observation checklist for anxiety symptoms. Their scores are very similar. What is being established?

a)

Construct Validity

b)

Test-retest Reliability

c)

Parallel-Form Reliability

d)

Inter-rater Reliability

23.

A professor creates a final exam for a course on Philippine history. The exam questions cover the entirety of the course material, including topics from every lecture and reading assignment. Which type of validity is being established?

a)

Predictive Validity

b)

Construct Validity

c)

Face Validity

d)

Content Validity

24.

An educational psychologist develops two different versions of a math skills test, Test A and Test B, that are designed to be equivalent in content and difficulty. A group of students takes both tests, and their scores are highly correlated. What kind of reliability is being demonstrated?

a)

Internal Consistency

b)

Parallel-Form Reliability

c)

Test-retest Reliability

d)

Split-half Reliability

25.

A company uses a pre-employment test to screen job applicants. After six months, they compare the test scores of new hires to their job performance ratings. They find a strong positive correlation. What type of validity is this a measure of?

a)

Predictive Validity

b)

Construct Validity

c)

Face Validity

d)

Concurrent Validity

26.

A researcher is developing a new measure of 'extroversion'. She finds that the scores on her test are highly correlated with scores on a well-established and validated measure of extroversion. What type of validity is this demonstrating?

a)

Predictive Validity

b)

Content Validity

c)

Divergent Validity

d)

Convergent Validity

27.

A test has 100 items. To quickly estimate its reliability, a researcher divides the test into two equal halves, one with odd-numbered items and the other with even-numbered items. She then correlates the scores from the two halves. Which method is being used?

a)

Test-retest reliability

b)

Parallel-forms reliability

c)

Split-half reliability

d)

Inter-rater reliability

28.

A new self-esteem questionnaire is created. The developers want to ensure that it is not just measuring depression. They administer the new questionnaire and a well-established depression scale to the same group of people. They hope to find a very low or negative correlation between the two. What kind of validity are they trying to establish?

a)

Predictive validity

b)

Divergent validity

c)

Convergent validity

d)

Content validity

29.

A test for 'math anxiety' is given to a group of students. The scores are then compared to the students' grades in their current math class. The scores and grades are found to be highly correlated. Which type of validity is being established?

a)

Content validity

b)

Concurrent validity

c)

Predictive validity

d)

Face validity

30.

A test question is answered correctly by 95% of the students. What can be said about this item's difficulty index?

a)

The item is too difficult.

b)

The item is invalid.

c)

The item is too easy.

d)

The item has a moderate difficulty.

31.

A test item has a difficulty index of P=0.50 and a discrimination index of D=0.45. What is the best conclusion about this item?

a)

The item is too easy and effectively discriminates.

b)

The item is too easy and does not discriminate.

c)

The item is too difficult and does not discriminate.

d)

The item is of ideal difficulty and effectively discriminates.

32.

A researcher is developing a new anxiety scale for adolescents. He wants to demonstrate that the scale is related to other established measures of anxiety (like clinical diagnoses) and unrelated to measures of different constructs, like social desirability. Which type of validity is he trying to establish?

a)

Content validity

b)

Face validity

c)

Predictive validity

d)

Construct validity

33.

A teacher wants to make sure her new math quiz has items that 'look' like they are measuring math skills, even to her students. Which type of validity is she most concerned with?

a)

Face validity

b)

Content validity

c)

Construct validity

d)

Predictive validity

34.

An item on a test is answered correctly by most of the high-scoring students and incorrectly by most of the low-scoring students. What does this indicate about the item?

a)

The item has a poor difficulty index.

b)

The item has a good discrimination index.

c)

The item has a negative discrimination index.

d)

The item has a low discrimination index.

35.

A test developer calculates a Cronbach's alpha coefficient for a new scale. The alpha coefficient is very high (>0.90). Which type of reliability is the developer measuring?

a)

Inter-rater reliability

b)

Parallel-forms consistency

c)

Internal consistency

d)

Test-retest reliability

36.

A university is developing a new admissions test. They want to know if the test scores correlate with students' first-year college GPAs. Which type of validity are they most interested in?

a)

Concurrent validity

b)

Predictive validity

c)

Face validity

d)

Content validity

37.

A test item's difficulty index is found to be 0.05. What is the most appropriate action to take with this item?

a)

Keep the item as is, it's a good discriminator.

b)

Move the item to the beginning of the test.

c)

Make the item more difficult.

d)

Revise or remove the item.

38.

A test has a negative discrimination index. What does this mean?

a)

The item is answered correctly by more low-scoring students than high-scoring students.

b)

The item is of ideal difficulty.

c)

The item is answered correctly by more high-scoring students than low-scoring students.

d)

The item is too easy for the test takers.

39.

A test is designed to measure 'creativity,' a concept that is not directly observable. What type of validity is most crucial to establish for this test?

a)

Construct validity

b)

Face validity

c)

Content validity

d)

Predictive validity

40.

A teacher wants to know if a new math test she created is a good measure of student achievement. She compares the scores on her test to the students' recent scores on a standardized math test. She finds a high correlation. Which type of validity is she demonstrating?

a)

Face validity

b)

Construct validity

c)

Predictive validity

d)

Concurrent validity

41.

A test item on a history test is correctly answered by 85% of the students. The upper-group students (top 27%) answered it correctly at a rate of 90%, while the lower-group students (bottom 27%) answered it correctly at a rate of 80%. What is the discrimination index for this item?

a)

D=0.80

b)

D=0.85

c)

D=0.90

d)

D=0.10

42.

A researcher creates two different forms of a vocabulary test and administers them to the same group of students on the same day. The scores are not correlated. Which type of reliability is low?

a)

Test-retest reliability

b)

Inter-rater reliability

c)

Internal consistency

d)

Parallel-forms reliability

43.

A test item has a negative discrimination index. What should be done with this item?

a)

Move the item to the beginning of the test.

b)

Revise or remove the item.

c)

Keep the item as is, it's a good discriminator.

d)

Make the item more difficult.

44.

A teacher wants to create a parallel test to her current history quiz. What is the most important characteristic of the new test?

a)

The new test must have the same form, but different content.

b)

The new test must have the same questions as the original test.

c)

The new test must have the same number of questions.

d)

The new test must be administered at the same time as the original test.

45.

A test item has a difficulty index of P=0.95 and a discrimination index of D=0.05. What is the most appropriate conclusion?

a)

The item is too easy but has a good discrimination.

b)

The item is too easy and has a low discrimination.

c)

The item is of ideal difficulty and has a good discrimination.

d)

The item is too difficult and has a low discrimination.