WorksheetsMidterm Examination on Psychological Assessment
Total questions: 60
Worksheet time: 45mins
A researcher is tasked with making a new aptitude test and starts by writing items, setting scoring rules, and formatting the test. Which phase of test development does this describe?
Construction
Conceptualization
Pilot Work
Item Analysis
A test developer identifies that no existing tools measure mindfulness in middle school students. This step belongs to which phase of test development?
Revision
Conceptualization
Tryout
Construction
A psychologist notes that existing personality tests do not account for cultural variations in responses. This limitation falls under which stimulus for conceptualization?
Norm-Referenced Testing
Emerging Phenomena
Limitations in Existing Tests
Pilot Work
Before developing a test, a team asks, "Who is the intended user of this test?" What aspect of development does this relate to?
Conceptualization
Construction
Pilot Work
Item Analysis
A teacher administers a math test designed to compare a student's performance to a national average. What type of test is this?
Criterion-Referenced Test
Norm-Referenced Test
Ipsative Scoring Test
Guttman Scale Test
A certification test for mechanics evaluates whether test-takers can master specific tasks, such as diagnosing an engine problem. What type of test is Construction?
Criterion-Referenced Test
Norm-Referenced Test
Comparative Scaling Test
Selected-Response Test
A developer uses focus groups and open-ended interviews to gather insights for writing items for a new leadership assessment. This activity is part of:
Pilot Work
Test Construction
Item Analysis
Test Revision
A psychologist ensures that numerical scores from a test reflect the levels of traits being measured. This process is known as:
Scaling
Scoring Models
Item Branching
Item Construction
A psychologist designs a reading test where scores are interpreted based on typical performance for specific grades. What scaling method does this involve?
Stanine Scale
Grade-Based Scales
Rating Scales
Age-Based Scales
A standardized aptitude test reports results on a nine-point scale to simplify interpretation. What type of scale is being used?
Comparative Scaling
Likert Scale
Stanine Scale
Categorical Scaling
A psychologist is creating a test to measure both verbal and mathematical skills in a single administration. What type of scale should the test use?
Unidimensional Scale
Multidimensional Scale
Stanine Scale
Guttman Scale
A test-taker ranks five activities from least to most interesting. This is an example of:
Categorical Scaling
Comparative Scaling
Rating Scales
Guttman Scale
A psychologist groups test items into predefined categories to analyze responses within specific groups. What type of scaling is this?
Categorical Scaling
Likert Scale
Comparative Scaling
Ipsative Scoring
A survey asks respondents to rate their agreement with statements on a scale from "Strongly Disagree" to "Strongly Agree." This is an example of:
Guttman Scale
Stanine Scale
Likert Scale
Equal-Appearing Intervals
A scale uses equal-appearing intervals where judges place items on a continuum to represent intensity. This method is known as:
Guttman Scaling
Scalogram Analysis
Equal-Appearing Intervals
Comparative Scaling
A test is designed so that agreeing with the most extreme statement implies agreement with less extreme ones. This is an example of:
Guttman Scale
Likert Scale
Categorical Scaling
Rating Scales
A developer conducts scalogram analysis to confirm that responses follow a hierarchical pattern of agreement. This analysis is applied to which type of scale?
Likert Scale
Guttman Scale
Stanine Scale
Categorical Scaling
A multiple-choice test includes questions with a stem, correct answer, and distractors. This represents which item format?
Constructed-Response Format
Selected-Response Format
Ipsative Scoring
Rating Scale
A teacher asks students to complete sentences about a historical event. This represents which item format?
Constructed-Response Format
Selected-Response Format
Essay Item
Matching Type
A diagnostic tool categorizes test-takers into groups such as "at risk," "average," and "above average" based on their response patterns. What scoring model does this represent?
Cumulative Scoring
Class Scoring
Ipsative Scoring
Norm Scoring
A clinical psychologist uses a career interest inventory where a participant's responses are used to highlight their strengths relative to their own responses across domains. What scoring model is this?
Cumulative Scoring
Class Scoring
Ipsative Scoring
Norm Scoring
A self-report inventory helps individuals identify their dominant personality traits by comparing their answers across different sections of the test. This scoring model is:
Cumulative Scoring
Class Scoring
Ipsative Scoring
Norm Scoring
In a national standardized test, the total number of correct responses determines the test-taker's percentile rank. This scoring model is best described as:
Cumulative Scoring
Class Scoring
Ipsative Scoring
Norm Scoring
A survey on job satisfaction assigns numerical values to each response, and these values are summed to provide an overall satisfaction score. This scoring method is:
Cumulative Scoring
Class Scoring
Ipsative Scoring
Norm Scoring
A test is designed to group respondents into categories like "introverted" or "extroverted" based on how they answer specific sets of questions. Which scoring model is being used?
Cumulative Scoring
Class Scoring
Ipsative Scoring
Norm Scoring
A manager selects items from a database of questions organized by topic to create a test. This represents:
Item Banks
Test Items
Test Construction
Item Branching
An adaptive test directs a test-taker to more challenging questions after correct responses to earlier items. This process is known as:
Cumulative Scoring
Item Branching
Comparative Scaling
Guttman Scale
A psychologist administers a questionnaire with the following instructions and items: What type of scale is being used in this example?
Rating Scale
Stanine Scale
Likert Scale
Guttman Scale
A survey asks participants to arrange the following values in order of importance to them: What type of scale format does this represent?
Rating Scales
Comparative Scaling
Likert Scale
Guttman Scale
A researcher administers the following scale on environmental conservation to a group of participants: During analysis, the researcher finds a participant who agrees with statements 1, 2, 3, and 5 but disagrees with statement 4. What does this indicate about the participant’s responses?
The participant’s responses fit the Guttman scale’s structure.
The participant’s responses did not fit the Guttman scale’s structure.
The participant demonstrates a clear preference for environmental advocacy.
The participant's responses suggest the scale items are not hierarchical.
A psychologist administers a draft version of a newly developed test to a sample group. What phase of test development does this represent?
Test Tryout
Item Analysis
Test Revision
Test Construction
A test is administered to a sample that closely resembles the demographics of the intended users. This reflects:
Matching Real Test Conditions
Target Population Similarity
Quality Control
Rule of Thumb Sample Size
To ensure accurate results during a trial, a test developer administers the test in an environment similar to actual conditions. This step is called:
Target Population Similarity
Rule of Thumb Sample Size
Matching Real Test Conditions
Item Validity Index
A developer creates a test with 30 items. As a rule of thumb, there sample size for trying out the items should be at least?
30
60
100
150
The primary goal of item analysis is to:
Compare test-takers' scores with others
Assess individual item performance
Identify the most difficult test items
Ensure the test is culturally unbiased
An item on a vocabulary test is answered correctly by 60% of participants. This describes the:
Item-Discrimination Index
Item-Difficulty Index
Factor Analysis Result
Item-Endorsement Index
What is the optimal difficulty level for our current exam?
0.25
0.50
0.63
0.75
An item on a personality test is endorsed by 80% of test-takers. This is an example of:
Item-Alternative Analysis
Item-Endorsement Index
Discrimination
Item-Characteristic Curve
A researcher performs statistical analysis to identify underlying constructs within test items. This method is called:
Factor Analysis
Item Discrimination
Think-Aloud Technique
Inter-Item Consistency
Test developers ensure that items within a test measure the same construct. This involves assessing:
Item-Endorsement Index
Inter-Item Consistency
Test Tryout Conditions
Item-Difficulty Index
Which index evaluates how well an item predicts performance on the entire test?
Item-Validity Index
Item-Discrimination Index
Item-Difficulty Index
Item-Endorsement Index
The ability of an item to differentiate between high and low performers on a test is measured by the:
Item-Discrimination Index
Item-Difficulty Index
Inter-Item Consistency
Item-Endorsement Index
Study and analyze the table: Which item is the least effective in distinguishing high and low scorers?
Item 1
Item 2
Item 3
All items are effective
Bonus: "Success is not the key to happiness. Happiness is the key to success. If you love what you are doing, you will be successful."
Option A
Option B
Option C
None of the options
Study and analyze the table: Which option is functioning as the best distractor?
Option A
Option B
Option C
None of the options
A test developer observes the following item-characteristic curve (ICC) for a test item: At lower ability levels, the probability of a correct response is close to 0. As ability increases, the probability of a correct response gradually rises to near 1. What does this curve indicate about the item?
The item does not differentiate between ability levels.
The item is highly effective in differentiating ability levels.
The item has low difficulty.
The item needs revision.
During a test tryout, participants verbalize their thoughts while answering each item. What technique is being used?
Think-Aloud Technique
Factor Analysis
ICC Analysis
Inter-Item Consistency Check
A math test item is answered correctly by 75 out of 100 students. Calculate the item-difficulty index.
0.65
0.75
0.85
0.95
The following table shows the responses for a test item: What is the discrimination index of the item?
0.20
0.40
0.60
0.80
What does a discrimination index value of 0.40 indicate about an item's performance?
The item has no discrimination and is equally answered correctly by high and low performers.
The item has poor discrimination and does not distinguish between high and low performers.
The item has moderate discrimination, effectively differentiating between high and low performers.
The item has excellent discrimination, perfectly distinguishing between high and low performers.
During test revision, items that fail to differentiate between high and low performers are removed. What key action is being performed?
Eliminating Weak Items
Using Large Item Pools
Balancing Item Characteristics
Ensuring Theoretical Advancements
A test developer adjusts the proportion of easy, moderate, and difficult items to ensure the test measures abilities across a wide range. What is this process called?
Using Large Item Pools
Eliminating Weak Items
Balancing Item Characteristics
Ensuring Theoretical Advancements
A vocabulary test includes words that are rarely used in modern language, making it difficult for contemporary test-takers. What reason for test revision does this represent?
Outdated Stimulus Materials
Dated Vocabulary
Cultural Sensitivity
Theoretical Advancements
Certain test items are revised to remove language that may be offensive or irrelevant to diverse cultural groups. What is the primary reason for this revision?
Improving Reliability and Validity
Ensuring Theoretical Advancements
Addressing Cultural Sensitivity
Dated Vocabulary
Norms for an intelligence test are updated to reflect the performance of the current population. What reason for revision is being addressed?
Outdated Stimulus Materials
Improving Reliability and Validity
Addressing Inadequate Norms
Cross-Validation
A researcher revalidates a test with a new sample to confirm its generalizability. What method is being used?
Co-Validation
Improving Validity
Anchor Protocol
Cross-Validation
Two intelligence tests are validated simultaneously on the same sample to reduce costs and improve comparability. What is this process called?
Co-Validation
Cross-Validation
Using Large Item Pools
Balancing Item Characteristics
A scoring protocol is developed to maintain consistency in evaluating test responses over time. What quality assurance mechanism is this?
Resolver Role
Anchor Protocol
Data Entry Checks
Cross-Validation
A designated scorer resolves disagreements between two reviewers during the evaluation of test responses. What is this role called?
Resolver Role
Anchor Protocol
Data Entry Role
Co-Validation Specialist
A test developer identifies items that perform poorly in one cultural group despite strong performance in others. What process is being used?
Cross-Validation
Addressing Dated Vocabulary
Improving Reliability
DIF Analysis
