- In an experiment, how can researchers ensure that there are no biases when they assign participants to groups?
- Do researchers usually measure at two different points in time to estimate test-retest reliability?
- Does alpha estimate the consistency of scores over time?
- Do researchers usually measure at two different points in time to estimate interobserver reliability?
- Researchers need to use at least how many observers to determine interobserver reliability?
- What is the highest possible value for a split-half reliability coefficient?
- A test is said to be reliable if it yields what?
- In Table 1, which two types of validity are classified as solely empirical?
- Should a researcher expect a very short multiple-choice test to be highly reliable?
- Professional test makers tend to be more successful in achieving high validity than in achieving high reliability. Is this statement "true" or "false"?
- When there are two quantitative scores per participant, researchers can compute what statistic to describe reliability?
- Which type of validity is based on superficial inspection?
- Should researchers consider the types of skills required by achievement test items when judging content validity?
- If a test has no validity whatsoever, what value will its validity coefficient have?
- Overall, is "validity" or "reliability" more important when evaluating a measure?
- According to this topic, most published tests have reliability coefficients that are about how high?
- Does the split-half method require "one" or "two" administrations of a test?
- If a test is perfectly valid, what value will its validity coefficient have?
- If a researcher collects the criterion data at about the same time the test is being administered, he or she is examining what type of empirical validity?
- Does confirming a hypothesis in a construct validity study offer "direct" or "indirect" evidence on the validity of a test?