Practice question · Multiple choice
A depression questionnaire has excellent internal consistency and excellent test-retest reliability. A critic argues it may still be measuring the wrong thing. Why do the reliability statistics fail to answer that objection?
Hints
- What is each statistic comparing to what? Neither comparison involves depression itself.
- Imagine a scale whose twenty items all ask about sleep. How would it score?
Show the answer
C. Because both compare the measure to itself, not to the construct.
Why
Both statistics are internal to the instrument, the items agree with each other, and the scale agrees with itself a fortnight later. Twenty items about sleep would be beautifully consistent, stable, and correlated with depression, without measuring it. Construct validity needs outside evidence: convergence, discrimination, and movement when treatment works.
Practise Reliability and Validity
The app has 6 more questions on this lesson, and keeps your place in the course. Psychology I is free to start.
More questions on Reliability and Validity
- A scale reading 73.0, 73.1, 72.9 and 73.0 for a certified 70 kg weight is reliable but not valid.
- A bathroom scale reading 3 kg heavy every time is perfectly reliable and not valid at all. Why can…
- Select every statement that is TRUE.
- Before you can study happiness you must decide what counts as happiness, a rating scale, hours smiling, a…
- A new questionnaire has just been written. Order the checks a psychometrician runs on it, from the one that…