U2.8 Intelligence and Achievement
Master AP Psych 2.8: compare Spearman's g, Sternberg, Gardner, and emotional intelligence, plus test reliability, validity, standardization, and heritability.
What you'll do in this lesson
A voice-first session with the Crimsora tutor on U2.8 Intelligence and Achievement, then targeted practice and FRQs — with the tutor adapting to where you get stuck.
What this lesson covers
Theories of Intelligence
Other theorists broke intelligence apart. Robert Sternberg proposed the triarchic theory, dividing intelligence into three types: analytical (problem-solving, the kind schools test), creative (adapting to novel situations, generating new ideas), and practical ("street smarts," handling everyday tasks). Howard Gardner went further with multiple intelligences, proposing relatively independent domains such as linguistic, logical-mathematical, spatial, musical, bodily-kinesthetic, interpersonal, intrapersonal, and naturalist. Gardner pointed to savant syndrome and brain damage that spares some abilities while destroying others as evidence.
Emotional intelligence (Salovey, Mayer, popularized by Goleman) is the ability to perceive, understand, manage, and use emotions. It predicts social success but is controversial as a form of "intelligence."
| Theory | Key idea |
|---|---|
| Spearman | One general factor underlies all abilities |
| Sternberg | Three types: analytical, creative, practical |
| Gardner | Many independent intelligences |
| Emotional intelligence | Perceiving and managing emotions |
History of Intelligence Testing
Lewis Terman at Stanford adapted Binet's work for American use, producing the Stanford-Binet test. This context introduced the intelligence quotient, originally calculated as . Modern tests no longer use this formula; instead they compare scores to age peers.
David Wechsler developed the widely used Wechsler Adult Intelligence Scale (WAIS) and versions for children, providing an overall score plus subscores for verbal comprehension, perceptual reasoning, working memory, and processing speed.
Be careful to distinguish achievement tests, which measure what you have already learned (like a history final), from aptitude tests, which attempt to predict future performance or capacity to learn. The exam frequently tests this difference. A misconception is that IQ is a fixed, precise measure of worth — it is a statistical comparison to a norm group, not a fixed biological quantity.
Test Properties: Standardization, Reliability, Validity
Reliability means the test yields consistent results. Types include test-retest reliability (same score on repeat administrations), split-half reliability (two halves of the test agree), and inter-rater reliability (different scorers agree).
Validity means the test measures what it claims to measure. Content validity covers whether items sample the relevant domain; predictive (criterion) validity covers whether scores forecast the outcome they should — for example, whether an aptitude test predicts college grades.
| Property | Question it answers |
|---|---|
| Standardization | Compared to whom? Are there norms? |
| Reliability | Are results consistent? |
| Validity | Does it measure the right thing? |
Heritability and Group Differences
Twin and adoption studies suggest intelligence is substantially heritable — identical twins raised apart still show correlated scores. But environment matters enormously: enriched or impoverished environments, nutrition, education, and stress all shape measured intelligence.
Critically, heritability within groups says nothing about the causes of differences between groups. Two groups raised in systematically different environments can differ for entirely environmental reasons even if the trait is heritable within each. Score gaps between groups have historically been influenced by unequal access to education, test bias, and stereotype threat — the self-fulfilling anxiety that arises when a person fears confirming a negative stereotype about their group, which can lower performance.
A test can show test bias if it disadvantages a group due to cultural content unrelated to the ability being measured. The exam expects you to recognize that group differences in scores are not evidence of genetic differences between groups.
Key terms
- General intelligence ().
- Spearman's proposed single underlying factor that accounts for correlated performance across different mental tasks, identified through factor analysis.
- Triarchic theory.
- Sternberg's model dividing intelligence into analytical, creative, and practical abilities.
- Multiple intelligences.
- Gardner's theory that intelligence consists of several independent domains such as linguistic, spatial, musical, and interpersonal.
- Standardization.
- Establishing test norms by administering it to a representative sample so individual scores can be compared.
- Reliability.
- The consistency of a test's results across time, forms, or scorers.
- Validity.
- The extent to which a test measures or predicts what it is intended to measure.
- Heritability.
- The proportion of variation in a trait within a population attributable to genetic differences; not a statement about a single individual.
- Stereotype threat.
- Performance decline caused by anxiety about confirming a negative stereotype about one's group.
Worked example
Next, evaluate validity. Validity asks whether the test measures what it claims to measure. Here the scores fail to correlate with grades or with other reasoning measures. That means the test lacks predictive (criterion) validity and likely lacks construct validity — it is not measuring reasoning ability as intended.
Finally, connect the two. This case illustrates the exam's favorite principle: a test can be reliable without being valid. Consistency alone does not guarantee that a test measures the right thing. The bathroom-scale analogy applies — a scale reading 10 pounds heavy every time is reliable but invalid. The psychologist should not use this test to make decisions about students, because its consistent scores do not reflect the trait it claims to assess.
Practice questions
A researcher claims her intelligence test is highly reliable but critics argue it is not valid. Which finding would best support the critics' claim?
- Students receive nearly the same score when they retake the test
- The two halves of the test produce closely matching scores
- Test scores fail to predict any real-world academic or cognitive outcomes
- Different examiners scoring the test arrive at the same results
Answer: Test scores fail to predict any real-world academic or cognitive outcomes
Compare Spearman's concept of general intelligence with Gardner's theory of multiple intelligences, and explain one type of evidence Gardner used to support his view.
Answer: Spearman argued that a single general factor, , underlies performance across all mental tasks because scores on different abilities tend to correlate. Gardner rejected a single factor, proposing several relatively independent intelligences (such as linguistic, spatial, and musical). Gardner supported his view with evidence like savant syndrome and cases of brain damage, where one ability can be preserved or destroyed independently of others.
A standardized IQ test administered decades ago is re-normed today, and researchers notice average raw performance has risen over time. What phenomenon does this illustrate, and why does it require re-standardization?
Answer: This illustrates the Flynn effect, the observed rise in average intelligence test performance over the 20th century. Re-standardization is required because norms are based on how a representative sample performs; if overall performance rises, old norms would make today's average person appear above average, distorting comparisons. Updating the standardization sample keeps the mean anchored appropriately.
FAQ
- What is the difference between reliability and validity?
- Reliability is consistency — getting the same result repeatedly. Validity is accuracy — measuring what you actually intend to measure. A test can be reliable without being valid, like a scale that consistently reads 10 pounds too high. But a valid test must be reliable, so reliability is necessary but not sufficient for validity.
- Does high heritability mean intelligence is mostly genetic and can't change?
- No. Heritability describes how much of the variation among people in a population is linked to genes, not how fixed a trait is or how much of one person's intelligence is genetic. Traits with high heritability can still be strongly shaped by environment, and environmental improvements can raise scores.
- What's the difference between an aptitude test and an achievement test?
- An achievement test measures what you have already learned, like a final exam. An aptitude test attempts to predict your future performance or capacity to learn. The distinction can blur in practice, but on the exam, focus on past learning versus future prediction.
- Why do group differences in test scores not prove genetic differences between groups?
- Heritability within a group says nothing about causes of differences between groups. Groups raised in systematically different environments — unequal schooling, nutrition, or facing stereotype threat and test bias — can differ for purely environmental reasons, even when a trait is heritable within each group.
Learn this with a teacher, not a page
The Crimsora tutor teaches U2.8 Intelligence and Achievement live — explaining on a whiteboard, asking you questions, and adapting to where you get stuck.