PSYCHOLOGY • PERSONALITY & INDIVIDUAL DIFFERENCES

Intelligence & Measurement — I can describe how intelligence is defined and measured and identify limitations of IQ tests at a basic level.

Explore how psychologists define, measure, and debate the nature of human intelligence.

Historical Context & Motivation

For centuries, people have noticed that individuals differ in how quickly they learn, how well they solve problems, and how effectively they adapt to new situations. But it wasn't until the late 1800s that scientists began asking whether these differences could actually be measured in a systematic way. The quest to define and quantify intelligence — a person's capacity for reasoning, learning, and problem-solving — has shaped modern psychology and sparked ongoing debates about fairness, culture, and what it truly means to be "smart."

1905
Binet-Simon Scale
French psychologist Alfred Binet and Théodore Simon created the first practical intelligence test, designed to identify schoolchildren who needed extra academic support in Paris.
1916
Stanford-Binet IQ Test
Lewis Terman of Stanford University adapted Binet's test for American use and introduced the term Intelligence Quotient (IQ), calculated by comparing mental age to chronological age.
1939
Wechsler-Bellevue Intelligence Scale
David Wechsler developed a new test that measured multiple subtypes of intelligence and introduced deviation IQ, comparing a person's score to their same-age peers rather than using mental age. This test was later revised into the Wechsler intelligence scales still used today.
1983
Multiple Intelligences
Howard Gardner proposed that intelligence is not a single ability but a set of multiple intelligences — including musical, bodily-kinesthetic, and interpersonal — challenging the traditional single-score IQ model.
1985
Triarchic Theory
Robert Sternberg proposed three aspects of intelligence — analytical, creative, and practical — arguing that traditional IQ tests capture only analytical intelligence.

This history raises a central question that psychologists still wrestle with today: Is intelligence a single, measurable trait, or is it a collection of different abilities? And if we can measure it, how do we ensure our tools are fair, accurate, and meaningful for everyone?

Core Principles & Definitions

Before diving into how intelligence is measured, you need to understand several foundational ideas that psychologists use to frame the concept. These principles represent different perspectives on what intelligence actually is and how it can be studied scientifically.

1

General Intelligence (g factor)

Psychologist Charles Spearman proposed that a single underlying factor — called g — influences performance across all cognitive tasks. People who score well on one type of mental test tend to score well on others, suggesting a common intellectual capacity.
2

Fluid vs. Crystallized Intelligence

Fluid intelligence (Gf) is the ability to solve novel problems and think abstractly, independent of prior knowledge. Crystallized intelligence (Gc) is accumulated knowledge and skills gained through education and experience. Raymond Cattell distinguished these two dimensions in the 1960s.
3

Reliability & Validity

A good intelligence test must be reliable (producing consistent results over time) and valid (actually measuring what it claims to measure). Without both, test scores are meaningless.
4

Standardization

Standardization means administering a test to a large, representative sample to establish norms — typical score ranges for different age groups. This allows an individual's score to be compared meaningfully to the broader population.
5

Nature vs. Nurture

Intelligence is influenced by both genetics (heredity) and environment (education, nutrition, socioeconomic factors). Twin studies suggest that about 50–80% of intelligence variation is heritable, but environment plays a significant role in shaping outcomes.
KEY TAKEAWAY
Think of intelligence like athletic ability. Some people have a natural aptitude for sports in general (like the g factor), but athleticism also breaks down into specific skills — speed, endurance, flexibility — just as intelligence includes different capacities like fluid and crystallized abilities. And just like athletic performance depends on both natural talent and training, intelligence reflects both nature and nurture.

Theories of Intelligence — Visual Overview

This diagram compares three major theories of intelligence. On the left, Spearman's model shows a single general factor (g) that branches into specific abilities. In the center, Sternberg's triarchic theory divides intelligence into analytical, creative, and practical components. On the right, Gardner's model proposes eight distinct, independent intelligences.

Each of these theories reflects a fundamentally different assumption about the nature of intelligence. Spearman believed that all mental abilities share a common core, which is why people who excel at math often also excel at language — they have high g. Sternberg challenged this by arguing that street smarts and creative thinking matter just as much as book smarts but are ignored by traditional tests. Gardner went even further, suggesting that a gifted dancer or a skilled musician possesses a form of intelligence that IQ tests completely miss.

How Intelligence Is Measured

Intelligence tests attempt to quantify cognitive ability using standardized procedures. The original approach, developed by Binet and refined by Terman, used the concept of mental age — the level of intellectual performance typical for a given chronological age. A child who solved problems at the level of an average 12-year-old, regardless of their actual age, was said to have a mental age of 12.

ORIGINAL IQ FORMULA (RATIO IQ)
IQ = (Mental Age ÷ Chronological Age) × 100
Mental Age (MA) = the age group whose average performance matches the test-taker's score. Chronological Age (CA) = the person's actual age. A score of 100 means a person performs exactly at the average for their age.

This ratio method worked reasonably well for children, but it broke down for adults because mental development plateaus in early adulthood. If a 40-year-old performs at the same level as a 20-year-old, the ratio formula would unfairly give them a low score. That's why modern tests use deviation IQ, introduced by David Wechsler.

MODERN IQ FORMULA (DEVIATION IQ)
IQ = 100 + 15 × ((X − μ) ÷ σ)
X = the individual's raw test score. μ (mu) = the mean score for the person's age group. σ (sigma) = the standard deviation of scores in that age group. The result is a score centered at 100 with a standard deviation of 15.

Under the deviation IQ system, scores follow a normal distribution (bell curve). About 68% of the population scores between 85 and 115, about 95% falls between 70 and 130, and scores above 130 or below 70 are quite rare. This distribution allows psychologists to determine where any individual falls relative to the rest of the population.

💡 Why 100?
The average IQ score is set to 100 by design. It doesn't mean someone answered 100 questions correctly. It simply means the person scored at the exact average for their age group. Scores above 100 indicate above-average performance; scores below 100 indicate below-average performance.

Major Intelligence Tests & Score Distribution

Several widely used intelligence tests exist today, each designed for different populations and purposes. The two most prominent families are the Wechsler scales and the Stanford-Binet. Understanding what these tests measure — and what they don't — is essential for interpreting IQ scores responsibly.

Major intelligence tests used in psychological assessment
TestAge RangeSubtests / Areas MeasuredKey Feature
WAIS-IVAges 16–90Verbal Comprehension, Perceptual Reasoning, Working Memory, Processing SpeedMost widely used adult intelligence test
WISC-VAges 6–16Same four areas as WAIS plus Fluid ReasoningUsed to identify learning disabilities and giftedness in children
Stanford-Binet 5Ages 2–85Fluid Reasoning, Knowledge, Quantitative, Visual-Spatial, Working MemoryBroadest age range; evolved from Binet's original test
Raven's Progressive MatricesAges 5–adultNonverbal pattern recognition (fluid intelligence)Culture-fair; does not rely on language or specific knowledge
IQ scores follow a normal distribution with a mean of 100 and a standard deviation of 15. The shaded regions show that 68% of people score between 85 and 115 (within one standard deviation), and 95% score between 70 and 130 (within two standard deviations). Scores at the extremes — below 70 or above 130 — represent a small fraction of the population.

The bell curve is important because it shows that most people cluster around the average, and extreme scores are rare. This statistical pattern helps psychologists identify individuals who may benefit from specialized services — either those who score very low and may need additional support, or those who score very high and may benefit from advanced educational programs.

Worked Example — Calculating & Interpreting IQ

Let's walk through two examples: first using the original ratio IQ formula, then using the modern deviation IQ formula. These will help you understand how IQ scores are calculated and what they actually mean.

Example A — Ratio IQ (Historical Method)
1
Step 1 — Identify Given ValuesA 10-year-old child takes the Stanford-Binet test and solves problems at the level of a typical 12-year-old. So, Mental Age (MA) = 12 and Chronological Age (CA) = 10.
2
Step 2 — Apply the Ratio IQ FormulaIQ = (MA ÷ CA) × 100 = (12 ÷ 10) × 100 = 1.2 × 100
IQ = 120
3
Step 3 — Interpret the ScoreA score of 120 means this child's cognitive performance is above average for their age. Since the average is 100 and the standard deviation on the modern scale is 15, a score of 120 is 20 points above the mean, which works out to about 1.33 standard deviations above average (20 ÷ 15 ≈ 1.33). This child is performing at a level typical of someone two years older.
Example B — Deviation IQ (Modern Method)
1
Step 1 — Identify Given ValuesA 25-year-old takes the WAIS-IV. Her raw score (X) is 112 points. The mean (μ) for her age group is 100, and the standard deviation (σ) is 10 (based on the raw score distribution for this test).
2
Step 2 — Calculate the Z-ScoreZ = (X − μ) ÷ σ = (112 − 100) ÷ 10 = 12 ÷ 10 = 1.2
Z-score = 1.2
3
Step 3 — Convert to Deviation IQIQ = 100 + 15 × Z = 100 + 15 × 1.2 = 100 + 18
IQ = 118
4
Step 4 — Interpret the ScoreAn IQ of 118 places this person above average — about 1.2 standard deviations above the mean. Looking at the bell curve, she scores higher than approximately 88% of people her age. This falls within the "high average" to "superior" range.

Strengths & Limitations of IQ Tests

IQ tests are among the most rigorously developed tools in psychology, but they are far from perfect. Understanding both their strengths and their limitations is crucial for using them responsibly. A single IQ score should never be treated as a complete picture of a person's abilities or potential.

Strengths and limitations of IQ testing
StrengthsLimitations
High reliability: IQ scores tend to be consistent when the same person is tested repeatedly over time.Cultural bias: Test questions may reflect the knowledge and values of the dominant culture, putting people from different cultural backgrounds at a disadvantage.
Predictive validity: IQ scores moderately predict academic achievement, job performance, and income.Narrow scope: Traditional IQ tests primarily measure analytical and verbal abilities, missing creativity, emotional intelligence, practical skills, and motivation.
Standardization: Tests are administered under consistent conditions with well-established norms, enabling fair comparisons.Stereotype threat: Awareness of negative stereotypes about one's group can cause anxiety that lowers test performance, skewing results.
Diagnostic utility: Helps identify intellectual disabilities and giftedness, guiding educational placements and interventions.Socioeconomic factors: Access to quality education, nutrition, and enrichment activities affects scores, meaning IQ partially reflects opportunity rather than innate ability.
Research foundation: Decades of scientific study support the statistical properties and usefulness of well-designed IQ tests.Historical misuse: IQ tests have been used to justify eugenics, discriminatory immigration policies, and racial segregation — serious ethical concerns.
KEY TAKEAWAY
Think of an IQ test like a bathroom scale. A scale reliably measures weight, and weight tells you something real about a person's body. But weight alone doesn't capture fitness, health, diet quality, or muscle-to-fat ratio. Similarly, an IQ score measures one dimension of cognitive ability but doesn't capture creativity, emotional intelligence, motivation, or the full complexity of what it means to be intelligent.

Beyond IQ — Emotional Intelligence & the Flynn Effect

Modern psychology has expanded the concept of intelligence well beyond what traditional IQ tests measure. Two particularly important developments are emotional intelligence (EQ) and the Flynn Effect. These concepts challenge us to reconsider what intelligence means and whether it's fixed or changeable over time.

Traditional vs. expanded views of intelligence
ConceptTraditional IQ ViewExpanded / Advanced View
What counts as intelligence?Logical reasoning, verbal ability, processing speed, and working memory.Also includes emotional awareness, social skills, creativity, and practical problem-solving (EQ, Gardner, Sternberg).
Is intelligence fixed?Largely stable after childhood; IQ scores tend to remain consistent over a lifetime.The Flynn Effect shows average IQ scores have risen about 3 points per decade worldwide, suggesting environmental factors can shift intelligence.
Single score or multiple?One overall IQ score (sometimes with subscale scores like verbal and performance).Multiple distinct forms of intelligence that may be relatively independent from each other.
Role of cultureTests are standardized within a culture but may not translate well across cultures.Intelligence is partly culturally defined — what counts as 'smart' varies by society and context.

The Flynn Effect is particularly fascinating because it shows that average IQ scores have been rising steadily across generations since testing began. Researchers attribute this to better nutrition, improved education, more cognitively stimulating environments, and greater familiarity with testing. This finding suggests that intelligence, as measured by IQ tests, is not entirely fixed by genetics — the environment matters a great deal.

🔮 Looking Ahead
In AP Psychology and college-level courses, you'll explore concepts like factor analysis (the statistical method Spearman used to discover g), heritability studies using twin and adoption data, and the ethical debates around group differences in IQ scores. These topics build directly on the foundations covered in this lesson.

Practice Problems

PROBLEM 1CONCEPTUAL
Explain the difference between Spearman's g factor and Gardner's theory of multiple intelligences. Which view does a traditional IQ test best reflect, and why?
PROBLEM 2BASIC CALCULATION
An 8-year-old child takes an intelligence test and achieves a mental age of 10. Using the original ratio IQ formula, calculate this child's IQ score. What does this score indicate about their cognitive ability?
PROBLEM 3INTERMEDIATE
A student scores 78 on a raw intelligence test. The mean for her age group is 70, and the standard deviation is 8. Using the deviation IQ formula (IQ = 100 + 15 × ((X − μ) ÷ σ)), calculate her IQ. Then determine whether her score falls within the "average" range (85–115) on the bell curve.
PROBLEM 4APPLIED
A school district decides to use a single IQ test to place all students into either "gifted," "regular," or "remedial" tracks. The district has a diverse student body with students from many different cultural and socioeconomic backgrounds. Identify at least three specific concerns a psychologist might raise about this policy, using concepts from the lesson.
PROBLEM 5CRITICAL THINKING
The Flynn Effect shows that average IQ scores have been rising about 3 points per decade worldwide. If this trend reflects genuine increases in intelligence, what implications does that have for the nature vs. nurture debate? If it does not reflect genuine increases, what does that suggest about the limitations of IQ tests? Argue both sides.

Lesson Summary

Intelligence is a complex and debated concept in psychology, broadly defined as the capacity for reasoning, learning, and adapting to new situations. Spearman's g factor proposed a single general intelligence underlying all mental abilities, while Gardner's multiple intelligences and Sternberg's triarchic theory argued for multiple, distinct cognitive abilities. Fluid intelligence (novel problem-solving) and crystallized intelligence (accumulated knowledge) represent two key dimensions that develop differently across the lifespan.

Intelligence is measured using standardized tests such as the WAIS and Stanford-Binet, which produce IQ scores following a normal distribution with a mean of 100 and standard deviation of 15. While IQ tests are reliable and valid for certain purposes, they have significant limitations including cultural bias, stereotype threat, narrow scope, and susceptibility to socioeconomic influence. The Flynn Effect — the steady rise in average IQ scores over generations — reminds us that intelligence is shaped by both genetics and environment, and that no single test can capture the full spectrum of human cognitive ability.

Varsity Tutors • Psychology • Intelligence & Measurement