Psychology Quiz: Intelligence And Measurement
20 questions · exam conditions
0:00
Intelligence And MeasurementQuestion 1 of 20

A common circular argument in the history of intelligence measurement is that 'intelligence is whatever intelligence tests measure.' While this statement highlights a genuine challenge, what is the most significant scientific problem with using this as a definition?

It is an operational definition that is not grounded in a clear theoretical construct.
It prevents intelligence tests from ever being reliable or consistent.
It guarantees that intelligence tests will be inherently biased against some groups.
It contradicts the findings of factor analysis, which suggest a single g factor.
← Back to quizzes

Psychology Quiz

Psychology Quiz: Intelligence And Measurement

Practice Intelligence And Measurement in Psychology with focused quiz questions that help you check what you know, review explanations, and build confidence with test-style prompts.

What this quiz covers

This quiz focuses on Intelligence And Measurement, giving you a quick way to practice the rules, question types, and explanations that matter most for Psychology.

How to use this quiz

Try each quiz question before looking at the correct answer. Use the explanations to review missed ideas, then come back to similar questions until the pattern feels familiar.

All questions

Question 1

A common circular argument in the history of intelligence measurement is that 'intelligence is whatever intelligence tests measure.' While this statement highlights a genuine challenge, what is the most significant scientific problem with using this as a definition?

  1. It is an operational definition that is not grounded in a clear theoretical construct. (correct answer)
  2. It prevents intelligence tests from ever being reliable or consistent.
  3. It guarantees that intelligence tests will be inherently biased against some groups.
  4. It contradicts the findings of factor analysis, which suggest a single g factor.
Explanation: This statement is a purely operational definition. Good scientific measurement requires an operational definition (how you measure something) to be linked to a theoretical construct (what you are conceptually trying to measure). Without a solid theory of what intelligence is, it is difficult to evaluate whether the test truly measures it (i.e., to assess its construct validity). The definition doesn't inherently prevent reliability (B) or guarantee bias (C), and it doesn't contradict factor analysis (D), which is the method used to create the measurement in the first place.

Question 2

A common circular argument in the history of intelligence measurement is that 'intelligence is whatever intelligence tests measure.' While this statement highlights a genuine challenge, what is the most significant scientific problem with using this as a definition?

  1. It is an operational definition that is not grounded in a clear theoretical construct. (correct answer)
  2. It prevents intelligence tests from ever being reliable or consistent.
  3. It guarantees that intelligence tests will be inherently biased against some groups.
  4. It contradicts the findings of factor analysis, which suggest a single g factor.
Explanation: This statement is a purely operational definition. Good scientific measurement requires an operational definition (how you measure something) to be linked to a theoretical construct (what you are conceptually trying to measure). Without a solid theory of what intelligence is, it is difficult to evaluate whether the test truly measures it (i.e., to assess its construct validity). The definition doesn't inherently prevent reliability (B) or guarantee bias (C), and it doesn't contradict factor analysis (D), which is the method used to create the measurement in the first place.

Question 3

During the development of a new aptitude test, an item is flagged for review. Analysis shows that test-takers who scored high on the overall test were just as likely to get this particular item wrong as test-takers who scored low on the overall test. From a psychometric standpoint, what is the most significant problem with this item?

  1. The item is too difficult for the target population.
  2. The item has a near-zero item-total correlation. (correct answer)
  3. The item is likely to be culturally biased against a specific subgroup.
  4. The item lowers the test's overall test-retest reliability.
Explanation: A core principle of test construction is that individual items should discriminate between those with high and low levels of the trait being measured. This is often assessed by the item-total correlation. An item that high-scorers and low-scorers answer correctly at similar rates has a low or zero correlation with the total score, meaning it fails to discriminate and is considered a poor item. While it might be difficult or biased, the data given directly points to the problem of poor discrimination.

Question 4

An individual undergoing assessment with the Wechsler Adult Intelligence Scale (WAIS-IV) obtains a very high score on the Verbal Comprehension Index but a significantly lower score on the Perceptual Reasoning Index. Which of the following is the most plausible interpretation of this pattern?

  1. The individual likely has a specific learning disability in reading or language.
  2. The test has low internal consistency reliability for this individual.
  3. The individual has strong crystallized intelligence but may have relative weaknesses in fluid intelligence. (correct answer)
  4. The results indicate a high general intelligence factor (g) with minor, insignificant fluctuations.
Explanation: The Verbal Comprehension Index heavily relies on accumulated knowledge, vocabulary, and verbal reasoning, which are hallmarks of crystallized intelligence. The Perceptual Reasoning Index involves novel problem-solving, pattern recognition, and abstract thought, which are key aspects of fluid intelligence. A significant discrepancy between these scores often suggests a difference in the underlying abilities of crystallized versus fluid intelligence.

Question 5

A 70-year-old history professor is competing in a trivia contest against her 25-year-old graduate student. The professor excels on questions about historical facts and vocabulary, while the student is much faster at solving novel logic puzzles presented during the contest. This difference in performance most likely reflects established findings on...

  1. age-related changes, where crystallized intelligence is maintained while fluid intelligence tends to decline. (correct answer)
  2. the Flynn effect, which has resulted in superior problem-solving abilities in younger generations.
  3. a difference in their general intelligence factor (g), with the professor possessing a higher g overall.
  4. the impact of stereotype threat on the professor's confidence with novel tasks.
Explanation: This scenario perfectly illustrates the differing lifespan trajectories of crystallized and fluid intelligence. Crystallized intelligence (accumulated knowledge and facts, like history and vocabulary) is often maintained or even increases into older adulthood. Fluid intelligence (the ability to reason quickly and solve novel, abstract problems) tends to peak in young adulthood and gradually decline thereafter. The professor's strength is in crystallized intelligence, and the student's is in fluid intelligence.

Question 6

An individual is a world-renowned sculptor and can mentally rotate complex three-dimensional objects with exceptional skill. However, they struggle with abstract reasoning in mathematics and have only average verbal skills. This person's cognitive profile, characterized by a distinct peak in one area alongside average or lower abilities in others, is most effectively explained by which theory of intelligence?

  1. Charles Spearman's concept of a general intelligence factor (g).
  2. Howard Gardner's theory of multiple intelligences. (correct answer)
  3. Lewis Terman's longitudinal studies on giftedness.
  4. Robert Sternberg's concept of analytical intelligence.
Explanation: Howard Gardner's theory posits the existence of multiple, relatively independent intelligences, including spatial intelligence, which would be exceptionally high in a sculptor. This theory readily accounts for a 'jagged' cognitive profile where an individual can be a genius in one domain but average in others. Spearman's g would predict more correlated abilities, Terman studied giftedness but didn't propose this type of model, and Sternberg's analytical intelligence relates to the academic skills where this individual is not gifted.

Question 7

While IQ scores are well-established as being predictive of academic achievement, the correlation is typically in the range of .40 to .60. Which statement best describes a key limitation or nuance of the predictive validity of IQ tests suggested by this correlation value?

  1. The correlation is strong enough that IQ should be the sole factor in academic or career placement decisions.
  2. The predictive validity of IQ tests drops to nearly zero after an individual completes formal education.
  3. IQ tests predict potential but have almost no relationship with actual achieved success, which is determined by environment.
  4. IQ is only one of many factors, and non-cognitive traits like motivation and conscientiousness also significantly predict success. (correct answer)
Explanation: When you encounter questions about IQ test validity, focus on understanding what correlation coefficients actually tell us about predictive relationships. A correlation of .40 to .60 between IQ and academic achievement represents a moderate positive relationship - meaningful but far from perfect. This correlation range reveals that IQ accounts for roughly 16-36% of the variance in academic outcomes (since variance explained equals the correlation squared). This leaves 64-84% of academic success unexplained by IQ alone, pointing to the significant role of other factors. Research consistently shows that traits like conscientiousness, motivation, persistence, and emotional regulation are powerful predictors of academic and life success, often rivaling or exceeding IQ's predictive power in certain contexts. Looking at the incorrect options: (A) misinterprets the moderate correlation as justification for sole reliance on IQ, ignoring the substantial unexplained variance. (B) incorrectly claims IQ's predictive validity disappears after formal education, when research shows IQ continues to predict job performance and other outcomes throughout life. (C) creates a false dichotomy between potential and achievement, wrongly suggesting IQ has "almost no relationship" with success when the .40-.60 correlation clearly indicates a meaningful relationship. The correct answer is (D) because it accurately reflects what the correlation data shows: IQ is one important predictor among many, with non-cognitive factors playing equally crucial roles in determining success. Remember this pattern: when you see moderate correlations in psychology research, they typically indicate that multiple factors are at work, not that the measured variable is unimportant or all-determining.

Question 8

A psychologist develops a new intelligence test designed to measure fluid reasoning. For the standardization sample, the test is administered exclusively to undergraduate students at a highly selective engineering university. When this test is later used with the general population, which of the following describes the most significant limitation that will likely affect the interpretation of scores?

  1. The test will lack content validity because the sample group are not experts in psychometrics.
  2. The test will have low test-retest reliability due to the homogenous nature of the sample.
  3. The test's norms will be unrepresentative, leading to potentially deflated IQ scores for the general population. (correct answer)
  4. The test will be inherently biased against individuals with strong verbal skills because it was normed on an engineering population.
Explanation: Standardization involves administering a test to a representative sample to establish norms. The sample described is highly selective and likely has a higher-than-average fluid reasoning ability. When comparing individuals from the general population to this high-performing norm group, their scores will appear lower than they would if compared to a representative sample, resulting in deflated IQ scores. This is the most direct and significant limitation of the described methodology.

Question 9

A 10-year-old child takes an intelligence test and performs at the level of a typical 13-year-old. Using Lewis Terman's original formula for the intelligence quotient (IQ), the child's score would be 130. How would this child's performance most likely be represented on a modern test like the Wechsler Intelligence Scale for Children (WISC-V), which uses a deviation IQ?

  1. The child's score would be calculated as a ratio of their mental age (13) to chronological age (10), resulting in 130.
  2. The child's raw score would be meaningless until they could be compared to the adult standardization sample.
  3. The child would be assigned a mental age of 13, which is then converted to a percentile rank without a final IQ score.
  4. The child's score would be approximately two standard deviations above the mean for their specific age group. (correct answer)
Explanation: Modern intelligence tests, like the Wechsler scales, use a deviation IQ. Scores are standardized with a mean of 100 and a standard deviation of 15. A score of 130 is exactly two standard deviations above the mean (100 + 2*15 = 130). This method compares the child's performance to that of other 10-year-olds in the standardization sample, rather than using the concept of 'mental age' to calculate a ratio.

Question 10

A researcher creates a new test to measure 'creative intelligence.' The test requires participants to list as many uses for a paperclip as they can in two minutes. When administered to the same group of people on two separate occasions a month apart, the scores are highly correlated (r = .92). However, the scores on this test show no correlation with real-world measures of creative achievement, such as artistic awards or innovative patents. Which of the following statements best describes this new test?

  1. It has high reliability but low validity. (correct answer)
  2. It has low reliability but high validity.
  3. It has high predictive validity but low content validity.
  4. It has not been properly standardized on a representative sample.
Explanation: Reliability refers to the consistency of a measure. Since the scores are highly correlated over time, the test has high test-retest reliability. Validity refers to whether the test measures what it intends to measure. Since the scores do not correlate with real-world creative achievement, the test has low validity (specifically, low predictive or criterion validity). The information provided does not address standardization.

Question 11

A large school district has used the same version of a group intelligence test, standardized in 1995, for the past 25 years. Administrators observe that the average score of students has steadily increased from a mean of 101 in 1996 to a mean of 109 in 2021. Assuming the student population's underlying abilities have not changed dramatically, which phenomenon best accounts for this observation?

  1. Stereotype threat, as students from all backgrounds have become less anxious about testing over time.
  2. Regression to the mean, as initially low-scoring student groups have naturally improved their performance.
  3. The Flynn effect, which suggests that the test's original norms have become outdated over time. (correct answer)
  4. A high degree of predictive validity, indicating the test accurately measures scholastic aptitude.
Explanation: The Flynn effect is the observed, long-term, generation-over-generation increase in fluid and crystallized intelligence test scores. When a test with old norms is used for many years, the average score will appear to rise because the current population is being compared to a lower-scoring standardization sample from the past. The other options do not explain a population-wide, gradual increase in scores over decades.

Question 12

A researcher administers a difficult standardized math test to a group of female and male students who are all highly qualified in mathematics. In Condition 1, participants are required to indicate their gender on the test booklet before starting. In Condition 2, the gender question is omitted. Based on the concept of stereotype threat, what is the most likely outcome?

  1. Female students will score lower than male students in Condition 1 but similarly to them in Condition 2. (correct answer)
  2. Male students will score significantly higher than female students in both conditions due to inherent ability differences.
  3. Female students will score lower than male students in Condition 2 but not in Condition 1.
  4. There will be no significant score differences in either condition, as all participants are highly qualified.
Explanation: Stereotype threat is a situational predicament in which people feel themselves to be at risk of conforming to stereotypes about their social group. Making gender salient (Condition 1) can activate the stereotype that women are not as good at math, creating anxiety that depresses performance among female students. When the threat is removed (Condition 2), their performance is expected to be comparable to that of their equally qualified male peers.

Question 13

An item on an intelligence test asks, 'A symphony is to a composer as a sonnet is to a  .' The answer choices are 'musician,' 'poet,' 'sculptor,' and 'author.' This item has been criticized as being culturally biased. What is the primary reason for this criticism?

  1. The item demonstrates poor content validity because artistic knowledge is not a core component of general intelligence.
  2. The item requires specific cultural and educational knowledge that may not be equally available to all individuals. (correct answer)
  3. The item has low criterion-related validity as it does not predict success in academic settings.
  4. The vocabulary used in the question is too complex for the average test-taker in the intended age range.
Explanation: Cultural bias in testing occurs when test items draw on knowledge or experiences that are more familiar to one cultural, linguistic, or socioeconomic group than another. Knowledge about symphonies, composers, and sonnets is highly dependent on a person's educational background and cultural exposure. Therefore, the item may measure acculturation or specific knowledge rather than the intended construct of abstract reasoning (analogy).

Question 14

A manager is known for her exceptional ability to navigate complex workplace politics, understand the motivations of her colleagues, and find effective, common-sense solutions to everyday challenges. However, her performance on traditional academic tests is only average. According to Robert Sternberg's triarchic theory of intelligence, this manager demonstrates a high level of which type of intelligence?

  1. Analytical intelligence
  2. Creative intelligence
  3. Practical intelligence (correct answer)
  4. Fluid intelligence
Explanation: Robert Sternberg's triarchic theory proposes three types of intelligence. Practical intelligence involves the ability to solve real-world problems and adapt to everyday contexts, often described as 'street smarts' or business sense. This perfectly describes the manager's skills. Analytical intelligence relates to academic problem-solving, and creative intelligence involves generating novel ideas. Fluid intelligence is a concept from a different theory (Cattell-Horn-Carroll).

Question 15

A school psychologist is evaluating a new intelligence test. A committee of teachers complains that the test primarily contains abstract logic puzzles and spatial reasoning tasks but very few questions assessing vocabulary, comprehension, and general knowledge. They argue the test does not adequately sample the range of abilities relevant to school performance. This criticism is primarily questioning the test's...

  1. content validity. (correct answer)
  2. predictive validity.
  3. test-retest reliability.
  4. standardization procedures.
Explanation: Content validity refers to the extent to which a test's items are a representative sample of the entire domain the test purports to measure. The teachers' complaint is that the test's content is too narrow and does not adequately cover all facets of intelligence relevant to their curriculum. This is a direct challenge to the test's content validity.

Question 16

A 15-year-old is administered an IQ test and receives a Full Scale IQ score of 65. Based solely on this information, which of the following conclusions is most appropriate according to current diagnostic standards (e.g., DSM-5)?

  1. The individual meets the primary criterion for a diagnosis of mild intellectual disability.
  2. The diagnosis of an intellectual disability cannot be made without evidence of concurrent deficits in adaptive functioning. (correct answer)
  3. The score is likely invalid, as IQ tests are not considered reliable for individuals in their mid-teens.
  4. The individual's score falls within the 'borderline' range of intellectual functioning, not the range for intellectual disability.
Explanation: According to the DSM-5, a diagnosis of intellectual disability requires three criteria to be met: 1) deficits in intellectual functions (such as an IQ score of approximately 70 or below), 2) deficits in adaptive functioning (e.g., failure to meet developmental standards for personal independence and social responsibility), and 3) onset of these deficits during the developmental period. An IQ score alone is necessary but not sufficient for the diagnosis.

Question 17

While IQ scores are well-established as being predictive of academic achievement, the correlation is typically in the range of .40 to .60. Which statement best describes a key limitation or nuance of the predictive validity of IQ tests suggested by this correlation value?

  1. The correlation is strong enough that IQ should be the sole factor in academic or career placement decisions.
  2. The predictive validity of IQ tests drops to nearly zero after an individual completes formal education.
  3. IQ tests predict potential but have almost no relationship with actual achieved success, which is determined by environment.
  4. IQ is only one of many factors, and non-cognitive traits like motivation and conscientiousness also significantly predict success. (correct answer)
Explanation: When you encounter questions about IQ test validity, focus on understanding what correlation coefficients actually tell us about predictive relationships. A correlation of .40 to .60 between IQ and academic achievement represents a moderate positive relationship - meaningful but far from perfect. This correlation range reveals that IQ accounts for roughly 16-36% of the variance in academic outcomes (since variance explained equals the correlation squared). This leaves 64-84% of academic success unexplained by IQ alone, pointing to the significant role of other factors. Research consistently shows that traits like conscientiousness, motivation, persistence, and emotional regulation are powerful predictors of academic and life success, often rivaling or exceeding IQ's predictive power in certain contexts. Looking at the incorrect options: (A) misinterprets the moderate correlation as justification for sole reliance on IQ, ignoring the substantial unexplained variance. (B) incorrectly claims IQ's predictive validity disappears after formal education, when research shows IQ continues to predict job performance and other outcomes throughout life. (C) creates a false dichotomy between potential and achievement, wrongly suggesting IQ has "almost no relationship" with success when the .40-.60 correlation clearly indicates a meaningful relationship. The correct answer is (D) because it accurately reflects what the correlation data shows: IQ is one important predictor among many, with non-cognitive factors playing equally crucial roles in determining success. Remember this pattern: when you see moderate correlations in psychology research, they typically indicate that multiple factors are at work, not that the measured variable is unimportant or all-determining.

Question 18

A child's full-scale IQ score on the WISC-V is reported as 105, with a 95% confidence interval of [98, 112]. What is the most accurate way for a psychologist to interpret and communicate this result to the child's parents?

  1. 'The child's IQ is exactly 105, which is slightly above the average score of 100.'
  2. 'There is a 95% probability that the child's IQ will increase to 112 on a future retest.'
  3. 'The child scored better than 95% of their peers, as indicated by the 95% confidence level.'
  4. 'We can be 95% confident that the child's true score falls in the range of 98 to 112, which is solidly average.' (correct answer)
Explanation: An IQ score is an estimate, not a perfectly precise number. The confidence interval accounts for the standard error of measurement. The correct interpretation is that if the child were tested repeatedly, 95% of the time their true score would be expected to fall within the given range (98 to 112). This communicates the score's precision appropriately and places it within the correct context (the Average range). Choice A is wrong because it treats 105 as an exact score. Choices B and C completely misinterpret the meaning of a confidence interval.

Question 19

A large school district has used the same version of a group intelligence test, standardized in 1995, for the past 25 years. Administrators observe that the average score of students has steadily increased from a mean of 101 in 1996 to a mean of 109 in 2021. Assuming the student population's underlying abilities have not changed dramatically, which phenomenon best accounts for this observation?

  1. Stereotype threat, as students from all backgrounds have become less anxious about testing over time.
  2. Regression to the mean, as initially low-scoring student groups have naturally improved their performance.
  3. The Flynn effect, which suggests that the test's original norms have become outdated over time. (correct answer)
  4. A high degree of predictive validity, indicating the test accurately measures scholastic aptitude.
Explanation: The Flynn effect is the observed, long-term, generation-over-generation increase in fluid and crystallized intelligence test scores. When a test with old norms is used for many years, the average score will appear to rise because the current population is being compared to a lower-scoring standardization sample from the past. The other options do not explain a population-wide, gradual increase in scores over decades.

Question 20

An individual is a world-renowned sculptor and can mentally rotate complex three-dimensional objects with exceptional skill. However, they struggle with abstract reasoning in mathematics and have only average verbal skills. This person's cognitive profile, characterized by a distinct peak in one area alongside average or lower abilities in others, is most effectively explained by which theory of intelligence?

  1. Charles Spearman's concept of a general intelligence factor (g).
  2. Howard Gardner's theory of multiple intelligences. (correct answer)
  3. Lewis Terman's longitudinal studies on giftedness.
  4. Robert Sternberg's concept of analytical intelligence.
Explanation: Howard Gardner's theory posits the existence of multiple, relatively independent intelligences, including spatial intelligence, which would be exceptionally high in a sculptor. This theory readily accounts for a 'jagged' cognitive profile where an individual can be a genius in one domain but average in others. Spearman's g would predict more correlated abilities, Terman studied giftedness but didn't propose this type of model, and Sternberg's analytical intelligence relates to the academic skills where this individual is not gifted.