All questions
Question 1
A cross-sectional study finds a strong positive correlation between the number of hours teenagers spend on social media and their reported levels of loneliness. A newspaper article about the study is titled: "Social Media Overuse is Causing an Epidemic of Loneliness in Teens."
Besides a potential third variable, what is the most significant flaw in the newspaper's causal conclusion?
- The study should have included a wider age range than just teenagers.
- The term 'overuse' is subjective and not operationally defined in the headline.
- The relationship might be reversed: teenagers who are already lonely may turn to social media for connection. (correct answer)
- Self-reported loneliness can be an unreliable measure of a person's true emotional state.
Explanation: The correct answer is C. This describes the directionality problem in correlational research. A correlation between two variables does not tell us the direction of the causal relationship. While it's possible that social media use causes loneliness, it is equally plausible that pre-existing loneliness causes teens to spend more time on social media. The headline ignores this alternative causal pathway. A and D are valid methodological points, but C addresses the fundamental error in drawing a specific causal conclusion from the correlational data.
Question 2
In a large corporate survey, employees were asked, "Do you consistently treat all your colleagues with respect, regardless of their position?" Over 95% of respondents answered "Yes." The company's HR department cites this as evidence of a highly positive and respectful workplace culture.
The validity of HR's conclusion is most directly challenged by the likelihood of:
- sampling error, as not all employees may have responded to the survey.
- social desirability bias, as employees are likely to give answers that portray them in a positive light. (correct answer)
- ambiguity in the term "respect," which may mean different things to different people.
- the halo effect, where employees' positive feelings about the company influence their responses.
Explanation: The correct answer is B. Social desirability bias is the tendency of survey respondents to answer questions in a manner that will be viewed favorably by others. Few people would admit to not treating colleagues with respect, even if it were true. This makes the self-report data highly suspect and challenges the validity of the conclusion. A, C, and D are all potential issues, but the nature of the question makes social desirability bias the most prominent and powerful threat to the validity of the data itself.
Question 3
An advertisement for a dietary supplement features a famous and respected movie star. She shares a moving story about her struggles with 'brain fog' and explains how taking this supplement helped her regain her mental sharpness and succeed in her demanding career.
This advertisement's primary method of persuasion relies on:
- presenting data from controlled clinical trials to support the supplement's efficacy.
- a detailed explanation of the biochemical mechanisms through which the supplement's ingredients work.
- the use of an emotional testimonial from an authoritative but non-expert figure. (correct answer)
- a comparison of the supplement's effectiveness against several leading competitors.
Explanation: The correct answer is C. The advertisement uses two powerful but non-scientific persuasive techniques: an appeal to emotion (the moving personal story) and an appeal to authority. The authority here is the celebrity, who is admired and respected but is not a scientific expert on nutrition or cognitive science. This combination is designed to create a strong positive association with the product, bypassing the need for actual scientific evidence. The other options describe evidence-based methods of persuasion, which are explicitly absent from the scenario.
Question 4
Researchers find that laboratory rats exposed to a specific Mozart sonata for eight hours a day navigate mazes 15% faster than rats in a silent control group. A parenting blog reports on this study with the headline: 'Science Says: Listening to Mozart Can Make Your Child a Genius!'
The blog's headline is a significant distortion of the research primarily because it:
- exaggerates the 15% performance increase by using the word 'genius.'
- fails to mention the specific Mozart sonata that was used in the experiment.
- inappropriately generalizes findings from maze navigation in rats to intelligence in human children. (correct answer)
- assumes that the results would be the same for other classical composers besides Mozart.
Explanation: The correct answer is C. This is a classic example of flawed popular psychology claims: overgeneralizing from animal studies to humans. The physiology, cognitive processes, and environmental influences of rats are vastly different from those of human children. Furthermore, maze-solving ability in rats is not equivalent to 'genius' or general intelligence in humans. While A is also true (exaggeration), C identifies the more fundamental scientific error of cross-species generalization.
Question 5
A study published in a popular nutrition magazine surveyed its readers about their dietary habits and longevity. The study found that 90% of respondents who lived past age 90 reported consuming a tablespoon of olive oil daily. The magazine concluded that daily olive oil consumption is a key to longevity for the general population.
Which factor most severely limits the generalizability of the magazine's conclusion?
- The study relied on self-reported dietary habits, which can be prone to memory errors.
- The sample consisted of magazine subscribers, who are likely more health-conscious than the average person. (correct answer)
- The study did not investigate the specific type or quality of the olive oil consumed.
- The correlational nature of the data does not permit a causal conclusion about olive oil and longevity.
Explanation: The correct answer is B. The question asks about generalizability. The sample is drawn from subscribers to a nutrition magazine, which is a self-selected group that is not representative of the general population. They are likely to have different diets, exercise habits, and socioeconomic status compared to non-subscribers, making it impossible to generalize the findings. While D is a correct statement about the study's limitations, the question specifically asks about the factor limiting generalizability, which directly relates to the nature of the sample (B). A is a valid criticism of the measurement, but the sampling issue is a more profound threat to the conclusion's external validity.
Question 6
A viral video demonstrates a 'psychological hack': crossing one's arms firmly for 30 seconds is said to increase persistence on difficult tasks by activating 'bilateral hemispheric stimulation.'
What is the most compelling reason for a psychologist to be skeptical of the explanation provided for this 'hack'?
- The duration of 30 seconds is likely too short to create a meaningful neurological change.
- Any effect is more likely due to a change in posture and mindset rather than a specific neurological process.
- The claim is being made in a video rather than a peer-reviewed academic journal.
- The term 'bilateral hemispheric stimulation' is used imprecisely and does not correspond to a known mechanism for increasing persistence. (correct answer)
Explanation: When evaluating psychological claims, especially those made in popular media, you should always examine whether the proposed mechanisms are scientifically valid and precisely defined. This question tests your ability to identify pseudoscientific explanations that use impressive-sounding but meaningless terminology.
The correct answer is D because "bilateral hemispheric stimulation" is essentially meaningless jargon in this context. While bilateral stimulation is a real concept in some therapeutic approaches (like EMDR), there's no established neurological mechanism by which crossing your arms would create meaningful bilateral hemispheric stimulation, let alone one that specifically increases persistence. The term is being used to make the claim sound scientific without any actual scientific basis.
Choice A focuses on duration, but 30 seconds could theoretically be sufficient for some neurological changes, so this isn't the strongest criticism. Choice B actually offers a plausible alternative explanation—posture and mindset changes can indeed affect behavior—making this a reasonable possibility rather than a reason for skepticism. Choice C addresses the source rather than the content; while peer review is important, the location of a claim doesn't automatically invalidate the underlying science.
Watch for psychology questions that test your ability to distinguish between legitimate scientific explanations and pseudoscientific ones. The key red flag is when technical-sounding terms are used imprecisely or without clear connection to established mechanisms. Always ask: "Is this terminology being used accurately, and does the proposed mechanism actually exist?"
Question 7
In a small pilot study, 25 individuals with self-reported high stress levels were given a bracelet that emits a low-frequency electromagnetic field. After four weeks of wearing the bracelet, 18 of the 25 participants reported a significant reduction in their feelings of stress. The company is now marketing the bracelet as a 'scientifically-backed stress reduction device.'
Which phenomenon represents the most compelling alternative explanation for the reported stress reduction?
- Regression to the mean
- The placebo effect (correct answer)
- The Hawthorne effect
- Spontaneous remission
Explanation: The correct answer is B. The placebo effect occurs when a person experiences a real improvement following an inert treatment simply because they expect it to work. In this case, the participants' belief in the bracelet's effectiveness is the most likely cause of their self-reported stress reduction. A (Regression to the mean) is possible, as people may have entered the study when their stress was at a peak. C (Hawthorne effect) relates to changes in behavior due to awareness of being observed, which is less relevant to internal feelings of stress. D (Spontaneous remission) is a possibility, but the placebo effect is a more direct and powerful explanation for improvement after a specific intervention is introduced.
Question 8
A popular wellness blog features a post titled "How I Cured My Lifelong Anxiety in One Week." The author describes their personal experience of quitting all social media for seven days, after which they felt a profound sense of peace and clarity. The post receives thousands of positive comments from readers who are inspired to try the same 'digital detox.'
From a scientific perspective, what is the most significant reason to be skeptical of the author's claim that the digital detox cured their anxiety?
- The author is likely biased because they want to sell a book or course about their method.
- The experience is based on a single individual's account rather than controlled, systematic observation. (correct answer)
- The one-week duration is too short to produce any lasting psychological change.
- The claim ignores the possibility that other lifestyle changes during that week caused the improvement.
Explanation: The correct answer is B. The primary weakness of the claim is that it is based on an anecdote (a single person's experience). Anecdotal evidence is not a reliable basis for a general claim because it is unsystematic, cannot be replicated, and does not control for other variables. While the other options present valid concerns, the reliance on a single case study instead of robust research is the most fundamental flaw in the evidence provided. A is a plausible motive but doesn't address the evidence itself. C is an assumption that may not be true. D points to a lack of control for confounds, which is a key reason why single anecdotes (as mentioned in B) are unscientific.
Question 9
A company markets a new mobile app with games designed to improve cognitive function. To support their claims, they present data from 1,000 users who played the games for 15 minutes daily. After one month, the users' average scores on the in-app cognitive tests improved by 25%.
What is the most critical piece of information missing from this evidence that prevents concluding the app is effective?
- Data from a control group that did not use the app but took the same cognitive tests. (correct answer)
- Information about the demographic characteristics (e.g., age, education) of the 1,000 users.
- Evidence that the cognitive improvements transfer to real-world tasks outside of the app.
- A breakdown of whether the 25% improvement was consistent across all types of games in the app.
Explanation: The correct answer is A. Without a control group, it is impossible to know if the improvement was due to the app itself. The users might have improved simply due to the practice effect (getting better at the tests through repetition), the placebo effect (expecting to improve), or other external factors. A control group is essential for making a causal claim about the app's effectiveness. C (transfer to real-world tasks) is a very important next step, but it's secondary to first establishing that the app causes any improvement at all, which requires a control group. B and D are useful details but do not address the fundamental flaw in the study's design.
Question 10
An online article promotes a new herbal supplement for improving mood. It supports its claims by citing a study from the 'Global Journal of Applied Psychopharmacology.' A librarian's check reveals this journal is a 'predatory journal,' which charges authors high fees to publish without rigorous peer review and is not indexed in reputable academic databases like PsycINFO or PubMed.
Based on this information, the most critical reason to be skeptical of the study is that its findings likely lack:
- statistical significance.
- a large enough sample size.
- validity due to a lack of independent scientific scrutiny. (correct answer)
- a clear operational definition of 'mood.'
Explanation: The correct answer is C. The fact that the study was published in a predatory journal means it did not undergo the rigorous, independent peer-review process that is the hallmark of credible science. This process is designed to vet methodology, analysis, and conclusions. Without it, the study's results are not trustworthy, and its validity is highly questionable. There may be other flaws (A, B, D), but the lack of legitimate peer review is the most fundamental problem with the evidence's source.
Question 11
A university offers a voluntary workshop for students who scored in the bottom 5% on their midterm exams. After attending the workshop and taking the final exam, the average score for this group of students increased by 12 percentage points. The workshop organizer claims this shows the workshop was highly effective.
Beyond the workshop's content, which statistical phenomenon provides the most likely alternative explanation for the students' score increase?
- The placebo effect, where students improved because they believed the workshop would help.
- The Hawthorne effect, where students studied harder simply because they were singled out for attention.
- Regression to the mean, where an extreme score on one measurement tends to be closer to the average on a second measurement. (correct answer)
- Survivorship bias, where only the most motivated students from the bottom group attended the workshop.
Explanation: The correct answer is C. Regression to the mean is a statistical phenomenon where a variable that is extreme on its first measurement will tend to be closer to the average on its second measurement. The students were selected precisely because of their extreme (low) scores. Therefore, even with a useless workshop, it is statistically probable that their average score would increase on the next exam, moving closer to the class mean. This makes it a powerful alternative explanation. A and B are possible psychological effects, but regression to the mean is a statistical artifact inherent in this type of selection process.
Question 12
An advertisement for a meditation app claims, "Studies show that using our app for 10 minutes a day can significantly reduce anxiety." The ad does not provide any further details about these studies.
To critically evaluate the scientific basis of this claim, which of the following is the most important question to ask about the 'studies'?
- What was the total number of participants across all the studies?
- How did the studies operationally define and measure 'anxiety'?
- Were the studies funded by the company that created the meditation app?
- Did the studies use a randomized, controlled design with an active control group? (correct answer)
Explanation: When evaluating scientific claims about treatments or interventions, you need to assess whether the research design can actually support causal conclusions. The gold standard for establishing that an intervention causes an effect is the randomized controlled trial (RCT) with proper controls.
Answer D is correct because a randomized, controlled design with an active control group is essential for determining whether the meditation app actually reduces anxiety. Randomization ensures that participant characteristics are evenly distributed between groups, controlling for confounding variables. An active control group (like a different relaxation technique) is crucial because it controls for placebo effects, researcher attention, and the simple act of taking time to focus on wellness. Without this design, you can't know if any observed anxiety reduction is due to the app itself or other factors.
Answer A is wrong because while sample size affects statistical power, even a large study can't establish causation without proper experimental controls. Answer B is incorrect because although operational definitions matter for interpreting results, poor measurement doesn't prevent you from evaluating whether a causal relationship exists—it just affects how you interpret the magnitude of effects. Answer C addresses potential bias, which is concerning, but even industry-funded studies can provide valid evidence if they use rigorous methodology.
Remember: when evaluating treatment claims, always ask "Can this study design rule out alternative explanations?" Look for randomization and appropriate control groups as your first check for whether causal claims are scientifically justified.
Question 13
An online news site conducts a poll asking visitors, "Do you think that modern technology is making people less intelligent?" The poll garners over 500,000 responses, with 85% of respondents answering "Yes." The site concludes that the vast majority of people believe technology is harmful to intelligence.
The strongest argument against the validity of the site's conclusion is that:
- the question is biased because it uses the loaded term 'less intelligent.'
- the conclusion is invalid because it is based on opinion, not on objective data about intelligence.
- the sample size of 500,000 is too small to accurately represent the views of the entire world.
- the sample is subject to extreme self-selection bias and is not representative of the general population. (correct answer)
Explanation: When evaluating research conclusions, you need to assess whether the sample truly represents the population being studied. This question tests your understanding of sampling bias and external validity.
The correct answer is D because this poll suffers from severe self-selection bias. Only people who visit this particular news site and feel motivated to respond participated. This creates a sample that likely skews toward people with strong opinions about technology, potentially those who are already skeptical of it. Self-selected samples are notorious for producing unrepresentative results because participation depends on individual motivation rather than random selection.
Option A is incorrect because while the question wording could influence responses, the more fundamental flaw is who's responding, not how they're responding. Option B misses the point—polls measuring public opinion are valid research tools when conducted properly. The issue isn't that it's opinion-based, but that it's a biased sample of opinions. Option C demonstrates a misunderstanding of sample size principles. 500,000 is actually an enormous sample; the problem isn't quantity but quality of representation.
For psychology research questions, always prioritize sampling issues over other methodological concerns. A biased sample of millions is less valuable than a representative sample of hundreds. Remember: "garbage in, garbage out"—even massive datasets are worthless if they don't represent your target population. When you see online polls or convenience samples, immediately consider who's likely to participate versus who's being excluded.
Question 14
To combat student apathy, a high school requires all sophomores to participate in a new community service program. At the beginning of the year, 40% of students reported feeling 'disconnected' from their community. By the end of the year, this number dropped to 25%. The administration presents this as strong evidence for the program's success.
The administration's conclusion is questionable because the simple pre-test/post-test design fails to control for:
- the Hawthorne effect, where students' behavior changed because they knew their feelings were being measured.
- selection bias, as sophomores may be inherently different from students in other grades.
- experimenter bias, where the administration may have interpreted the results in a way that favored the program.
- maturation effects, as students may have naturally developed a greater sense of community connection over the course of a school year. (correct answer)
Explanation: When evaluating research designs in psychology, you need to identify what alternative explanations could account for observed changes beyond the intended intervention. This question tests your understanding of internal validity threats in experimental design.
The administration's conclusion is problematic because they used a simple pre-test/post-test design without a control group. The most significant threat here is maturation effects (D) – the natural developmental changes that occur over time. Sophomores are typically 15-16 years old, an age when cognitive and social development naturally leads to greater community awareness and connection. The observed 15% decrease in feeling "disconnected" could simply reflect normal adolescent development over an entire school year, not the community service program's effectiveness.
Let's examine why the other options don't represent the primary concern: The Hawthorne effect (A) involves behavior change due to awareness of being observed, but this study measured attitudes, not behaviors, and there's no indication students knew their responses were being monitored. Selection bias (B) is irrelevant here since all sophomores participated – there's no comparison between grade levels that would create this bias. Experimenter bias (C) could affect result interpretation, but the administration used objective survey data showing clear percentage changes, limiting subjective interpretation.
Study tip: When you see pre-test/post-test designs without control groups, immediately ask "What else could cause this change over time?" Maturation effects are especially likely when studying developmental populations like children or adolescents over extended periods.
Question 15
A website selling expensive crystals claims, "Ancient civilizations have known for millennia that rose quartz opens the heart chakra, promoting emotional healing. This timeless wisdom is the foundation of our product's power."
The primary logical fallacy used to support this claim is the:
- appeal to tradition, which asserts that a belief is correct simply because it is old or has been long-held. (correct answer)
- bandwagon effect, which suggests a claim is true because many people believe it.
- appeal to ignorance, which argues that a claim must be true because it has not been proven false.
- false cause fallacy, which incorrectly assumes that one event causes another.
Explanation: The correct answer is A. The claim's justification rests entirely on 'ancient wisdom' and what civilizations have 'known for millennia.' This is a classic example of the appeal to tradition (or appeal to antiquity) fallacy. The argument assumes that because a belief is old, it must be true. This is fallacious because the age of a belief is not evidence of its empirical validity. The other options describe different logical fallacies that are not the primary one being used in this specific claim.
Question 16
A new self-help program, 'Quantum Entrainment Therapy,' claims to help individuals manifest their desires by 'harmonizing their personal bio-field with cosmic energy frequencies.' The program's website is filled with terms like 'neuro-vibrational resonance' and 'subatomic manifestation,' but it does not cite any peer-reviewed empirical studies.
A critical evaluation of this program's claims should primarily identify its reliance on:
- the placebo effect, as users may feel better simply because they expect to.
- unfalsifiable claims and pseudoscientific language that mimics scientific terminology. (correct answer)
- anecdotal testimonials from users who have had positive experiences.
- the fallacy of appealing to authority, by implying a deep, hidden knowledge.
Explanation: The correct answer is B. The core issue with the program is its use of 'psychobabble' or pseudoscientific jargon. Terms like 'neuro-vibrational resonance' and 'cosmic energy frequencies' have no clear, measurable, or scientific meaning. They are designed to sound impressive and scientific but are ultimately unfalsifiable—they cannot be tested or proven wrong. This is a key characteristic of pseudoscience. While A, C, and D are likely also true of the program, the use of this specific type of language is the most fundamental problem with the claim's scientific basis.
Question 17
An online article promotes a new herbal supplement for improving mood. It supports its claims by citing a study from the 'Global Journal of Applied Psychopharmacology.' A librarian's check reveals this journal is a 'predatory journal,' which charges authors high fees to publish without rigorous peer review and is not indexed in reputable academic databases like PsycINFO or PubMed.
Based on this information, the most critical reason to be skeptical of the study is that its findings likely lack:
- statistical significance.
- a large enough sample size.
- validity due to a lack of independent scientific scrutiny. (correct answer)
- a clear operational definition of 'mood.'
Explanation: The correct answer is C. The fact that the study was published in a predatory journal means it did not undergo the rigorous, independent peer-review process that is the hallmark of credible science. This process is designed to vet methodology, analysis, and conclusions. Without it, the study's results are not trustworthy, and its validity is highly questionable. There may be other flaws (A, B, D), but the lack of legitimate peer review is the most fundamental problem with the evidence's source.
Question 18
A university offers a voluntary workshop for students who scored in the bottom 5% on their midterm exams. After attending the workshop and taking the final exam, the average score for this group of students increased by 12 percentage points. The workshop organizer claims this shows the workshop was highly effective.
Beyond the workshop's content, which statistical phenomenon provides the most likely alternative explanation for the students' score increase?
- The placebo effect, where students improved because they believed the workshop would help.
- The Hawthorne effect, where students studied harder simply because they were singled out for attention.
- Regression to the mean, where an extreme score on one measurement tends to be closer to the average on a second measurement. (correct answer)
- Survivorship bias, where only the most motivated students from the bottom group attended the workshop.
Explanation: The correct answer is C. Regression to the mean is a statistical phenomenon where a variable that is extreme on its first measurement will tend to be closer to the average on its second measurement. The students were selected precisely because of their extreme (low) scores. Therefore, even with a useless workshop, it is statistically probable that their average score would increase on the next exam, moving closer to the class mean. This makes it a powerful alternative explanation. A and B are possible psychological effects, but regression to the mean is a statistical artifact inherent in this type of selection process.
Question 19
An advertisement for a meditation app claims, "Studies show that using our app for 10 minutes a day can significantly reduce anxiety." The ad does not provide any further details about these studies.
To critically evaluate the scientific basis of this claim, which of the following is the most important question to ask about the 'studies'?
- What was the total number of participants across all the studies?
- How did the studies operationally define and measure 'anxiety'?
- Were the studies funded by the company that created the meditation app?
- Did the studies use a randomized, controlled design with an active control group? (correct answer)
Explanation: When evaluating scientific claims about treatments or interventions, you need to assess whether the research design can actually support causal conclusions. The gold standard for establishing that an intervention causes an effect is the randomized controlled trial (RCT) with proper controls.
Answer D is correct because a randomized, controlled design with an active control group is essential for determining whether the meditation app actually reduces anxiety. Randomization ensures that participant characteristics are evenly distributed between groups, controlling for confounding variables. An active control group (like a different relaxation technique) is crucial because it controls for placebo effects, researcher attention, and the simple act of taking time to focus on wellness. Without this design, you can't know if any observed anxiety reduction is due to the app itself or other factors.
Answer A is wrong because while sample size affects statistical power, even a large study can't establish causation without proper experimental controls. Answer B is incorrect because although operational definitions matter for interpreting results, poor measurement doesn't prevent you from evaluating whether a causal relationship exists—it just affects how you interpret the magnitude of effects. Answer C addresses potential bias, which is concerning, but even industry-funded studies can provide valid evidence if they use rigorous methodology.
Remember: when evaluating treatment claims, always ask "Can this study design rule out alternative explanations?" Look for randomization and appropriate control groups as your first check for whether causal claims are scientifically justified.
Question 20
A best-selling book argues that all people are fundamentally one of two learning types: 'Sequential-Analytical' or 'Holistic-Intuitive.' The author claims that for people to succeed, they must identify their type and seek educational and career environments that match it.
From the perspective of modern psychology, a primary scientific criticism of this 'learning types' model is that it:
- presents a false dichotomy by categorizing a complex, continuous trait into two distinct boxes. (correct answer)
- fails to provide a reliable and valid test for individuals to determine their specific type.
- overlooks the role of genetics in determining an individual's learning preferences.
- is difficult to apply in a practical classroom setting with many students.
Explanation: The correct answer is A. Most psychological traits, including cognitive styles and learning preferences, exist on a continuum, not in discrete categories. Such 'typing' models are criticized for creating a false dichotomy, which oversimplifies the reality of human personality and cognition. People are not one thing or the other; they possess a mix of traits. While B, C, and D are valid concerns, the fundamental scientific flaw is the imposition of an artificial binary classification on a complex spectrum of human behavior.