Welcome to Paper 3: Synoptic Review of Studies

Welcome to one of the most rewarding parts of your Pearson Edexcel A Level Psychology journey! In Paper 3 (Psychological Skills - 9PS0/03), Section B is dedicated entirely to the Synoptic Review of Studies. This section carries 24 marks out of the 80 marks available on Paper 3.

The word synoptic simply means "taking a broad, overall view." Instead of looking at studies in isolated topic boxes, you will now bring together your knowledge across the course, comparing classic pieces of research with modern contemporary studies using a shared evaluative toolkit.

Don't worry if comparing multiple studies feels challenging at first. Once you learn the universal framework (GRAVE-SO), you will be able to evaluate and contrast any pair of studies with confidence.

---

1. The Master Evaluative Framework: GRAVE-SO

Whenever you are asked to evaluate, contrast, or review studies in Section B, structure your thinking using the GRAVE-SO criteria. This ensures you cover all the methodological standards required by the Pearson Edexcel specification.

Breakdown of the GRAVE-SO Parameters:

G — Generalisability:
Can the findings be applied to the wider target population?
• Look at sample size, age, gender, culture, and whether the sample is ethnocentric (culture-biased) or androcentric (male-biased).
• Consider population validity (does the sample represent the general public?) and ecological validity (does the setting reflect everyday life?).

R — Reliability:
Can the study be replicated to find consistent results?
• Look for high levels of standardisation (identical instructions, timings, environments).
• Consider inter-rater reliability (do multiple observers or raters agree on the scores?).

A — Application:
How useful are the findings in the real world?
• Can they help improve school learning, reform courtroom procedures, reduce prejudice in communities, or inform clinical therapies for mental disorders?

V — Validity:
Did the study measure what it actually claimed to measure?
Internal validity: Were extraneous and confounding variables controlled?
Mundane realism: Did the experimental task mirror an activity people do in everyday life?
Historical validity: Do the findings still apply today, or were they a product of their time?

E — Ethics:
Did the researchers follow the BPS Code of Human Research Ethics?
• Consider informed consent, protection from physical and psychological harm, the right to withdraw, confidentiality, deception, and thorough debriefing.

S — Scientific Credibility:
Is the research objectively testable, falsifiable, empirical, and subject to peer review?

O — Objectivity vs. Subjectivity:
Did the researchers gather quantitative, unbiased numerical data (objective), or did they rely on qualitative, personal interpretation (subjective)?

Key Takeaway: Whenever you see a Section B question asking you to evaluate or compare, build your paragraphs directly around G-R-A-V-E-S-O criteria.

---

2. Topic-by-Topic Synoptic Review of Core Studies

Below is your complete guide to the compulsory Classic and Contemporary studies across Topics 1 to 5. Notice how each pair explores similar themes using different methodological approaches.

---

Topic 1: Social Psychology

Classic Study: Sherif et al. (1954/1961) — Intergroup Conflict and Cooperation: The Robbers Cave Experiment

Aim & Core Focus: To investigate whether intergroup conflict and prejudice occur due to competition for scarce resources (Realistic Conflict Theory) and whether cooperation on superordinate goals (goals that require both groups to work together) can reduce that friction.
Procedure Summary: 22 eleven-to-twelve-year-old white, middle-class Protestant American boys were placed in an artificial summer camp at Robbers Cave State Park, Oklahoma. They were split into two groups (the Rattlers and the Eagles). Stage 1 formed in-group norms; Stage 2 introduced friction via competitive tournaments (e.g., baseball, tug-of-war); Stage 3 introduced superordinate goals (e.g., fixing the camp's broken water supply system, pulling a broken-down truck).
Key Evaluative Points (GRAVE):
- Generalisability: Low. Ethnocentric and androcentric sample (all boys, same age, same cultural background).
- Reliability: Difficult to replicate fully due to the natural field setting, though Sherif used standardised observation schedules.
- Application: Powerful real-world application for resolving community conflict, desegregating schools, and promoting international diplomacy through shared goals.
- Validity: High ecological validity (boys believed it was a genuine summer camp), but high risk of researcher bias since researchers acted as camp counsellors.
- Ethics: Deception was used; boys experienced genuine distress and hostility; informed consent was obtained from parents, not the boys themselves.

Contemporary Study: Burger (2009) — Replicating Milgram: Would People Still Obey Today?

Aim & Core Focus: To investigate whether obedience levels in a modern sample would match Milgram's classic 1963 findings, while introducing rigorous ethical safeguards.
Procedure Summary: Used a two-step screening process to exclude individuals with emotional vulnerability. Tested participants up to a 150V maximum ("the point of no return" where Milgram's learner first screamed). Included two conditions: the Base Condition and Condition 2 (Model Refusal), where a confederate refused to continue at 90V.
Key Evaluative Points (GRAVE):
- Generalisability: Better than Milgram's original; included diverse adult men and women across various educational and ethnic backgrounds.
- Reliability: Highly standardised laboratory procedure (identical script, prods, and apparatus), allowing exact replication.
- Application: Proves blind obedience to authority remains remarkably stable over time, informing training for police, medical personnel, and military units.
- Validity: High internal validity due to tight laboratory controls, but lower mundane realism (administering electric shocks to learn word pairs is not a daily task).
- Ethics: Significantly improved over Milgram. Screened out vulnerable participants, gave 3 explicit written reminders of the right to withdraw, kept the maximum shock to 150V, and provided immediate debriefing.

---

Topic 2: Cognitive Psychology

Classic Study: Baddeley (1966b) — Acoustic and Semantic Similarity in LTM

Aim & Core Focus: To determine whether encoding in Short-Term Memory (STM) and Long-Term Memory (LTM) relies on acoustic (sound) or semantic (meaning) codes.
Procedure Summary: Participants learned word sequences from four distinct lists: Acoustically Similar (e.g., man, can, cat), Acoustically Dissimilar (e.g., pit, day, cow), Semantically Similar (e.g., great, large, big), and Semantically Dissimilar (e.g., hot, old, late). LTM was tested after an interference task and a 20-minute delay.
Key Evaluative Points (GRAVE):
- Generalisability: British university/volunteer sample may not represent how all age groups or non-English speakers process information.
- Reliability: Exceptionally high standardisation (controlled presentation times, fixed interference tasks), making it easily replicable.
- Application: Informs revision techniques; students learn that LTM relies on semantic encoding, meaning understanding concepts works better than rote repetition.
- Validity: High internal validity, but low mundane realism because recalling artificial word lists does not reflect everyday memory demands.
- Ethics: Highly ethical; no deception, no distress, full consent obtained.

Contemporary Study: Sebastian and Hernández-Gil (2012) — Developmental Pattern of Digit Span for Verbal Data

Aim & Core Focus: To investigate the developmental trajectory of the phonological loop (digit span) in Spanish children aged 5 to 17, and compare their performance against adult and clinical dementia cohorts.
Procedure Summary: 570 native Spanish children with no hearing or cognitive impairments were tested individually using standardised digit span tasks (recalling sequences of increasing length). Results showed digit span increases with age up to age 17 (reaching \(5.91\)), which is lower than the English average (\(\approx 7\)) due to word length differences in Spanish number words.
Key Evaluative Points (GRAVE):
- Generalisability: Large sample (\(n = 570\)) across broad age bands, but specific to Spanish speakers.
- Reliability: Standardised digit span protocol allows straightforward replication.
- Application: Establishes baseline norms for identifying developmental language disorders and tracking cognitive decline in Alzheimer's patients.
- Validity: High construct validity for testing the sub-components of the Working Memory Model (phonological loop).
- Ethics: Preserved child protection guidelines, parental informed consent, and comfortable testing environments.

---

Topic 3: Biological Psychology

Classic Study: Raine, Buchsbaum & LaCasse (1997) — Brain Abnormalities in Murderers Indicated by PET

Aim & Core Focus: To investigate whether individuals charged with murder but pleading Not Guilty by Reason of Insanity (NGRI) show localized brain dysfunction compared to matched controls.
Procedure Summary: 41 NGRI murderers (39 male, 2 female) were matched with 41 control participants on age, sex, and diagnosis of schizophrenia where applicable. Participants performed a Continuous Performance Task (CPT) focused on target recognition for 32 minutes before and during fluorodeoxyglucose (FDG) injection, followed by PET scans.
Key Findings: NGRI participants showed significantly reduced glucose metabolism in the prefrontal cortex, corpus callosum, and asymmetrical activity in the amygdala and medial temporal lobe.
Key Evaluative Points (GRAVE):
- Generalisability: Sample restricted to NGRI murderers; cannot be generalised to all violent offenders or the general population.
- Reliability: Standardised CPT task and objective PET imaging techniques provide high test-retest reliability.
- Application: Cautiously informs neurobiological theories of impulse control, though Raine emphasizes PET scans cannot be used as sole evidence in courtrooms.
- Validity: PET provides objective physiological data, but the study is correlational—brain abnormalities do not prove direct causation of violent behaviour.
- Ethics: The NGRI participants were already undergoing PET for legal defense, reducing ethical harm; however, radioactive tracer injection carries minor physical risk.

Contemporary Study: Brendgen et al. (2005) — Genetic and Environmental Influences on Physical and Social Aggression in 6-Year-Old Twins

Aim & Core Focus: To examine the relative contributions of genetic and environmental factors to physical aggression versus social aggression.
Procedure Summary: 234 twin pairs (monozygotic [MZ] and dizygotic [DZ]) recruited from the Quebec Newborn Twin Study at age 6. Aggression was rated independently by both teachers (using standardised questionnaires) and peers (using photographic nomination booklets).
Key Findings: Physical aggression was heavily genetically mediated (higher correlation in MZ twins), whereas social aggression was primarily influenced by non-shared environmental factors.
Key Evaluative Points (GRAVE):
- Generalisability: Twin studies can sometimes have limited generalisability to single-born children, but the large cohort strengthened population validity.
- Reliability: High inter-rater agreement between independent teacher and peer ratings.
- Application: Suggests early childhood interventions should target physical aggression biologically/developmentally and social aggression via social learning and peer dynamics.
- Validity: Naturalistic observations within school settings avoid artificial laboratory bias.
- Ethics: Peer nomination methods require careful handling to avoid creating peer friction or stigma in classrooms.

---

Topic 4: Learning Theories

Classic Study: Watson & Rayner (1920) — Conditioned Emotional Reactions (Little Albert)

Aim & Core Focus: To demonstrate that human emotional responses, such as fear and phobias, can be acquired through classical conditioning.
Procedure Summary: 11-month-old infant "Little Albert" was shown a white rat (Neutral Stimulus). When he reached for it, researchers struck a suspended steel bar with a hammer behind his head (Unconditioned Stimulus), producing fear (Unconditioned Response). After repeated pairings, the rat alone produced fear (Conditioned Response). Fear was also shown to generalise to similar stimuli (a rabbit, a fur coat, a Santa Claus mask).
Key Evaluative Points (GRAVE):
- Generalisability: Single participant (\(n = 1\)), unique temperament; cannot reliably represent all infants.
- Reliability: The conditioning procedure was documented with clear baseline tests and stimulus presentations, though precise controls were loose by modern standards.
- Application: Formed the empirical foundation for understanding phobia acquisition and led to the development of counter-conditioning therapies.
- Validity: High internal validity regarding immediate association, but lacks ecological validity due to the artificial laboratory setup.
- Ethics: Highly unethical by modern standards. Caused severe psychological distress with no systematic de-conditioning carried out before Albert left the hospital.

Contemporary Study: Capafóns et al. (1998) — Systematic Desensitisation in the Treatment of Fear of Flying

Aim & Core Focus: To evaluate the therapeutic effectiveness of Systematic Desensitisation (SD) for aerophobia (fear of flying).
Procedure Summary: 41 participants with aerophobia were assigned to an experimental treatment group (\(n = 20\)) or a waiting-list control group (\(n = 21\)). Treatment involved relaxation training combined with graduated exposure (in vivo and video simulations). Outcome was measured using self-report diagnostic scales and physiological indices (heart rate, muscle tension, finger temperature) during simulated flight scenes.
Key Evaluative Points (GRAVE):
- Generalisability: Good adult clinical sample, though specific only to aerophobia.
- Reliability: Standardised treatment protocol (fixed number of sessions and uniform exposure material).
- Application: Shows that behavioral therapy reliably reduces both cognitive anxiety and physiological fear responses, enabling people to resume air travel.
- Validity: Use of both objective physiological measures and subjective self-reports reduced bias and increased construct validity.
- Ethics: Waiting-list design ensured control participants received therapy after the study concluded, maintaining ethical fairness.

---

Topic 5: Clinical Psychology

Classic Study: Rosenhan (1973) — On Being Sane in Insane Places

Aim & Core Focus: To investigate whether psychiatric staff can reliably distinguish between the sane and the insane, and to expose the power of psychiatric labelling.
Procedure Summary: 8 pseudopatients gained admission to 12 different psychiatric hospitals across 5 US states by claiming to hear a voice saying "empty, hollow, thud." Immediately upon admission, they ceased simulating any symptoms, acted normally, and recorded their observations in notebooks.
Key Findings: All pseudopatients were admitted; 11 were diagnosed with Schizophrenia and 1 with Manic-Depressive Psychosis. None were detected by staff (though many real patients suspected them). Average stay was 19 days. Normal behaviors (e.g., note-taking, arriving early for lunch) were interpreted as symptoms of illness ("stickiness of diagnostic labels").
Key Evaluative Points (GRAVE):
- Generalisability: High institutional generalisability (12 varied hospitals in 5 states, ranging from private to underfunded public clinics).
- Reliability: Standardised symptom report at intake, though field notes relied on individual observation.
- Application: Sparked major reforms in psychiatric diagnosis and accelerated the revision of the DSM diagnostic criteria (DSM-III onwards).
- Validity: High ecological validity (covert field experiment in real hospital settings).
- Ethics: Heavy deception of hospital staff; staff time was taken away from genuine patients; pseudopatients could not leave instantly upon request.

Contemporary Study: Carlsson et al. (2000) — Network Interactions in Schizophrenia: Dopamine and Glutamate

Aim & Core Focus: To review neurochemical evidence and reconsider the traditional dopamine hypothesis by examining how glutamate deficiency (NMDA receptor hypofunction) interacts with dopamine pathways in schizophrenia.
Procedure Summary: Meta-analytic literature review drawing upon brain imaging, animal models (e.g., using NMDA antagonists like PCP and ketamine), and post-mortem tissue analyses.
Key Evaluative Points (GRAVE):
- Generalisability: Draws on diverse studies across human and animal models, making conclusions widely applicable.
- Reliability: Relies on published, peer-reviewed biochemical and neuroimaging studies.
- Application: Facilitated the development of novel atypical antipsychotic medications that target both dopamine and glutamate systems with fewer motor side effects.
- Validity: High theoretical validity, synthesising multiple streams of empirical evidence to form a more complete biological explanation.
- Ethics: Secondary review of literature involves no direct human participant distress.

---

3. Classic vs. Contemporary Comparison Matrix

Use this quick-reference summary table to contrast studies when tackling comparative exam questions in Section B:

Topic 1: Social Psychology
Classic: Sherif et al. (1954/1961) — Field experiment; high ecological validity; limited modern ethical safeguards; focuses on group-level conflict.
Contemporary: Burger (2009) — Laboratory experiment; lower ecological validity; strict modern BPS ethical safeguards; focuses on individual obedience.

Topic 2: Cognitive Psychology
Classic: Baddeley (1966b) — Word sequence recall; identified acoustic STM vs. semantic LTM encoding in British adults.
Contemporary: Sebastian & Hernández-Gil (2012) — Digit span tasks across developmental age groups (\(5\text{--}17\)); showed phonological loop capacity expands with age and is influenced by word length in Spanish.

Topic 3: Biological Psychology
Classic: Raine et al. (1997) — PET neuroimaging; investigated prefrontal and subcortical localized brain abnormalities in NGRI murderers.
Contemporary: Brendgen et al. (2005) — Twin study methodology; separated genetic influence (physical aggression) from environmental influence (social aggression).

Topic 4: Learning Theories
Classic: Watson & Rayner (1920) — Single-case infant laboratory study; conditioned a phobia using classical conditioning; low ethics.
Contemporary: Capafóns et al. (1998) — Controlled clinical trial; treated phobias using Systematic Desensitisation with physiological and self-report checks; high ethics.

Topic 5: Clinical Psychology
Classic: Rosenhan (1973) — Covert field study; exposed diagnostic invalidity, institutional depersonalisation, and the power of psychiatric labels.
Contemporary: Carlsson et al. (2000) — Neurochemical review/meta-analysis; expanded biological models beyond dopamine to include glutamate interactions.

---

4. Common Pitfalls and How to Avoid Them

Examiners frequently highlight the following common mistakes in Section B responses:

1. Pure Description Without Evaluation (AO1 vs. AO3 Trap):
The Mistake: Writing two pages describing the exact steps of two studies without directly comparing them.
The Fix: Balance your writing. State the feature briefly (AO1), then immediately evaluate or compare it using a GRAVE-SO criterion (AO3).

2. Confusing Ecological Validity with Mundane Realism:
The Mistake: Saying "Baddeley lacked ecological validity because memorising word lists is not realistic."
The Fix: Memorising word lists is an artificial task (lacks mundane realism). The laboratory room is an artificial setting (lacks ecological validity). Keep these two terms distinct!

3. Making Generic, Vague Criticisms:
The Mistake: Writing "Burger was more ethical than Milgram" or "Raine had good controls."
The Fix: Include exact operational details. Write: "Burger improved ethics by screening participants across a two-step psychological exclusion process and capping the voltage at 150V."

4. Historical Anachronism:
The Mistake: Judging Watson & Rayner (1920) as if they were operating under 2020s BPS guidelines.
The Fix: Acknowledge historical context. State that while Watson & Rayner's study breached modern BPS guidelines regarding protection from harm, formal ethical codes did not exist in 1920, and the study helped establish the need for modern safeguards.

---

5. Section Summary and Revision Checklist

Before sitting Paper 3 Section B, make sure you can answer the following questions with confidence:

☐ Can I evaluate every classic and contemporary study using Generalisability, Reliability, Application, Validity, and Ethics?
☐ Can I explain how modern contemporary studies build upon, replicate, or refine classic studies?
☐ Can I compare the quantitative objectivity of Raine or Baddeley with the qualitative observations of Rosenhan or Sherif?
☐ Can I explain the specific ethical improvements in Burger (2009) compared to historical obedience research?

Final Tip for Success: Whenever you revise a classic study, always revise its contemporary partner right alongside it. Thinking about them in pairs will make your synoptic comparisons flow naturally in the exam!