Methods of Language Analysis: Phonetics, Phonology and Prosodics

Welcome to your complete revision guide for Phonetics, Phonology and Prosodics for AQA A Level English Language (7702). Whether you are analysing a spoken conversation transcript, exploring how children learn to talk, or investigating regional accents across the UK, understanding sound is one of your most powerful analytical tools.

Don't worry if these terms seem technical at first! Think of spoken language like music: you have individual musical notes (phonetics and phonology), and you have the rhythm, volume, and tempo with which the musician plays them (prosodics). By the end of these notes, you will be able to identify, describe, and explain these features with confidence for AO1.


1. The Core Distinctions: Three Key Levels of Sound

To master this topic, you first need to understand the distinct roles of the three branches:

1. Phonetics: The physical study of speech sounds. It focuses on the mechanics of how human speech organs (lips, tongue, vocal cords) physically produce sounds (articulatory phonetics), how sound waves travel through the air (acoustic phonetics), and how the human ear receives them (auditory phonetics).

2. Phonology: The study of the sound system of a language. Instead of just looking at the physical breath and muscle movements, phonology looks at how sounds are organised into meaningful patterns, rules, and structures within English to create meaning.

3. Prosodics (Prosody): The study of suprasegmental features of speech. "Suprasegmental" simply means features that stretch across units larger than an individual sound segment (such as across whole syllables, words, or sentences). This includes pitch, intonation, stress, volume, tempo, and pauses.

Analogy to Remember:
Phonetics: The physical keys and vibrating strings of a piano.
Phonology: The musical scale and chords that make up a song.
Prosodics: The volume, tempo, rhythm, and emotional expression added by the pianist.

Key Takeaway: Phonetics is about physical sound production, phonology is about the linguistic system of sounds, and prosodics is about the overarching rhythm, tone, and delivery.


2. The Segmental Sound System: Phonemes, IPA, Consonants, and Vowels

Phonemes and the International Phonetic Alphabet (IPA)

A phoneme is the smallest contrastive unit of sound in a language system that can distinguish one word from another. In English linguistics, phonemes are always transcribed between forward slashes, such as /p/ or /b/. Changing a single phoneme changes the meaning of a word (e.g., swapping /p/ in pat to /b/ creates bat).

Because the English spelling system (orthography) does not always match the way words sound, linguists use the International Phonetic Alphabet (IPA). The IPA provides a standardised, universal symbol for every distinct speech sound, ensuring clear and accurate transcription.

Classifying Consonants: The Three Criteria

Every consonant sound in English is classified by three essential characteristics:

1. Place of Articulation (Where the sound is made):
Bilabial: Using both lips (e.g., /p/, /b/, /m/).
Labiodental: Using the lower lip and upper teeth (e.g., /f/, /v/).
Dental / Interdental: Tongue between or behind the teeth (e.g., the 'th' sounds in thin /θ/ and this /ð/).
Alveolar: Tongue touching the alveolar ridge just behind the upper front teeth (e.g., /t/, /d/, /s/, /z/, /n/, /l/).
Post-alveolar: Tongue just behind the alveolar ridge (e.g., /ʃ/ as in ship, /ʒ/ as in measure).
Palatal: Tongue against the hard palate on the roof of the mouth (e.g., /j/ as in yes).
Velar: Back of the tongue against the soft palate or velum (e.g., /k/, /ɡ/, /ŋ/ as in sing).
Glottal: In the space between the vocal cords (the glottis) (e.g., /h/ or the glottal stop [ʔ]).

2. Manner of Articulation (How the sound is made):
Plosive (Stop): Complete blockage of the airflow followed by an explosive release (e.g., /p/, /b/, /t/, /d/, /k/, /ɡ/).
Fricative: Air forced through a narrow gap, creating continuous friction/hissing (e.g., /f/, /v/, /s/, /z/, /ʃ/, /θ/).
Affricate: A combination of a plosive release into a fricative (e.g., /tʃ/ as in church, /dʒ/ as in judge).
Nasal: Airflow diverted through the nasal cavity while the mouth is blocked (e.g., /m/, /n/, /ŋ/).
Lateral / Approximant: Air escapes smoothly around the sides of the tongue or with minimal constriction (e.g., /l/, /r/, /w/, /j/).

3. Voicing (Vocal cord vibration):
Voiced: The vocal cords vibrate during production (e.g., /b/, /d/, /ɡ/, /z/, /v/).
Voiceless: The vocal cords remain relaxed and open, so they do not vibrate (e.g., /p/, /t/, /k/, /s/, /f/).

Quick Physical Check: Place your fingers gently against the front of your throat (your larynx/Adam's apple). Say a prolonged "zzzzz" sound—you will feel a strong vibration (voiced). Now switch to a prolonged "sssss" sound—the vibration stops completely (voiceless)!

Vowels: Monophthongs, Diphthongs, and the Schwa

Unlike consonants, vowels are produced with an open vocal tract without audible friction.

Monophthongs: Pure, unchanging vowel sounds. These can be short vowels (e.g., /ɪ/ in bit) or long vowels (e.g., /iː/ in beat).
Diphthongs: Glide vowels where the tongue moves smoothly from one vowel position to another within a single syllable (e.g., /aɪ/ in bite, /eɪ/ in day, /ɔɪ/ in boy).
The Schwa (/ə/): The most common sound in spoken English. It is a neutral, central, unstressed vowel sound heard in the first syllable of about, the second syllable of butter, or in unstressed grammatical words like a and the.

Key Takeaway: Consonants are defined by place, manner, and voicing. Vowels are either pure monophthongs or gliding diphthongs, with the unstressed schwa (/ə/) serving as the workhorse of fluent English speech.


3. Connected Speech and Phonological Processes

In natural conversation, we do not speak word-by-word like robots. Words blend together in continuous streams known as connected speech. When sounds meet, several regular phonological processes occur:

1. Assimilation: When a sound changes to become more similar to a neighbouring sound.
Example: Saying ten boys as /tem bɔɪz/, where the alveolar nasal /n/ shifts to the bilabial nasal /m/ in anticipation of the bilabial plosive /b/.

2. Elision (Deletion): The omission or dropping of an unstressed sound, vowel, or consonant in connected speech to make pronunciation faster and easier.
Example: Saying friendship as /ˈfrenʃɪp/ (dropping the /d/), or saying camera as /ˈkæmrə/ (eliding the middle vowel).

3. Insertion / Epenthesis (Addition): Inserting an extra sound between words or within a word to ease articulation.
Example: The linking /r/ (pronouncing the 'r' in four apples) or the intrusive /r/ (inserting an /r/ sound between words ending and starting with vowels, such as saying law and order as "law-r-and order").

4. Glottalisation / Glottal Stop ([ʔ]): The replacement of a plosive consonant (most famously /t/) with a momentary closure of the vocal cords.
Example: Pronouncing butter as [ˈbʌʔə] or water as [ˈwɔːʔə]. Glottal stopping is common across many UK sociolects and regional accents.

5. Weak Forms: The reduction of grammatical function words (like pronouns, prepositions, and conjunctions) to unstressed versions featuring a schwa in fluent speech.
Example: In isolation, to is pronounced with a long vowel /tuː/, but in connected speech (e.g., I want to go), it reduces to the weak form /tə/.

Key Takeaway: Connected speech processes (assimilation, elision, insertion, glottalisation, and weak forms) are standard, efficient mechanisms of human communication—never label them as "errors" or "sloppy speech" in your exam essays!


4. Sound Patterning and Stylistics

Writers, advertisers, public speakers, and poets deliberately use sound devices to create mood, emphasis, and memorable representations. When analysing written texts that represent speech or use creative acoustic styling, look for these patterns:

Alliteration: The repetition of identical initial consonant sounds in adjacent or nearby words (e.g., big bold banners).
Sibilance: A specific form of alliteration/consonance involving the repetition of hissing or hushing fricative and affricate sounds, such as /s/, /z/, /ʃ/, and /tʃ/ (e.g., the soft sea sighs). Sibilance often creates calming, sinister, or whispery tones.
Assonance: The repetition of identical or similar vowel sounds within nearby words (e.g., the light of the fire /aɪ/).
Consonance: The repetition of identical consonant patterns across or at the ends of words (e.g., the fluck and fluck of the lock).
Onomatopoeia: Words that phonetically resemble the sound they describe.
  – Lexical onomatopoeia: Standard dictionary words (e.g., crash, buzz, whisper, bang).
  – Non-lexical onomatopoeia: Invented, non-standard spelling representations of sounds (e.g., vroom, ker-ching, sploosh).
Sound Iconicity (Phonological Symbolism): The direct associative link between speech sounds and physical qualities. For instance, high front vowels like /iː/ (in tiny, mini, wee) often evoke smallness, lightness, or delicacy, while heavy voiced plosives like /b/ and /d/ often evoke weight, impact, or force.

Key Takeaway: In literary, advertising, or persuasive texts, sound patterning works alongside semantics to guide the reader's emotions, imagery, and reactions.


5. Suprasegmental Features: Prosodics

When analysing spoken transcripts (such as in Paper 1 Section A, Child Language Development in Section B, or Paper 2 Diversity), you must evaluate how speakers use prosody to manage interaction, signal emotion, and structure their turns.

1. Pitch and Intonation

Pitch: The perceived highness or lowness of a speaker's voice.
Intonation: The pitch contour or melodic movement across an utterance (e.g., a rising tone, falling tone, or fall-rise).
Function: Intonation distinguishes declarative statements (falling pitch) from questions (rising pitch). It also signals attitude (e.g., sarcasm or enthusiasm) and turn-taking in conversation.
High Rising Terminal (HRT / Uptalk): A rising intonation contour at the end of a declarative statement. This can signal insecurity, seek validation, or check that the listener is following along.

2. Stress and Accentuation

Word Stress: Emphasising a specific syllable within a word (e.g., present [noun] vs. present [verb]).
Tonic Prominence (Sentence Stress): Placing heavy emphasis on a key word in an utterance to signal contrast or pragmatic importance (e.g., "I didn't steal the money" vs. "I didn't steal the money").

3. Volume (Dynamics)

• The loudness or softness of speech.
Function: Increased volume can establish dominance, assert authority, or signal excitement/anger. Decreased volume (whispering) may create intimacy, secrecy, or signal hesitation.

4. Tempo (Pace)

• The speed of spoken delivery.
Function: Accelerated tempo can signal excitement, urgency, high conversational involvement, or nervousness. Decelerated (slow) tempo is often used for emphasis, clarity, solemnity, or careful cognitive formulation.

5. Pauses and Timing

In AQA spoken transcripts, pauses are represented with specific notation:
Micro-pause: Indicated by a full stop within parentheses: (.) (a brief pause of less than 0.5 seconds).
Timed pause: Indicated by numbers within parentheses: (1.5) (a measured pause lasting 1.5 seconds).

Why do speakers pause?
1. Cognitive planning: The speaker is formulating thoughts or searching for vocabulary.
2. Hesitation / Nervousness: Reflecting uncertainty or emotional tension.
3. Conversational Turn-Taking: Marking transition-relevance places where another speaker may begin speaking.
4. Dramatic / Rhetorical Effect: Allowing key points to sink in with an audience.
5. Child Language Milestones: In language acquisition, pauses often show a child processing complex grammatical or phonetic combinations.

Key Takeaway: Prosodics transforms raw words into dynamic communication. Always explain how pitch, stress, tempo, volume, and pauses shape meaning and power dynamics in spoken transcripts.


6. Common Student Pitfalls and Examiner Insights

Avoid these common traps highlighted in AQA examiner reports:

1. Confusing Orthography (Spelling) with Phonetics:
Common Mistake: Claiming that "phone" and "party" alliterate because they both start with the letter p.
The Fix: Always listen to the sound, not the spelling! Phone begins with the labiodental fricative /f/, whereas party begins with the bilabial plosive /p/. They do NOT alliterate.

2. Feature-Spotting Without Explaining Function:
Common Mistake: Writing: "The text contains sibilance and plosives." (This earns very few marks).
The Fix: Link the linguistic feature directly to meaning, context, and effect: "The speaker's use of repeated alveolar plosives (/t/, /d/) creates a harsh, authoritative delivery, reinforcing their dominant stance in the debate."

3. Vague Labelling of Pauses:
Common Mistake: Writing: "The speaker pauses at (.) which shows they stop talking."
The Fix: Explain the pragmatic or developmental reason for the pause (e.g., turn allocation, cognitive load, hesitation, or dramatic emphasis).

4. Judging Regional Variation as "Incorrect" or "Lazy":
Common Mistake: Describing glottal stopping or h-dropping as "bad English" or "poor pronunciation."
The Fix: Maintain objective linguistic neutrality. Treat regional sound patterns as systematic phonological variations that reflect identity, dialect, and social group membership.

5. Forcing Phonological Analysis on Inappropriate Texts:
Common Mistake: Analysing sound patterns in a formal, purely informational written report where acoustic effects play no deliberate role.
The Fix: Apply phonetics and prosody where they are relevant—spoken transcripts, poetry, advertising slogans, speeches, and child language data!


Quick Review Summary Checklist

Before sitting your exam, make sure you can:
• Distinguish clearly between phonetics, phonology, and prosodics.
• Classify consonants using place, manner, and voicing.
• Identify monophthongs, diphthongs, and the unstressed schwa (/ə/).
• Explain connected speech processes: assimilation, elision, insertion, glottalisation, and weak forms.
• Spot and evaluate stylistic sound patterns (alliteration, sibilance, assonance, consonance, onomatopoeia).
• Accurately analyse transcript prosody including pitch, stress, volume, tempo, and pauses (.) / (1.5) for full AO1 credit.