Understanding KS2 SATs Scaled Score Conversion: Raw Marks to National Benchmarks

When Year 6 mock papers or official statutory test results come home, parents are often met with a confusing dual system: raw marks and scaled scores. For parents navigating the end of Key Stage 2, grasping how ks2 sats scaled score conversion works is essential for making sense of academic progress and preparing for the transition to secondary school.

In the National Curriculum assessments, raw marks—the actual number of marks a pupil scores on a test paper—are transformed into a standardised scaled score ranging from \(80\) to \(120\). A scaled score of \(100\) represents the statutory Expected Standard (EXS) set by the Department for Education, while a score of \(110\) or above signifies Working at Greater Depth (GDS). A score below \(100\) indicates that a pupil has not yet reached the expected standard and requires structured intervention to bridge foundational subject gaps before Year 7.

How Raw Marks Map Across Key Stage 2 Papers

To understand where your child sits on the \(80\)–\(120\) spectrum, it helps to examine the raw mark architecture of the three externally marked Key Stage 2 subjects:

1. Mathematics (Total: \(110\) Marks)

The mathematics assessment is split into three distinct papers: Paper 1 (Arithmetic, \(40\) marks in \(30\) minutes), Paper 2 (Reasoning, \(35\) marks in \(40\) minutes), and Paper 3 (Reasoning, \(35\) marks in \(40\) minutes). Raw marks from all three papers are combined out of \(110\) before being converted into the single scaled score. Historically, securing around \(55\%\) to \(58\%\) of raw marks (approx. \(58\)–\(61\) marks) has aligned with the \(100\) expected standard threshold, whereas Greater Depth typically demands around \(85\%\)+ (approx. \(94\)–\(96\) marks).

2. Reading (Total: \(50\) Marks)

The reading comprehension paper consists of three unrelated texts of increasing difficulty with \(50\) raw marks available in \(60\) minutes. Because comprehension text complexity varies annually, raw mark thresholds fluctuate. Meeting the expected standard generally requires roughly \(26\)–\(29\) marks out of \(50\), while Greater Depth usually requires \(38\)–\(41\) marks.

3. Grammar, Punctuation and Spelling (GPS / SPaG) (Total: \(70\) Marks)

The GPS assessment combines Paper 1 (short-answer grammar and punctuation questions, \(50\) marks in \(45\) minutes) and Paper 2 (a \(20\)-word spelling test, \(20\) marks). The expected standard benchmark generally sits around \(35\)–\(38\) marks out of \(70\), with Greater Depth reaching upwards of \(54\)–\(56\) marks.

Why Raw Marks Fluctuate: Equating and Cohort Standards

Parents often wonder why a raw mark of \(30/50\) in Reading might yield a scaled score of \(103\) in one year but \(101\) in another. This variation is intentional. The Standards and Testing Agency (STA) employs a psychometric process called equating. Even with rigorous pre-testing, minor variations in question difficulty occur between years.

Converting raw marks to a fixed \(80\)–\(120\) scale ensures that a score of \(100\) always represents the exact same standard of curriculum mastery, irrespective of whether a particular year's test was marginally harder or easier. While this maintains national comparability across cohorts, it can lead parents into two common cognitive traps: false confidence and false alarm.

    The False Confidence Trap: A pupil scoring \(102\) has technically met the expected standard, but this may mask critical vulnerabilities in foundational arithmetic or structural grammar that will hinder them in Key Stage 3.

    The False Alarm Trap: A raw mark drop from \(75\%\) to \(62\)% on a mock paper often reflects an unusually demanding reasoning text or complex multi-step fractions paper rather than actual regression.

Diagnostic Error Triage: Moving Beyond the Number

Looking only at the scaled score tells you where your child landed, but nothing about why. To drive meaningful improvement, parents should work through test papers using a three-tier diagnostic framework:

1. Fluency & Cognitive Load Slips

These occur when core mental retrieval falters under timed conditions. Common examples include basic multiplication table errors in long division, misreading unit conversions (e.g., grams to kilograms), or missing capital letters in GPS sentence transformations. These do not require re-teaching the whole topic; they require short, high-frequency retrieval practice drills.

2. Conceptual Misconceptions

These are genuine gaps in understanding. In maths, this frequently involves dividing fractions by whole numbers (\(\frac{2}{5} \div 3\)), working with multi-step word problems involving ratio, or calculating missing angles in composite shapes. In reading, it manifests as confusing literal retrieval with 3-mark authorial intent questions. Conceptual errors require Socratic deconstruction and visual models before returning to mock questions.

3. Question Decentering and Exam Craft

Often, a child understands the underlying concept but loses marks due to test mechanics—such as failing to provide two distinct pieces of textual evidence for a 2-mark reading question, or omitting units in area calculations. Pinpointing these formatting slips unlocks rapid raw mark gains without curriculum anxiety.

Building a Post-Exam Action Plan with AI-Powered Practice

Rather than purchasing stacks of generic commercial workbooks that repeat concepts your child has already mastered, targeted post-exam preparation should be surgical. Analysing where raw marks were dropped enables you to design micro-study routines that specifically target high-yield objectives.

This is where modern technology transforms home revision. With Thinka's interactive learning platform, families can move from static mock results to automated diagnostic mastery. Instead of spending hours categorising errors manually, parents can use adaptive AI practice on Thinka to identify the exact sub-skills—such as relative clauses, equivalent fractions, or inference vocabulary—that will move a child from a scaled score of \(96\) across the \(100\) threshold, or propel a \(106\) pupil to Greater Depth (\(110+\)).

Complementing interactive revision with structured revision notes—such as these curated KS2 study notes and revision resources—gives Year 6 learners clear explanations alongside targeted questions. Educators and tutors can also utilise classroom assessment tools for teachers to track cohort trends against National Curriculum descriptors and generate bespoke mock papers.

Empowering Your Year 6 Child for the Road Ahead

The primary scaled score is not a ceiling on your child's academic capability; it is a snapshot of curriculum fluency at a single point in time. By understanding how raw marks convert into scaled benchmarks and diagnosing the precise error patterns behind lost marks, you can turn assessment data into an empowering, stress-free roadmap that builds confidence for SATs week and sets a resilient foundation for Year 7.