Sample paper · GCSE Higher · grades 4–9
GCSE Higher sample Paper 1 (non-calculator)
The real Paper 1 is 1.5 hour 30 minutes and 80 marks, non-calculator, and all three papers carry equal weight. This sample is 20 original questions in the same content proportions as the Higher qualification — number 15%, algebra 30%, ratio, proportion and rates of change 20%, geometry and measures 20%, probability 7.5%, statistics 7.5% — with no calculator-only items. AO1 / AO2 / AO3 at this tier: 40% / 30% / 30%.
Non-calculatorGCSE Higher
Answer key: GCSE Higher sample Paper 1 (non-calculator)
MathsUKwww.geekhero.co.uk
- (c) An under-estimate, by 8 — Method: work out the exact product, then compare it with the estimate; an estimate that is smaller than the exact value is an under-estimate, and the difference between them is the size of the error. Working: 48 × 21 = 48 × 20 + 48 = 960 + 48 = 1,008, and 1,008 − 1,000 = 8, so the estimate falls short. Answer: an under-estimate, by 8. The distractors: an over-estimate by 8 has the size of the error right but the direction wrong, and comes from assuming that rounding 48 up to 50 must push the estimate above the exact value, without allowing for 21 being rounded down; an over-estimate by 19 comes from working out 48 × 21 as 48 × 20 + 21 = 981, adding a 21 where another 48 belongs; the claim that the estimate is exactly right comes from arguing that one number was rounded up and the other down, so the two changes must cancel.
- (c) Fastest at t = 2 min — steepest gradient. — The rate of cooling is given by the size (magnitude) of the gradient, ignoring its sign — the steeper the tangent, the faster the temperature is changing. Of −8, −3 and −0.5, the gradient −8 has the greatest magnitude, so the tea is cooling fastest at t = 2 minutes. 'Fastest at t = 20 min — largest gradient' confuses the signed value with the size of the rate: −0.5 is the largest NUMBER of the three, but it's the smallest in magnitude, meaning the tea is barely cooling at all by then. 'Cools at the same rate throughout' ignores that the three gradients are different sizes, not just all negative. 'Fastest at t = 10 min — the middle reading' isn't a mathematical reason at all — the gradients themselves have to be compared, not their position in the list.
- (c) 30 — Method: use y = kx and find k from the given pair of values, then substitute x = 12. Working: k = 20 ÷ 8 = 2.5, so y = 2.5 × 12 = 30. Answer: 30. 24 comes from treating the relationship as additive, adding the increase in x (12 − 8 = 4) straight onto y (20 + 4 = 24), instead of multiplying by k. 14.5 comes from finding k correctly (2.5) but then adding it to x instead of multiplying (12 + 2.5 = 14.5). 4.8 comes from finding k upside down, 8 ÷ 20 = 0.4, and multiplying by x: 12 × 0.4 = 4.8.
- (d) (3/4)a + (1/4)c — Method: OX = OA + AX, and since AX is a third of XC, AX is 1/4 of the whole of AC, with AC = c − a. Working: OX = a + 1/4(c − a) = a − (1/4)a + (1/4)c = (3/4)a + (1/4)c. Answer: OX = (3/4)a + (1/4)c. Measuring 1/4 of AC from C's end instead of A's swaps the fractions round, giving (1/4)a + (3/4)c; adding (1/4)c onto the whole of a without subtracting a inside the bracket first gives a + (1/4)c; and treating the ratio as though AX and XC were equal gives the midpoint, (1/2)a + (1/2)c. Convert the ratio to a fraction of AC measured from A, subtract before you scale, and then add the result to OA.
- (c) 30 — The numbers less than 4 are 1, 2 and 3, so the probability of that event is 3/6, and the expected count in 90 rolls is 90 × 3/6 = 45. The probability of rolling a 6 is 1/6, and the expected count is 90 × 1/6 = 15. The difference between the two expected counts is 45 − 15 = 30. A candidate who answers 45 has given the expected count for 'less than 4' only, forgetting to subtract the other expected count. A candidate who answers 15 has given the expected count for '6' only. A candidate who answers 36 has used a dice with 5 possible numbers instead of 6, giving 90 × 3/5 = 54 and 90 × 1/5 = 18, a difference of 36.
- (c) No — 8 from one class is too small to represent the school. — Method: judge reliability by asking whether the sample is both large enough, and spread across the population, relative to what it is meant to represent. Working: 8 pupils is a tiny fraction of the school's 1,000 pupils, and all 8 come from a single class rather than a range of year groups, so the sample is both too small and too narrow to represent the whole school reliably. She is not right. Saying any sample size gives an equally reliable estimate ignores that reliability generally improves with a larger, more representative sample. Saying the method is unreliable because it was not done online is not a reason connected to sample size or representativeness at all. Saying 8 is reliable because it is more than half her class compares the sample to the wrong population — the school has 1,000 pupils, not one class. Always judge a sample's size against the population it is meant to represent, not against a smaller group within it.
- (b) 30 — Method: the smallest matching total is the lowest common multiple of the two pack sizes. Working: multiples of 6 are 6, 12, 18, 24, 30 …; multiples of 10 are 10, 20, 30 …. The lowest common multiple is 30. 60 comes from working out 6 × 10 = 60, the product of the pack sizes rather than their lowest common multiple. 16 comes from working out 6 + 10 = 16, which is not a common multiple at all. 2 is the highest common factor of 6 and 10, not a total of tickets. Answer: 30.
- (a) 38 — Method: substitute the position number into the rule and follow the order of operations, so the squaring is carried out before the 2 is added. Working: n = 6 gives 6² + 2; 6² means 6 × 6 = 36, and then 2 is added to 36. Answer: 38. The distractors: 14 comes from multiplying the position by 2 instead of squaring it, 6 × 2 + 2; 36 comes from squaring correctly and then forgetting to add the 2; 64 comes from adding the 2 first and squaring afterwards, (6 + 2)².
- (a) 3 hours — Method: for a fixed pool the rate of flow multiplied by the time taken is constant, so multiplying the rate by a factor divides the time by that same factor. Working: tap B's rate is 2 times tap A's rate, so tap B's time is 6 ÷ 2 = 3 hours. Answer: 3 hours. The distractors: 12 hours comes from multiplying the time by 2 as well, which treats the time as directly proportional to the rate and has the faster tap taking longer; 4 hours comes from reading ‘twice as fast’ additively, as two hours quicker, and working out 6 − 2 instead of scaling the time by a factor of 2; 1.5 hours comes from applying the factor of 2 twice, halving 6 to 3 and then halving again.
- (a) 104° — Method: find each inscribed angle separately using the angle-at-the-centre theorem: for a point on the arc NOT cut off by the given central angle, halve that central angle; for a point on the OTHER arc, halve the REFLEX central angle instead; then subtract the smaller from the larger. Working: for C on the major arc, angle ACB = 76 ÷ 2 = 38 degrees. For D on the minor arc, D sees the reflex angle at the centre, 360 − 76 = 284 degrees, so angle ADB = 284 ÷ 2 = 142 degrees. The difference is 142 − 38 = 104 degrees. Answer: 104°. Both inscribed angles need the theorem applied separately, using the correct arc's central angle each time (the reflex angle for D), and the question asks for the DIFFERENCE between the two, not either angle on its own and not their sum, 180°, which is simply the opposite-angle total for the cyclic quadrilateral ACBD.
- (c) 100 — There are 3 even numbers on a fair dice (2, 4 and 6), so the probability of landing on an even number is 3/6 = 1/2, and 300 × 1/2 = 150. The probability of landing on a six is 1/6, so 300 × 1/6 = 50. The dice is expected to land on an even number 150 − 50 = 100 more times than on a six. Writing 50 is wrong because that is just the expected number of sixes on its own, without comparing it to the expected number of evens. Writing 150 is wrong because that is just the expected number of evens on its own, without subtracting the sixes. Writing 200 is wrong because it adds the two expected frequencies together (150 + 50 = 200) instead of finding the difference between them. The dice is expected to land on an even number 100 more times than on a six.
- (a) 12 — Method: for any two numbers, their highest common factor multiplied by their lowest common multiple equals the product of the two numbers. This holds because the HCF collects every prime factor the two numbers share, and the LCM collects every prime factor that appears in either number, so between them they use each prime factor of the two numbers exactly once — the same primes as the product. Working: 4 × 60 = 240, and 240 ÷ 20 = 12. 15 comes from working out 60 ÷ 4 = 15, dividing the wrong pair of numbers. 16 comes from working out 20 − 4 = 16, subtracting the highest common factor instead of using the product rule. 240 is 4 × 60, the product of the highest common factor and the lowest common multiple, left un-divided by 20. Answer: 12.
- (b) 6 — Method: for any point that lies on y = k/x, the value of k is found by multiplying the x-coordinate and the y-coordinate together, since k = x × y. Working: k = 2 × 3 = 6. Answer: k = 6. Distractor refutation: 1.5 comes from dividing the y-coordinate by the x-coordinate instead of multiplying them. 5 comes from adding the two coordinates instead of multiplying them. 9 comes from misreading the point's x-coordinate as 3 instead of 2, then multiplying 3 × 3.
- (c) 45% — Method: adding water changes the total volume but not the amount of fruit juice, so find the juice, find the new total volume, and write the first as a percentage of the second. Working: 3 × 0.6 = 1.8 litres of fruit juice; the new volume is 3 + 1 = 4 litres; 1.8 ÷ 4 = 0.45, which is 45%. Answer: 45%. The distractors: 60% is the strength before the water goes in, and assumes that adding water leaves the strength unchanged; 15% comes from dividing the 60% by the 4 litres of mixture instead of dividing the 1.8 litres of juice by the 4 litres; 75% is the fraction of the new mixture that came out of the original jug, 3 litres out of 4, which ignores that only 60% of that 3 litres was juice.
- (d) 10 — Method: the distance between two points is the hypotenuse of a right-angled triangle whose shorter sides are the horizontal and vertical gaps, so work out both gaps first, handling the negative coordinates carefully, and then apply Pythagoras' theorem. Working: the horizontal gap is 5 − (−3) = 5 + 3 = 8 and the vertical gap is 4 − (−2) = 4 + 2 = 6. Then d² = 8² + 6² = 64 + 36 = 100, so d = √100 = 10. Answer: 10. The distractors: 14 comes from adding the two gaps, 8 + 6, instead of adding their squares and taking the root; 100 comes from stopping at the sum of the squares and never taking the square root; 50 comes from reaching 100 correctly and then halving it instead of taking its square root, a candidate who has read the last step as “halve” rather than “root”.
- (a) y = (4/3)x + 25/3 — The radius from (0, 0) to (−4, 3) has gradient 3 ÷ (−4) = −3/4. The tangent is perpendicular to the radius, so its gradient is the negative reciprocal, 4/3. Using y − y₁ = m(x − x₁) with the point (−4, 3): y − 3 = (4/3)(x + 4), so y = (4/3)x + 16/3 + 3 = (4/3)x + 25/3. Using the radius's own gradient, −3/4, instead of taking the perpendicular gradient, gives y − 3 = (−3/4)(x + 4), which simplifies to y = −(3/4)x once the −3 and +3 in the constant cancel out. Taking the reciprocal of the radius's gradient but keeping the wrong sign, using −4/3 instead of 4/3, gives y = −(4/3)x − 7/3. Correctly finding the gradient 4/3 and expanding the bracket, but forgetting to add the y-coordinate 3 at the end, gives y = (4/3)x + 16/3.
- (b) 4 — Method: find the length scale factor by taking the cube root of the volume scale factor, then square it to get the area scale factor. Working: 8 = 2³, so the length scale factor is 2, and the area scale factor is 2² = 4. Answer: 4. 8 comes from using the volume scale factor itself as if it were the area scale factor. 64 comes from squaring the volume scale factor, 8² = 64, instead of first taking its cube root. 2 comes from correctly finding the length scale factor but then forgetting to square it.
- (a) A smaller circle — Since the cutting plane is parallel to the circular base, the cross-section is also a circle, but smaller than the base because the cone narrows as it rises towards the apex, so 'a smaller circle' is correct. 'A triangle' wrongly describes the outline seen from the side of the cone, not a horizontal cross-section. 'An ellipse' would only result from a cut made at an angle to the base, not one parallel to it. 'The same size circle as the base' wrongly ignores that the cone tapers, so any parallel cross-section above the base must be smaller.
- (c) y = 4x − 5 — Parallel lines have the same gradient, so the new line has gradient 4; since it passes through (0, −5), its y-intercept is −5, giving y = 4x − 5. A candidate who drops the negative sign on the y-intercept would write y = 4x + 5. A candidate who changes the sign of the gradient, instead of keeping it the same for a parallel line, would write y = −4x − 5. A candidate who confuses m and c, using the y-intercept of the first line (3) as the gradient of the second, would write y = 3x − 5.
- (c) (6n + 4)/2 — Six full boxes hold 6 lots of n pencils, which is 6n, and the 4 loose pencils are added on, so the shop has 6n + 4 pencils altogether. Sharing them equally between 2 classes divides that whole total by 2, and brackets are what show that the division applies to all of it: (6n + 4)/2. Without the brackets, 6n + 4/2 halves only the loose pencils; 6(n + 4)/2 adds the loose pencils to every box before the division; 2(6n + 4) doubles the total instead of halving it.
How the 20 questions are shared out
- Number — 3 questions (15% of the qualification)
- Algebra — 6 questions (30% of the qualification)
- Ratio, proportion and rates of change — 4 questions (20% of the qualification)
- Geometry and measures — 4 questions (20% of the qualification)
- Probability — 2 questions (7.5% of the qualification)
- Statistics — 1 question (7.5% of the qualification)
Where an area has fewer printable questions than its share, the shortfall is filled from the other areas. These are original questions, not past papers.