Printable · GCSE Higher · ages 14-16
Statistics worksheet — GCSE Higher
Fifteen questions across the statistics statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Non-calculator
Statistics worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- 1.A survey of 25 pupils in Derby records how many siblings each has: 0 siblings — 6 pupils, 1 sibling — 10 pupils, 2 siblings — 6 pupils, 3 siblings — 3 pupils. Calculate the mean number of siblings.
- 2.A scatter graph has 50 points. Most of them lie close to a rising line of best fit, but two of them lie a long way from that line. Write down how those two points should be treated.
- 3.A school has 1,500 pupils. The head teacher takes a random sample of 150 of them from the school register and asks how long they spend on homework. Rory says the sample is too small for the result to mean anything. Is Rory right? Give a reason for your answer.
- 4.A histogram shows the times, t minutes, taken by 120 visitors to complete an escape room. The bar for 0 ≤ t < 10 has a frequency density of 5 visitors per minute, the bar for 10 ≤ t < 20 has a frequency density of 2 visitors per minute, the bar for 20 ≤ t < 40 has a frequency density of 1.5 visitors per minute, and the bar for 40 ≤ t < 60 has a frequency density of 1 visitor per minute. Work out which class contains the median time.
- 5.A box plot for the ages of 40 members of a gym is drawn from this five-number summary: minimum 15, lower quartile 22, median 29, upper quartile 38, maximum 61. Work out the interquartile range shown by this box plot.
- 6.Harry counted the coins he found on each of six days: 5, 7, 8, 9, 11, 30. Write down the outlier.
- 7.A teacher records the number of pets owned by each of 25 pupils in a class; each pupil owns 0, 1, 2, 3 or 4 pets. The teacher wants to show how many pupils own each number of pets. Write down the most suitable type of chart for this data, and give a reason for your answer.
- 8.A scatter graph shows the number of days, x, that each of 16 tomato plants was watered and its height, y cm. The line of best fit has equation y = 1.5x + 4. Write down what the 1.5 in this equation tells you about the plants.y = 1.5x + 4
- 9.The marks scored by 11 pupils in a test are given in order: 12, 15, 18, 21, 24, 27, 30, 33, 36, 39, 42. Work out the lower quartile of these marks.
- 10.A scientist has grouped the lifetimes, in hours, of 300 batteries into classes of unequal width. She wants a diagram in which the number of batteries in a class is given by the area of its bar. Write down the type of diagram she should draw.
- 11.Two classes sat the same test. The 30 pupils in Class A had a mean mark of 72. The 20 pupils in Class B had a mean mark of 82. Work out the mean mark of all 50 pupils.
- 12.The times, t minutes, taken by 80 people to travel to work are grouped like this: 0 ≤ t < 10, 6 people; 10 ≤ t < 20, 14 people; 20 ≤ t < 30, 25 people; 30 ≤ t < 40, 20 people; 40 ≤ t < 50, 15 people. Work out the cumulative frequency for t < 30.
- 13.A dual bar chart shows the number of hours of rain recorded in Leeds and in Bristol on each of four days. Leeds: Monday 3 hours, Tuesday 5 hours, Wednesday 2 hours, Thursday 4 hours. Bristol: Monday 4 hours, Tuesday 4 hours, Wednesday 6 hours, Thursday 2 hours. Work out the greatest amount, in hours, by which Bristol's rainfall exceeded Leeds's rainfall on a single day.
- 14.A scatter graph of the number of hours, x, that pupils revised against their test score, y, has the line of best fit y = 2.5x + 15. Amelia wants a score of at least 80. Work out the least whole number of hours of revision the line of best fit suggests she needs.y = 2.5x + 15
- 15.A bus company runs two routes into the centre of Exeter. On ten weekdays the journey time on Route 1 was, in minutes: 22, 23, 24, 24, 25, 25, 26, 26, 27 and 28. On Route 2 it was: 18, 19, 20, 20, 21, 22, 26, 30, 36 and 38. A commuter must reach the centre on time every day. Work out the mean and the range for each route, and write down which route she should take.
Answer key
- (a) 1.24 — Method: for data given as a frequency table, the mean is Σfx ÷ Σf — multiply each value by its frequency, add the results, then divide by the total frequency. Working: 0 × 6 = 0. 1 × 10 = 10. 2 × 6 = 12. 3 × 3 = 9. So Σfx = 0 + 10 + 12 + 9 = 31. The total frequency is Σf = 6 + 10 + 6 + 3 = 25. Mean = 31 ÷ 25 = 1.24 siblings. Averaging the frequency column itself, (6 + 10 + 6 + 3) ÷ 4 = 6.25, mixes up the frequencies with the values they belong to. Writing down 1, the number of siblings with the highest frequency, gives the mode, not the mean. Writing down 31 stops after finding Σfx and forgets to divide by the total frequency, 25. Always divide Σfx by Σf — never stop at the top of the fraction.
- (b) Treat them as outliers and check them before deciding — Method: a point lying a long way from the pattern the rest of the data make is called an outlier, and an outlier is investigated before anything is done with it, because it may be an error in the data or it may be a genuine but unusual case. Working: 48 of the 50 points lie close to the rising line of best fit, so the trend is set by those 48; the two remaining points do not follow it, so they are identified as outliers and checked — a mistake in measuring or recording would be corrected, while a genuine reading would be kept and reported. Answer: treat them as outliers and check them before deciding what to do with them. The distractors: deleting them at once assumes that every point far from the line must be an error, which throws away real data; moving the line so that it passes through them assumes a line of best fit must touch particular points, when it is drawn to follow all 50; taking them as proof that there is no correlation lets two points overturn the pattern that the other 48 agree on.
- (a) No, 150 pupils are a tenth of the school, chosen at random — Method: judge a sample on two things, whether every member of the population had the same chance of being chosen, and whether the sample is large enough to carry a pattern. Working: the 150 pupils were drawn from the register of every pupil in the school, so no year group or set is shut out and no pupil chooses to take part; and 150 ÷ 1,500 = 0.1, so one pupil in ten has been asked. A random sample of that share is ample for an estimate of how long the school's pupils spend on homework. Answer: no, because 150 pupils are a tenth of the school and were chosen at random. The distractors: saying a random sample always gives the exact school figure reaches the same verdict for a reason that is false, since a second random sample of 150 would give a slightly different mean; saying 150 pupils cannot be picked at random from 1,500 treats randomness as something only a whole population can have, when drawing names from the register is exactly how a random sample is taken; saying that only asking all 1,500 could show anything rejects sampling altogether, which would leave no way to study any population too large to count.
- (c) The class with times from 10 up to 20 — Method: to find the median class from a histogram, first turn each bar's frequency density into a frequency using density × class width, build up the cumulative frequency, and find the first class whose cumulative frequency reaches or passes n ÷ 2. Working: the four classes have widths 10, 10, 20 and 20, so their frequencies are 5 × 10 = 50, 2 × 10 = 20, 1.5 × 20 = 30 and 1 × 20 = 20, which add to the 120 visitors stated. The median sits at position 120 ÷ 2 = 60. The cumulative frequency is 50 after the first class and 50 + 20 = 70 after the second, so the 60th visitor is reached during the second class. Answer: the median lies in the class 10 ≤ t < 20. Watch which class each shortcut lands on: the tallest bar belongs to the first class, with the highest frequency density, 5 — but the tallest bar shows where visitors are packed most densely, not where the middle visitor falls, and picking it lands one class too early, at 0 ≤ t < 10; taking half of the total TIME span instead of half of the total NUMBER of visitors, 60 minutes ÷ 2 = 30 minutes, lands in the class 20 ≤ t < 40, confusing a value on the horizontal axis with a position in the data; and using the full 120 visitors as the target position, rather than 120 ÷ 2 = 60, reaches all the way to the last class, 40 ≤ t < 60, treating the whole data set's size as though it were the position of a single middle value.
- (d) 16 — Method: on a box plot the interquartile range is the width of the box itself, upper quartile take away lower quartile. Working: the upper quartile is 38 and the lower quartile is 22, so 38 − 22 = 16. Answer: the interquartile range is 16 years. Watch which part of the box plot you are reading: the whole line from whisker to whisker gives the range, 61 − 15 = 46; the left half of the box alone gives median take away lower quartile, 29 − 22 = 7; and the right half of the box alone gives upper quartile take away median, 38 − 29 = 9 — neither half is the interquartile range on its own.
- (a) 30 — Method: an outlier is a value that lies far away from the pattern set by the rest of the data, so compare each value with the group the others form. Working: five of the counts, 5, 7, 8, 9 and 11, lie within 6 of one another and the steps between them are 2, 1, 1 and 2; the remaining count of 30 is 19 above the nearest of them, so it is the value that does not belong to the pattern. Answer: 30. The distractors: 5 comes from picking the smallest value, on the idea that the odd one out must be at the bottom of the list; 11 comes from ordering the data and stopping one value short, taking the largest of the counts that sit close together; 8.5 comes from working out the median, (8 + 9) ÷ 2, and giving a measure of centre where a value standing apart was asked for.
- (a) A vertical line chart (discrete numerical data) — The number of pets is discrete numerical data — whole-number values such as 0, 1, 2, 3 or 4 — recorded for one variable, so a vertical line chart is the chart specified for this kind of data. A bar chart is used for categorical data, such as favourite colour, not numerical values counted like this. A pie chart shows proportions of a whole and does not show the frequency of each separate value. A scatter graph compares two different variables against each other, and only one variable, the number of pets, is recorded here.
- (b) On average a plant grew 1.5 cm taller for each extra day — Method: in the equation of a line, the number multiplying x is the gradient, and a gradient states the change in y produced by an increase of 1 in x, read in the units of the two axes. Working: here x is measured in days and y in centimetres, so the gradient 1.5 carries the units centimetres per day. Testing it on the line, 5 days gives 1.5 × 5 + 4 = 11.5 cm and 6 days gives 1.5 × 6 + 4 = 13 cm, a rise of 1.5 cm for the one extra day. Answer: on average a plant grew 1.5 cm taller for each extra day of watering. The distractors: 1.5 cm as the height before any watering is the value of y when x is 0, which is the other number in the equation, 4 cm, so this swaps the gradient and the intercept; 1.5 cm as the gap between the tallest and the shortest plant reads the gradient as a range, when a range is a difference between two of the 16 plants and a gradient is a rate; 1.5 days for each extra centimetre inverts the rate, dividing days by centimetres instead of centimetres by days, and the line gives 1 cm of growth in two thirds of a day.
- (c) 18 — Method: for n ordered values, GCSE convention places the lower quartile at position (n + 1) ÷ 4, counting from the smallest value. Working: there are 11 marks, so n + 1 = 11 + 1 = 12 and 12 ÷ 4 = 3, so the lower quartile is the 3rd value in the ordered list 12, 15, 18, 21, 24, 27, 30, 33, 36, 39, 42, which is 18. Answer: the lower quartile is 18 marks. Watch the position you count to: dividing 11 ÷ 4 = 2.75 without adding 1 first, then rounding down, lands on the 2nd value, 15, not the 3rd; reaching for the middle of the whole list instead gives the median, 27, a different statistic; and averaging the 3rd and 4th values, 18 + 21 = 39 and 39 ÷ 2 = 19.5, borrows a method for an even split where it is not needed here.
- (c) A histogram, with frequency density up the vertical axis — Method: decide which diagram makes area stand for frequency, which is the property the question asks for. Working: on a histogram the vertical axis is frequency density, so the area of a bar is frequency density × class width, and that product is the frequency; this is exactly what is wanted, and it is what allows classes of unequal width to be shown fairly. Answer: a histogram, with frequency density up the vertical axis. The distractors: a bar chart plots frequency as the height, so with unequal widths a wide class would cover far more area than a narrow class holding the same number of batteries, and area would measure nothing; a cumulative frequency diagram plots running totals against upper class boundaries, so a point on it gives how many lie below a value rather than how many lie in a class; a pie chart shows each class as a share of the whole 300 and loses the class widths entirely, so no area on it is tied to a scale of hours.
- (c) 76 marks — Method: a mean of means only works when the groups are the same size, so rebuild each class's total mark, add the totals and divide by the number of pupils altogether. Working: Class A scored 30 × 72 = 2160 marks and Class B scored 20 × 82 = 1640 marks, giving 2160 + 1640 = 3800 marks between 50 pupils, so the overall mean is 3800 ÷ 50 = 76 marks. Answer: 76 marks. The distractors: 77 marks comes from averaging the two class means, (72 + 82) ÷ 2, which ignores the different class sizes; 78 marks comes from attaching each mean to the other class's size, (30 × 82 + 20 × 72) ÷ 50; 3800 marks comes from stopping at the combined total and never dividing by 50.
- (d) 45 — Method: a cumulative frequency is a running total — it counts everybody in every class up to and including the one that ends at the value given. Working: the classes that lie wholly below 30 minutes are 0 ≤ t < 10, 10 ≤ t < 20 and 20 ≤ t < 30, with frequencies 6, 14 and 25, so the running total is 6 + 14 = 20 and then 20 + 25 = 45. Answer: 45 people took less than 30 minutes. The distractors: 25 comes from quoting the frequency of the class 20 ≤ t < 30 on its own instead of the running total; 65 comes from accumulating one class too many and including 30 ≤ t < 40, which is 45 + 20; 35 comes from accumulating from the top downwards, 15 + 20, which counts the people who took 30 minutes or more rather than fewer.
- (b) 4 — The difference, Bristol minus Leeds, on each day is: Monday 4 − 3 = 1, Tuesday 4 − 5 = −1, Wednesday 6 − 2 = 4, Thursday 2 − 4 = −2. The greatest amount by which Bristol exceeded Leeds is 4 hours, on Wednesday. Choosing 1 takes Monday's smaller positive difference instead of the greatest one. Choosing 2 takes the size of Thursday's difference, but that is the amount by which Leeds exceeded Bristol, the opposite direction to the one asked for. Choosing 6 takes Bristol's raw figure on Wednesday without subtracting Leeds's 2 hours first.
- (a) 26 — Method: a line of best fit lets one quantity be predicted from the other, so the score is substituted into the equation of the line and the resulting inequality is solved for the number of hours. Working: a score of at least 80 means 2.5x + 15 ≥ 80; taking 15 from both sides gives 2.5x ≥ 65, and dividing both sides by 2.5 gives x ≥ 26, so the least whole number of hours is 26. Checking, 2.5 × 26 + 15 = 80, which does reach the target. Answer: 26 hours — and this is only an estimate, because a line of best fit predicts a trend rather than an individual result, and a prediction made outside the range of hours the pupils actually revised for would be an extrapolation and less reliable still. The distractors: 27 comes from reaching 26 and then rounding up again, although 26 hours already gives a score of exactly 80; 32 comes from 80 ÷ 2.5, which ignores the 15 in the equation of the line; 38 comes from (80 + 15) ÷ 2.5, that is from adding the 15 instead of subtracting it when rearranging.
- (d) Route 1, as its times vary by 6 minutes rather than 20 — Method: work out an average and a measure of spread for each route, then decide which matters to a commuter who must arrive on time every day. Working: for Route 1, 22 + 23 + 24 + 24 + 25 + 25 + 26 + 26 + 27 + 28 = 250 and 250 ÷ 10 = 25, so the mean is 25 minutes, and the range is 28 − 22 = 6 minutes. For Route 2, 18 + 19 + 20 + 20 + 21 + 22 + 26 + 30 + 36 + 38 = 250 and 250 ÷ 10 = 25, so the mean is also 25 minutes, but the range is 38 − 18 = 20 minutes. The means give no reason to prefer either route; the spreads do, because a commuter who must never be late has to allow for the worst day, which is 28 minutes on Route 1 and 38 minutes on Route 2. Answer: Route 1, as its times vary by 6 minutes rather than 20. The distractors: saying Route 2 has the lower mean assumes that its quicker-looking early times must pull the average down, when both routes total 250 minutes over the ten days; choosing Route 2 for its fastest journey of 18 minutes judges a route by its best day, and the commuter has to survive its worst; saying either route will do uses the equal means and ignores the spread altogether, which is the one thing that separates the two routes.
Build your own mix at the worksheet builder.