Printable · GCSE Foundation · ages 14-16
Statistics worksheet — GCSE Foundation
Fifteen questions across the statistics statements at Foundation tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Non-calculator
Statistics worksheet — GCSE Foundation
MathsUKwww.geekhero.co.uk
- 1.A scatter graph shows the number of guests, x, at a wedding and the length of buffet table needed, y metres. The line of best fit is y = 0.5x + 2. Write down what the 2 in this equation tells you about the buffet table.y = 0.5x + 2
- 2.The lowest temperature in a garden in Sheffield was recorded on each of seven days: 3 °C, 5 °C, 2 °C, 7 °C, 4 °C, 6 °C and 3 °C. Write down what the range of these seven temperatures describes.
- 3.The age, in years, and the score in a reaction test are recorded for eight members of a sports club: (14, 92), (18, 88), (23, 85), (27, 80), (31, 78), (36, 74), (42, 70), (49, 65). The eight pairs are plotted on a scatter graph. Describe the correlation between age and score.
- 4.A straight line is drawn on a scatter graph to show the trend of the points. Write down the name given to this line.
- 5.A dual bar chart shows the number of hours of rain recorded in Leeds and in Bristol on each of four days. Leeds: Monday 3 hours, Tuesday 5 hours, Wednesday 2 hours, Thursday 4 hours. Bristol: Monday 4 hours, Tuesday 4 hours, Wednesday 6 hours, Thursday 2 hours. Work out the greatest amount, in hours, by which Bristol's rainfall exceeded Leeds's rainfall on a single day.
- 6.Two classes at a school in Coventry sit the same maths test, out of 20 marks. Class A has a mean mark of 14 and a range of 6. Class B has a mean mark of 14 and a range of 14. Write a sentence comparing the two classes, using the mean and the range.
- 7.On a scatter graph the horizontal axis shows height in centimetres and the vertical axis shows mass in kilograms. One point is plotted at (170, 65). Write down the height and the mass of that person.
- 8.A café in York counts the number of customers in each of the nine hours it is open on one day: 4, 4, 4, 11, 13, 15, 18, 22 and 25. The owner says that a typical hour has about 4 customers, because 4 is the mode. Is the owner right? Give a reason for your answer.
- 9.A company makes 50,000 light bulbs a day and wants to check how long they last before they fail. Testing a bulb to find out how long it lasts destroys it. Give a reason why the company should test a sample of bulbs rather than every bulb it makes.
- 10.A charity shop in Bath holds 2,000 books. Volunteer A checks a random sample of 50 books and finds 35 paperbacks. Volunteer B checks a different random sample of 50 books and finds 31 paperbacks. Work out the estimate each sample gives for the whole stock, and write down what the shop should do next.
- 11.A café sold 10 sandwiches on each of four days: 10, 10, 10, 10. Work out the range of the numbers sold.
- 12.The mean of the numbers x, 20 and 30 is equal to the mean of the numbers 15 and 25. Work out the value of x.
- 13.Noah wrote down the numbers 9, 3, 7, 1, 5 and said, “The median is 7, because 7 is in the middle of my list.” Is Noah right? Give a reason for your answer.
- 14.A random sample of 50 pupils at a school were asked whether they prefer sport to music. 30 of the 50 pupils said they prefer sport. The school has 500 pupils altogether. Work out an estimate for the number of the 500 pupils who prefer sport.
- 15.Write down the statement that correctly describes the difference between correlation and causation.
Answer key
- (b) At 0 guests, the model predicts 2 m of table — The y-intercept of a line of best fit y = mx + c is the value of y when x = 0. Here y = 0.5 × 0 + 2 = 2, so the line predicts a table length of 2 m when there are 0 guests. The 2 m does not grow as more guests arrive — that role belongs to the gradient, 0.5 — so an option saying each extra guest adds 2 m has swapped the two numbers around. The 2 is a length in metres, not a number of guests, so an option requiring 2 guests before set-up has misread its units. And the table length does change with x, since it is 0.5x + 2 and not a fixed value, so an option claiming the table is always 2 m ignores the 0.5x term completely.
- (d) How much the temperatures varied over the seven days — Method: the range of a set of values is the largest value take away the smallest, so it is built from two values only and it measures the gap they leave between them. Working: the largest of the seven readings is 7 °C and the smallest is 2 °C, so the range is 7 − 2 = 5 °C. That figure says the week's readings covered a band 5 °C wide; it names no particular day and no particular reading. Answer: the range describes how much the temperatures varied over the seven days. The distractors: the temperature that occurred most often is the mode, which here is 3 °C, and a mode counts repeats instead of measuring a gap; the temperature typical of the week is an average, and the range is not an average, since it throws away every value lying between the two extremes; the number of different temperatures recorded is 6, a count of how many distinct values appear, while the range is a difference between two of them.
- (c) Strong negative correlation — Method: correlation is described by two things — the direction the points take as the graph is read from left to right, and how closely the points lie to a single straight line. Working: reading the pairs in order of age, the ages rise 14, 18, 23, 27, 31, 36, 42, 49 while the scores fall 92, 88, 85, 80, 78, 74, 70, 65; the score falls at every single step, with no reversal anywhere, so the points fall from left to right and lie close to a straight line. Answer: strong negative correlation — negative for the falling direction, strong because every point follows the pattern. The distractors: strong positive correlation comes from noticing a clear pattern and calling any clear pattern positive, without checking the direction; weak negative correlation comes from reading the direction correctly but judging points that do not lie exactly on a straight line to be only loosely related, when these eight fall without a single exception; no correlation comes from reading a falling trend as though it showed no relationship at all, when a falling trend is itself a relationship.
- (c) A line of best fit — Method: the straight line drawn on a scatter graph is named from the job it does — it is chosen so that it follows the whole set of points as closely as possible. Working: the line passes through the middle of the points, with roughly as many points above it as below it, and it need not pass through any of the plotted points at all; the name given to the straight line chosen in that way is a line of best fit. Answer: a line of best fit. The distractors: a line of symmetry comes from confusing a trend with symmetry, which is a property of a shape rather than of a set of data; a horizontal line through the mean comes from thinking the trend is shown by an average, when a horizontal line would say that the vertical quantity does not change and so show no correlation; a line joining the first and last points comes from thinking the line must join the two extreme points, which lets two points decide a trend that all of the points should share in.
- (b) 4 — The difference, Bristol minus Leeds, on each day is: Monday 4 − 3 = 1, Tuesday 4 − 5 = −1, Wednesday 6 − 2 = 4, Thursday 2 − 4 = −2. The greatest amount by which Bristol exceeded Leeds is 4 hours, on Wednesday. Choosing 1 takes Monday's smaller positive difference instead of the greatest one. Choosing 2 takes the size of Thursday's difference, but that is the amount by which Leeds exceeded Bristol, the opposite direction to the one asked for. Choosing 6 takes Bristol's raw figure on Wednesday without subtracting Leeds's 2 hours first.
- (a) Equal means; Class A is more consistent, smaller range. — Method: when two data sets share a measure of location, compare a measure of spread to say more about consistency. Working: both classes have the same mean mark, 14, so on average they performed equally well. Class A has the smaller range, 6, so its marks are more tightly grouped around 14 than Class B's marks, which vary by as much as 14. So Class A's marks were more consistent, even though neither class did better on average. Saying Class B did better because it has the bigger range confuses a wide spread with a high score — a big range describes variability, not performance. Saying Class A did better because it has the smaller range makes the same mistake in the other direction: the two classes are tied on the mean, so neither one 'did better'. Saying the classes cannot be compared because their means are equal misses the whole point of also comparing the range. Always compare both an average AND a spread before describing two data sets — either one alone tells only half the story.
- (c) Height 170 cm, mass 65 kg — Method: a point on a scatter graph is written as a pair of coordinates in which the horizontal value is written first and the vertical value second, so each value is matched to the quantity named on its own axis. Working: in (170, 65) the value 170 is the horizontal coordinate and the horizontal axis shows height in centimetres, so the height is 170 cm; the value 65 is the vertical coordinate and the vertical axis shows mass in kilograms, so the mass is 65 kg. Answer: height 170 cm, mass 65 kg, each with the unit named on its own axis. The distractors: height 65 cm and mass 170 kg come from reading the pair the wrong way round, which would describe an impossible person; height 170 cm and mass 170 kg come from reading the horizontal coordinate for both quantities and never using the second number; height 235 cm and mass 105 kg come from combining the two coordinates, 170 + 65 and 170 − 65, instead of reading them separately.
- (c) No, the mode here is the lowest value of the nine — Method: an average is meant to stand for the data as a whole, so test any proposed average by asking how many values it sits near. Working: the value 4 appears three times and every other count appears once, so 4 is indeed the mode. But those three hours are the quiet ones at the start of the day, and the other six counts run from 11 up to 25; putting the nine counts in order, the middle one is the fifth, which is 13. So the mode sits at the very bottom of the data, with six of the nine hours far above it. Answer: no, because the mode here is the lowest value of the nine, so it describes the quiet opening hours rather than a typical hour. The distractors: saying the mode can only be used when no value repeats reverses the definition, since a mode exists only because a value does repeat; saying the mode is the value that occurs most often is a correct definition, but being the commonest value does not make a value typical when it lies at one end of the data; saying the mode is the best average for any list of numbers ignores the fact that mean, median and mode each describe a population well in different circumstances.
- (d) Testing destroys bulbs, so testing all leaves none to sell. — Method: testing every item in a population instead of a sample is a census — sensible only when testing does not use up or destroy what is being tested. Working: here, testing a bulb to find its lifespan destroys it, so testing all 50,000 bulbs would leave nothing left to sell — a sample lets the company estimate the typical lifespan without destroying its whole stock. Extra electricity used in testing is not the real reason a census is avoided here — it is the destruction of the product that matters. Saying a sample is always more accurate than a full census is the wrong way round: a census, if it could be carried out, gives the exact figure for the whole population — it is testing being destructive, not a lack of accuracy, that rules it out here. There is no law against testing every item a company makes — nothing in the question suggests that. When testing destroys the item being tested, sampling is necessary, not just convenient.
- (a) 1,400 and 1,240, so combine the samples for one estimate — Method: scale each sample up to the whole stock, then use the fact that a larger sample gives a more reliable estimate than a smaller one. Working: the first sample gives 35 ÷ 50 = 0.7 and 0.7 × 2,000 = 1,400 paperbacks; the second gives 31 ÷ 50 = 0.62 and 0.62 × 2,000 = 1,240 paperbacks. Two random samples of the same size are expected to differ a little, so neither estimate is wrong. Putting the two together gives 35 + 31 = 66 paperbacks in 100 books, and 66 ÷ 100 = 0.66 with 0.66 × 2,000 = 1,320, an estimate resting on twice as many books as either volunteer checked. Answer: 1,400 and 1,240, so combine the samples for one estimate. The distractors: keeping 1,400 because it is larger picks an estimate by its size, when both samples held 50 books and neither has a stronger claim; saying a volunteer must have miscounted assumes two random samples ought to agree exactly, which is precisely what random sampling does not promise; 1,750 and 1,550 come from 35 × 50 = 1,750 and 31 × 50 = 1,550, multiplying each count by the size of the sample instead of scaling by 2,000 ÷ 50.
- (d) 0 — Method: the range is the largest value minus the smallest value, whatever those two values turn out to be. Working: every value is 10, so the largest value is 10 and the smallest value is 10 as well, and the range is 10 − 10 = 0. Answer: 0 — a range of nothing says the data do not vary at all. The distractors: 10 comes from writing down the repeated value itself instead of the difference between the extremes; 20 comes from adding the largest and the smallest, 10 + 10, instead of subtracting; 40 comes from adding all four values, which gives the total sold and not a measure of spread.
- (c) 10 — Method: work out the mean that can be found straight away, then use total = mean × number of values on the group of three to find the missing number. Working: the mean of 15 and 25 is (15 + 25) ÷ 2 = 40 ÷ 2 = 20, so the group of three must also have a mean of 20; three numbers with a mean of 20 have a total of 20 × 3 = 60, and 20 + 30 = 50 of that total is already accounted for, so x = 60 − 50 = 10. Answer: 10, and checking, (10 + 20 + 30) ÷ 3 = 20. The distractors: 20 comes from working out the mean the two groups share and writing that down as x; −10 comes from dividing the group of three by 2 instead of by 3, which gives x + 50 = 40; 70 comes from reading the total 15 + 25 = 40 as the mean of the pair, which sets the target total at 120 and leaves x = 70.
- (b) No — in order the numbers are 1, 3, 5, 7, 9, so the median is 5. — Method: the median is the middle value of the data in order of size, so the data must be sorted before any position is read off. Working: Noah's list 9, 3, 7, 1, 5 is not in order; sorted it becomes 1, 3, 5, 7, 9, and with 5 values the middle position is the third, which now holds 5 rather than 7. Noah has read the third value of the unsorted list. Answer: no — in order the numbers are 1, 3, 5, 7, 9, so the median is 5. The distractors: the reply giving 3 as the median sorts the data correctly but then reads the value in the second place instead of the third; the reply that 7 is the third number he wrote accepts a position in the unsorted list, which is exactly the mistake the question is about; the reply using the mean claims a value of 7 for it, but the mean is 25 ÷ 5 = 5, so that reasoning is false as well.
- (d) 300 pupils — Method: an estimate for a whole population is made by finding the proportion in the sample and applying that same proportion to the population. Working: in the sample 30 of the 50 pupils prefer sport, a proportion of 30 ÷ 50 = 0.6, and applying that proportion to the school gives 0.6 × 500 = 300 pupils. Answer: 300 pupils, and it is only an estimate, because a different random sample of 50 would give a slightly different figure. The distractors: 200 pupils comes from scaling up the 20 pupils in the sample who did not prefer sport, 20 × 10, which answers the opposite question; 150 pupils comes from reading 30 out of 50 as 30% and taking 30% of 500; 60 pupils comes from working out the proportion correctly as 60% and then writing the 60 down as a number of pupils instead of applying it to the 500.
- (d) Correlation is a link; causation is one causing the other — Method: the two words describe different claims — one is about a pattern in the data, the other is about what produced that pattern. Working: correlation says only that two quantities tend to change together, which is something a scatter graph can display; causation says that a change in one quantity actually brings about the change in the other, which needs evidence a scatter graph cannot supply, because a third quantity may be driving both. Answer: correlation is a link between the quantities, while causation is one quantity causing the change in another. The distractors: the statement giving causation as the link and correlation as the cause simply swaps the two words over; the statement that the words mean the same thing is the classic error of reading a correlation as proof of cause; the statement that a scatter graph shows causation but not correlation reverses what a scatter graph can do, since the pattern it displays is exactly the correlation.
Build your own mix at the worksheet builder.