Printable · GCSE Foundation · ages 14-16
Statistics worksheet — GCSE Foundation
Fifteen questions across the statistics statements at Foundation tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Non-calculator
Answer key: Statistics worksheet — GCSE Foundation
MathsUKwww.geekhero.co.uk
- (d) 29 — Method: with an even number of values there is no single middle value, so the median is the mean of the two values either side of the middle. Working: the six numbers are already in order and 6 ÷ 2 = 3, so the middle pair are the third and fourth values, 22 and 36; their mean is (22 + 36) ÷ 2 = 58 ÷ 2 = 29. Answer: 29, which lies between the two middle values as a median of an even data set must. The distractors: 22 comes from taking the lower of the two middle values and stopping there instead of averaging the pair; 36 comes from taking the larger value of that pair because it sits just past the halfway point of the list; 34 comes from working out the range, 47 − 13, instead of a measure of centre.
- (d) About 510 of the 600 bulbs are likely to have flowered — Method: the proportion found in a random sample is used as an estimate of the proportion in the whole population, and the conclusion is stated as an estimate, never as a fact about every member. Working: 17 of the 20 bulbs dug up had flowered, so the sample proportion is 17 ÷ 20 = 0.85, and applying that proportion to the whole planting gives 0.85 × 600 = 510 bulbs. A different random sample of 20 would very probably give a slightly different figure, so 510 is an estimate. Answer: about 510 of the 600 bulbs are likely to have flowered. The distractors: saying exactly 510 have flowered takes an estimate from a sample of 20 as a count of all 600, which no sample can deliver; saying exactly 17 of the 600 have flowered reports the sample count as though it were the population count, leaving the other 580 bulbs out of the answer altogether; saying about 20 have flowered uses the size of the sample as the estimate, when 20 is the number of bulbs she dug up rather than a number that flowered.
- (c) 7 — Method: the median is the middle value when the data are written in order of size, and with an odd number of values there is exactly one middle value. Working: the numbers are already in order, 2, 4, 7, 12, 26, and there are 5 of them, so the middle position is the third and the value sitting there is 7. Answer: 7, with two values below it and two above it. The distractors: 10.2 comes from working out the mean, 51 ÷ 5, instead of the median; 14 comes from taking the value halfway between the smallest and the largest, (2 + 26) ÷ 2; 24 comes from working out the range, 26 − 2, which measures spread rather than centre.
- (a) Negative correlation — As the age of the car increases, the points fall towards a lower value, so the value decreases as the age increases. This falling pattern is a negative correlation. A positive correlation would show the points rising together instead. No correlation would apply only if the points showed no pattern at all, and correlation is not the same as causation — strong causation is not a type of correlation.
- (d) Correlation is a link; causation is one causing the other — Method: the two words describe different claims — one is about a pattern in the data, the other is about what produced that pattern. Working: correlation says only that two quantities tend to change together, which is something a scatter graph can display; causation says that a change in one quantity actually brings about the change in the other, which needs evidence a scatter graph cannot supply, because a third quantity may be driving both. Answer: correlation is a link between the quantities, while causation is one quantity causing the change in another. The distractors: the statement giving causation as the link and correlation as the cause simply swaps the two words over; the statement that the words mean the same thing is the classic error of reading a correlation as proof of cause; the statement that a scatter graph shows causation but not correlation reverses what a scatter graph can do, since the pattern it displays is exactly the correlation.
- (d) 75 — Method: to find a total from a bar chart, add together the height of every bar. Working: 12 + 18 + 9 + 15 + 21 = 75 books. Leaving out Wednesday's bar by mistake, 12 + 18 + 15 + 21 = 66, misses one day out of the total. Giving 21 states only Friday's bar, the tallest one, not the total of all five days. Dividing the total by the number of days, 75 ÷ 5 = 15, finds the mean number of books per day, not the total borrowed. Add up every single bar — do not stop at the biggest one, and do not divide once you have added them all.
- (a) No, the size of the fire affects both of the quantities — Method: correlation says that two quantities change together; a claim that one of them produces the other is a further claim, and it needs evidence that a scatter graph on its own cannot give. Working: the graph does show strong positive correlation, so more engines did go with greater damage. But neither quantity was set by the researchers: both were decided by how large the fire was. A large blaze brings many appliances and also destroys a great deal, while a small one brings few and destroys little, so a third quantity is driving both of the recorded ones. Answer: no, because the size of the fire affects both of the quantities. The distractors: saying the correlation is negative contradicts the graph, which shows the two quantities rising together, and reaching the right verdict from a false reading of the data is not the reason the mark is for; saying that strong positive correlation shows one quantity causes the other is the assumption the question exists to test, and no strength of correlation can establish cause; saying the points lie close to the line of best fit describes how strong the correlation is, and strength and cause are different matters entirely.
- (c) Strong negative correlation — Method: correlation is described by two things — the direction the points take as the graph is read from left to right, and how closely the points lie to a single straight line. Working: reading the pairs in order of age, the ages rise 14, 18, 23, 27, 31, 36, 42, 49 while the scores fall 92, 88, 85, 80, 78, 74, 70, 65; the score falls at every single step, with no reversal anywhere, so the points fall from left to right and lie close to a straight line. Answer: strong negative correlation — negative for the falling direction, strong because every point follows the pattern. The distractors: strong positive correlation comes from noticing a clear pattern and calling any clear pattern positive, without checking the direction; weak negative correlation comes from reading the direction correctly but judging points that do not lie exactly on a straight line to be only loosely related, when these eight fall without a single exception; no correlation comes from reading a falling trend as though it showed no relationship at all, when a falling trend is itself a relationship.
- (a) 23 minutes — Method: the equation of a line of best fit converts a value of x into a predicted value of y, so substitute the known number of pages for x and evaluate. Working: x is the number of pages, so put x = 10 into y = 2x + 3. Multiplication is carried out before addition, so 2 × 10 + 3 = 23. Answer: 23 minutes, and it is a prediction of the trend rather than a promise about any one chapter. The distractors: 20 minutes comes from working out 2 × 10 and stopping there, leaving out the 3 that the line adds; 26 minutes comes from reading the equation as y = 2(x + 3), adding first and then doubling, so 2 × 13 = 26; 13 minutes comes from adding 10 and 3 and never using the gradient at all, which treats the 2 as though it were not there.
- (c) 33 — Method: multiply the mean by the number of tests to get the total marks, then subtract the marks that are already known. Working: four tests with a mean of 29 give a total of 29 × 4 = 116 marks; the first three marks total 31 + 26 + 26 = 83; so the fourth mark is 116 − 83 = 33. Answer: 33, and checking, (31 + 26 + 26 + 33) ÷ 4 = 116 ÷ 4 = 29. The distractors: 116 comes from stopping at the total for all four tests; 29 comes from assuming the missing mark must be the mean itself; 4 comes from multiplying the mean by 3, the number of marks given, leaving 87 − 83 = 4.
- (a) No correlation — Shoe size has no real relationship with spelling ability, and the points here are scattered with no rising or falling trend, so this is no correlation. A positive correlation would show the points rising together, and a negative correlation would show them falling as one increases; neither pattern is present here. Strong correlation is not correct either, since strength only applies once a positive or negative trend exists, and there isn't one.
- (c) 72 — Method: when a pie chart is divided into equal sectors, each sector stands for the same share of the people asked, so the fraction of the sectors that are shaded is also the fraction of the people. Working: 6 sectors out of 10 are purple, which is the fraction 6/10 of the whole pie chart; one tenth of the 120 people is 120 ÷ 10 = 12 people, so six tenths is 6 × 12 = 72 people. Answer: 72 people, a count of people rather than a number of sectors. The distractors: 48 comes from working with the 4 sectors that are not purple, 4 × 12, and so answering for the wrong part of the chart; 60 comes from turning the fraction 6/10 into 60% and then writing the 60 down as though it were a number of people; 6 comes from writing down the number of purple sectors instead of the number of people those sectors stand for.
- (d) 6 — Method: the mode, or modal value, is the value that occurs most often in the data set, and it is a value from the data rather than a count. Working: size 4 occurs twice, size 5 occurs once, size 6 occurs three times and size 9 occurs once, so the highest frequency is three and the size it belongs to is 6. Answer: 6. The distractors: 3 comes from writing down the frequency of the most common size instead of the size itself; 9 comes from picking the largest size in the list, which confuses the mode with the maximum; 4 comes from stopping at the first size that repeats rather than checking which size repeats most often.
- (a) The relationship between two variables — Method: what a diagram shows is decided by what has to be known before a single mark can be plotted on it. Working: every point on a scatter graph is plotted from a pair of measurements taken from the same person or object, one read on the horizontal axis and one on the vertical axis; having two measurements for each point is what makes it possible to look for a pattern between them, and the pattern between two variables is what the graph displays. Answer: a scatter graph shows the relationship between two variables. The distractors: the frequency of each single value is what a bar chart or a vertical line chart shows, and it needs only one list of values; how a total is shared between categories is what a pie chart shows; how one quantity changes over time is what a time series line graph shows, in which one of the two axes is always time.
- (d) How much the temperatures varied over the seven days — Method: the range of a set of values is the largest value take away the smallest, so it is built from two values only and it measures the gap they leave between them. Working: the largest of the seven readings is 7 °C and the smallest is 2 °C, so the range is 7 − 2 = 5 °C. That figure says the week's readings covered a band 5 °C wide; it names no particular day and no particular reading. Answer: the range describes how much the temperatures varied over the seven days. The distractors: the temperature that occurred most often is the mode, which here is 3 °C, and a mode counts repeats instead of measuring a gap; the temperature typical of the week is an average, and the range is not an average, since it throws away every value lying between the two extremes; the number of different temperatures recorded is 6, a count of how many distinct values appear, while the range is a difference between two of them.
Build your own mix at the worksheet builder.