Printable · GCSE Foundation · ages 14-16
Statistics worksheet — GCSE Foundation
Fifteen questions across the statistics statements at Foundation tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Non-calculator
Statistics worksheet — GCSE Foundation
MathsUKwww.geekhero.co.uk
- 1.Two classes sat the same maths test, both marked out of 100. Class A had a mean mark of 70 and a range of 30 marks. Class B had a mean mark of 70 and a range of 10 marks. Compare the marks of the two classes.
- 2.On a scatter graph the horizontal axis shows height in centimetres and the vertical axis shows mass in kilograms. One point is plotted at (170, 65). Write down the height and the mass of that person.
- 3.Isla wants to find out the favourite television programme of the pupils at her school. She asks only her own close group of friends. Write down what is wrong with her sample.
- 4.A school has 1200 pupils. A teacher wants to take a random sample of 60 of them. Write down which of these methods gives a random sample.
- 5.A scatter graph shows the height, x cm, and the mass, y kg, of 20 pupils in Year 10. The heights on the graph run from 150 cm to 180 cm, and the line of best fit is y = 0.9x − 85. Nadia puts x = 90 into this equation to estimate the mass of a two-year-old child who is 90 cm tall. Is her estimate reliable? Give a reason for your answer.y = 0.9x − 85
- 6.A café in York counts the number of customers in each of the nine hours it is open on one day: 4, 4, 4, 11, 13, 15, 18, 22 and 25. The owner says that a typical hour has about 4 customers, because 4 is the mode. Is the owner right? Give a reason for your answer.
- 7.Grace asked 12 children how many brothers and sisters they have. Her results, in order, were 0, 0, 1, 1, 1, 1, 2, 2, 3, 3, 4, 6. Work out the median number of brothers and sisters.
- 8.A random sample of 50 pupils at a school were asked whether they prefer sport to music. 30 of the 50 pupils said they prefer sport. The school has 500 pupils altogether. Work out an estimate for the number of the 500 pupils who prefer sport.
- 9.A scatter graph shows the number of pages, x, in a chapter and the time, y minutes, a pupil took to read it, for 15 chapters. The line of best fit has equation y = 2x + 3. Use the line of best fit to estimate the time taken to read a chapter of 10 pages.y = 2x + 3
- 10.A survey of 200 car owners in Bristol records the colour of each car: silver 74, black 52, blue 40 and red 34. Write down which average should be used to describe a typical car in this survey, and give a reason for your answer.
- 11.Write down the statement that correctly describes the difference between correlation and causation.
- 12.A pupil writes this question for a school survey: “Do you agree that learning matters and that we should be set more homework?” Write down what is wrong with the survey question.
- 13.The mean of 5 numbers is 8. Work out the total of the 5 numbers.
- 14.Five friends have heights, in cm, of 150, 152, 155, 158 and 160. A sixth friend, with a height of 170 cm, joins the group. Write down what happens to the mean and the range of the heights once this sixth friend is included.
- 15.Four books have 120, 200, 160 and 80 pages. Work out the mean number of pages.
Answer key
- (a) The means are equal, and Class B's marks are the more consistent because its range is smaller. — Method: comparing two distributions needs two things — a measure of average and a measure of spread — and each must be put into the context of the question. Working: both classes have a mean mark of 70, so on average the two classes scored the same; the range measures spread, and Class A's range of 30 marks is three times Class B's range of 10 marks, so Class B's marks sit closely around the mean while Class A's are far more spread out. Answer: the means are equal, and Class B's marks are the more consistent because its range is smaller. The distractors: the reply crediting Class A with more consistency reverses the meaning of the range, treating a larger range as tighter data when a larger range means more spread; the reply that Class A's mean mark is higher compares the wrong pair of figures, reading the range of 30 as an average; the reply that Class B's mean mark is higher reads the spread correctly but its claim about the means is false, since both means are 70.
- (c) Height 170 cm, mass 65 kg — Method: a point on a scatter graph is written as a pair of coordinates in which the horizontal value is written first and the vertical value second, so each value is matched to the quantity named on its own axis. Working: in (170, 65) the value 170 is the horizontal coordinate and the horizontal axis shows height in centimetres, so the height is 170 cm; the value 65 is the vertical coordinate and the vertical axis shows mass in kilograms, so the mass is 65 kg. Answer: height 170 cm, mass 65 kg, each with the unit named on its own axis. The distractors: height 65 cm and mass 170 kg come from reading the pair the wrong way round, which would describe an impossible person; height 170 cm and mass 170 kg come from reading the horizontal coordinate for both quantities and never using the second number; height 235 cm and mass 105 kg come from combining the two coordinates, 170 + 65 and 170 − 65, instead of reading them separately.
- (c) It is not representative, as she picked her own friends — Method: judge a sample by asking whether the pupils in it were chosen in a way that gives the whole school a fair chance of being heard. Working: Isla's friends are a group she formed herself, and friends tend to share tastes, so their favourite programme is likely to match hers rather than the school's, and pupils in other year groups and other friendship groups had no chance at all of being asked; the fault lies in how the pupils were selected, not in how many of them there were. Answer: it is not representative, as she picked her own friends. The distractors: the reply blaming the size claims the pupils were picked at random, which is false here, and it is the common mistake of thinking a biased sample can be cured by making it bigger; the reply calling the sample too large is false in the other direction, as a survey is never spoilt by collecting more replies; the reply that the sample is fine treats attending the school as enough, which would make any group of pupils in the building a fair sample.
- (a) Drawing 60 names at random from a list of all 1200 pupils — Method: a sample is random when every member of the population has the same chance of being chosen and nobody, including the pupils themselves, can influence who ends up in it; test each method against that. Working: drawing names from a list of all 1200 pupils gives each pupil the same chance, 60 out of 1200, whatever their year group, class or opinion, so the method is random. Answer: drawing 60 names at random from a list of all 1200 pupils. The distractors: asking the pupils who volunteer is self-selection, and the pupils with the strongest views volunteer first, so they decide the sample; asking the pupils nearest the door is convenience sampling, which reaches only those who happen to be in one place at one time; asking two Year 10 classes samples a cluster, so every pupil in the other year groups has no chance of being chosen at all.
- (c) No, 90 cm is far outside the heights on the graph — Method: a line of best fit describes the trend only across the stretch of data it was drawn through; predicting beyond that stretch is extrapolation, and nothing in the data supports it. Working: the heights used to draw this line run from 150 cm to 180 cm, all of them Year 10 pupils, while 90 cm is 60 cm below the shortest of them and belongs to a two-year-old child, whose build follows no trend the graph has measured. Substituting anyway gives 0.9 × 90 − 85 = −4, a mass of −4 kg, which cannot exist. Answer: no, because 90 cm is far outside the heights on the graph. The distractors: saying a line of best fit cannot be used to predict at all throws away its main purpose, since a prediction made between the plotted values is perfectly sound; saying the line passes through all 20 points misdescribes a line of best fit, which is drawn to follow the trend of the points and will normally pass through few of them; saying the equation works for any value put into it treats an equation fitted to Year 10 heights as a law of nature, and the mass of −4 kg shows what that assumption produces.
- (c) No, the mode here is the lowest value of the nine — Method: an average is meant to stand for the data as a whole, so test any proposed average by asking how many values it sits near. Working: the value 4 appears three times and every other count appears once, so 4 is indeed the mode. But those three hours are the quiet ones at the start of the day, and the other six counts run from 11 up to 25; putting the nine counts in order, the middle one is the fifth, which is 13. So the mode sits at the very bottom of the data, with six of the nine hours far above it. Answer: no, because the mode here is the lowest value of the nine, so it describes the quiet opening hours rather than a typical hour. The distractors: saying the mode can only be used when no value repeats reverses the definition, since a mode exists only because a value does repeat; saying the mode is the value that occurs most often is a correct definition, but being the commonest value does not make a value typical when it lies at one end of the data; saying the mode is the best average for any list of numbers ignores the fact that mean, median and mode each describe a population well in different circumstances.
- (a) 1.5 — Method: with an even number of values the median is the mean of the two middle values, which for 12 values are the 6th and the 7th once the data are in order. Working: the results are already in order, and 12 ÷ 2 = 6, so the middle pair are the 6th value, 1, and the 7th value, 2; the median is (1 + 2) ÷ 2 = 1.5. Answer: 1.5 brothers and sisters. The distractors: 1 comes from reading the 6th value and stopping there instead of averaging the middle pair; 2 comes from working out the mean, 24 ÷ 12, instead of the median; 6 comes from working out the range, 6 − 0, which measures spread rather than centre.
- (d) 300 pupils — Method: an estimate for a whole population is made by finding the proportion in the sample and applying that same proportion to the population. Working: in the sample 30 of the 50 pupils prefer sport, a proportion of 30 ÷ 50 = 0.6, and applying that proportion to the school gives 0.6 × 500 = 300 pupils. Answer: 300 pupils, and it is only an estimate, because a different random sample of 50 would give a slightly different figure. The distractors: 200 pupils comes from scaling up the 20 pupils in the sample who did not prefer sport, 20 × 10, which answers the opposite question; 150 pupils comes from reading 30 out of 50 as 30% and taking 30% of 500; 60 pupils comes from working out the proportion correctly as 60% and then writing the 60 down as a number of pupils instead of applying it to the 500.
- (a) 23 minutes — Method: the equation of a line of best fit converts a value of x into a predicted value of y, so substitute the known number of pages for x and evaluate. Working: x is the number of pages, so put x = 10 into y = 2x + 3. Multiplication is carried out before addition, so 2 × 10 + 3 = 23. Answer: 23 minutes, and it is a prediction of the trend rather than a promise about any one chapter. The distractors: 20 minutes comes from working out 2 × 10 and stopping there, leaving out the 3 that the line adds; 26 minutes comes from reading the equation as y = 2(x + 3), adding first and then doubling, so 2 × 13 = 26; 13 minutes comes from adding 10 and 3 and never using the gradient at all, which treats the 2 as though it were not there.
- (d) The mode, because colours cannot be added or ordered — Method: an average can only be used on data that supports the operation it needs. A mean needs the values to be added and divided, a median needs them to be placed in order, and a range needs one value to be taken away from another; a mode needs only counting, so it is the average available when the data are categories rather than numbers. Working: the data collected here are colours, silver, black, blue and red. The numbers 74, 52, 40 and 34 count the cars of each colour, they do not measure them, and 74 + 52 + 40 + 34 = 200 simply returns the size of the survey. No colour can be added to another, and there is no order that puts blue before red, so of the four averages only the one found by counting survives. Answer: the mode, because colours cannot be added or ordered, and the mode is silver. The distractors: the mean is said to use all 200 colours, and a mean of the four frequencies, 200 ÷ 4 = 50, is a number of cars rather than a colour, so it describes nothing about a typical car; the median is said to put the colours in order, but ordering the frequencies 34, 40, 52, 74 orders the counts, not the colours, and gives 46, again a number of cars; the range is not an average at all, and 74 − 34 = 40 measures the gap between the commonest and rarest counts, which is a measure of spread.
- (d) Correlation is a link; causation is one causing the other — Method: the two words describe different claims — one is about a pattern in the data, the other is about what produced that pattern. Working: correlation says only that two quantities tend to change together, which is something a scatter graph can display; causation says that a change in one quantity actually brings about the change in the other, which needs evidence a scatter graph cannot supply, because a third quantity may be driving both. Answer: correlation is a link between the quantities, while causation is one quantity causing the change in another. The distractors: the statement giving causation as the link and correlation as the cause simply swaps the two words over; the statement that the words mean the same thing is the classic error of reading a correlation as proof of cause; the statement that a scatter graph shows causation but not correlation reverses what a scatter graph can do, since the pattern it displays is exactly the correlation.
- (d) It asks two things at once and invites agreement — Method: a survey question is faulty when a reply to it cannot be read as evidence about one single thing, so check how many claims it contains and whether its wording pushes the reader one way. Working: the question joins two separate claims, that learning matters and that more homework should be set, so a reply of yes could mean either of them or both and cannot be counted as evidence about homework; the opening words “Do you agree” also invite agreement instead of leaving the reader free to say no. Answer: it asks two things at once and invites agreement. The distractors: the reply calling it too short mistakes length for clarity, when the fault is that too much has been packed in rather than too little; the reply about long words is false, since every word in the question is an everyday one and the fault lies in what is being asked rather than in the vocabulary used to ask it; the reply that the question is fine takes a yes or no answer as proof that the question works, which is exactly what a double question defeats.
- (a) 40 — Method: the mean is the total divided by how many values there are, so rearranging gives total = mean × number of values. Working: the mean is 8 and there are 5 numbers, so the total is 8 × 5 = 40. Answer: 40, and checking, 40 ÷ 5 = 8, which is the mean given. The distractors: 13 comes from adding the mean and the count, 8 + 5, instead of multiplying them; 1.6 comes from dividing the mean by the count, 8 ÷ 5, which reverses the relationship; 8 comes from quoting the mean itself as the total, which is only true when there is a single number.
- (d) Both the mean and the range increase. — The original mean is 150 + 152 + 155 + 158 + 160 = 775, and 775 ÷ 5 = 155 cm; the original range is 160 − 150 = 10 cm. Including the new height of 170 cm gives a new total of 775 + 170 = 945, and 945 ÷ 6 = 157.5 cm, which is higher than 155 cm, and a new range of 170 − 150 = 20 cm, which is higher than 10 cm, so both the mean and the range increase. Saying the range stays the same ignores that 170 cm is a new, higher maximum than the old 160 cm. Saying the mean stays the same ignores that 170 cm is above the original mean of 155 cm, which pulls the average up. Saying both decrease is the opposite of what happens here.
- (c) 140 — Method: the mean shares the total equally between the items, so add the values and then divide by how many there are. Working: the total is 120 + 200 + 160 + 80 = 560 pages and there are 4 books, so the mean is 560 ÷ 4 = 140 pages. Answer: 140. The distractors: 560 comes from stopping at the total number of pages and never dividing by 4; 280 comes from dividing the total by 2 instead of by the 4 books; 120 comes from working out the range, 200 − 80, which measures spread instead of centre.
Build your own mix at the worksheet builder.