Printable · GCSE Foundation · ages 14-16
Statistics worksheet — GCSE Foundation
Fifteen questions across the statistics statements at Foundation tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Non-calculator
Statistics worksheet — GCSE Foundation
MathsUKwww.geekhero.co.uk
- 1.A charity shop in Bath holds 2,000 books. Volunteer A checks a random sample of 50 books and finds 35 paperbacks. Volunteer B checks a different random sample of 50 books and finds 31 paperbacks. Work out the estimate each sample gives for the whole stock, and write down what the shop should do next.
- 2.The number of pets owned by each of 19 pupils in a class is recorded: 0 pets — 7 pupils, 1 pet — 3 pupils, 2 pets — 4 pupils, 3 pets — 5 pupils. Work out the median number of pets.
- 3.A council in Leeds wants to know what local people think about letting shops stay open later in the evening. It rings landline telephone numbers between 10 am and 2 pm on a Tuesday. Write down which group is most likely to be under-represented in the sample, and give a reason for your answer.
- 4.A study found that people who drink more coffee tend to concentrate better at work. A coffee company says that this shows that drinking coffee improves concentration. Give the reason why this conclusion cannot be drawn.
- 5.A Year 10 class has 20 boys with a mean height of 150 cm and 10 girls with a mean height of 168 cm. Work out the mean height of all 30 pupils in the class.
- 6.A town council wants to find out about the eating habits of the people who live in the town. It asks only the members of a local sports club. Give a reason why this sample is biased.
- 7.The rainfall, in millimetres, was recorded in Cambridge on six days: 12, 5, 9, 15, 3 and 11. Work out the range of the rainfall.
- 8.A vertical line chart shows the number of goals scored by a football team in each of its 20 matches: 0 goals in 4 matches, 1 goal in 7 matches, 2 goals in 6 matches and 3 goals in 3 matches. Write down the modal number of goals.
- 9.A sector of a pie chart stands for 25% of the data. Work out the angle at the centre of that sector.
- 10.In a spelling test the 20 pupils in Group A had a mean mark of 80, and the 30 pupils in Group B had a mean mark of 70. Work out the mean mark of all 50 pupils.
- 11.A factory makes 4,000 light bulbs a day. In a random sample of 80 of one day's bulbs, 3 were faulty. Work out an estimate for the number of faulty bulbs the factory makes in a day.
- 12.Ben and Chloe each sat five maths tests. Ben's marks were 62, 64, 65, 66 and 68. Chloe's marks were 40, 52, 65, 78 and 90. Both pupils have a mean mark of 65. Their teacher says the mean on its own does not describe the two sets of marks well. Give a reason why the teacher is right.
- 13.A table shows the time each of four pupils took to run 100 m: Oliver 12 seconds, Grace 15 seconds, Ethan 10 seconds, Freya 14 seconds. Write down the name of the fastest runner and the time taken.
- 14.Two classes sat the same test. The 30 pupils in Class A had a mean mark of 72. The 20 pupils in Class B had a mean mark of 82. Work out the mean mark of all 50 pupils.
- 15.The mean of the numbers x, 20 and 30 is equal to the mean of the numbers 15 and 25. Work out the value of x.
Answer key
- (a) 1,400 and 1,240, so combine the samples for one estimate — Method: scale each sample up to the whole stock, then use the fact that a larger sample gives a more reliable estimate than a smaller one. Working: the first sample gives 35 ÷ 50 = 0.7 and 0.7 × 2,000 = 1,400 paperbacks; the second gives 31 ÷ 50 = 0.62 and 0.62 × 2,000 = 1,240 paperbacks. Two random samples of the same size are expected to differ a little, so neither estimate is wrong. Putting the two together gives 35 + 31 = 66 paperbacks in 100 books, and 66 ÷ 100 = 0.66 with 0.66 × 2,000 = 1,320, an estimate resting on twice as many books as either volunteer checked. Answer: 1,400 and 1,240, so combine the samples for one estimate. The distractors: keeping 1,400 because it is larger picks an estimate by its size, when both samples held 50 books and neither has a stronger claim; saying a volunteer must have miscounted assumes two random samples ought to agree exactly, which is precisely what random sampling does not promise; 1,750 and 1,550 come from 35 × 50 = 1,750 and 31 × 50 = 1,550, multiplying each count by the size of the sample instead of scaling by 2,000 ÷ 50.
- (b) 1 — Method: for data in a frequency table, find the position of the median using (n + 1) ÷ 2, then read off the value at that position from the cumulative frequencies. Working: there are 19 pupils, so the median is the 10th value. The cumulative frequencies are 7 (up to 0 pets), 10 (up to 1 pet), 14 (up to 2 pets) and 19 (up to 3 pets). The 10th value falls at the end of the '1 pet' group, so the median is 1 pet. Giving 0 pets is the mode — the category with the highest frequency, 7 — not the median. Giving 3, the highest number of pets minus the lowest, finds the range, a different statistic entirely. Giving 19 states the total number of pupils, not a number of pets at all. Find the middle POSITION first, then read off the value it belongs to — do not confuse it with the mode, the range or the total.
- (d) Full-time workers, as most are at work at that time — Method: a sample is biased when the method of contact makes part of the population much less likely to be reached, so test each group against where its members actually are between 10 am and 2 pm on a weekday, and test each stated reason against the facts. Working: those hours are the middle of the working day, so people in full-time employment are at work and not beside a landline telephone, while people who are retired and people who are unemployed are far more likely to be at home and are reached at the usual rate; the method therefore collects far fewer replies from full-time workers than their share of the adult population the council is consulting. Answer: full-time workers, as most are at work at that time. The distractors: the reply naming retired people rests on the false claim that most retired people are at work in the daytime, when in fact a daytime call reaches them more easily than anyone; the reply naming unemployed people rests on the false claim that they are out during the day, when they too are among the easiest people to reach by a daytime call; the reply naming children rests on the false claim that children are at home at 11 am on a Tuesday, when they are at school and so are not reached by the call at all, and school-age children are in any case not the adults whose views the council is collecting.
- (c) The data show a link only; a third factor may affect both — Method: a study of this kind measures two quantities and reports how they change together; deciding that one of them produces the other is a further claim, and it needs evidence that the measurements alone cannot give. Working: the study shows that more coffee goes with better concentration, which is a positive correlation; but a third factor that was never measured, such as how motivated someone is, could raise both the coffee drinking and the concentration, and the concentration could equally be what leads to the extra coffee. Answer: the data show a link only, because a third factor may be affecting both quantities, so no claim about cause can be made. The distractors: calling the conclusion safe because the correlation is positive treats the direction of a correlation as proof of cause, which no direction can give; calling it wrong because the correlation is negative misreads the direction of the relationship, since the study reports both quantities rising together; saying the two quantities are not linked denies the correlation the study actually found, when what fails is only the claim about cause.
- (c) 156 cm — Method: to combine two groups' means, multiply each group's mean by its own number of pupils, add the two totals together, then divide by the total number of pupils in both groups. Working: 20 × 150 = 3,000 cm for the boys and 10 × 168 = 1,680 cm for the girls, giving a combined total of 3,000 + 1,680 = 4,680 cm. Dividing by all 30 pupils gives 4,680 ÷ 30 = 156 cm. Giving 159 cm averages the two means, (150 + 168) ÷ 2, treating the two groups as if they had the same number of pupils, when there are twice as many boys as girls. Giving 4,680 cm finds the correct combined total height but stops there, forgetting the final division by the 30 pupils. Giving 234 cm divides the combined total by 20, the number of boys only, forgetting that the total also includes the 10 girls. Always weight each mean by its own group size, and always divide by the TOTAL number of pupils in both groups combined.
- (c) Club members probably eat differently from most people — Method: a sample is biased when the group it is drawn from differs from the population in the very thing the survey is measuring, so compare the subgroup with the population on that quantity. Working: the survey measures eating habits, and people who join a sports club take more exercise than average and are known to eat differently from the town as a whole, so their replies pull the results away from the true picture for the town however many of them are asked. Answer: club members probably eat differently from most people. The distractors: the reply about the number of members treats bias as a question of size, but a large biased sample is still biased; the reply that the members were picked at random is false, since the council picked a club rather than picking residents, and it confuses bias with non-response; the reply that everyone asked lives in the town notes something true of the members but draws the false conclusion that the sample therefore covers the town, when a sample must reflect a population and not merely be taken from inside it.
- (a) 12 mm — Method: the range is the highest value minus the lowest value. Working: the highest rainfall is 15 mm and the lowest is 3 mm, so the range is 15 − 3 = 12 mm. Giving 15 mm alone states the highest value, not the range. Giving 3 mm alone states the lowest value, not the range. Sorting the six values, 3, 5, 9, 11, 12 and 15, and averaging the middle two, (9 + 11) ÷ 2 = 10 mm, finds the median, a completely different statistic. The range always needs BOTH the highest and the lowest value — never just one of them.
- (b) 1 — The four frequencies are 4, 7, 6 and 3 matches, and the largest of these is 7, which corresponds to 1 goal, so the modal number of goals is 1. Choosing 7 confuses the frequency, how many matches, with the number of goals itself. Choosing 2 uses the second-largest frequency, 6 matches, instead of the largest. Choosing 3 uses the smallest frequency, which belongs to the fewest matches, not the most.
- (c) 90° — Method: the sectors of a pie chart share the 360° at the centre of the chart in the same proportion as the data, so a sector's angle is its share of the data multiplied by 360°. Working: a share of 25% is the fraction 25/100, which is one quarter of the whole chart, and one quarter of 360° is 360 ÷ 4 = 90°. Answer: 90°, an angle in degrees rather than a percentage. The distractors: 25° comes from sharing out 100 instead of 360, so the percentage is written straight down as a number of degrees; 14.4° comes from dividing 360 by 25 instead of multiplying 360 by the fraction 25/100, which is the division done the wrong way round; 45° comes from taking a quarter of 180° instead of a quarter of 360°, treating the pie chart as a semicircle.
- (b) 74 marks — Method: the two groups are different sizes, so their means cannot simply be averaged — rebuild each group's total mark, add the totals and divide by all 50 pupils. Working: Group A scored 20 × 80 = 1600 marks and Group B scored 30 × 70 = 2100 marks, giving 1600 + 2100 = 3700 marks altogether, so the overall mean is 3700 ÷ 50 = 74 marks. Answer: 74 marks, which sits nearer to 70 than to 80 because the larger group scored 70. The distractors: 75 marks comes from averaging the two group means, (80 + 70) ÷ 2, as though the groups were the same size; 76 marks comes from attaching each mean to the other group's size, (20 × 70 + 30 × 80) ÷ 50; 150 marks comes from adding the two means together and never dividing at all.
- (a) 150 bulbs — Method: assume the proportion faulty in a random sample is the proportion faulty in the whole day's output, and scale the sample up to the population. Working: the sample of 80 has to be scaled up to 4,000 bulbs, and 4,000 ÷ 80 = 50, so the day's output is 50 sample-sized batches. Each batch is expected to contain the same 3 faulty bulbs, so the estimate is 3 × 50 = 150. Answer: 150 bulbs, and it is an estimate, because another sample of 80 would probably contain a different number of faulty bulbs. The distractors: 50 bulbs is the scale factor 4,000 ÷ 80 written down as though it were the answer, so it reports how many batches there are rather than how many faulty bulbs; 120 bulbs comes from reading 3 out of 80 as 3%, then taking 0.03 × 4,000 = 120, but 3 out of 80 is 3.75%; 240 bulbs comes from 3 × 80 = 240, multiplying the faulty bulbs by the size of the sample instead of by the scale factor, which uses the 80 twice and the 4,000 not at all.
- (a) Chloe's marks are far more spread out than Ben's — Method: a mean reports where a set of values sits, and two sets can sit in the same place while behaving quite differently, so a measure of spread has to be worked out as well. Working: Ben's marks add to 62 + 64 + 65 + 66 + 68 = 325 and 325 ÷ 5 = 65; Chloe's add to 40 + 52 + 65 + 78 + 90 = 325 and 325 ÷ 5 = 65, so the two means agree, as the question says. The ranges do not: Ben's is 68 − 62 = 6 marks, while Chloe's is 90 − 40 = 50 marks. Ben's five marks all sit within 3 marks of 65; Chloe's lowest is 25 marks below it and her highest 25 marks above it. Answer: Chloe's marks are far more spread out than Ben's, which is exactly what the mean cannot show. The distractors: saying Ben's marks are more spread out comes from subtracting in the order the values are written, 62 − 68 = −6 against 40 − 90 = −50, and then reading −6 as the larger spread; saying Chloe scored far more marks in total assumes a wider set of marks must add to more, when both totals are 325; saying the two sets vary by the same amount assumes that equal means force equal spread, when the two ranges are 6 and 50.
- (b) Ethan, 10 seconds — Method: over the same distance the fastest runner is the one who takes the least time, so the smallest time in the table is found first and the name is then read from the same row. Working: the four times are 12 seconds, 15 seconds, 10 seconds and 14 seconds; in order of size these are 10, 12, 14 and 15, so the least time is 10 seconds, and the row holding 10 seconds is the row for Ethan. Answer: Ethan, 10 seconds — the time is in seconds, and a smaller time means a faster runner. The distractors: Grace with 15 seconds comes from taking the largest number in the table to mean the fastest runner, which reverses the relationship between time and speed over a fixed distance; Oliver with 12 seconds comes from writing down the first row of the table without comparing the four times; Ethan with 15 seconds comes from identifying the right runner but then reading the time from a different row of the table.
- (c) 76 marks — Method: a mean of means only works when the groups are the same size, so rebuild each class's total mark, add the totals and divide by the number of pupils altogether. Working: Class A scored 30 × 72 = 2160 marks and Class B scored 20 × 82 = 1640 marks, giving 2160 + 1640 = 3800 marks between 50 pupils, so the overall mean is 3800 ÷ 50 = 76 marks. Answer: 76 marks. The distractors: 77 marks comes from averaging the two class means, (72 + 82) ÷ 2, which ignores the different class sizes; 78 marks comes from attaching each mean to the other class's size, (30 × 82 + 20 × 72) ÷ 50; 3800 marks comes from stopping at the combined total and never dividing by 50.
- (c) 10 — Method: work out the mean that can be found straight away, then use total = mean × number of values on the group of three to find the missing number. Working: the mean of 15 and 25 is (15 + 25) ÷ 2 = 40 ÷ 2 = 20, so the group of three must also have a mean of 20; three numbers with a mean of 20 have a total of 20 × 3 = 60, and 20 + 30 = 50 of that total is already accounted for, so x = 60 − 50 = 10. Answer: 10, and checking, (10 + 20 + 30) ÷ 3 = 20. The distractors: 20 comes from working out the mean the two groups share and writing that down as x; −10 comes from dividing the group of three by 2 instead of by 3, which gives x + 50 = 40; 70 comes from reading the total 15 + 25 = 40 as the mean of the pair, which sets the target total at 120 and leaves x = 70.
Build your own mix at the worksheet builder.