Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- 1.A clinic recorded 300 booked appointments using a frequency tree. The first branch splits them into 210 adult appointments and the rest child appointments. Of the adult appointments, 189 were attended and the rest were missed. Of the child appointments, 81 were attended. Work out the probability that a booked appointment, chosen at random from the 300, was missed. Give your answer as a fraction in its simplest form.
- 2.A seed company tests germination using results from three greenhouses. Greenhouse 1 plants 200 seeds and 172 germinate. Greenhouse 2 plants 150 seeds and 126 germinate. Greenhouse 3 plants 250 seeds and 212 germinate. Using the combined results from all three greenhouses, work out the best estimate of the number of seeds, out of a new batch of 4000 seeds, that would be expected to germinate.
- 3.A quality inspector examines a sample of 60 items from a production line and finds that 8 are faulty. Using this sample's proportion, work out how many faulty items would be expected in a new batch of 750 items.
- 4.A grower knows that the probability that one of their seeds germinates is 0.6. The grower wants to expect 300 of the seeds to germinate. Work out how many seeds the grower should plant.
- 5.A train company runs 25 trains a day, every day. The probability that any one train is late is 0.08. Work out how many late trains the company should expect over a period of 4 weeks.
- 6.A council surveyed 140 households about recycling. 85 of the households are in Zone A. 52 of the Zone A households recycle glass. 21 of the Zone B households do not recycle glass. Work out the probability that a household, chosen at random from the 140, recycles glass. Give your answer as a fraction in its simplest form.
- 7.A two-way table records how 130 pupils travel to school. 70 of the pupils are girls and the rest are boys. 42 of the girls walk to school and the rest of the girls cycle. 33 of the boys walk to school. Work out what fraction of the girls walk to school.
- 8.A four-colour spinner (red, blue, green, yellow) is spun repeatedly, and the relative frequency of landing on green is recorded as the number of spins increases: after 20 spins it is 0.350; after 200 spins it is 0.290; after 2000 spins it is 0.251. Using the result from 2000 spins as the best estimate of the probability, work out the number of times the spinner would be expected to land on green in a further 3000 spins.
- 9.A fair coin is flipped again and again. After the first 10 flips there have been 7 heads. After 1000 flips there have been 528 heads. Which statement best describes what these results show?
- 10.At a fête, a game costs £3 to play. The probability of winning is 0.1, and each win pays out £20. 150 people play the game. Work out the fête's expected profit from the game.
- 11.Dice A is a fair six-sided dice. Dice B is biased so that P(6) = 0.3. Dice A is rolled 150 times and Dice B is rolled 150 times. Work out how many more sixes you would expect from Dice B than from Dice A.
- 12.A survey of 90 people records which of two apps, X and Y, they use. 55 people use app X, 42 people use app Y, and 20 people use neither app. Work out the probability that a randomly chosen person uses both apps, giving your answer as a fraction in its simplest form.
- 13.A biased spinner is spun 40 times and lands on red 16 times. It is then spun a further 60 times and lands on red 21 times. Work out the best estimate of the probability that the spinner lands on red, using the results of all 100 spins together.
- 14.A doctors' surgery has 400 patients. 3 in every 10 of the patients are over 65 years old. 90 of the patients over 65 and 70 of the patients aged 65 or under had a flu jab. One of the patients who had a flu jab is picked at random. Work out the probability that this patient is over 65.
- 15.A survey found that, of 250 shoppers questioned, 68% said they had used a self-checkout in the past month. A different store expects 1400 shoppers this week. Using this relative frequency, work out how many of the 1400 shoppers would be expected to have used a self-checkout in the past month.
Answer key
- (c) 1/10 — On the adult branch, 210 − 189 = 21 appointments were missed. There are 300 − 210 = 90 child appointments, and 90 − 81 = 9 of those were missed. In total, 21 + 9 = 30 appointments were missed, out of 300: 30/300 = 1/10. Writing 7/100 is wrong because 21/300 simplifies to 7/100, and 21 only counts the adult branch, leaving out the 9 missed child appointments. Writing 3/100 is wrong because 9/300 simplifies to 3/100, and 9 only counts the child branch, leaving out the 21 missed adult appointments. Writing 1/9 is wrong because it divides the 30 missed appointments by the 270 that were attended (300 − 30) instead of by the whole 300 booked. The probability is 1/10.
- (d) 3400 — Combining all three greenhouses gives 200 + 150 + 250 = 600 seeds planted in total, and 172 + 126 + 212 = 510 germinated, so the combined estimate of the germination probability is 510/600 = 0.85. Out of a new batch of 4000 seeds, the expected number to germinate is 4000 × 0.85 = 3400. Writing 3440 is wrong because it uses only Greenhouse 1's rate, 172/200 = 0.86, instead of the combined rate from all three: 4000 × 0.86 = 3440. Writing 3360 is wrong because it uses only Greenhouse 2's rate, 126/150 = 0.84: 4000 × 0.84 = 3360. Writing 510 is wrong because that is the total number that germinated in the ORIGINAL trial, not scaled up to the new batch of 4000 seeds at all. The best estimate is 3400 seeds.
- (b) 100 — The sample shows a proportion of 8/60 = 2/15 faulty. Apply that proportion to the new batch of 750: 750 × 2/15 = 100. Flipping the ratio, calculating 8/750 × 60 instead of 8/60 × 750, gives 0.64, which rounds to about 1. Assuming the same number of faulty items applies to the new batch, without scaling for its larger size, just repeats the sample's count of 8. Rounding the proportion 8/60 = 0.1333... down to 0.1 before multiplying gives 750 × 0.1 = 75.
- (c) 500 — Method: the expected number of successes is the number of trials multiplied by the probability, so to find the number of trials, divide the expected number by the probability. Working: let n be the number of seeds planted. Then n multiplied by 0.6 must come to 300, so n = 300 ÷ 0.6 = 500. Answer: the grower should plant 500 seeds. The distractors: 180 comes from multiplying instead of dividing, 300 × 0.6 = 180, which answers how many of 300 seeds would germinate; 750 comes from dividing by the probability of not germinating, 300 ÷ 0.4 = 750; 120 comes from multiplying by that same 0.4, 300 × 0.4 = 120.
- (d) 56 — Method: count the trials over the whole period first, then multiply the number of trials by the probability. Working: 4 weeks is 4 × 7 = 28 days, and at 25 trains a day that is 25 × 28 = 700 trains. The expected number of late trains is 700 × 0.08 = 56. Answer: about 56 late trains over the 4 weeks. The distractors: 2 is the expected number for a single day, 25 × 0.08 = 2, with the 28 days never brought in; 14 uses one week instead of four, 25 × 7 × 0.08 = 14; 644 is 700 − 56 and counts the trains expected to be on time.
- (a) 43/70 — There are 140 − 85 = 55 Zone B households, and 55 − 21 = 34 of them recycle glass. In total, 52 + 34 = 86 households recycle glass, out of 140: 86/140 = 43/70. Writing 13/35 is wrong because 52/140 simplifies to 13/35, and 52 only counts Zone A, leaving out the 34 Zone B recyclers. Writing 17/70 is wrong because 34/140 simplifies to 17/70, and 34 only counts Zone B, leaving out the 52 Zone A recyclers. Writing 73/140 is wrong because it adds the 52 Zone A recyclers to the 21 Zone B households that do NOT recycle, mixing up two different groups instead of adding the two recycling groups. The probability is 43/70.
- (b) 3/5 — The question asks about the girls only, so use the girls' total of 70 as the denominator: 42 out of 70 girls walk, giving 42/70 = 3/5. Choosing 21/65 comes from using the whole survey of 130 pupils as the denominator instead of just the 70 girls, 42/130 = 21/65. Choosing 33/70 comes from using the boys' walking count, 33, over the girls' total of 70, mixing up the two rows of the table. Choosing 2/5 comes from using the number of girls who CYCLE, 70 − 42 = 28, instead of the number who walk, giving 28/70 = 2/5.
- (a) 753 — The estimate from 2000 spins is the most reliable, since it comes from the largest sample size, so the best estimate of the probability is 0.251. Over a further 3000 spins, the expected number landing on green is 3000 × 0.251 = 753. Writing 1050 is wrong because 3000 × 0.350 = 1050 uses the estimate from only 20 spins, the LEAST reliable of the three. Writing 870 is wrong because 3000 × 0.290 = 870 uses the estimate from 200 spins rather than the more reliable 2000-spin estimate. Writing 750 is wrong because 3000 × 0.25 = 750 ignores the recorded data completely and simply assumes each of the 4 colours is equally likely. The best estimate is 753 expected green spins.
- (b) The relative frequency is settling near 0.5 — Method: turn each result into a relative frequency before comparing them, because it is the relative frequency, and not the difference between the two counts, that tends towards the theoretical probability. Working: after 10 flips the relative frequency of a head is 7 ÷ 10 = 0.7, which is a long way from 0.5. After 1000 flips it is 528 ÷ 1000 = 0.528, which is much closer to 0.5. Meanwhile the gap between the two counts has grown rather than shrunk: it was 7 − 3 = 4 after 10 flips and is 528 − 472 = 56 after 1000 flips. Answer: the relative frequency is settling near 0.5, which is what an unbiased experiment does as the sample grows. The distractors: saying the counts are levelling out is the usual form of this idea and the figures contradict it, since the gap went from 4 to 56; saying the coin is biased treats 28 extra heads in 1000 flips as proof, when 0.528 sits close to 0.5 and a fair coin gives results like this often; saying the next flip is more likely to be a tail is the gambler's fallacy, since each flip stays at 1/2 whatever came before.
- (a) £150 — Each game, the expected payout is 0.1 × £20 = £2, so the fête's expected profit per game is the £3 charged minus the £2 expected payout, £1. Over 150 games, that is 150 × £1 = £150. Writing £300 is wrong because 150 × £2 = £300 is the total expected PAYOUT, not the profit — it has not been subtracted from the entry fees. Writing £450 is wrong because 150 × £3 = £450 is the total money taken in entry fees, without accounting for what is expected to be paid out in prizes. Writing £1 is wrong because that is only the expected profit for ONE game — it has not been scaled up to all 150 games. The fête's expected profit is £150.
- (a) 20 — Dice A is fair, so its expected number of sixes is 150 × 1/6 = 25. Dice B has P(6) = 0.3, so its expected number of sixes is 150 × 0.3 = 45. The difference is 45 − 25 = 20. Adding the two expected values instead of subtracting them gives 25 + 45 = 70. Reporting Dice B's expected sixes on their own, without comparing to Dice A, gives 45. Using the fair probability 1/6 for Dice B as well as Dice A ignores the bias altogether, giving 150 × 1/6 = 25 for both dice and a difference of 0.
- (b) 3/10 — The number who use at least one app is 90 − 20 = 70. Since 55 + 42 double-counts the overlap, n(X ∩ Y) = 55 + 42 − 70 = 27, so P(both) = 27/90 = 3/10. Forgetting to subtract the 20 who use neither, and using the full 90 as the union, gives 55 + 42 − 90 = 7, so 7/90. Reporting the probability of using X or Y (or both), 70/90 = 7/9, answers a different question about the union, not the overlap. Reporting the probability of using neither app, 20/90 = 2/9, is the complement of the union, not the intersection.
- (c) 0.37 — Method: pool the two runs into one combined set of results, then find the relative frequency of red across all of the spins together. Working: total reds = 16 + 21 = 37. Total spins = 40 + 60 = 100. Relative frequency = 37 ÷ 100 = 0.37. Answer: 0.37. Watch out: writing down 0.40 uses only the first run, 16 ÷ 40, and throws away the extra evidence from the second 60 spins. Writing down 0.35 uses only the second run, 21 ÷ 60, and throws away the first run instead. And writing down 0.375 averages the two runs' separate rates, (0.40 + 0.35) ÷ 2, which treats a run of 40 spins and a run of 60 spins as equally weighted, when pooling the actual counts gives the larger run its fair share of influence.
- (c) 9/16 — Method: two steps. Total the patients who had a flu jab, since the patient picked is known to be one of them, then divide the over 65s who had a jab by that total. Working: 90 patients over 65 and 70 patients aged 65 or under had a jab, so 160 patients had one. The over 65s give 90/160, and dividing the numerator and the denominator by 10 gives 9/16. Answer: the probability is 9/16. The distractors: 7/16 is 70/160, the probability that the patient picked is aged 65 or under, which is the other part of the same restricted group; 3/4 is 90/120, the probability that a patient had a jab given that they are over 65, which is the condition and the event the wrong way round and needs the 120 patients over 65; 9/40 is 90/400, dividing by every patient on the list instead of by the 160 who had a jab.
- (a) 952 — 68% = 0.68. The relative frequency from the survey applies to the new group of 1400 shoppers, so the expected number is 0.68 × 1400 = 952. Working out 1 − 0.68 = 0.32 and applying that instead, 0.32 × 1400 = 448, finds the number who have NOT used a self-checkout, not the number who have. Applying 68% to the original sample size of 250 instead of the new total of 1400 gives 0.68 × 250 = 170. Applying the complement percentage to the original sample size, 0.32 × 250 = 80, compounds both mistakes.
Build your own mix at the worksheet builder.