Printable · GCSE Foundation · ages 14-16
Probability worksheet — GCSE Foundation
Fifteen questions across the probability statements at Foundation tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Foundation
MathsUKwww.geekhero.co.uk
- 1.An ordinary six-sided dice, numbered 1 to 6, is rolled 30 times and lands on a 6 seven times. Ravi says the theoretical probability of rolling a 6 and the relative frequency of rolling a 6 in this trial are the same number. Is Ravi right?
- 2.A garden centre tests 250 packets of seeds for germination, recording the results on a frequency tree. 68% of the 250 packets germinated successfully, and the rest did not. Work out how many packets germinated successfully.
- 3.A survey found that, of 250 shoppers questioned, 68% said they had used a self-checkout in the past month. A different store expects 1400 shoppers this week. Using this relative frequency, work out how many of the 1400 shoppers would be expected to have used a self-checkout in the past month.
- 4.A padlock code is formed by arranging three different digits chosen from 2, 3, 4 and 5, with no digit repeated. Work out how many different three-digit codes can be made.
- 5.A survey asked 80 people to name their favourite drink. The table shows the results: Tea 20 people, Coffee 32 people, Juice 16 people, Water 12 people. Work out the percentage of people whose favourite drink was Coffee.
- 6.Two goalkeepers face penalty kicks. Elin saves 34% of the penalties she faces. Noah saves 11/32 of the penalties he faces. Which statement correctly compares them?
- 7.A fair six-sided dice is rolled once. Work out the probability that the score is an odd number. Give your answer as a decimal.
- 8.A frequency tree records the results of 160 patients who took a new medicine. It splits them into those who reported side effects and those who did not. 15% of the patients reported side effects. Work out how many of the 160 patients did not report side effects.
- 9.A four-colour spinner (red, blue, green, yellow) is spun repeatedly, and the relative frequency of landing on green is recorded as the number of spins increases: after 20 spins it is 0.350; after 200 spins it is 0.290; after 2000 spins it is 0.251. Using the result from 2000 spins as the best estimate of the probability, work out the number of times the spinner would be expected to land on green in a further 3000 spins.
- 10.A Venn diagram shows two sets, P and Q, inside a universal set. n(P) = 34, n(Q) = 27, n(P ∩ Q) = 11, and n(ξ) = 90, where ξ is the universal set. Work out n((P ∪ Q)′), the number of elements in neither P nor Q.
- 11.A spinner can land on red, blue, green or yellow. The probability that it lands on red is 0.24 and the probability that it lands on yellow is 0.16. The probabilities of landing on blue and on green are equal, and each is called x. Work out the value of x.
- 12.Two games are offered at a school fair, and every one of the 40 pupils in a class plays each game once. In Game A, the probability of winning a prize is 0.2 and the prize is worth £5. In Game B, the probability of winning a prize is 0.1 and the prize is worth £12. Work out which game gives a higher expected total value of prizes for the class, and by how much.
- 13.A seed company tests germination using results from three greenhouses. Greenhouse 1 plants 200 seeds and 172 germinate. Greenhouse 2 plants 150 seeds and 126 germinate. Greenhouse 3 plants 250 seeds and 212 germinate. Using the combined results from all three greenhouses, work out the best estimate of the number of seeds, out of a new batch of 4000 seeds, that would be expected to germinate.
- 14.A library recorded loans for 250 books over a week, using a frequency tree. The first branch splits the books into 160 fiction and 90 non-fiction. Of the fiction books, 112 were returned on time and the rest were returned late. All the non-fiction books were returned on time. Work out the probability that a book, chosen at random from the 250, was returned late. Give your answer as a fraction in its simplest form.
- 15.A council surveyed 140 households about recycling. 85 of the households are in Zone A. 52 of the Zone A households recycle glass. 21 of the Zone B households do not recycle glass. Work out the probability that a household, chosen at random from the 140, recycles glass. Give your answer as a fraction in its simplest form.
Answer key
- (c) No — 7/30 is the relative frequency; theory stays 1/6. — The theoretical probability of rolling a 6 on an ordinary dice is fixed at 1/6, worked out from the number of equally likely outcomes, and does not change however the dice is actually rolled. The relative frequency from this trial is 7/30, found from what happened in these particular 30 rolls. Since 7/30 and 1/6 are different numbers, the correct statement is 'No — 7/30 is the relative frequency; theory stays 1/6.' Assuming the two values must always match because they describe the same event gives 'Yes — relative frequency always equals theory.' Believing that an observed result redefines the theoretical probability gives 'Yes — the theoretical probability has now become 7/30.' Refusing to work out either value at all gives 'Neither can be found — 30 rolls is too few to tell', which ignores that both numbers CAN be calculated from the information given.
- (c) 170 — 68% of the 250 packets germinated: 250 × 0.68 = 170. Using the complement, the 32% that did NOT germinate, gives 250 × 0.32 = 80. Shifting the decimal point, using 0.068 instead of 0.68, gives 250 × 0.068 = 17. Rounding 68% up to 70% before multiplying gives 250 × 0.70 = 175.
- (a) 952 — 68% = 0.68. The relative frequency from the survey applies to the new group of 1400 shoppers, so the expected number is 0.68 × 1400 = 952. Working out 1 − 0.68 = 0.32 and applying that instead, 0.32 × 1400 = 448, finds the number who have NOT used a self-checkout, not the number who have. Applying 68% to the original sample size of 250 instead of the new total of 1400 gives 0.68 × 250 = 170. Applying the complement percentage to the original sample size, 0.32 × 250 = 80, compounds both mistakes.
- (d) 24 — There are 4 choices for the first digit. Once that digit is used, 3 digits remain for the second position, and then 2 digits remain for the third position: 4 × 3 × 2 = 24 codes. Choosing 64 comes from allowing a digit to be reused at every position, 4 × 4 × 4 = 64, which is not allowed here since no digit repeats. Choosing 12 comes from multiplying only the first two positions, 4 × 3 = 12, and forgetting that a third digit is also chosen from the digits that remain. Choosing 6 comes from counting only the arrangements of one single set of three digits, 3 × 2 × 1 = 6, and forgetting that there are 4 different sets of three digits that can be chosen from 2, 3, 4 and 5.
- (b) 40% — 32 out of the 80 people chose Coffee, so the percentage is (32 ÷ 80) × 100 = 40%. Choosing 32% comes from treating 32 as if it were already a percentage out of 100, instead of dividing by the actual total of 80 people surveyed. Choosing 25% comes from using the Tea row instead of the Coffee row, (20 ÷ 80) × 100 = 25%. Choosing 20% comes from using the Juice row instead of the Coffee row, (16 ÷ 80) × 100 = 20%.
- (c) Noah — 11/32 = 0.34375, above 34%. — Converting 11/32 to a decimal gives 11 ÷ 32 = 0.34375, which is greater than 34% (0.34), so Noah has the better save rate: 'Noah — 11/32 = 0.34375, above 34%.' Comparing the raw numbers 34 and 11 directly, without converting the fraction to the same form, gives 'Elin — 34 is bigger than 11.' Treating a larger denominator as meaning a bigger value, rather than smaller equal shares, gives 'Noah — 32 is a bigger denominator.' Rounding 34.375% to 34% to the nearest whole percent hides the difference and gives 'Equal — both round to 34% to the nearest percent.'
- (c) 0.5 — Method: count the favourable outcomes, write them over the total number of equally likely outcomes and then divide to turn the fraction into a decimal. Working: the odd scores are 1, 3 and 5, which is 3 of the 6 equally likely scores, so the probability is 3/6, and 3 ÷ 6 = 0.5. Answer: 0.5, the middle of the 0 to 1 probability scale. The distractors: 0.33 comes from listing only 3 and 5 as odd and working out 2 ÷ 6; 0.17 comes from giving the probability of one particular odd score, 1 ÷ 6; 0.3 comes from writing '3 out of 6' as 0.3, reading the 3 straight off as tenths instead of dividing.
- (d) 136 — 15% of 160 = 0.15 × 160 = 24 patients reported side effects, so 160 − 24 = 136 did not. Stopping after finding the number who reported side effects, 24, answers the wrong question — it is not the number who did NOT report them. Misreading '15%' as a raw count of 15 patients, rather than a percentage, gives 160 − 15 = 145. Subtracting 15% of 160 twice, 160 − 24 − 24 = 112, double-counts the side-effect group.
- (a) 753 — The estimate from 2000 spins is the most reliable, since it comes from the largest sample size, so the best estimate of the probability is 0.251. Over a further 3000 spins, the expected number landing on green is 3000 × 0.251 = 753. Writing 1050 is wrong because 3000 × 0.350 = 1050 uses the estimate from only 20 spins, the LEAST reliable of the three. Writing 870 is wrong because 3000 × 0.290 = 870 uses the estimate from 200 spins rather than the more reliable 2000-spin estimate. Writing 750 is wrong because 3000 × 0.25 = 750 ignores the recorded data completely and simply assumes each of the 4 colours is equally likely. The best estimate is 753 expected green spins.
- (b) 40 — n(P ∪ Q) = n(P) + n(Q) − n(P ∩ Q) = 34 + 27 − 11 = 50. The complement is everyone outside both sets: n((P ∪ Q)′) = 90 − 50 = 40. Adding P and Q without subtracting the overlap gives 34 + 27 = 61, so 90 − 61 = 29 double-subtracts the 11 who are in both. Reporting n(P ∪ Q) itself, 50, forgets to take the complement at all. Subtracting only n(P) from the universal set, 90 − 34 = 56, ignores set Q altogether.
- (a) 0.30 — Red, blue, green and yellow are exhaustive, so all four probabilities sum to 1: 0.24 + 0.16 + x + x = 1, so 2x + 0.40 = 1, giving 2x = 0.60 and x = 0.30. Stopping at 2x = 0.60 without dividing by 2 leaves 0.60, the combined probability of both blue and green together, not the value of x on its own. Sharing the 0.60 across all four colours instead of just the two unknown ones gives 0.60 ÷ 4 = 0.15. Leaving out the 0.16 for yellow gives 2x + 0.24 = 1, so 2x = 0.76 and x = 0.38.
- (a) Game B, by £8 — Game A's expected total is 40 × 0.2 × £5 = £40. Game B's expected total is 40 × 0.1 × £12 = £48. Game B is higher, by £48 − £40 = £8. Writing 'Game A, by £8' is wrong because it has the right difference but the wrong game — Game A's total (£40) is actually LOWER than Game B's, not higher. Writing 'Game B, by £48' is wrong because £48 is Game B's whole expected total, not the DIFFERENCE between the two games. Writing 'Game A, by £40' is wrong in the same way, using Game A's whole total as if it were the margin, and naming the wrong game as the winner. Game B gives the higher expected total, by £8.
- (d) 3400 — Combining all three greenhouses gives 200 + 150 + 250 = 600 seeds planted in total, and 172 + 126 + 212 = 510 germinated, so the combined estimate of the germination probability is 510/600 = 0.85. Out of a new batch of 4000 seeds, the expected number to germinate is 4000 × 0.85 = 3400. Writing 3440 is wrong because it uses only Greenhouse 1's rate, 172/200 = 0.86, instead of the combined rate from all three: 4000 × 0.86 = 3440. Writing 3360 is wrong because it uses only Greenhouse 2's rate, 126/150 = 0.84: 4000 × 0.84 = 3360. Writing 510 is wrong because that is the total number that germinated in the ORIGINAL trial, not scaled up to the new batch of 4000 seeds at all. The best estimate is 3400 seeds.
- (b) 24/125 — On the fiction branch, 160 − 112 = 48 books were returned late. None of the non-fiction books were late, so the total number of late books is 48, out of 250 books altogether: 48/250 = 24/125. Writing 3/10 is wrong because it divides the 48 late fiction books by the fiction total (160) instead of the whole library total (250). Writing 56/125 is wrong because 112/250 simplifies to 56/125, and 112 is the number of fiction books returned ON TIME, not late. Writing 9/25 is wrong because 90/250 simplifies to 9/25, and 90 is simply the number of non-fiction books, which has nothing to do with late returns. The probability is 24/125.
- (a) 43/70 — There are 140 − 85 = 55 Zone B households, and 55 − 21 = 34 of them recycle glass. In total, 52 + 34 = 86 households recycle glass, out of 140: 86/140 = 43/70. Writing 13/35 is wrong because 52/140 simplifies to 13/35, and 52 only counts Zone A, leaving out the 34 Zone B recyclers. Writing 17/70 is wrong because 34/140 simplifies to 17/70, and 34 only counts Zone B, leaving out the 52 Zone A recyclers. Writing 73/140 is wrong because it adds the 52 Zone A recyclers to the 21 Zone B households that do NOT recycle, mixing up two different groups instead of adding the two recycling groups. The probability is 43/70.
Build your own mix at the worksheet builder.