Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- 1.Two games are offered at a school fair, and every one of the 40 pupils in a class plays each game once. In Game A, the probability of winning a prize is 0.2 and the prize is worth £5. In Game B, the probability of winning a prize is 0.1 and the prize is worth £12. Work out which game gives a higher expected total value of prizes for the class, and by how much.
- 2.A survey of 90 people records which of two apps, X and Y, they use. 55 people use app X, 42 people use app Y, and 20 people use neither app. Work out the probability that a randomly chosen person uses both apps, giving your answer as a fraction in its simplest form.
- 3.Two pupils each flip the same coin to estimate the probability of heads. Leah flips it 40 times and gets 24 heads. Ben flips it 60 times and gets 33 heads. By combining both pupils' results, work out the relative frequency of heads, giving your answer as a fraction in its simplest form.
- 4.At a fun run, each runner finishes, retires or is disqualified, and cannot do more than one of these. The probability that a runner finishes is 68.5% and the probability that a runner retires is 24.75%. Work out the probability, as a percentage, that a runner is disqualified.
- 5.A spinner can land on red, blue, green or yellow. The probability that it lands on red is 0.24 and the probability that it lands on yellow is 0.16. The probabilities of landing on blue and on green are equal, and each is called x. Work out the value of x.
- 6.A library recorded loans for 250 books over a week, using a frequency tree. The first branch splits the books into 160 fiction and 90 non-fiction. Of the fiction books, 112 were returned on time and the rest were returned late. All the non-fiction books were returned on time. Work out the probability that a book, chosen at random from the 250, was returned late. Give your answer as a fraction in its simplest form.
- 7.A clinic recorded 300 booked appointments using a frequency tree. The first branch splits them into 210 adult appointments and the rest child appointments. Of the adult appointments, 189 were attended and the rest were missed. Of the child appointments, 81 were attended. Work out the probability that a booked appointment, chosen at random from the 300, was missed. Give your answer as a fraction in its simplest form.
- 8.A market stall sells umbrellas. Over the last 250 days, it rained on 70 of them. Using this as an estimate of the probability of rain, work out how many rainy days would be expected in the next 365 days.
- 9.A four-colour spinner (red, blue, green, yellow) is spun repeatedly, and the relative frequency of landing on green is recorded as the number of spins increases: after 20 spins it is 0.350; after 200 spins it is 0.290; after 2000 spins it is 0.251. Using the result from 2000 spins as the best estimate of the probability, work out the number of times the spinner would be expected to land on green in a further 3000 spins.
- 10.A phone network sends automatic text alerts to customers. On average, 1,500 alerts are sent each day, and the probability that a customer replies 'STOP' to an alert is 0.18. Work out how many replies of 'STOP' the network should expect over a 30-day month.
- 11.A frequency tree records the results of 160 patients who took a new medicine. It splits them into those who reported side effects and those who did not. 15% of the patients reported side effects. Work out how many of the 160 patients did not report side effects.
- 12.A fair six-sided dice is rolled 90 times. Work out how many more times you would expect it to land on a number less than 4 than on a 6.
- 13.A quality inspector tests 145 light bulbs and finds that 33 are faulty. Work out the relative frequency of a bulb being faulty, as a percentage correct to 1 decimal place.
- 14.A two-way table records how 180 students at a school travel: by bus or on foot, split by year group. There are 84 students in Year 11, of whom 38 travel by bus and the rest walk. The rest of the 180 students are in Year 10, and 42 of the Year 10 students travel by bus. Work out the probability that a randomly chosen Year 10 student walks to school. Give your answer as a fraction in its simplest form.
- 15.A factory tests components from a large batch in which 6% are defective. Two components are selected at random, and the batch is large enough that the selections can be treated as independent. Given that at least one of the two components is defective, work out the probability that both are defective.
Answer key
- (a) Game B, by £8 — Game A's expected total is 40 × 0.2 × £5 = £40. Game B's expected total is 40 × 0.1 × £12 = £48. Game B is higher, by £48 − £40 = £8. Writing 'Game A, by £8' is wrong because it has the right difference but the wrong game — Game A's total (£40) is actually LOWER than Game B's, not higher. Writing 'Game B, by £48' is wrong because £48 is Game B's whole expected total, not the DIFFERENCE between the two games. Writing 'Game A, by £40' is wrong in the same way, using Game A's whole total as if it were the margin, and naming the wrong game as the winner. Game B gives the higher expected total, by £8.
- (b) 3/10 — The number who use at least one app is 90 − 20 = 70. Since 55 + 42 double-counts the overlap, n(X ∩ Y) = 55 + 42 − 70 = 27, so P(both) = 27/90 = 3/10. Forgetting to subtract the 20 who use neither, and using the full 90 as the union, gives 55 + 42 − 90 = 7, so 7/90. Reporting the probability of using X or Y (or both), 70/90 = 7/9, answers a different question about the union, not the overlap. Reporting the probability of using neither app, 20/90 = 2/9, is the complement of the union, not the intersection.
- (b) 57/100 — Pooling both trials: total heads = 24 + 33 = 57, total flips = 40 + 60 = 100, so the combined relative frequency is 57/100, which is already in its simplest form since 57 and 100 share no common factor. Averaging the two separate relative frequencies instead, (24/40 + 33/60) ÷ 2 = (0.6 + 0.55) ÷ 2 = 0.575 = 23/40, treats the two trials as equally weighted even though Ben made more flips, which is not correct. Using only Leah's data gives 24/40 = 3/5. Using only Ben's data gives 33/60 = 11/20.
- (a) 6.75% — Finishing, retiring and being disqualified are exhaustive, so the three percentages sum to 100%: 100% − 68.5% − 24.75% = 6.75%. Adding the two given percentages instead of subtracting them from 100% gives 68.5% + 24.75% = 93.25%, the combined probability of finishing or retiring, not of being disqualified. Subtracting only the retiring percentage from 100% and forgetting the finishing percentage gives 100% − 24.75% = 75.25%. Subtracting only the finishing percentage and forgetting the retiring percentage gives 100% − 68.5% = 31.50%.
- (a) 0.30 — Red, blue, green and yellow are exhaustive, so all four probabilities sum to 1: 0.24 + 0.16 + x + x = 1, so 2x + 0.40 = 1, giving 2x = 0.60 and x = 0.30. Stopping at 2x = 0.60 without dividing by 2 leaves 0.60, the combined probability of both blue and green together, not the value of x on its own. Sharing the 0.60 across all four colours instead of just the two unknown ones gives 0.60 ÷ 4 = 0.15. Leaving out the 0.16 for yellow gives 2x + 0.24 = 1, so 2x = 0.76 and x = 0.38.
- (b) 24/125 — On the fiction branch, 160 − 112 = 48 books were returned late. None of the non-fiction books were late, so the total number of late books is 48, out of 250 books altogether: 48/250 = 24/125. Writing 3/10 is wrong because it divides the 48 late fiction books by the fiction total (160) instead of the whole library total (250). Writing 56/125 is wrong because 112/250 simplifies to 56/125, and 112 is the number of fiction books returned ON TIME, not late. Writing 9/25 is wrong because 90/250 simplifies to 9/25, and 90 is simply the number of non-fiction books, which has nothing to do with late returns. The probability is 24/125.
- (c) 1/10 — On the adult branch, 210 − 189 = 21 appointments were missed. There are 300 − 210 = 90 child appointments, and 90 − 81 = 9 of those were missed. In total, 21 + 9 = 30 appointments were missed, out of 300: 30/300 = 1/10. Writing 7/100 is wrong because 21/300 simplifies to 7/100, and 21 only counts the adult branch, leaving out the 9 missed child appointments. Writing 3/100 is wrong because 9/300 simplifies to 3/100, and 9 only counts the child branch, leaving out the 21 missed adult appointments. Writing 1/9 is wrong because it divides the 30 missed appointments by the 270 that were attended (300 − 30) instead of by the whole 300 booked. The probability is 1/10.
- (c) 102 — Method: turn the past record into a relative frequency, then use it as an estimate of the probability of rain and multiply by the number of days being predicted for. Working: relative frequency of rain = 70 ÷ 250 = 0.28. Expected rainy days in 365 days = 365 × 0.28 = 102.2, which rounds to about 102 days. Answer: about 102 days. Watch out: writing down 48 swaps which number is the sample and which is the target, working out 70 ÷ 365 × 250 instead of 70 ÷ 250 × 365. Writing down 70 just repeats the original count of rainy days without scaling it up to the new, longer period at all. And writing down 110 comes from rounding the relative frequency to 0.3 before multiplying, 365 × 0.3 = 109.5, when 70 ÷ 250 is exactly 0.28 and needs no rounding at all.
- (a) 753 — The estimate from 2000 spins is the most reliable, since it comes from the largest sample size, so the best estimate of the probability is 0.251. Over a further 3000 spins, the expected number landing on green is 3000 × 0.251 = 753. Writing 1050 is wrong because 3000 × 0.350 = 1050 uses the estimate from only 20 spins, the LEAST reliable of the three. Writing 870 is wrong because 3000 × 0.290 = 870 uses the estimate from 200 spins rather than the more reliable 2000-spin estimate. Writing 750 is wrong because 3000 × 0.25 = 750 ignores the recorded data completely and simply assumes each of the 4 colours is equally likely. The best estimate is 753 expected green spins.
- (b) 8,100 — First find the total number of alerts sent in the month: 1,500 × 30 = 45,000. Then apply the probability of a 'STOP' reply: 45,000 × 0.18 = 8,100. Stopping after finding only one day's expected replies, 1,500 × 0.18 = 270, forgets to scale up to the whole month. Multiplying the number of days by the probability instead of by the daily total of alerts gives 30 × 0.18 = 5.4, which rounds to 5. Shifting the decimal point in the probability, using 0.018 instead of 0.18, gives 45,000 × 0.018 = 810.
- (d) 136 — 15% of 160 = 0.15 × 160 = 24 patients reported side effects, so 160 − 24 = 136 did not. Stopping after finding the number who reported side effects, 24, answers the wrong question — it is not the number who did NOT report them. Misreading '15%' as a raw count of 15 patients, rather than a percentage, gives 160 − 15 = 145. Subtracting 15% of 160 twice, 160 − 24 − 24 = 112, double-counts the side-effect group.
- (c) 30 — The numbers less than 4 are 1, 2 and 3, so the probability of that event is 3/6, and the expected count in 90 rolls is 90 × 3/6 = 45. The probability of rolling a 6 is 1/6, and the expected count is 90 × 1/6 = 15. The difference between the two expected counts is 45 − 15 = 30. A candidate who answers 45 has given the expected count for 'less than 4' only, forgetting to subtract the other expected count. A candidate who answers 15 has given the expected count for '6' only. A candidate who answers 36 has used a dice with 5 possible numbers instead of 6, giving 90 × 3/5 = 54 and 90 × 1/5 = 18, a difference of 36.
- (b) 22.8% — Relative frequency as a percentage is the faulty count divided by the total, then multiplied by 100: 33 ÷ 145 × 100 = 22.76, which rounds to 22.8%. Giving 33.0% as the answer uses the frequency, 33, directly as a percentage without dividing by the total 145 at all. Rounding 22.76 down to 22.7% instead of up applies the wrong rounding direction at the first decimal place. Finding the relative frequency of the bulbs that were NOT faulty first: 145 − 33 = 112, and 112 ÷ 145 × 100 = 77.24, answers the opposite question and rounds to 77.2%.
- (c) 9/16 — There are 180 students in total and 84 are in Year 11, so Year 10 has 180 − 84 = 96 students. Of those 96, 42 travel by bus, so 96 − 42 = 54 walk. P(Year 10 student walks) = 54/96 = 9/16. Using the whole school of 180 as the denominator instead of just the 96 Year 10 students gives 54/180 = 3/10. Using the bus count, 42, as if it were the number who walk gives 42/96 = 7/16, the wrong branch of the Year 10 row. Working out the probability for Year 11 instead of Year 10 — 46 walkers out of 84 — gives 46/84 = 23/42.
- (b) 0.0309 — Method: P(both defective | at least one defective) = P(both defective) ÷ P(at least one defective). Find each using independence: P(both) = 0.06², P(at least one) = 1 − P(neither) = 1 − 0.94². Working: P(both) = 0.06² = 0.0036. P(neither) = 0.94² = 0.8836, so P(at least one) = 1 − 0.8836 = 0.1164. P(both | at least one) = 0.0036 ÷ 0.1164 = 0.0309 (3 s.f.). Answer: 0.0309. Watch out: leaving the answer as 0.0036 gives P(both defective) itself, not the probability once you already know at least one is defective — you still need to divide by P(at least one defective). Giving 0.0600 answers with the single-component defect rate, ignoring the condition altogether. And 0.5000 assumes that 'at least one' makes the outcomes 'exactly one defective' and 'both defective' equally likely, which is not how these probabilities combine.
Build your own mix at the worksheet builder.