Printable · GCSE Foundation · ages 14-16
Probability worksheet — GCSE Foundation
Fifteen questions across the probability statements at Foundation tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Foundation
MathsUKwww.geekhero.co.uk
- 1.At a school car park, each car is a saloon, an estate or a hatchback, and no car is more than one of these. Among 200 cars, the probability that a car chosen at random is a saloon is 0.28 and the probability that it is an estate is 0.37. Work out how many of the 200 cars are hatchbacks.
- 2.Two goalkeepers face penalty kicks. Elin saves 34% of the penalties she faces. Noah saves 11/32 of the penalties he faces. Which statement correctly compares them?
- 3.A survey of 90 people records which of two apps, X and Y, they use. 55 people use app X, 42 people use app Y, and 20 people use neither app. Work out the probability that a randomly chosen person uses both apps, giving your answer as a fraction in its simplest form.
- 4.A leisure centre has 150 members. 80 of the members are male and the rest are female. Every member uses either the pool or the gym, but not both. 66 members use the pool, and 35 of those pool users are male. Work out how many female members use the gym.
- 5.A drawing pin is dropped many times and lands either point up or point down. The relative frequency of landing point up is recorded as the experiment goes on: after 50 drops it is 0.720, after 200 drops it is 0.665, and after 1000 drops it is 0.638. The pin is to be dropped a further 2000 times. Work out the best estimate of the number of times it will land point up.
- 6.Ellie drops a bottle top and records whether it lands open end up. In her first 20 drops it lands open end up 13 times. She carries on, and after 200 drops in total it has landed open end up 84 times. Work out the best estimate of the probability that the bottle top lands open end up.
- 7.A council surveyed 140 households about recycling. 85 of the households are in Zone A. 52 of the Zone A households recycle glass. 21 of the Zone B households do not recycle glass. Work out the probability that a household, chosen at random from the 140, recycles glass. Give your answer as a fraction in its simplest form.
- 8.A train company runs 25 trains a day, every day. The probability that any one train is late is 0.08. Work out how many late trains the company should expect over a period of 4 weeks.
- 9.A factory finds that the probability a randomly chosen light bulb is defective is 0.035. In a batch of 4,000 bulbs, work out how many bulbs you would expect to work correctly.
- 10.A seed company tests germination using results from three greenhouses. Greenhouse 1 plants 200 seeds and 172 germinate. Greenhouse 2 plants 150 seeds and 126 germinate. Greenhouse 3 plants 250 seeds and 212 germinate. Using the combined results from all three greenhouses, work out the best estimate of the number of seeds, out of a new batch of 4000 seeds, that would be expected to germinate.
- 11.A padlock code is formed by arranging three different digits chosen from 2, 3, 4 and 5, with no digit repeated. Work out how many different three-digit codes can be made.
- 12.A fair six-sided dice is rolled 150 times. The table shows how many times each number came up: 1 came up 22 times, 2 came up 27 times, 3 came up 24 times, 4 came up 34 times, 5 came up 21 times and 6 came up 22 times. The theoretical probability of each number is 1/6. Which number is most over-represented compared with its theoretical probability?
- 13.A survey found that, of 250 shoppers questioned, 68% said they had used a self-checkout in the past month. A different store expects 1400 shoppers this week. Using this relative frequency, work out how many of the 1400 shoppers would be expected to have used a self-checkout in the past month.
- 14.A raffle sells 400 tickets at 50p each. There is one prize of £45. Aisha buys 8 tickets. Work out how much money Aisha should expect to lose from playing, giving your answer in pounds.
- 15.A garden centre tests 250 packets of seeds for germination, recording the results on a frequency tree. 68% of the 250 packets germinated successfully, and the rest did not. Work out how many packets germinated successfully.
Answer key
- (b) 70 — Saloon, estate and hatchback are exhaustive, so their probabilities sum to 1: the probability of a hatchback is 1 − 0.28 − 0.37 = 0.35. The number of hatchbacks is 0.35 × 200 = 70. Treating the SUM of the other two probabilities, 0.28 + 0.37 = 0.65, as the probability of a hatchback instead of its complement gives 0.65 × 200 = 130. Multiplying the correct probability, 0.35, by 100 instead of the 200 cars actually surveyed gives 35. Averaging the two given probabilities, (0.28 + 0.37) ÷ 2 = 0.325, instead of subtracting them from 1, and then multiplying by 200 gives 65.
- (c) Noah — 11/32 = 0.34375, above 34%. — Converting 11/32 to a decimal gives 11 ÷ 32 = 0.34375, which is greater than 34% (0.34), so Noah has the better save rate: 'Noah — 11/32 = 0.34375, above 34%.' Comparing the raw numbers 34 and 11 directly, without converting the fraction to the same form, gives 'Elin — 34 is bigger than 11.' Treating a larger denominator as meaning a bigger value, rather than smaller equal shares, gives 'Noah — 32 is a bigger denominator.' Rounding 34.375% to 34% to the nearest whole percent hides the difference and gives 'Equal — both round to 34% to the nearest percent.'
- (b) 3/10 — The number who use at least one app is 90 − 20 = 70. Since 55 + 42 double-counts the overlap, n(X ∩ Y) = 55 + 42 − 70 = 27, so P(both) = 27/90 = 3/10. Forgetting to subtract the 20 who use neither, and using the full 90 as the union, gives 55 + 42 − 90 = 7, so 7/90. Reporting the probability of using X or Y (or both), 70/90 = 7/9, answers a different question about the union, not the overlap. Reporting the probability of using neither app, 20/90 = 2/9, is the complement of the union, not the intersection.
- (b) 39 — Method: put the counts into a two-way table and fill each missing cell by subtracting along a row or down a column. Working: the number of female members is 150 − 80 = 70. The pool column holds 66 members and 35 of them are male, so the number of female pool users is 66 − 35 = 31. Subtracting along the female row, 70 − 31 = 39 female members use the gym. Answer: 39 female members use the gym. The distractors: 45 comes from subtracting along the male row instead, 80 − 35 = 45, which counts male gym users; 31 is the female pool cell, written down one step before the gym cell; 84 is 150 − 66 and counts every gym user, male and female together.
- (a) 1276 — Method: an unbiased relative frequency tends towards the theoretical probability as the number of trials increases, so use the record resting on the most trials, then multiply by the number of new trials. Working: the three records rest on 50, 200 and 1000 drops, so the most reliable is the one after 1000 drops, namely 0.638, and the run is indeed settling as the trials increase. The expected number of point up landings in 2000 further drops is 2000 × 0.638 = 1276. Answer: about 1276 times. The distractors: 1440 uses the earliest record, which rests on only 50 drops, giving 2000 × 0.720 = 1440; 1330 uses the middle record, treating 200 drops as a safe compromise when 1000 drops is better still, giving 2000 × 0.665 = 1330; 1348 comes from averaging the three records, since 0.720 + 0.665 + 0.638 = 2.023 and 2.023 ÷ 3 = 0.674, then 2000 × 0.674 = 1348, which gives the 50 drop record the same weight as the 1000 drop record.
- (d) 0.420 — Method: use the relative frequency worked out from the larger number of trials as the best estimate of the probability, since a bigger sample tends to sit closer to the true probability. Working: 84 out of 200 drops land open end up, so the relative frequency from the total is 84 ÷ 200 = 0.420. Answer: 0.420. Watch out: writing down 0.650 comes from 13 ÷ 20, using only the first, much smaller sample instead of the total. Writing down 0.535 comes from averaging 0.650 and 0.420, treating the 20-drop run and the 200-drop run as equally reliable instead of using the larger sample on its own. And writing down 0.580 comes from 1 − 0.420, working out the probability that the bottle top lands the other way up instead of open end up.
- (a) 43/70 — There are 140 − 85 = 55 Zone B households, and 55 − 21 = 34 of them recycle glass. In total, 52 + 34 = 86 households recycle glass, out of 140: 86/140 = 43/70. Writing 13/35 is wrong because 52/140 simplifies to 13/35, and 52 only counts Zone A, leaving out the 34 Zone B recyclers. Writing 17/70 is wrong because 34/140 simplifies to 17/70, and 34 only counts Zone B, leaving out the 52 Zone A recyclers. Writing 73/140 is wrong because it adds the 52 Zone A recyclers to the 21 Zone B households that do NOT recycle, mixing up two different groups instead of adding the two recycling groups. The probability is 43/70.
- (d) 56 — Method: count the trials over the whole period first, then multiply the number of trials by the probability. Working: 4 weeks is 4 × 7 = 28 days, and at 25 trains a day that is 25 × 28 = 700 trains. The expected number of late trains is 700 × 0.08 = 56. Answer: about 56 late trains over the 4 weeks. The distractors: 2 is the expected number for a single day, 25 × 0.08 = 2, with the 28 days never brought in; 14 uses one week instead of four, 25 × 7 × 0.08 = 14; 644 is 700 − 56 and counts the trains expected to be on time.
- (a) 3,860 — The probability a bulb works correctly is the complement of being defective: 1 − 0.035 = 0.965. Expected number working correctly = 0.965 × 4,000 = 3,860. Using the probability of being defective instead of its complement gives 4,000 × 0.035 = 140, the expected number of DEFECTIVE bulbs, not working ones. Shifting the decimal point in the complement, using 0.0965 instead of 0.965, gives 4,000 × 0.0965 = 386. Assuming every bulb works, ignoring the 0.035 probability altogether, gives the full batch of 4,000.
- (d) 3400 — Combining all three greenhouses gives 200 + 150 + 250 = 600 seeds planted in total, and 172 + 126 + 212 = 510 germinated, so the combined estimate of the germination probability is 510/600 = 0.85. Out of a new batch of 4000 seeds, the expected number to germinate is 4000 × 0.85 = 3400. Writing 3440 is wrong because it uses only Greenhouse 1's rate, 172/200 = 0.86, instead of the combined rate from all three: 4000 × 0.86 = 3440. Writing 3360 is wrong because it uses only Greenhouse 2's rate, 126/150 = 0.84: 4000 × 0.84 = 3360. Writing 510 is wrong because that is the total number that germinated in the ORIGINAL trial, not scaled up to the new batch of 4000 seeds at all. The best estimate is 3400 seeds.
- (d) 24 — There are 4 choices for the first digit. Once that digit is used, 3 digits remain for the second position, and then 2 digits remain for the third position: 4 × 3 × 2 = 24 codes. Choosing 64 comes from allowing a digit to be reused at every position, 4 × 4 × 4 = 64, which is not allowed here since no digit repeats. Choosing 12 comes from multiplying only the first two positions, 4 × 3 = 12, and forgetting that a third digit is also chosen from the digits that remain. Choosing 6 comes from counting only the arrangements of one single set of three digits, 3 × 2 × 1 = 6, and forgetting that there are 4 different sets of three digits that can be chosen from 2, 3, 4 and 5.
- (b) 4 — With 150 rolls and probability 1/6 for each number, the expected count is 150 ÷ 6 = 25. Comparing each actual count with 25: 1 is 22 (3 below), 2 is 27 (2 above), 3 is 24 (1 below), 4 is 34 (9 above), 5 is 21 (4 below) and 6 is 22 (3 below). Number 4 is furthest above its expected count, so it is the most over-represented. Number 2 is also above its expected count, but by only 2, far less than 4's 9. Number 3's count of 24 is below the expected 25, so it is under-represented, not over. Number 6's count of 22 is also below the expected 25, so it too is under-represented.
- (a) 952 — 68% = 0.68. The relative frequency from the survey applies to the new group of 1400 shoppers, so the expected number is 0.68 × 1400 = 952. Working out 1 − 0.68 = 0.32 and applying that instead, 0.32 × 1400 = 448, finds the number who have NOT used a self-checkout, not the number who have. Applying 68% to the original sample size of 250 instead of the new total of 1400 gives 0.68 × 250 = 170. Applying the complement percentage to the original sample size, 0.32 × 250 = 80, compounds both mistakes.
- (c) £3.10 — Aisha's tickets cost 8 × 50p = £4.00. Her expected winnings are (8/400) × £45 = £0.90, since she holds 8 of the 400 tickets. Her expected loss is the cost minus the expected winnings: £4.00 − £0.90 = £3.10. Writing £4.00 is wrong because it is only the cost of her tickets, with no account taken of the expected winnings she might get back. Writing £0.90 is wrong because that is her expected WINNINGS, not her loss — the cost has not been subtracted. Writing £3.89 is wrong because it uses 1 ticket instead of her actual 8 tickets when working out the expected winnings: (1/400) × £45 = £0.1125, giving £4.00 − £0.11 = £3.89. Aisha should expect to lose £3.10.
- (c) 170 — 68% of the 250 packets germinated: 250 × 0.68 = 170. Using the complement, the 32% that did NOT germinate, gives 250 × 0.32 = 80. Shifting the decimal point, using 0.068 instead of 0.68, gives 250 × 0.068 = 17. Rounding 68% up to 70% before multiplying gives 250 × 0.70 = 175.
Build your own mix at the worksheet builder.