Printable · GCSE Foundation · ages 14-16
Probability worksheet — GCSE Foundation
Fifteen questions across the probability statements at Foundation tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Answer key: Probability worksheet — GCSE Foundation
MathsUKwww.geekhero.co.uk
- (b) likely — On the probability scale, 'certain' is reserved for a probability of exactly 1, and 'evens' describes a probability of exactly 0.5. A probability of 0.9 is high but not equal to 1, so the correct word is 'likely'. Choosing 'certain' treats a probability close to 1 as if it were exactly 1, which it is not. Choosing 'evens' misjudges 0.9 as being close to the midpoint of the scale, when it is close to the 'certain' end instead. Choosing 'unlikely' reads the scale the wrong way round, as if a high probability meant a low chance of happening.
- (d) 3/16 — Method: work out each draw's own probability first, then multiply them together since the two draws are independent. Working: there are 8 tickets in all, 3 of them blue, so P(blue) = 3/8. P(spinner number greater than 2) = 2/4 = 1/2, since 3 and 4 qualify. Multiplying gives 3/8 × 1/2, which comes to 3/16. Answer: 3/16. Watch out: writing down 3/8 stops after the first draw and never brings in the spinner at all. Writing down 1/2 does the opposite, using only the spinner and ignoring the ticket draw. And writing down 5/16 uses 5/8, the probability of a RED ticket, instead of 3/8 for blue — reading the wrong colour off the raffle.
- (b) 500 — Method: when a dice is known to be fair, the theoretical probability is the best thing to work from, and the more trials there are the closer the results tend to it. Working: for a fair dice the probability of a six is 1/6, so the expected number of sixes in 3000 rolls is 3000 × 1 ÷ 6 = 500. The class experiment gave a relative frequency of 14/60, but 60 trials is far too few to overturn a known theoretical value, and the school's 3000 rolls will tend towards 1/6 in any case. Answer: about 500 sixes. The distractors: 700 comes from using the class relative frequency instead of the theory, 3000 × 14 ÷ 60 = 700; 600 comes from splitting the difference between the two, since 1/6 is about 0.167 and 14/60 is about 0.233, whose mean is 0.2, and 3000 × 0.2 = 600; 2500 uses 5/6 instead of 1/6 and counts the rolls expected not to be a six.
- (a) 60 — Method: because the counter goes back each time, every draw is the same experiment, so the expected number of reds is the number of draws multiplied by P(red). Working: there are 8 red counters out of 20, so P(red) = 8 ÷ 20 = 0.4. Over 150 draws the expected number of reds is 150 × 0.4 = 60. Answer: about 60 red counters would be expected. The distractors: 90 uses P(blue) by mistake, 150 × 12 ÷ 20 = 90; 100 comes from writing the probability from the red to blue ratio of 8 to 12, giving 150 × 8 ÷ 12 = 100; 75 comes from treating red and blue as equally likely because there are two colours, which gives 150 ÷ 2 = 75.
- (b) 39 — Method: put the counts into a two-way table and fill each missing cell by subtracting along a row or down a column. Working: the number of female members is 150 − 80 = 70. The pool column holds 66 members and 35 of them are male, so the number of female pool users is 66 − 35 = 31. Subtracting along the female row, 70 − 31 = 39 female members use the gym. Answer: 39 female members use the gym. The distractors: 45 comes from subtracting along the male row instead, 80 − 35 = 45, which counts male gym users; 31 is the female pool cell, written down one step before the gym cell; 84 is 150 − 66 and counts every gym user, male and female together.
- (d) No, because 3/8 + 5/12 + 1/6 = 23/24 — Using a common denominator of 24: 3/8 = 9/24, 5/12 = 10/24 and 1/6 = 4/24. Adding these numerators gives 9 + 10 + 4 = 23, so the three probabilities sum to 23/24, which is less than 1 — Zara is not correct. Adding the original numerators (3 + 5 + 1 = 9) over a denominator of 12 instead of converting each fraction properly gives 9/12 = 3/4, still less than 1 but the wrong fraction. Converting 1/6 to 5/24 instead of 4/24 (using the wrong scaling) makes the total 9/24 + 10/24 + 5/24 = 24/24 = 1, wrongly suggesting the probabilities are valid. Judging validity from the fact that each individual fraction lies between 0 and 1 ignores that an exhaustive set must sum to exactly 1, not merely contain valid individual values.
- (b) 24/125 — On the fiction branch, 160 − 112 = 48 books were returned late. None of the non-fiction books were late, so the total number of late books is 48, out of 250 books altogether: 48/250 = 24/125. Writing 3/10 is wrong because it divides the 48 late fiction books by the fiction total (160) instead of the whole library total (250). Writing 56/125 is wrong because 112/250 simplifies to 56/125, and 112 is the number of fiction books returned ON TIME, not late. Writing 9/25 is wrong because 90/250 simplifies to 9/25, and 90 is simply the number of non-fiction books, which has nothing to do with late returns. The probability is 24/125.
- (a) 0.48 — There are two ways to score exactly one throw: scoring on the first and missing the second, 0.6 × 0.4 = 0.24, or missing the first and scoring the second, 0.4 × 0.6 = 0.24. Adding these gives 0.24 + 0.24 = 0.48. Choosing 0.24 comes from working out only one of the two paths and forgetting the other one also gives exactly one score. Choosing 0.36 comes from working out the probability of scoring BOTH throws, 0.6 × 0.6 = 0.36, instead of exactly one. Choosing 0.84 comes from working out the probability of scoring AT LEAST one throw, 1 − 0.4 × 0.4 = 0.84, instead of exactly one.
- (a) 753 — The estimate from 2000 spins is the most reliable, since it comes from the largest sample size, so the best estimate of the probability is 0.251. Over a further 3000 spins, the expected number landing on green is 3000 × 0.251 = 753. Writing 1050 is wrong because 3000 × 0.350 = 1050 uses the estimate from only 20 spins, the LEAST reliable of the three. Writing 870 is wrong because 3000 × 0.290 = 870 uses the estimate from 200 spins rather than the more reliable 2000-spin estimate. Writing 750 is wrong because 3000 × 0.25 = 750 ignores the recorded data completely and simply assumes each of the 4 colours is equally likely. The best estimate is 753 expected green spins.
- (b) 4 — With 150 rolls and probability 1/6 for each number, the expected count is 150 ÷ 6 = 25. Comparing each actual count with 25: 1 is 22 (3 below), 2 is 27 (2 above), 3 is 24 (1 below), 4 is 34 (9 above), 5 is 21 (4 below) and 6 is 22 (3 below). Number 4 is furthest above its expected count, so it is the most over-represented. Number 2 is also above its expected count, but by only 2, far less than 4's 9. Number 3's count of 24 is below the expected 25, so it is under-represented, not over. Number 6's count of 22 is also below the expected 25, so it too is under-represented.
- (a) 75% — 'Percentage of the women' restricts the group to the 80 women, of whom 60 attend yoga: 60/80 = 0.75 = 75%. Dividing by the number of men (200 − 80 = 120) instead of the number of women gives 60/120 = 0.5 = 50%. Dividing by all 200 members instead of just the 80 women gives 60/200 = 0.3 = 30%. Using the 20 women who do NOT attend yoga (80 − 60) as the numerator instead of the 60 who do gives 20/80 = 0.25 = 25%.
- (c) 0.37 — Method: pool the two runs into one combined set of results, then find the relative frequency of red across all of the spins together. Working: total reds = 16 + 21 = 37. Total spins = 40 + 60 = 100. Relative frequency = 37 ÷ 100 = 0.37. Answer: 0.37. Watch out: writing down 0.40 uses only the first run, 16 ÷ 40, and throws away the extra evidence from the second 60 spins. Writing down 0.35 uses only the second run, 21 ÷ 60, and throws away the first run instead. And writing down 0.375 averages the two runs' separate rates, (0.40 + 0.35) ÷ 2, which treats a run of 40 spins and a run of 60 spins as equally weighted, when pooling the actual counts gives the larger run its fair share of influence.
- (b) 40% — 32 out of the 80 people chose Coffee, so the percentage is (32 ÷ 80) × 100 = 40%. Choosing 32% comes from treating 32 as if it were already a percentage out of 100, instead of dividing by the actual total of 80 people surveyed. Choosing 25% comes from using the Tea row instead of the Coffee row, (20 ÷ 80) × 100 = 25%. Choosing 20% comes from using the Juice row instead of the Coffee row, (16 ÷ 80) × 100 = 20%.
- (b) 8,100 — First find the total number of alerts sent in the month: 1,500 × 30 = 45,000. Then apply the probability of a 'STOP' reply: 45,000 × 0.18 = 8,100. Stopping after finding only one day's expected replies, 1,500 × 0.18 = 270, forgets to scale up to the whole month. Multiplying the number of days by the probability instead of by the daily total of alerts gives 30 × 0.18 = 5.4, which rounds to 5. Shifting the decimal point in the probability, using 0.018 instead of 0.18, gives 45,000 × 0.018 = 810.
- (a) 0.30 — Red, blue, green and yellow are exhaustive, so all four probabilities sum to 1: 0.24 + 0.16 + x + x = 1, so 2x + 0.40 = 1, giving 2x = 0.60 and x = 0.30. Stopping at 2x = 0.60 without dividing by 2 leaves 0.60, the combined probability of both blue and green together, not the value of x on its own. Sharing the 0.60 across all four colours instead of just the two unknown ones gives 0.60 ÷ 4 = 0.15. Leaving out the 0.16 for yellow gives 2x + 0.24 = 1, so 2x = 0.76 and x = 0.38.
Build your own mix at the worksheet builder.