Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Answer key: Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- (b) 40 — n(P ∪ Q) = n(P) + n(Q) − n(P ∩ Q) = 34 + 27 − 11 = 50. The complement is everyone outside both sets: n((P ∪ Q)′) = 90 − 50 = 40. Adding P and Q without subtracting the overlap gives 34 + 27 = 61, so 90 − 61 = 29 double-subtracts the 11 who are in both. Reporting n(P ∪ Q) itself, 50, forgets to take the complement at all. Subtracting only n(P) from the universal set, 90 − 34 = 56, ignores set Q altogether.
- (a) 3/8 — There are 4 × 6 = 24 equally likely outcomes. The outcomes with no 3 at all have Spinner E showing 1, 2 or 4 and Spinner F showing 1, 2, 4, 5 or 6, giving 3 × 5 = 15 outcomes. So the outcomes with at least one 3 are 24 − 15 = 9, and the probability is 9/24 = 3/8. Choosing 5/12 comes from adding the two individual probabilities of a 3, 1/4 + 1/6, which counts the outcome where both spinners show 3 twice over. Choosing 1/4 comes from only counting the case where Spinner E shows 3, and forgetting the outcomes where Spinner F shows 3 instead. Choosing 5/8 comes from working out the probability of getting no 3 at all, 15/24 = 5/8, and giving that as the final answer instead of subtracting it from 1.
- (a) 9/25 — Method: P(badminton | tennis) = n(tennis and badminton) ÷ n(tennis) — restrict to the tennis-players, then find what fraction of them also play badminton. Working: n(tennis and badminton) = 18, n(tennis) = 50, so P(badminton | tennis) = 18/50 = 9/25. Answer: 9/25. Watch out: dividing by 40 (the badminton total) finds P(tennis | badminton) instead of P(badminton | tennis) — the wrong direction. Dividing by 90 (all the members named in the question) ignores that you already know the member plays tennis. And dividing by 72 (50 + 40 − 18, the number who play at least one of the two sports) answers a question about the union, not the condition you were given.
- (d) 5 — Being red and not being red are exhaustive, so their probabilities sum to 1: the probability of red is 1 − 0.8 = 0.2. The number of red counters is 0.2 × 25 = 5. Using 0.8 directly as the probability of red, without taking the complement, gives 0.8 × 25 = 20 — the number of counters that are NOT red. Sharing the 25 counters equally between the three colours, ignoring the given probability altogether, gives 25 ÷ 3 ≈ 8. Misreading the total as 20 counters instead of 25 gives 0.2 × 20 = 4.
- (d) 3400 — Combining all three greenhouses gives 200 + 150 + 250 = 600 seeds planted in total, and 172 + 126 + 212 = 510 germinated, so the combined estimate of the germination probability is 510/600 = 0.85. Out of a new batch of 4000 seeds, the expected number to germinate is 4000 × 0.85 = 3400. Writing 3440 is wrong because it uses only Greenhouse 1's rate, 172/200 = 0.86, instead of the combined rate from all three: 4000 × 0.86 = 3440. Writing 3360 is wrong because it uses only Greenhouse 2's rate, 126/150 = 0.84: 4000 × 0.84 = 3360. Writing 510 is wrong because that is the total number that germinated in the ORIGINAL trial, not scaled up to the new batch of 4000 seeds at all. The best estimate is 3400 seeds.
- (b) 50 — There are 180 − 100 = 80 south-plot gardeners. 30 of them do not grow organically, so the rest do: 80 − 30 = 50. Writing 64 is wrong because that is the number of NORTH-plot gardeners who grow organically, not south. Writing 30 again is wrong because that is the number of south-plot gardeners who do NOT grow organically — the question asks for those who do. Writing 80 is wrong because that is the whole south-plot total, without subtracting the 30 who do not grow organically. The answer is 50.
- (a) 2/3 — Method: 'in A or in B' means every score that belongs to at least one of the two sets; a score that belongs to both is still only one outcome, so it is listed once. Working: the even scores are 2, 4 and 6; the scores greater than 4 are 5 and 6. Listing the scores that appear in either set gives 2, 4, 5 and 6, with 6 written once. That is 4 of the 6 faces, or 4/6. Answer: the probability is 2/3. The distractors: 5/6 comes from adding the sizes of the two sets, 3 + 2, so that the score 6 is counted in both and appears twice; 1/2 comes from using the even scores alone; 1/6 comes from giving the probability that the score is in both sets, which is the single score 6, rather than in either of them.
- (b) 3 times as likely — Method: to say how many times as likely one event is as another, divide the larger probability by the smaller one; subtracting them gives the gap between the two probabilities, not the multiple. Working: both probabilities are counted in tenths, so 6/10 ÷ 2/10 compares 6 tenths with 2 tenths, and 6 ÷ 2 = 3. Answer: winning at the hoopla stall is 3 times as likely, which is why 6/10 sits three times as far along the 0 to 1 scale as 2/10. The distractors: 4 times as likely comes from subtracting the two counts, 6 − 2, instead of dividing them, which measures the gap rather than the multiple; 6 times as likely comes from reading the larger probability's 6 tenths straight off as the multiple without ever comparing it with the 2 tenths at the other stall; 12 times as likely comes from multiplying the two counts, 6 × 2, instead of dividing one by the other.
- (b) 0.0309 — Method: P(both defective | at least one defective) = P(both defective) ÷ P(at least one defective). Find each using independence: P(both) = 0.06², P(at least one) = 1 − P(neither) = 1 − 0.94². Working: P(both) = 0.06² = 0.0036. P(neither) = 0.94² = 0.8836, so P(at least one) = 1 − 0.8836 = 0.1164. P(both | at least one) = 0.0036 ÷ 0.1164 = 0.0309 (3 s.f.). Answer: 0.0309. Watch out: leaving the answer as 0.0036 gives P(both defective) itself, not the probability once you already know at least one is defective — you still need to divide by P(at least one defective). Giving 0.0600 answers with the single-component defect rate, ignoring the condition altogether. And 0.5000 assumes that 'at least one' makes the outcomes 'exactly one defective' and 'both defective' equally likely, which is not how these probabilities combine.
- (a) 1200 — Method: take the estimate from the larger sample, because an unbiased relative frequency tends towards the true probability as the sample grows, then multiply by the number of bulbs made in a week. Working: Inspector B tested 500 bulbs, far more than Inspector A's 40, so use B's relative frequency: 30 ÷ 500 = 0.06. A week's production is 4000 × 5 = 20000 bulbs. The expected number of faulty bulbs is 20000 × 0.06 = 1200. Answer: about 1200 faulty bulbs a week. The distractors: 2000 uses Inspector A's estimate, 4 ÷ 40 = 0.1, giving 20000 × 0.1 = 2000, and so rests on a sample of only 40 bulbs; 1600 comes from averaging the two estimates of 0.1 and 0.06 to get 0.08, and 20000 × 0.08 = 1600, which gives the small sample equal weight with the large one; 240 uses the right estimate but stops at a single day, 4000 × 0.06 = 240.
- (b) 12/25 — List every outcome as an ordered pair (first dice, second dice) out of the 25 equally likely outcomes. The pairs with a difference of exactly 2 are (1, 3), (3, 1), (2, 4), (4, 2), (3, 5) and (5, 3), the pairs with a difference of 3 are (1, 4), (4, 1), (2, 5) and (5, 2), and the pairs with a difference of 4 are (1, 5) and (5, 1), giving 6 + 4 + 2 = 12 outcomes with a difference of at least 2, so the probability is 12/25. Choosing 13/25 comes from finding the probability that the two scores differ by LESS than 2 instead, using the remaining 13 outcomes, the opposite of what was asked. Choosing 10/25 comes from forgetting the two outcomes where the difference is 4, (1, 5) and (5, 1), and adding only the difference-2 and difference-3 outcomes, 6 + 4 = 10 out of 25. Choosing 6/25 comes from counting only the pairs with a difference of exactly 2, forgetting that differences of 3 and 4 also count as at least 2.
- (a) 70 — Method: two linked steps. Find the expected number of billing calls first, then take the 35% of those, because the 35% is quoted for billing calls only. Working: 40% of 500 is 200 billing calls. 35% of 200 is 70 calls. Answer: you would expect 70 calls. The distractors: 200 stops after the first step and gives the billing calls, forgetting that only some of them are dealt with quickly; 175 is 35% of 500, applying the quick response rate to every call the centre takes rather than to the billing calls only; 375 comes from adding 40% and 35% to get 75% and taking 75% of 500, which treats two stages of one journey as separate outcomes to be added.
- (c) No — 7/30 is the relative frequency; theory stays 1/6. — The theoretical probability of rolling a 6 on an ordinary dice is fixed at 1/6, worked out from the number of equally likely outcomes, and does not change however the dice is actually rolled. The relative frequency from this trial is 7/30, found from what happened in these particular 30 rolls. Since 7/30 and 1/6 are different numbers, the correct statement is 'No — 7/30 is the relative frequency; theory stays 1/6.' Assuming the two values must always match because they describe the same event gives 'Yes — relative frequency always equals theory.' Believing that an observed result redefines the theoretical probability gives 'Yes — the theoretical probability has now become 7/30.' Refusing to work out either value at all gives 'Neither can be found — 30 rolls is too few to tell', which ignores that both numbers CAN be calculated from the information given.
- (c) 325 — Method: first find the relative frequency of NOT landing on red from the 200 spins, then scale that up to 500 spins. Working: non-red results = 50 + 80 = 130, out of 200 spins, so P(not red) = 130 ÷ 200 = 0.65. Expected non-red results in 500 spins = 500 × 0.65 = 325. Answer: 325. Watch out: writing down 175 finds the expected number of RED results instead, 70 ÷ 200 × 500 = 175, answering the opposite of what was asked. Writing down 250 assumes landing red or not landing red must be a fair 50-50 split, but the spinner is biased and the actual results do not split evenly. And writing down 130 stops after finding how many of the 200 spins were non-red and forgets to scale that figure up to the 500 spins asked for.
- (c) 80 — Since 180 calls are 0.75 of all the technical support calls, the technical support total is 180 ÷ 0.75 = 240. The billing calls make up the rest of the 320 calls, so 320 − 240 = 80. Choosing 240 comes from stopping after finding the technical support total and forgetting the question asks for the billing calls, which are the rest. Choosing 185 comes from multiplying 180 × 0.75 = 135 instead of dividing, then working out 320 − 135 = 185. Choosing 140 comes from using 180 directly as the whole technical support total, ignoring the probability altogether, then working out 320 − 180 = 140.
Build your own mix at the worksheet builder.