Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- 1.A Venn diagram shows two sets, P and Q, inside a universal set. n(P) = 34, n(Q) = 27, n(P ∩ Q) = 11, and n(ξ) = 90, where ξ is the universal set. Work out n((P ∪ Q)′), the number of elements in neither P nor Q.
- 2.Spinner E has 4 equal sections, numbered 1 to 4. Spinner F has 6 equal sections, numbered 1 to 6. Both spinners are spun once. Work out the probability of getting at least one 3.
- 3.At a sports club, the numbers of members who play tennis and badminton are: 50 members play tennis, 40 members play badminton, and 18 members play both tennis and badminton. A member who plays tennis is chosen at random. Work out the probability that this member also plays badminton.
- 4.A bag contains counters that are red, blue or green only. The probability that a counter taken at random is not red is 0.8. The bag contains 25 counters in total. Work out how many of the counters are red.
- 5.A seed company tests germination using results from three greenhouses. Greenhouse 1 plants 200 seeds and 172 germinate. Greenhouse 2 plants 150 seeds and 126 germinate. Greenhouse 3 plants 250 seeds and 212 germinate. Using the combined results from all three greenhouses, work out the best estimate of the number of seeds, out of a new batch of 4000 seeds, that would be expected to germinate.
- 6.180 gardeners at an allotment were asked whether they grow vegetables organically. 100 of the gardeners are on the north plots. 64 of the north-plot gardeners grow organically. 30 of the south-plot gardeners do not grow organically. Work out how many south-plot gardeners grow organically.
- 7.An ordinary fair dice is rolled once. Set A is the set of even scores and set B is the set of scores greater than 4. Work out the probability that the score rolled is in set A or in set B or in both.
- 8.At a summer fair, the probability of winning at the hoopla stall is 6/10 and the probability of winning at the coconut shy is 2/10. Work out how many times as likely a player is to win at the hoopla stall as at the coconut shy.
- 9.A factory tests components from a large batch in which 6% are defective. Two components are selected at random, and the batch is large enough that the selections can be treated as independent. Given that at least one of the two components is defective, work out the probability that both are defective.
- 10.A machine makes 4000 light bulbs a day and runs 5 days a week. Two inspectors test bulbs from this machine. Inspector A tests 40 bulbs and finds 4 faulty. Inspector B tests 500 bulbs and finds 30 faulty. Using the better of the two estimates, work out how many faulty bulbs the machine is expected to make in one week.
- 11.Two fair five-sided dice, numbered 1 to 5, are rolled. Work out the probability that the two scores differ by at least 2.
- 12.A call centre finds that 40% of its calls are about billing. 35% of the calls about billing are dealt with in under five minutes. The centre takes 500 calls on Monday. Work out how many of Monday's calls you would expect to be about billing and dealt with in under five minutes.
- 13.An ordinary six-sided dice, numbered 1 to 6, is rolled 30 times and lands on a 6 seven times. Ravi says the theoretical probability of rolling a 6 and the relative frequency of rolling a 6 in this trial are the same number. Is Ravi right?
- 14.A biased spinner is spun 200 times. It lands on red 70 times, on blue 50 times, and on green 80 times. Using these results, work out the expected number of times the spinner does NOT land on red, in 500 spins of the same spinner.
- 15.A frequency tree records 320 calls to a helpline. The calls first split into calls about billing and calls about technical support. Of the technical support calls, the probability that a call was resolved on the first contact is 0.75. If 180 technical support calls were resolved on the first contact, work out how many calls were about billing.
Answer key
- (b) 40 — n(P ∪ Q) = n(P) + n(Q) − n(P ∩ Q) = 34 + 27 − 11 = 50. The complement is everyone outside both sets: n((P ∪ Q)′) = 90 − 50 = 40. Adding P and Q without subtracting the overlap gives 34 + 27 = 61, so 90 − 61 = 29 double-subtracts the 11 who are in both. Reporting n(P ∪ Q) itself, 50, forgets to take the complement at all. Subtracting only n(P) from the universal set, 90 − 34 = 56, ignores set Q altogether.
- (a) 3/8 — There are 4 × 6 = 24 equally likely outcomes. The outcomes with no 3 at all have Spinner E showing 1, 2 or 4 and Spinner F showing 1, 2, 4, 5 or 6, giving 3 × 5 = 15 outcomes. So the outcomes with at least one 3 are 24 − 15 = 9, and the probability is 9/24 = 3/8. Choosing 5/12 comes from adding the two individual probabilities of a 3, 1/4 + 1/6, which counts the outcome where both spinners show 3 twice over. Choosing 1/4 comes from only counting the case where Spinner E shows 3, and forgetting the outcomes where Spinner F shows 3 instead. Choosing 5/8 comes from working out the probability of getting no 3 at all, 15/24 = 5/8, and giving that as the final answer instead of subtracting it from 1.
- (a) 9/25 — Method: P(badminton | tennis) = n(tennis and badminton) ÷ n(tennis) — restrict to the tennis-players, then find what fraction of them also play badminton. Working: n(tennis and badminton) = 18, n(tennis) = 50, so P(badminton | tennis) = 18/50 = 9/25. Answer: 9/25. Watch out: dividing by 40 (the badminton total) finds P(tennis | badminton) instead of P(badminton | tennis) — the wrong direction. Dividing by 90 (all the members named in the question) ignores that you already know the member plays tennis. And dividing by 72 (50 + 40 − 18, the number who play at least one of the two sports) answers a question about the union, not the condition you were given.
- (d) 5 — Being red and not being red are exhaustive, so their probabilities sum to 1: the probability of red is 1 − 0.8 = 0.2. The number of red counters is 0.2 × 25 = 5. Using 0.8 directly as the probability of red, without taking the complement, gives 0.8 × 25 = 20 — the number of counters that are NOT red. Sharing the 25 counters equally between the three colours, ignoring the given probability altogether, gives 25 ÷ 3 ≈ 8. Misreading the total as 20 counters instead of 25 gives 0.2 × 20 = 4.
- (d) 3400 — Combining all three greenhouses gives 200 + 150 + 250 = 600 seeds planted in total, and 172 + 126 + 212 = 510 germinated, so the combined estimate of the germination probability is 510/600 = 0.85. Out of a new batch of 4000 seeds, the expected number to germinate is 4000 × 0.85 = 3400. Writing 3440 is wrong because it uses only Greenhouse 1's rate, 172/200 = 0.86, instead of the combined rate from all three: 4000 × 0.86 = 3440. Writing 3360 is wrong because it uses only Greenhouse 2's rate, 126/150 = 0.84: 4000 × 0.84 = 3360. Writing 510 is wrong because that is the total number that germinated in the ORIGINAL trial, not scaled up to the new batch of 4000 seeds at all. The best estimate is 3400 seeds.
- (b) 50 — There are 180 − 100 = 80 south-plot gardeners. 30 of them do not grow organically, so the rest do: 80 − 30 = 50. Writing 64 is wrong because that is the number of NORTH-plot gardeners who grow organically, not south. Writing 30 again is wrong because that is the number of south-plot gardeners who do NOT grow organically — the question asks for those who do. Writing 80 is wrong because that is the whole south-plot total, without subtracting the 30 who do not grow organically. The answer is 50.
- (a) 2/3 — Method: 'in A or in B' means every score that belongs to at least one of the two sets; a score that belongs to both is still only one outcome, so it is listed once. Working: the even scores are 2, 4 and 6; the scores greater than 4 are 5 and 6. Listing the scores that appear in either set gives 2, 4, 5 and 6, with 6 written once. That is 4 of the 6 faces, or 4/6. Answer: the probability is 2/3. The distractors: 5/6 comes from adding the sizes of the two sets, 3 + 2, so that the score 6 is counted in both and appears twice; 1/2 comes from using the even scores alone; 1/6 comes from giving the probability that the score is in both sets, which is the single score 6, rather than in either of them.
- (b) 3 times as likely — Method: to say how many times as likely one event is as another, divide the larger probability by the smaller one; subtracting them gives the gap between the two probabilities, not the multiple. Working: both probabilities are counted in tenths, so 6/10 ÷ 2/10 compares 6 tenths with 2 tenths, and 6 ÷ 2 = 3. Answer: winning at the hoopla stall is 3 times as likely, which is why 6/10 sits three times as far along the 0 to 1 scale as 2/10. The distractors: 4 times as likely comes from subtracting the two counts, 6 − 2, instead of dividing them, which measures the gap rather than the multiple; 6 times as likely comes from reading the larger probability's 6 tenths straight off as the multiple without ever comparing it with the 2 tenths at the other stall; 12 times as likely comes from multiplying the two counts, 6 × 2, instead of dividing one by the other.
- (b) 0.0309 — Method: P(both defective | at least one defective) = P(both defective) ÷ P(at least one defective). Find each using independence: P(both) = 0.06², P(at least one) = 1 − P(neither) = 1 − 0.94². Working: P(both) = 0.06² = 0.0036. P(neither) = 0.94² = 0.8836, so P(at least one) = 1 − 0.8836 = 0.1164. P(both | at least one) = 0.0036 ÷ 0.1164 = 0.0309 (3 s.f.). Answer: 0.0309. Watch out: leaving the answer as 0.0036 gives P(both defective) itself, not the probability once you already know at least one is defective — you still need to divide by P(at least one defective). Giving 0.0600 answers with the single-component defect rate, ignoring the condition altogether. And 0.5000 assumes that 'at least one' makes the outcomes 'exactly one defective' and 'both defective' equally likely, which is not how these probabilities combine.
- (a) 1200 — Method: take the estimate from the larger sample, because an unbiased relative frequency tends towards the true probability as the sample grows, then multiply by the number of bulbs made in a week. Working: Inspector B tested 500 bulbs, far more than Inspector A's 40, so use B's relative frequency: 30 ÷ 500 = 0.06. A week's production is 4000 × 5 = 20000 bulbs. The expected number of faulty bulbs is 20000 × 0.06 = 1200. Answer: about 1200 faulty bulbs a week. The distractors: 2000 uses Inspector A's estimate, 4 ÷ 40 = 0.1, giving 20000 × 0.1 = 2000, and so rests on a sample of only 40 bulbs; 1600 comes from averaging the two estimates of 0.1 and 0.06 to get 0.08, and 20000 × 0.08 = 1600, which gives the small sample equal weight with the large one; 240 uses the right estimate but stops at a single day, 4000 × 0.06 = 240.
- (b) 12/25 — List every outcome as an ordered pair (first dice, second dice) out of the 25 equally likely outcomes. The pairs with a difference of exactly 2 are (1, 3), (3, 1), (2, 4), (4, 2), (3, 5) and (5, 3), the pairs with a difference of 3 are (1, 4), (4, 1), (2, 5) and (5, 2), and the pairs with a difference of 4 are (1, 5) and (5, 1), giving 6 + 4 + 2 = 12 outcomes with a difference of at least 2, so the probability is 12/25. Choosing 13/25 comes from finding the probability that the two scores differ by LESS than 2 instead, using the remaining 13 outcomes, the opposite of what was asked. Choosing 10/25 comes from forgetting the two outcomes where the difference is 4, (1, 5) and (5, 1), and adding only the difference-2 and difference-3 outcomes, 6 + 4 = 10 out of 25. Choosing 6/25 comes from counting only the pairs with a difference of exactly 2, forgetting that differences of 3 and 4 also count as at least 2.
- (a) 70 — Method: two linked steps. Find the expected number of billing calls first, then take the 35% of those, because the 35% is quoted for billing calls only. Working: 40% of 500 is 200 billing calls. 35% of 200 is 70 calls. Answer: you would expect 70 calls. The distractors: 200 stops after the first step and gives the billing calls, forgetting that only some of them are dealt with quickly; 175 is 35% of 500, applying the quick response rate to every call the centre takes rather than to the billing calls only; 375 comes from adding 40% and 35% to get 75% and taking 75% of 500, which treats two stages of one journey as separate outcomes to be added.
- (c) No — 7/30 is the relative frequency; theory stays 1/6. — The theoretical probability of rolling a 6 on an ordinary dice is fixed at 1/6, worked out from the number of equally likely outcomes, and does not change however the dice is actually rolled. The relative frequency from this trial is 7/30, found from what happened in these particular 30 rolls. Since 7/30 and 1/6 are different numbers, the correct statement is 'No — 7/30 is the relative frequency; theory stays 1/6.' Assuming the two values must always match because they describe the same event gives 'Yes — relative frequency always equals theory.' Believing that an observed result redefines the theoretical probability gives 'Yes — the theoretical probability has now become 7/30.' Refusing to work out either value at all gives 'Neither can be found — 30 rolls is too few to tell', which ignores that both numbers CAN be calculated from the information given.
- (c) 325 — Method: first find the relative frequency of NOT landing on red from the 200 spins, then scale that up to 500 spins. Working: non-red results = 50 + 80 = 130, out of 200 spins, so P(not red) = 130 ÷ 200 = 0.65. Expected non-red results in 500 spins = 500 × 0.65 = 325. Answer: 325. Watch out: writing down 175 finds the expected number of RED results instead, 70 ÷ 200 × 500 = 175, answering the opposite of what was asked. Writing down 250 assumes landing red or not landing red must be a fair 50-50 split, but the spinner is biased and the actual results do not split evenly. And writing down 130 stops after finding how many of the 200 spins were non-red and forgets to scale that figure up to the 500 spins asked for.
- (c) 80 — Since 180 calls are 0.75 of all the technical support calls, the technical support total is 180 ÷ 0.75 = 240. The billing calls make up the rest of the 320 calls, so 320 − 240 = 80. Choosing 240 comes from stopping after finding the technical support total and forgetting the question asks for the billing calls, which are the rest. Choosing 185 comes from multiplying 180 × 0.75 = 135 instead of dividing, then working out 320 − 135 = 185. Choosing 140 comes from using 180 directly as the whole technical support total, ignoring the probability altogether, then working out 320 − 180 = 140.
Build your own mix at the worksheet builder.