Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Answer key: Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- (c) 500 — Method: the expected number of successes is the number of trials multiplied by the probability, so to find the number of trials, divide the expected number by the probability. Working: let n be the number of seeds planted. Then n multiplied by 0.6 must come to 300, so n = 300 ÷ 0.6 = 500. Answer: the grower should plant 500 seeds. The distractors: 180 comes from multiplying instead of dividing, 300 × 0.6 = 180, which answers how many of 300 seeds would germinate; 750 comes from dividing by the probability of not germinating, 300 ÷ 0.4 = 750; 120 comes from multiplying by that same 0.4, 300 × 0.4 = 120.
- (b) 37/75 — There are 150 − 84 = 66 men. 84 − 50 = 34 women prefer weight training, and 40 men prefer weight training, so 34 + 40 = 74 people in total prefer weight training, out of 150: 74/150 = 37/75. Writing 4/15 is wrong because it only counts the men who prefer weight training (40/150, simplified), leaving out the 34 women. Writing 17/75 is wrong because it only counts the women who prefer weight training (34/150, simplified), leaving out the 40 men. Writing 37/42 is wrong because it uses the number of women (84) as the denominator instead of the whole gym (150) — 74/84 simplifies to 37/42, but that is not a probability out of the whole group. The probability is 37/75.
- (b) 100 — The sample shows a proportion of 8/60 = 2/15 faulty. Apply that proportion to the new batch of 750: 750 × 2/15 = 100. Flipping the ratio, calculating 8/750 × 60 instead of 8/60 × 750, gives 0.64, which rounds to about 1. Assuming the same number of faulty items applies to the new batch, without scaling for its larger size, just repeats the sample's count of 8. Rounding the proportion 8/60 = 0.1333... down to 0.1 before multiplying gives 750 × 0.1 = 75.
- (b) 500 — Method: when a dice is known to be fair, the theoretical probability is the best thing to work from, and the more trials there are the closer the results tend to it. Working: for a fair dice the probability of a six is 1/6, so the expected number of sixes in 3000 rolls is 3000 × 1 ÷ 6 = 500. The class experiment gave a relative frequency of 14/60, but 60 trials is far too few to overturn a known theoretical value, and the school's 3000 rolls will tend towards 1/6 in any case. Answer: about 500 sixes. The distractors: 700 comes from using the class relative frequency instead of the theory, 3000 × 14 ÷ 60 = 700; 600 comes from splitting the difference between the two, since 1/6 is about 0.167 and 14/60 is about 0.233, whose mean is 0.2, and 3000 × 0.2 = 600; 2500 uses 5/6 instead of 1/6 and counts the rolls expected not to be a six.
- (b) 3/10 — The number who use at least one app is 90 − 20 = 70. Since 55 + 42 double-counts the overlap, n(X ∩ Y) = 55 + 42 − 70 = 27, so P(both) = 27/90 = 3/10. Forgetting to subtract the 20 who use neither, and using the full 90 as the union, gives 55 + 42 − 90 = 7, so 7/90. Reporting the probability of using X or Y (or both), 70/90 = 7/9, answers a different question about the union, not the overlap. Reporting the probability of using neither app, 20/90 = 2/9, is the complement of the union, not the intersection.
- (c) 100 — There are 3 even numbers on a fair dice (2, 4 and 6), so the probability of landing on an even number is 3/6 = 1/2, and 300 × 1/2 = 150. The probability of landing on a six is 1/6, so 300 × 1/6 = 50. The dice is expected to land on an even number 150 − 50 = 100 more times than on a six. Writing 50 is wrong because that is just the expected number of sixes on its own, without comparing it to the expected number of evens. Writing 150 is wrong because that is just the expected number of evens on its own, without subtracting the sixes. Writing 200 is wrong because it adds the two expected frequencies together (150 + 50 = 200) instead of finding the difference between them. The dice is expected to land on an even number 100 more times than on a six.
- (c) 325 — Method: first find the relative frequency of NOT landing on red from the 200 spins, then scale that up to 500 spins. Working: non-red results = 50 + 80 = 130, out of 200 spins, so P(not red) = 130 ÷ 200 = 0.65. Expected non-red results in 500 spins = 500 × 0.65 = 325. Answer: 325. Watch out: writing down 175 finds the expected number of RED results instead, 70 ÷ 200 × 500 = 175, answering the opposite of what was asked. Writing down 250 assumes landing red or not landing red must be a fair 50-50 split, but the spinner is biased and the actual results do not split evenly. And writing down 130 stops after finding how many of the 200 spins were non-red and forgets to scale that figure up to the 500 spins asked for.
- (c) 9/16 — There are 180 students in total and 84 are in Year 11, so Year 10 has 180 − 84 = 96 students. Of those 96, 42 travel by bus, so 96 − 42 = 54 walk. P(Year 10 student walks) = 54/96 = 9/16. Using the whole school of 180 as the denominator instead of just the 96 Year 10 students gives 54/180 = 3/10. Using the bus count, 42, as if it were the number who walk gives 42/96 = 7/16, the wrong branch of the Year 10 row. Working out the probability for Year 11 instead of Year 10 — 46 walkers out of 84 — gives 46/84 = 23/42.
- (d) 56 — Method: count the trials over the whole period first, then multiply the number of trials by the probability. Working: 4 weeks is 4 × 7 = 28 days, and at 25 trains a day that is 25 × 28 = 700 trains. The expected number of late trains is 700 × 0.08 = 56. Answer: about 56 late trains over the 4 weeks. The distractors: 2 is the expected number for a single day, 25 × 0.08 = 2, with the 28 days never brought in; 14 uses one week instead of four, 25 × 7 × 0.08 = 14; 644 is 700 − 56 and counts the trains expected to be on time.
- (b) 4 — With 150 rolls and probability 1/6 for each number, the expected count is 150 ÷ 6 = 25. Comparing each actual count with 25: 1 is 22 (3 below), 2 is 27 (2 above), 3 is 24 (1 below), 4 is 34 (9 above), 5 is 21 (4 below) and 6 is 22 (3 below). Number 4 is furthest above its expected count, so it is the most over-represented. Number 2 is also above its expected count, but by only 2, far less than 4's 9. Number 3's count of 24 is below the expected 25, so it is under-represented, not over. Number 6's count of 22 is also below the expected 25, so it too is under-represented.
- (b) 12/19 — Method: turn the percentages into expected frequencies out of 1000, total the faulty pens, then divide supplier Y's faulty pens by that total, because the pen picked is known to be faulty. Working: supplier X provided 700 pens and 2% of them are faulty, which is 14 pens. Supplier Y provided 300 pens and 8% of them are faulty, which is 24 pens. Altogether 38 pens are faulty, so the probability is 24/38, and dividing the numerator and the denominator by 2 gives 12/19. Answer: the probability is 12/19. The distractors: 3/10 is supplier Y's share of the stock, the answer before the faulty information is used at all; 3/125 is 24/1000, the probability that a pen is from supplier Y and faulty, which stops at the joint probability and never divides by the probability of a fault; 4/5 is 8 divided by 2 + 8, comparing the two fault rates as though the suppliers provided equal numbers of pens, so the 70 to 30 split is thrown away.
- (c) 15/23 — Method: find P(rough and delayed) and the overall P(delayed) using the tree, then divide. Working: P(rough and delayed) = 0.2 × 0.75 = 0.15. P(calm and delayed) = 0.8 × 0.1 = 0.08. P(delayed) = 0.15 + 0.08 = 0.23. P(rough | delayed) = 0.15 ÷ 0.23 = 15/23. Answer: 15/23. Watch out: leaving the answer as 0.15 (3/20) gives P(rough and delayed) itself, without dividing by the overall probability that a crossing is delayed. Giving 0.75 (3/4) is the probability you were told to start with — that a crossing is delayed GIVEN the sea is rough — which is the reverse of what's being asked. And 0.2 (1/5) is just the original probability that the sea is rough, before you take the fact that the crossing was delayed into account.
- (a) 70 — Method: two linked steps. Find the expected number of billing calls first, then take the 35% of those, because the 35% is quoted for billing calls only. Working: 40% of 500 is 200 billing calls. 35% of 200 is 70 calls. Answer: you would expect 70 calls. The distractors: 200 stops after the first step and gives the billing calls, forgetting that only some of them are dealt with quickly; 175 is 35% of 500, applying the quick response rate to every call the centre takes rather than to the billing calls only; 375 comes from adding 40% and 35% to get 75% and taking 75% of 500, which treats two stages of one journey as separate outcomes to be added.
- (b) 0.225 — The relative frequency of rain is the number of rainy days out of all days recorded: 9 ÷ 40 = 0.225, which is noticeably less than the forecaster's claimed 0.3. Using the number of dry days, 40 − 9 = 31, as the denominator instead of the total of 40 gives 9 ÷ 31 = 0.29 (2 d.p.). Simply reporting the forecaster's claimed value, 0.3, without calculating anything from the data at all, ignores the recorded results completely. Misplacing the decimal point, treating 9 out of 40 as 9%, gives 0.09 instead of 0.225.
- (d) 136 — 15% of 160 = 0.15 × 160 = 24 patients reported side effects, so 160 − 24 = 136 did not. Stopping after finding the number who reported side effects, 24, answers the wrong question — it is not the number who did NOT report them. Misreading '15%' as a raw count of 15 patients, rather than a percentage, gives 160 − 15 = 145. Subtracting 15% of 160 twice, 160 − 24 − 24 = 112, double-counts the side-effect group.
Build your own mix at the worksheet builder.