Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- 1.A library recorded loans for 250 books over a week, using a frequency tree. The first branch splits the books into 160 fiction and 90 non-fiction. Of the fiction books, 112 were returned on time and the rest were returned late. All the non-fiction books were returned on time. Work out the probability that a book, chosen at random from the 250, was returned late. Give your answer as a fraction in its simplest form.
- 2.A school buys pens from two suppliers and has 1000 pens in stock. Supplier X provided 70% of the pens and supplier Y provided the other 30%. 2% of supplier X's pens are faulty and 8% of supplier Y's pens are faulty. A pen picked at random from the stock is found to be faulty. Work out the probability that it came from supplier Y. Give your answer as a fraction.
- 3.Two ordinary fair dice are rolled and the two scores are multiplied together. Work out the probability that the product is odd.
- 4.A biased spinner is spun 200 times. It lands on red 70 times, on blue 50 times, and on green 80 times. Using these results, work out the expected number of times the spinner does NOT land on red, in 500 spins of the same spinner.
- 5.A survey found that, of 250 shoppers questioned, 68% said they had used a self-checkout in the past month. A different store expects 1400 shoppers this week. Using this relative frequency, work out how many of the 1400 shoppers would be expected to have used a self-checkout in the past month.
- 6.A frequency tree records 320 calls to a helpline. The calls first split into calls about billing and calls about technical support. Of the technical support calls, the probability that a call was resolved on the first contact is 0.75. If 180 technical support calls were resolved on the first contact, work out how many calls were about billing.
- 7.A weather station records whether it rains each day for 40 days: it rains on 9 of the days and does not rain on the rest. A local forecaster claims that the probability of rain on any day is 0.3. Work out the relative frequency of rain from the recorded data.
- 8.A phone network sends automatic text alerts to customers. On average, 1,500 alerts are sent each day, and the probability that a customer replies 'STOP' to an alert is 0.18. Work out how many replies of 'STOP' the network should expect over a 30-day month.
- 9.A fair coin is flipped again and again. After the first 10 flips there have been 7 heads. After 1000 flips there have been 528 heads. Which statement best describes what these results show?
- 10.A box contains 9 red balls and 11 green balls. Two balls are taken out at random, one after the other, without being replaced. Given that both balls taken out are the same colour, work out the probability that both balls are red.
- 11.A ferry company finds that on 20% of days the sea is rough. If the sea is rough, the probability that a crossing is delayed is 0.75. If the sea is calm, the probability that a crossing is delayed is 0.1. Given that a crossing was delayed, work out the probability that the sea was rough that day.
- 12.A factory checked 400 items. A frequency tree splits them into 250 items made by machine A and 150 items made by machine B. On machine A's branch, 15 of the items were faulty. On machine B's branch, 5 of the items were faulty. One of the 400 items is picked at random. Work out the probability that it is faulty. Give your answer as a fraction in its simplest form.
- 13.At a school car park, each car is a saloon, an estate or a hatchback, and no car is more than one of these. Among 200 cars, the probability that a car chosen at random is a saloon is 0.28 and the probability that it is an estate is 0.37. Work out how many of the 200 cars are hatchbacks.
- 14.A grower knows that the probability that one of their seeds germinates is 0.6. The grower wants to expect 300 of the seeds to germinate. Work out how many seeds the grower should plant.
- 15.A clinic recorded 300 booked appointments using a frequency tree. The first branch splits them into 210 adult appointments and the rest child appointments. Of the adult appointments, 189 were attended and the rest were missed. Of the child appointments, 81 were attended. Work out the probability that a booked appointment, chosen at random from the 300, was missed. Give your answer as a fraction in its simplest form.
Answer key
- (b) 24/125 — On the fiction branch, 160 − 112 = 48 books were returned late. None of the non-fiction books were late, so the total number of late books is 48, out of 250 books altogether: 48/250 = 24/125. Writing 3/10 is wrong because it divides the 48 late fiction books by the fiction total (160) instead of the whole library total (250). Writing 56/125 is wrong because 112/250 simplifies to 56/125, and 112 is the number of fiction books returned ON TIME, not late. Writing 9/25 is wrong because 90/250 simplifies to 9/25, and 90 is simply the number of non-fiction books, which has nothing to do with late returns. The probability is 24/125.
- (b) 12/19 — Method: turn the percentages into expected frequencies out of 1000, total the faulty pens, then divide supplier Y's faulty pens by that total, because the pen picked is known to be faulty. Working: supplier X provided 700 pens and 2% of them are faulty, which is 14 pens. Supplier Y provided 300 pens and 8% of them are faulty, which is 24 pens. Altogether 38 pens are faulty, so the probability is 24/38, and dividing the numerator and the denominator by 2 gives 12/19. Answer: the probability is 12/19. The distractors: 3/10 is supplier Y's share of the stock, the answer before the faulty information is used at all; 3/125 is 24/1000, the probability that a pen is from supplier Y and faulty, which stops at the joint probability and never divides by the probability of a fault; 4/5 is 8 divided by 2 + 8, comparing the two fault rates as though the suppliers provided equal numbers of pens, so the 70 to 30 split is thrown away.
- (d) 1/4 — Method: a product is odd only when BOTH factors are odd, so list the ordered pairs where both scores are odd and divide by 36. Working: the odd scores on a dice are 1, 3 and 5, so there are 3 × 3 = 9 ordered pairs where both scores are odd, out of the 36 equally likely pairs, cancelling down to 1/4. Answer: 1/4. Watch out: writing down 3/4 finds the probability that AT LEAST ONE score is odd, 1 minus the probability both are even, which is a different, easier condition to meet than both being odd. Considering only the first dice's score and ignoring the second gives 1/2, since 3 of the first dice's 6 scores are odd — but the product also depends on what the second dice shows. And writing down 1/12 comes from counting only the pairs where the SAME odd number appears twice, (1, 1), (3, 3) and (5, 5), missing pairs like (1, 3) and (5, 1) where the two odd scores differ.
- (c) 325 — Method: first find the relative frequency of NOT landing on red from the 200 spins, then scale that up to 500 spins. Working: non-red results = 50 + 80 = 130, out of 200 spins, so P(not red) = 130 ÷ 200 = 0.65. Expected non-red results in 500 spins = 500 × 0.65 = 325. Answer: 325. Watch out: writing down 175 finds the expected number of RED results instead, 70 ÷ 200 × 500 = 175, answering the opposite of what was asked. Writing down 250 assumes landing red or not landing red must be a fair 50-50 split, but the spinner is biased and the actual results do not split evenly. And writing down 130 stops after finding how many of the 200 spins were non-red and forgets to scale that figure up to the 500 spins asked for.
- (a) 952 — 68% = 0.68. The relative frequency from the survey applies to the new group of 1400 shoppers, so the expected number is 0.68 × 1400 = 952. Working out 1 − 0.68 = 0.32 and applying that instead, 0.32 × 1400 = 448, finds the number who have NOT used a self-checkout, not the number who have. Applying 68% to the original sample size of 250 instead of the new total of 1400 gives 0.68 × 250 = 170. Applying the complement percentage to the original sample size, 0.32 × 250 = 80, compounds both mistakes.
- (c) 80 — Since 180 calls are 0.75 of all the technical support calls, the technical support total is 180 ÷ 0.75 = 240. The billing calls make up the rest of the 320 calls, so 320 − 240 = 80. Choosing 240 comes from stopping after finding the technical support total and forgetting the question asks for the billing calls, which are the rest. Choosing 185 comes from multiplying 180 × 0.75 = 135 instead of dividing, then working out 320 − 135 = 185. Choosing 140 comes from using 180 directly as the whole technical support total, ignoring the probability altogether, then working out 320 − 180 = 140.
- (b) 0.225 — The relative frequency of rain is the number of rainy days out of all days recorded: 9 ÷ 40 = 0.225, which is noticeably less than the forecaster's claimed 0.3. Using the number of dry days, 40 − 9 = 31, as the denominator instead of the total of 40 gives 9 ÷ 31 = 0.29 (2 d.p.). Simply reporting the forecaster's claimed value, 0.3, without calculating anything from the data at all, ignores the recorded results completely. Misplacing the decimal point, treating 9 out of 40 as 9%, gives 0.09 instead of 0.225.
- (b) 8,100 — First find the total number of alerts sent in the month: 1,500 × 30 = 45,000. Then apply the probability of a 'STOP' reply: 45,000 × 0.18 = 8,100. Stopping after finding only one day's expected replies, 1,500 × 0.18 = 270, forgets to scale up to the whole month. Multiplying the number of days by the probability instead of by the daily total of alerts gives 30 × 0.18 = 5.4, which rounds to 5. Shifting the decimal point in the probability, using 0.018 instead of 0.18, gives 45,000 × 0.018 = 810.
- (b) The relative frequency is settling near 0.5 — Method: turn each result into a relative frequency before comparing them, because it is the relative frequency, and not the difference between the two counts, that tends towards the theoretical probability. Working: after 10 flips the relative frequency of a head is 7 ÷ 10 = 0.7, which is a long way from 0.5. After 1000 flips it is 528 ÷ 1000 = 0.528, which is much closer to 0.5. Meanwhile the gap between the two counts has grown rather than shrunk: it was 7 − 3 = 4 after 10 flips and is 528 − 472 = 56 after 1000 flips. Answer: the relative frequency is settling near 0.5, which is what an unbiased experiment does as the sample grows. The distractors: saying the counts are levelling out is the usual form of this idea and the figures contradict it, since the gap went from 4 to 56; saying the coin is biased treats 28 extra heads in 1000 flips as proof, when 0.528 sits close to 0.5 and a fair coin gives results like this often; saying the next flip is more likely to be a tail is the gambler's fallacy, since each flip stays at 1/2 whatever came before.
- (d) 36/91 — Method: P(both red | same colour) = P(both red) ÷ P(same colour), where P(same colour) = P(both red) + P(both green). Working: P(both red) = 9/20 × 8/19 = 72/380 = 18/95. P(both green) = 11/20 × 10/19 = 110/380 = 11/38. P(same colour) = 18/95 + 11/38 = 36/190 + 55/190 = 91/190. P(both red | same colour) = (36/190) ÷ (91/190) = 36/91. Answer: 36/91. Watch out: stopping at 18/95 gives P(both red) itself, without dividing by the probability that the colours matched at all. Working out 55/91 finds the same-colour probability for green instead of red — check which colour's count you are putting on top. And 9/20 is just the chance the first ball drawn is red, which ignores the second draw and the without-replacement condition completely.
- (c) 15/23 — Method: find P(rough and delayed) and the overall P(delayed) using the tree, then divide. Working: P(rough and delayed) = 0.2 × 0.75 = 0.15. P(calm and delayed) = 0.8 × 0.1 = 0.08. P(delayed) = 0.15 + 0.08 = 0.23. P(rough | delayed) = 0.15 ÷ 0.23 = 15/23. Answer: 15/23. Watch out: leaving the answer as 0.15 (3/20) gives P(rough and delayed) itself, without dividing by the overall probability that a crossing is delayed. Giving 0.75 (3/4) is the probability you were told to start with — that a crossing is delayed GIVEN the sea is rough — which is the reverse of what's being asked. And 0.2 (1/5) is just the original probability that the sea is rough, before you take the fact that the crossing was delayed into account.
- (d) 1/20 — Method: add the counts on the faulty end branches, divide by the total number of items in the experiment, then cancel. Working: the faulty items number 15 + 5 = 20, and 400 items were checked, so the probability is 20/400. Dividing the top and the bottom by 20 gives 1/20. Answer: the probability is 1/20. The distractors: 3/50 is 15/250 and comes from dividing machine A's faults by machine A's output, which is that machine's own fault rate rather than the probability for the whole batch; 1/30 is 5/150 and does the same on machine B's branch; 19/20 is 380/400 and gives the probability that the item picked is not faulty.
- (b) 70 — Saloon, estate and hatchback are exhaustive, so their probabilities sum to 1: the probability of a hatchback is 1 − 0.28 − 0.37 = 0.35. The number of hatchbacks is 0.35 × 200 = 70. Treating the SUM of the other two probabilities, 0.28 + 0.37 = 0.65, as the probability of a hatchback instead of its complement gives 0.65 × 200 = 130. Multiplying the correct probability, 0.35, by 100 instead of the 200 cars actually surveyed gives 35. Averaging the two given probabilities, (0.28 + 0.37) ÷ 2 = 0.325, instead of subtracting them from 1, and then multiplying by 200 gives 65.
- (c) 500 — Method: the expected number of successes is the number of trials multiplied by the probability, so to find the number of trials, divide the expected number by the probability. Working: let n be the number of seeds planted. Then n multiplied by 0.6 must come to 300, so n = 300 ÷ 0.6 = 500. Answer: the grower should plant 500 seeds. The distractors: 180 comes from multiplying instead of dividing, 300 × 0.6 = 180, which answers how many of 300 seeds would germinate; 750 comes from dividing by the probability of not germinating, 300 ÷ 0.4 = 750; 120 comes from multiplying by that same 0.4, 300 × 0.4 = 120.
- (c) 1/10 — On the adult branch, 210 − 189 = 21 appointments were missed. There are 300 − 210 = 90 child appointments, and 90 − 81 = 9 of those were missed. In total, 21 + 9 = 30 appointments were missed, out of 300: 30/300 = 1/10. Writing 7/100 is wrong because 21/300 simplifies to 7/100, and 21 only counts the adult branch, leaving out the 9 missed child appointments. Writing 3/100 is wrong because 9/300 simplifies to 3/100, and 9 only counts the child branch, leaving out the 21 missed adult appointments. Writing 1/9 is wrong because it divides the 30 missed appointments by the 270 that were attended (300 − 30) instead of by the whole 300 booked. The probability is 1/10.
Build your own mix at the worksheet builder.