Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- 1.A charity bake sale tracked 240 items sold using a frequency tree. The first branch splits sales into 150 cake sales and 90 biscuit sales. Of the cake sales, 100 were bought by adults and the rest by children. Of the biscuit sales, 54 were bought by children. Each item sold for £1.20. Work out the total amount of money raised from sales to children.
- 2.A factory checked 400 items. A frequency tree splits them into 250 items made by machine A and 150 items made by machine B. On machine A's branch, 15 of the items were faulty. On machine B's branch, 5 of the items were faulty. One of the 400 items is picked at random. Work out the probability that it is faulty. Give your answer as a fraction in its simplest form.
- 3.A frequency tree records 320 calls to a helpline. The calls first split into calls about billing and calls about technical support. Of the technical support calls, the probability that a call was resolved on the first contact is 0.75. If 180 technical support calls were resolved on the first contact, work out how many calls were about billing.
- 4.A quality inspector examines a sample of 60 items from a production line and finds that 8 are faulty. Using this sample's proportion, work out how many faulty items would be expected in a new batch of 750 items.
- 5.Two games are offered at a school fair, and every one of the 40 pupils in a class plays each game once. In Game A, the probability of winning a prize is 0.2 and the prize is worth £5. In Game B, the probability of winning a prize is 0.1 and the prize is worth £12. Work out which game gives a higher expected total value of prizes for the class, and by how much.
- 6.A quality inspector tests 145 light bulbs and finds that 33 are faulty. Work out the relative frequency of a bulb being faulty, as a percentage correct to 1 decimal place.
- 7.A weather station records whether it rains each day for 40 days: it rains on 9 of the days and does not rain on the rest. A local forecaster claims that the probability of rain on any day is 0.3. Work out the relative frequency of rain from the recorded data.
- 8.At a school fête, a tombola stall costs £1.50 to play. The probability of winning is 0.2, and the prize is worth £6. Work out the stall's expected profit, on average, from each game played.
- 9.A fair six-sided dice is rolled 300 times. Work out how many more times you would expect it to land on an even number than on a six.
- 10.A four-colour spinner (red, blue, green, yellow) is spun repeatedly, and the relative frequency of landing on green is recorded as the number of spins increases: after 20 spins it is 0.350; after 200 spins it is 0.290; after 2000 spins it is 0.251. Using the result from 2000 spins as the best estimate of the probability, work out the number of times the spinner would be expected to land on green in a further 3000 spins.
- 11.A bag is known to contain red, blue and green counters in equal numbers, so the theoretical probability of taking each colour is 1/3. In 90 trials with replacement, red was taken 38 times, blue was taken 26 times and green was taken 26 times. Which colour is over-represented compared with its theoretical probability?
- 12.A Venn diagram shows two sets, P and Q, inside a universal set. n(P) = 34, n(Q) = 27, n(P ∩ Q) = 11, and n(ξ) = 90, where ξ is the universal set. Work out n((P ∪ Q)′), the number of elements in neither P nor Q.
- 13.A box contains 9 red balls and 11 green balls. Two balls are taken out at random, one after the other, without being replaced. Given that both balls taken out are the same colour, work out the probability that both balls are red.
- 14.A charity raffle sells 250 tickets. A group of friends buy 40 of the tickets between them. Three prizes are awarded, each to a different randomly drawn ticket, and no ticket can win more than one prize. Work out the expected number of prizes won by the group. Give your answer as a fraction in its simplest form.
- 15.Dice A is a fair six-sided dice. Dice B is biased so that P(6) = 0.3. Dice A is rolled 150 times and Dice B is rolled 150 times. Work out how many more sixes you would expect from Dice B than from Dice A.
Answer key
- (a) £124.80 — On the cake branch, 150 − 100 = 50 cakes were bought by children. Adding the 54 biscuits bought by children gives 50 + 54 = 104 items sold to children in total, and at £1.20 each that raises 104 × £1.20 = £124.80. Writing £60.00 is wrong because 50 × £1.20 = £60.00 only counts the cake sales to children and leaves out the 54 biscuits. Writing £163.20 is wrong because it uses the ADULT sales instead of children's: 100 cake adults plus 90 − 54 = 36 biscuit adults gives 136 × £1.20 = £163.20. Writing £136.80 is wrong because it finds the cake children's number by subtracting the wrong branch (150 − 90 = 60 instead of 150 − 100 = 50), giving 60 + 54 = 114 items and 114 × £1.20 = £136.80. The total raised from sales to children is £124.80.
- (d) 1/20 — Method: add the counts on the faulty end branches, divide by the total number of items in the experiment, then cancel. Working: the faulty items number 15 + 5 = 20, and 400 items were checked, so the probability is 20/400. Dividing the top and the bottom by 20 gives 1/20. Answer: the probability is 1/20. The distractors: 3/50 is 15/250 and comes from dividing machine A's faults by machine A's output, which is that machine's own fault rate rather than the probability for the whole batch; 1/30 is 5/150 and does the same on machine B's branch; 19/20 is 380/400 and gives the probability that the item picked is not faulty.
- (c) 80 — Since 180 calls are 0.75 of all the technical support calls, the technical support total is 180 ÷ 0.75 = 240. The billing calls make up the rest of the 320 calls, so 320 − 240 = 80. Choosing 240 comes from stopping after finding the technical support total and forgetting the question asks for the billing calls, which are the rest. Choosing 185 comes from multiplying 180 × 0.75 = 135 instead of dividing, then working out 320 − 135 = 185. Choosing 140 comes from using 180 directly as the whole technical support total, ignoring the probability altogether, then working out 320 − 180 = 140.
- (b) 100 — The sample shows a proportion of 8/60 = 2/15 faulty. Apply that proportion to the new batch of 750: 750 × 2/15 = 100. Flipping the ratio, calculating 8/750 × 60 instead of 8/60 × 750, gives 0.64, which rounds to about 1. Assuming the same number of faulty items applies to the new batch, without scaling for its larger size, just repeats the sample's count of 8. Rounding the proportion 8/60 = 0.1333... down to 0.1 before multiplying gives 750 × 0.1 = 75.
- (a) Game B, by £8 — Game A's expected total is 40 × 0.2 × £5 = £40. Game B's expected total is 40 × 0.1 × £12 = £48. Game B is higher, by £48 − £40 = £8. Writing 'Game A, by £8' is wrong because it has the right difference but the wrong game — Game A's total (£40) is actually LOWER than Game B's, not higher. Writing 'Game B, by £48' is wrong because £48 is Game B's whole expected total, not the DIFFERENCE between the two games. Writing 'Game A, by £40' is wrong in the same way, using Game A's whole total as if it were the margin, and naming the wrong game as the winner. Game B gives the higher expected total, by £8.
- (b) 22.8% — Relative frequency as a percentage is the faulty count divided by the total, then multiplied by 100: 33 ÷ 145 × 100 = 22.76, which rounds to 22.8%. Giving 33.0% as the answer uses the frequency, 33, directly as a percentage without dividing by the total 145 at all. Rounding 22.76 down to 22.7% instead of up applies the wrong rounding direction at the first decimal place. Finding the relative frequency of the bulbs that were NOT faulty first: 145 − 33 = 112, and 112 ÷ 145 × 100 = 77.24, answers the opposite question and rounds to 77.2%.
- (b) 0.225 — The relative frequency of rain is the number of rainy days out of all days recorded: 9 ÷ 40 = 0.225, which is noticeably less than the forecaster's claimed 0.3. Using the number of dry days, 40 − 9 = 31, as the denominator instead of the total of 40 gives 9 ÷ 31 = 0.29 (2 d.p.). Simply reporting the forecaster's claimed value, 0.3, without calculating anything from the data at all, ignores the recorded results completely. Misplacing the decimal point, treating 9 out of 40 as 9%, gives 0.09 instead of 0.225.
- (b) £0.30 profit for the stall — The stall keeps the £1.50 entry fee whatever happens, and expects to pay out prize × probability of winning = £6 × 0.2 = £1.20 on average. So its expected profit per game is £1.50 − £1.20 = £0.30. Reporting the expected pay-out of £1.20 itself as the profit forgets that the stall also keeps the entry fee. Assuming the player always wins gives an expected cost of £6 − £1.50 = £4.50, treated as a loss for the stall. Using the probability of NOT winning, 0.8, to find the expected pay-out gives £6 × 0.8 = £4.80, and £1.50 − £4.80 = −£3.30, a £3.30 loss.
- (c) 100 — There are 3 even numbers on a fair dice (2, 4 and 6), so the probability of landing on an even number is 3/6 = 1/2, and 300 × 1/2 = 150. The probability of landing on a six is 1/6, so 300 × 1/6 = 50. The dice is expected to land on an even number 150 − 50 = 100 more times than on a six. Writing 50 is wrong because that is just the expected number of sixes on its own, without comparing it to the expected number of evens. Writing 150 is wrong because that is just the expected number of evens on its own, without subtracting the sixes. Writing 200 is wrong because it adds the two expected frequencies together (150 + 50 = 200) instead of finding the difference between them. The dice is expected to land on an even number 100 more times than on a six.
- (a) 753 — The estimate from 2000 spins is the most reliable, since it comes from the largest sample size, so the best estimate of the probability is 0.251. Over a further 3000 spins, the expected number landing on green is 3000 × 0.251 = 753. Writing 1050 is wrong because 3000 × 0.350 = 1050 uses the estimate from only 20 spins, the LEAST reliable of the three. Writing 870 is wrong because 3000 × 0.290 = 870 uses the estimate from 200 spins rather than the more reliable 2000-spin estimate. Writing 750 is wrong because 3000 × 0.25 = 750 ignores the recorded data completely and simply assumes each of the 4 colours is equally likely. The best estimate is 753 expected green spins.
- (c) Red — Theoretical probability is 1/3 ≈ 0.333 for each colour. Red's relative frequency is 38/90 ≈ 0.422, above 1/3, so red is over-represented. Blue's relative frequency is 26/90 ≈ 0.289, below 1/3, so blue is under-represented, not over. Green's relative frequency is also 26/90 ≈ 0.289, below 1/3 for the same reason. Since red's relative frequency clearly exceeds 1/3, it is not true that none of the colours are over-represented.
- (b) 40 — n(P ∪ Q) = n(P) + n(Q) − n(P ∩ Q) = 34 + 27 − 11 = 50. The complement is everyone outside both sets: n((P ∪ Q)′) = 90 − 50 = 40. Adding P and Q without subtracting the overlap gives 34 + 27 = 61, so 90 − 61 = 29 double-subtracts the 11 who are in both. Reporting n(P ∪ Q) itself, 50, forgets to take the complement at all. Subtracting only n(P) from the universal set, 90 − 34 = 56, ignores set Q altogether.
- (d) 36/91 — Method: P(both red | same colour) = P(both red) ÷ P(same colour), where P(same colour) = P(both red) + P(both green). Working: P(both red) = 9/20 × 8/19 = 72/380 = 18/95. P(both green) = 11/20 × 10/19 = 110/380 = 11/38. P(same colour) = 18/95 + 11/38 = 36/190 + 55/190 = 91/190. P(both red | same colour) = (36/190) ÷ (91/190) = 36/91. Answer: 36/91. Watch out: stopping at 18/95 gives P(both red) itself, without dividing by the probability that the colours matched at all. Working out 55/91 finds the same-colour probability for green instead of red — check which colour's count you are putting on top. And 9/20 is just the chance the first ball drawn is red, which ignores the second draw and the without-replacement condition completely.
- (d) 12/25 — The group holds 40 of the 250 tickets, so for any one prize the probability the group wins it is 40/250 = 4/25. There are 3 prizes and the group has the same chance at each one, so the expected number won is 3 × 4/25 = 12/25. Writing 4/25 is wrong because it is the chance of winning just ONE prize, without multiplying by the 3 prizes available. Writing 4/75 is wrong because it divides by the 3 prizes instead of multiplying (4/25 ÷ 3 = 4/75), which would mean the group did worse the more prizes were on offer. Writing 64/15625 is wrong because it multiplies the single-prize probability by itself three times, (4/25)³, as though all three prizes had to be won together, instead of adding up the expected number across the three separate prizes. The expected number of prizes won by the group is 12/25.
- (a) 20 — Dice A is fair, so its expected number of sixes is 150 × 1/6 = 25. Dice B has P(6) = 0.3, so its expected number of sixes is 150 × 0.3 = 45. The difference is 45 − 25 = 20. Adding the two expected values instead of subtracting them gives 25 + 45 = 70. Reporting Dice B's expected sixes on their own, without comparing to Dice A, gives 45. Using the fair probability 1/6 for Dice B as well as Dice A ignores the bias altogether, giving 150 × 1/6 = 25 for both dice and a difference of 0.
Build your own mix at the worksheet builder.