Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- 1.A council surveyed 140 households about recycling. 85 of the households are in Zone A. 52 of the Zone A households recycle glass. 21 of the Zone B households do not recycle glass. Work out the probability that a household, chosen at random from the 140, recycles glass. Give your answer as a fraction in its simplest form.
- 2.A machine makes 4000 light bulbs a day and runs 5 days a week. Two inspectors test bulbs from this machine. Inspector A tests 40 bulbs and finds 4 faulty. Inspector B tests 500 bulbs and finds 30 faulty. Using the better of the two estimates, work out how many faulty bulbs the machine is expected to make in one week.
- 3.In a survey of 200 households, 120 have a garden and 80 own a dog. 54 of the households have a garden and own a dog. Work out the probability that a household owns a dog given that it has a garden, and compare it with the probability that a household picked from the whole survey owns a dog.
- 4.A two-way table records 200 members of a gym. 80 of the members are women and the rest are men. 60 of the women attend yoga classes. Work out the percentage of the women who attend yoga classes.
- 5.A phone network sends automatic text alerts to customers. On average, 1,500 alerts are sent each day, and the probability that a customer replies 'STOP' to an alert is 0.18. Work out how many replies of 'STOP' the network should expect over a 30-day month.
- 6.A factory checked 400 items. A frequency tree splits them into 250 items made by machine A and 150 items made by machine B. On machine A's branch, 15 of the items were faulty. On machine B's branch, 5 of the items were faulty. One of the 400 items is picked at random. Work out the probability that it is faulty. Give your answer as a fraction in its simplest form.
- 7.A market stall sells umbrellas. Over the last 250 days, it rained on 70 of them. Using this as an estimate of the probability of rain, work out how many rainy days would be expected in the next 365 days.
- 8.A quality inspector examines a sample of 60 items from a production line and finds that 8 are faulty. Using this sample's proportion, work out how many faulty items would be expected in a new batch of 750 items.
- 9.A frequency tree records the results of 160 patients who took a new medicine. It splits them into those who reported side effects and those who did not. 15% of the patients reported side effects. Work out how many of the 160 patients did not report side effects.
- 10.A survey of 90 people records which of two apps, X and Y, they use. 55 people use app X, 42 people use app Y, and 20 people use neither app. Work out the probability that a randomly chosen person uses both apps, giving your answer as a fraction in its simplest form.
- 11.150 people at a gym were asked whether they prefer weight training or cardio; each person chose exactly one. 84 of the people are women. 50 of the women prefer cardio. 40 of the men prefer weight training. Work out the probability that a person, chosen at random from the 150, prefers weight training. Give your answer as a fraction in its simplest form.
- 12.A box contains 9 red balls and 11 green balls. Two balls are taken out at random, one after the other, without being replaced. Given that both balls taken out are the same colour, work out the probability that both balls are red.
- 13.A bag is known to contain red, blue and green counters in equal numbers, so the theoretical probability of taking each colour is 1/3. In 90 trials with replacement, red was taken 38 times, blue was taken 26 times and green was taken 26 times. Which colour is over-represented compared with its theoretical probability?
- 14.At a fun run, a raffle stall charges £1.50 per ticket. The probability that any one ticket wins a prize worth £8 is 0.12, and a losing ticket wins nothing. Nadia buys 25 tickets. Work out how much money Nadia should expect to lose in total.
- 15.At a school fête, a tombola stall costs £1.50 to play. The probability of winning is 0.2, and the prize is worth £6. Work out the stall's expected profit, on average, from each game played.
Answer key
- (a) 43/70 — There are 140 − 85 = 55 Zone B households, and 55 − 21 = 34 of them recycle glass. In total, 52 + 34 = 86 households recycle glass, out of 140: 86/140 = 43/70. Writing 13/35 is wrong because 52/140 simplifies to 13/35, and 52 only counts Zone A, leaving out the 34 Zone B recyclers. Writing 17/70 is wrong because 34/140 simplifies to 17/70, and 34 only counts Zone B, leaving out the 52 Zone A recyclers. Writing 73/140 is wrong because it adds the 52 Zone A recyclers to the 21 Zone B households that do NOT recycle, mixing up two different groups instead of adding the two recycling groups. The probability is 43/70.
- (a) 1200 — Method: take the estimate from the larger sample, because an unbiased relative frequency tends towards the true probability as the sample grows, then multiply by the number of bulbs made in a week. Working: Inspector B tested 500 bulbs, far more than Inspector A's 40, so use B's relative frequency: 30 ÷ 500 = 0.06. A week's production is 4000 × 5 = 20000 bulbs. The expected number of faulty bulbs is 20000 × 0.06 = 1200. Answer: about 1200 faulty bulbs a week. The distractors: 2000 uses Inspector A's estimate, 4 ÷ 40 = 0.1, giving 20000 × 0.1 = 2000, and so rests on a sample of only 40 bulbs; 1600 comes from averaging the two estimates of 0.1 and 0.06 to get 0.08, and 20000 × 0.08 = 1600, which gives the small sample equal weight with the large one; 240 uses the right estimate but stops at a single day, 4000 × 0.06 = 240.
- (a) 0.45, different from 0.4 for all the households — Method: work out the probability inside the restricted group of garden owners, then work out the probability across the whole survey, and compare the two. Working: 54 of the 120 households with a garden own a dog, so the conditional probability is 54 divided by 120, which is 0.45. Across the whole survey 80 of the 200 households own a dog, which is 0.4. Since 0.45 is not 0.4, having a garden changes the chance of owning a dog and the two events are not independent. Answer: 0.45, different from 0.4 for all the households. The distractors: 0.27 is 54/200, dividing the households with both by the whole survey instead of by the 120 with a garden; 0.675 is 54/80, the probability that a household has a garden given that it owns a dog, which is the condition and the event the wrong way round; 0.4 is 80/200, the probability of owning a dog with the garden information never used, which is why that route also reports no difference.
- (a) 75% — 'Percentage of the women' restricts the group to the 80 women, of whom 60 attend yoga: 60/80 = 0.75 = 75%. Dividing by the number of men (200 − 80 = 120) instead of the number of women gives 60/120 = 0.5 = 50%. Dividing by all 200 members instead of just the 80 women gives 60/200 = 0.3 = 30%. Using the 20 women who do NOT attend yoga (80 − 60) as the numerator instead of the 60 who do gives 20/80 = 0.25 = 25%.
- (b) 8,100 — First find the total number of alerts sent in the month: 1,500 × 30 = 45,000. Then apply the probability of a 'STOP' reply: 45,000 × 0.18 = 8,100. Stopping after finding only one day's expected replies, 1,500 × 0.18 = 270, forgets to scale up to the whole month. Multiplying the number of days by the probability instead of by the daily total of alerts gives 30 × 0.18 = 5.4, which rounds to 5. Shifting the decimal point in the probability, using 0.018 instead of 0.18, gives 45,000 × 0.018 = 810.
- (d) 1/20 — Method: add the counts on the faulty end branches, divide by the total number of items in the experiment, then cancel. Working: the faulty items number 15 + 5 = 20, and 400 items were checked, so the probability is 20/400. Dividing the top and the bottom by 20 gives 1/20. Answer: the probability is 1/20. The distractors: 3/50 is 15/250 and comes from dividing machine A's faults by machine A's output, which is that machine's own fault rate rather than the probability for the whole batch; 1/30 is 5/150 and does the same on machine B's branch; 19/20 is 380/400 and gives the probability that the item picked is not faulty.
- (c) 102 — Method: turn the past record into a relative frequency, then use it as an estimate of the probability of rain and multiply by the number of days being predicted for. Working: relative frequency of rain = 70 ÷ 250 = 0.28. Expected rainy days in 365 days = 365 × 0.28 = 102.2, which rounds to about 102 days. Answer: about 102 days. Watch out: writing down 48 swaps which number is the sample and which is the target, working out 70 ÷ 365 × 250 instead of 70 ÷ 250 × 365. Writing down 70 just repeats the original count of rainy days without scaling it up to the new, longer period at all. And writing down 110 comes from rounding the relative frequency to 0.3 before multiplying, 365 × 0.3 = 109.5, when 70 ÷ 250 is exactly 0.28 and needs no rounding at all.
- (b) 100 — The sample shows a proportion of 8/60 = 2/15 faulty. Apply that proportion to the new batch of 750: 750 × 2/15 = 100. Flipping the ratio, calculating 8/750 × 60 instead of 8/60 × 750, gives 0.64, which rounds to about 1. Assuming the same number of faulty items applies to the new batch, without scaling for its larger size, just repeats the sample's count of 8. Rounding the proportion 8/60 = 0.1333... down to 0.1 before multiplying gives 750 × 0.1 = 75.
- (d) 136 — 15% of 160 = 0.15 × 160 = 24 patients reported side effects, so 160 − 24 = 136 did not. Stopping after finding the number who reported side effects, 24, answers the wrong question — it is not the number who did NOT report them. Misreading '15%' as a raw count of 15 patients, rather than a percentage, gives 160 − 15 = 145. Subtracting 15% of 160 twice, 160 − 24 − 24 = 112, double-counts the side-effect group.
- (b) 3/10 — The number who use at least one app is 90 − 20 = 70. Since 55 + 42 double-counts the overlap, n(X ∩ Y) = 55 + 42 − 70 = 27, so P(both) = 27/90 = 3/10. Forgetting to subtract the 20 who use neither, and using the full 90 as the union, gives 55 + 42 − 90 = 7, so 7/90. Reporting the probability of using X or Y (or both), 70/90 = 7/9, answers a different question about the union, not the overlap. Reporting the probability of using neither app, 20/90 = 2/9, is the complement of the union, not the intersection.
- (b) 37/75 — There are 150 − 84 = 66 men. 84 − 50 = 34 women prefer weight training, and 40 men prefer weight training, so 34 + 40 = 74 people in total prefer weight training, out of 150: 74/150 = 37/75. Writing 4/15 is wrong because it only counts the men who prefer weight training (40/150, simplified), leaving out the 34 women. Writing 17/75 is wrong because it only counts the women who prefer weight training (34/150, simplified), leaving out the 40 men. Writing 37/42 is wrong because it uses the number of women (84) as the denominator instead of the whole gym (150) — 74/84 simplifies to 37/42, but that is not a probability out of the whole group. The probability is 37/75.
- (d) 36/91 — Method: P(both red | same colour) = P(both red) ÷ P(same colour), where P(same colour) = P(both red) + P(both green). Working: P(both red) = 9/20 × 8/19 = 72/380 = 18/95. P(both green) = 11/20 × 10/19 = 110/380 = 11/38. P(same colour) = 18/95 + 11/38 = 36/190 + 55/190 = 91/190. P(both red | same colour) = (36/190) ÷ (91/190) = 36/91. Answer: 36/91. Watch out: stopping at 18/95 gives P(both red) itself, without dividing by the probability that the colours matched at all. Working out 55/91 finds the same-colour probability for green instead of red — check which colour's count you are putting on top. And 9/20 is just the chance the first ball drawn is red, which ignores the second draw and the without-replacement condition completely.
- (c) Red — Theoretical probability is 1/3 ≈ 0.333 for each colour. Red's relative frequency is 38/90 ≈ 0.422, above 1/3, so red is over-represented. Blue's relative frequency is 26/90 ≈ 0.289, below 1/3, so blue is under-represented, not over. Green's relative frequency is also 26/90 ≈ 0.289, below 1/3 for the same reason. Since red's relative frequency clearly exceeds 1/3, it is not true that none of the colours are over-represented.
- (b) £13.50 — The total cost of Nadia's 25 tickets is 25 × £1.50 = £37.50. The expected number of winning tickets is 25 × 0.12 = 3, so the expected prize money is 3 × £8 = £24.00. Nadia's expected loss is the cost minus the expected prize money: £37.50 − £24.00 = £13.50. A candidate who answers £24.00 has given the expected prize money and mistaken it for the loss. A candidate who answers £37.50 has given the total cost of the tickets, forgetting to subtract the expected prize money. A candidate who answers £34.50 has subtracted the expected number of wins, 3, from the cost instead of first converting it to prize money by multiplying by £8.
- (b) £0.30 profit for the stall — The stall keeps the £1.50 entry fee whatever happens, and expects to pay out prize × probability of winning = £6 × 0.2 = £1.20 on average. So its expected profit per game is £1.50 − £1.20 = £0.30. Reporting the expected pay-out of £1.20 itself as the profit forgets that the stall also keeps the entry fee. Assuming the player always wins gives an expected cost of £6 − £1.50 = £4.50, treated as a loss for the stall. Using the probability of NOT winning, 0.8, to find the expected pay-out gives £6 × 0.8 = £4.80, and £1.50 − £4.80 = −£3.30, a £3.30 loss.
Build your own mix at the worksheet builder.