Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- 1.A two-way table records 200 members of a gym. 80 of the members are women and the rest are men. 60 of the women attend yoga classes. Work out the percentage of the women who attend yoga classes.
- 2.Dice A is a fair six-sided dice. Dice B is biased so that P(6) = 0.3. Dice A is rolled 150 times and Dice B is rolled 150 times. Work out how many more sixes you would expect from Dice B than from Dice A.
- 3.An archer hits the bullseye with probability 0.24 on any one shot. She wants to know how many shots she must take to expect to hit the bullseye 12 times. Work out how many shots this is.
- 4.A fair six-sided dice is rolled 150 times. The table shows how many times each number came up: 1 came up 22 times, 2 came up 27 times, 3 came up 24 times, 4 came up 34 times, 5 came up 21 times and 6 came up 22 times. The theoretical probability of each number is 1/6. Which number is most over-represented compared with its theoretical probability?
- 5.A leisure centre has 150 members. 80 of the members are male and the rest are female. Every member uses either the pool or the gym, but not both. 66 members use the pool, and 35 of those pool users are male. Work out how many female members use the gym.
- 6.A two-way table records how 180 students at a school travel: by bus or on foot, split by year group. There are 84 students in Year 11, of whom 38 travel by bus and the rest walk. The rest of the 180 students are in Year 10, and 42 of the Year 10 students travel by bus. Work out the probability that a randomly chosen Year 10 student walks to school. Give your answer as a fraction in its simplest form.
- 7.A two-way table records 150 customers at a café. 90 of the customers bought a hot drink and the rest did not. 54 of the hot-drink customers also bought a cake. 21 of the customers who did not buy a hot drink bought a cake. Work out the percentage of all 150 customers who bought a cake.
- 8.Two goalkeepers face penalty kicks. Elin saves 34% of the penalties she faces. Noah saves 11/32 of the penalties he faces. Which statement correctly compares them?
- 9.A drawing pin is dropped many times and lands either point up or point down. The relative frequency of landing point up is recorded as the experiment goes on: after 50 drops it is 0.720, after 200 drops it is 0.665, and after 1000 drops it is 0.638. The pin is to be dropped a further 2000 times. Work out the best estimate of the number of times it will land point up.
- 10.A council surveyed 140 households about recycling. 85 of the households are in Zone A. 52 of the Zone A households recycle glass. 21 of the Zone B households do not recycle glass. Work out the probability that a household, chosen at random from the 140, recycles glass. Give your answer as a fraction in its simplest form.
- 11.A market stall sells umbrellas. Over the last 250 days, it rained on 70 of them. Using this as an estimate of the probability of rain, work out how many rainy days would be expected in the next 365 days.
- 12.A test for a medical condition is given to 1000 people. 50 of the people have the condition and 950 do not. The test is positive for 45 of the 50 people who have the condition, and it is also positive for 95 of the 950 people who do not have the condition. One of the people whose test is positive is picked at random. Work out the probability that this person has the condition.
- 13.A grower knows that the probability that one of their seeds germinates is 0.6. The grower wants to expect 300 of the seeds to germinate. Work out how many seeds the grower should plant.
- 14.A fair coin is flipped again and again. After the first 10 flips there have been 7 heads. After 1000 flips there have been 528 heads. Which statement best describes what these results show?
- 15.A library recorded loans for 250 books over a week, using a frequency tree. The first branch splits the books into 160 fiction and 90 non-fiction. Of the fiction books, 112 were returned on time and the rest were returned late. All the non-fiction books were returned on time. Work out the probability that a book, chosen at random from the 250, was returned late. Give your answer as a fraction in its simplest form.
Answer key
- (a) 75% — 'Percentage of the women' restricts the group to the 80 women, of whom 60 attend yoga: 60/80 = 0.75 = 75%. Dividing by the number of men (200 − 80 = 120) instead of the number of women gives 60/120 = 0.5 = 50%. Dividing by all 200 members instead of just the 80 women gives 60/200 = 0.3 = 30%. Using the 20 women who do NOT attend yoga (80 − 60) as the numerator instead of the 60 who do gives 20/80 = 0.25 = 25%.
- (a) 20 — Dice A is fair, so its expected number of sixes is 150 × 1/6 = 25. Dice B has P(6) = 0.3, so its expected number of sixes is 150 × 0.3 = 45. The difference is 45 − 25 = 20. Adding the two expected values instead of subtracting them gives 25 + 45 = 70. Reporting Dice B's expected sixes on their own, without comparing to Dice A, gives 45. Using the fair probability 1/6 for Dice B as well as Dice A ignores the bias altogether, giving 150 × 1/6 = 25 for both dice and a difference of 0.
- (a) 50 — To find the number of shots needed for an expected 12 hits, divide the number of hits wanted by the probability of a hit: 12 ÷ 0.24 = 50. Multiplying the number of hits by the probability instead of dividing gives 12 × 0.24 = 2.88, which rounds to 3 shots. Rounding 0.24 to 0.25 before dividing gives 12 ÷ 0.25 = 48. Using the probability of missing, 1 − 0.24 = 0.76, instead of the probability of hitting, gives 12 ÷ 0.76 = 15.79, which rounds to 16.
- (b) 4 — With 150 rolls and probability 1/6 for each number, the expected count is 150 ÷ 6 = 25. Comparing each actual count with 25: 1 is 22 (3 below), 2 is 27 (2 above), 3 is 24 (1 below), 4 is 34 (9 above), 5 is 21 (4 below) and 6 is 22 (3 below). Number 4 is furthest above its expected count, so it is the most over-represented. Number 2 is also above its expected count, but by only 2, far less than 4's 9. Number 3's count of 24 is below the expected 25, so it is under-represented, not over. Number 6's count of 22 is also below the expected 25, so it too is under-represented.
- (b) 39 — Method: put the counts into a two-way table and fill each missing cell by subtracting along a row or down a column. Working: the number of female members is 150 − 80 = 70. The pool column holds 66 members and 35 of them are male, so the number of female pool users is 66 − 35 = 31. Subtracting along the female row, 70 − 31 = 39 female members use the gym. Answer: 39 female members use the gym. The distractors: 45 comes from subtracting along the male row instead, 80 − 35 = 45, which counts male gym users; 31 is the female pool cell, written down one step before the gym cell; 84 is 150 − 66 and counts every gym user, male and female together.
- (c) 9/16 — There are 180 students in total and 84 are in Year 11, so Year 10 has 180 − 84 = 96 students. Of those 96, 42 travel by bus, so 96 − 42 = 54 walk. P(Year 10 student walks) = 54/96 = 9/16. Using the whole school of 180 as the denominator instead of just the 96 Year 10 students gives 54/180 = 3/10. Using the bus count, 42, as if it were the number who walk gives 42/96 = 7/16, the wrong branch of the Year 10 row. Working out the probability for Year 11 instead of Year 10 — 46 walkers out of 84 — gives 46/84 = 23/42.
- (b) 50% — The total number of customers who bought a cake is 54 + 21 = 75, combining both hot-drink and non-hot-drink customers. As a percentage of all 150 customers, this is (75 ÷ 150) × 100 = 50%. Choosing 36% comes from only counting the hot-drink customers who bought a cake, (54 ÷ 150) × 100 = 36%, and forgetting the 21 non-hot-drink customers who also bought a cake. Choosing 14% comes from only counting the non-hot-drink customers who bought a cake, (21 ÷ 150) × 100 = 14%, and forgetting the 54 hot-drink customers who also bought a cake. Choosing 60% comes from dividing by the hot-drink total of 90 instead of the grand total of 150, (54 ÷ 90) × 100 = 60%.
- (c) Noah — 11/32 = 0.34375, above 34%. — Converting 11/32 to a decimal gives 11 ÷ 32 = 0.34375, which is greater than 34% (0.34), so Noah has the better save rate: 'Noah — 11/32 = 0.34375, above 34%.' Comparing the raw numbers 34 and 11 directly, without converting the fraction to the same form, gives 'Elin — 34 is bigger than 11.' Treating a larger denominator as meaning a bigger value, rather than smaller equal shares, gives 'Noah — 32 is a bigger denominator.' Rounding 34.375% to 34% to the nearest whole percent hides the difference and gives 'Equal — both round to 34% to the nearest percent.'
- (a) 1276 — Method: an unbiased relative frequency tends towards the theoretical probability as the number of trials increases, so use the record resting on the most trials, then multiply by the number of new trials. Working: the three records rest on 50, 200 and 1000 drops, so the most reliable is the one after 1000 drops, namely 0.638, and the run is indeed settling as the trials increase. The expected number of point up landings in 2000 further drops is 2000 × 0.638 = 1276. Answer: about 1276 times. The distractors: 1440 uses the earliest record, which rests on only 50 drops, giving 2000 × 0.720 = 1440; 1330 uses the middle record, treating 200 drops as a safe compromise when 1000 drops is better still, giving 2000 × 0.665 = 1330; 1348 comes from averaging the three records, since 0.720 + 0.665 + 0.638 = 2.023 and 2.023 ÷ 3 = 0.674, then 2000 × 0.674 = 1348, which gives the 50 drop record the same weight as the 1000 drop record.
- (a) 43/70 — There are 140 − 85 = 55 Zone B households, and 55 − 21 = 34 of them recycle glass. In total, 52 + 34 = 86 households recycle glass, out of 140: 86/140 = 43/70. Writing 13/35 is wrong because 52/140 simplifies to 13/35, and 52 only counts Zone A, leaving out the 34 Zone B recyclers. Writing 17/70 is wrong because 34/140 simplifies to 17/70, and 34 only counts Zone B, leaving out the 52 Zone A recyclers. Writing 73/140 is wrong because it adds the 52 Zone A recyclers to the 21 Zone B households that do NOT recycle, mixing up two different groups instead of adding the two recycling groups. The probability is 43/70.
- (c) 102 — Method: turn the past record into a relative frequency, then use it as an estimate of the probability of rain and multiply by the number of days being predicted for. Working: relative frequency of rain = 70 ÷ 250 = 0.28. Expected rainy days in 365 days = 365 × 0.28 = 102.2, which rounds to about 102 days. Answer: about 102 days. Watch out: writing down 48 swaps which number is the sample and which is the target, working out 70 ÷ 365 × 250 instead of 70 ÷ 250 × 365. Writing down 70 just repeats the original count of rainy days without scaling it up to the new, longer period at all. And writing down 110 comes from rounding the relative frequency to 0.3 before multiplying, 365 × 0.3 = 109.5, when 70 ÷ 250 is exactly 0.28 and needs no rounding at all.
- (c) 9/28 — Method: two linked steps. Total everyone whose test is positive, since the person picked is known to be one of them, then divide the positive tests that belong to people with the condition by that total. Working: 45 positive tests come from people who have the condition and 95 come from people who do not, so 140 tests are positive. The people with the condition give 45/140, and dividing the numerator and the denominator by 5 gives 9/28. Answer: the probability is 9/28. The distractors: 9/10 is 45/50, the probability of a positive test given that the person has the condition, which is the condition and the event the wrong way round and is the figure a candidate quotes when the two are confused; 9/200 is 45/1000, dividing by everyone tested rather than by the 140 who tested positive; 1/20 is 50/1000, the probability that a person has the condition before the test result is used at all.
- (c) 500 — Method: the expected number of successes is the number of trials multiplied by the probability, so to find the number of trials, divide the expected number by the probability. Working: let n be the number of seeds planted. Then n multiplied by 0.6 must come to 300, so n = 300 ÷ 0.6 = 500. Answer: the grower should plant 500 seeds. The distractors: 180 comes from multiplying instead of dividing, 300 × 0.6 = 180, which answers how many of 300 seeds would germinate; 750 comes from dividing by the probability of not germinating, 300 ÷ 0.4 = 750; 120 comes from multiplying by that same 0.4, 300 × 0.4 = 120.
- (b) The relative frequency is settling near 0.5 — Method: turn each result into a relative frequency before comparing them, because it is the relative frequency, and not the difference between the two counts, that tends towards the theoretical probability. Working: after 10 flips the relative frequency of a head is 7 ÷ 10 = 0.7, which is a long way from 0.5. After 1000 flips it is 528 ÷ 1000 = 0.528, which is much closer to 0.5. Meanwhile the gap between the two counts has grown rather than shrunk: it was 7 − 3 = 4 after 10 flips and is 528 − 472 = 56 after 1000 flips. Answer: the relative frequency is settling near 0.5, which is what an unbiased experiment does as the sample grows. The distractors: saying the counts are levelling out is the usual form of this idea and the figures contradict it, since the gap went from 4 to 56; saying the coin is biased treats 28 extra heads in 1000 flips as proof, when 0.528 sits close to 0.5 and a fair coin gives results like this often; saying the next flip is more likely to be a tail is the gambler's fallacy, since each flip stays at 1/2 whatever came before.
- (b) 24/125 — On the fiction branch, 160 − 112 = 48 books were returned late. None of the non-fiction books were late, so the total number of late books is 48, out of 250 books altogether: 48/250 = 24/125. Writing 3/10 is wrong because it divides the 48 late fiction books by the fiction total (160) instead of the whole library total (250). Writing 56/125 is wrong because 112/250 simplifies to 56/125, and 112 is the number of fiction books returned ON TIME, not late. Writing 9/25 is wrong because 90/250 simplifies to 9/25, and 90 is simply the number of non-fiction books, which has nothing to do with late returns. The probability is 24/125.
Build your own mix at the worksheet builder.