Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Answer key: Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- (c) 80 — Since 180 calls are 0.75 of all the technical support calls, the technical support total is 180 ÷ 0.75 = 240. The billing calls make up the rest of the 320 calls, so 320 − 240 = 80. Choosing 240 comes from stopping after finding the technical support total and forgetting the question asks for the billing calls, which are the rest. Choosing 185 comes from multiplying 180 × 0.75 = 135 instead of dividing, then working out 320 − 135 = 185. Choosing 140 comes from using 180 directly as the whole technical support total, ignoring the probability altogether, then working out 320 − 180 = 140.
- (b) 41/160 — In total, 27 + 14 = 41 of the 160 employees cycle to work, so the probability is 41/160 (41 and 160 share no common factor, so this is already in its simplest form). Writing 27/160 is wrong because it only counts the full-time cyclists and leaves out the 14 part-time cyclists. Writing 41/90 is wrong because it uses the full-time total (90) as the denominator instead of the whole survey (160). Writing 1/5 is wrong because it only uses the part-time branch, simplifying 14/70 to 1/5 and ignoring the full-time cyclists completely. The probability is 41/160.
- (c) No — 7/30 is the relative frequency; theory stays 1/6. — The theoretical probability of rolling a 6 on an ordinary dice is fixed at 1/6, worked out from the number of equally likely outcomes, and does not change however the dice is actually rolled. The relative frequency from this trial is 7/30, found from what happened in these particular 30 rolls. Since 7/30 and 1/6 are different numbers, the correct statement is 'No — 7/30 is the relative frequency; theory stays 1/6.' Assuming the two values must always match because they describe the same event gives 'Yes — relative frequency always equals theory.' Believing that an observed result redefines the theoretical probability gives 'Yes — the theoretical probability has now become 7/30.' Refusing to work out either value at all gives 'Neither can be found — 30 rolls is too few to tell', which ignores that both numbers CAN be calculated from the information given.
- (b) The relative frequency is settling near 0.5 — Method: turn each result into a relative frequency before comparing them, because it is the relative frequency, and not the difference between the two counts, that tends towards the theoretical probability. Working: after 10 flips the relative frequency of a head is 7 ÷ 10 = 0.7, which is a long way from 0.5. After 1000 flips it is 528 ÷ 1000 = 0.528, which is much closer to 0.5. Meanwhile the gap between the two counts has grown rather than shrunk: it was 7 − 3 = 4 after 10 flips and is 528 − 472 = 56 after 1000 flips. Answer: the relative frequency is settling near 0.5, which is what an unbiased experiment does as the sample grows. The distractors: saying the counts are levelling out is the usual form of this idea and the figures contradict it, since the gap went from 4 to 56; saying the coin is biased treats 28 extra heads in 1000 flips as proof, when 0.528 sits close to 0.5 and a fair coin gives results like this often; saying the next flip is more likely to be a tail is the gambler's fallacy, since each flip stays at 1/2 whatever came before.
- (a) 753 — The estimate from 2000 spins is the most reliable, since it comes from the largest sample size, so the best estimate of the probability is 0.251. Over a further 3000 spins, the expected number landing on green is 3000 × 0.251 = 753. Writing 1050 is wrong because 3000 × 0.350 = 1050 uses the estimate from only 20 spins, the LEAST reliable of the three. Writing 870 is wrong because 3000 × 0.290 = 870 uses the estimate from 200 spins rather than the more reliable 2000-spin estimate. Writing 750 is wrong because 3000 × 0.25 = 750 ignores the recorded data completely and simply assumes each of the 4 colours is equally likely. The best estimate is 753 expected green spins.
- (b) 57/100 — Pooling both trials: total heads = 24 + 33 = 57, total flips = 40 + 60 = 100, so the combined relative frequency is 57/100, which is already in its simplest form since 57 and 100 share no common factor. Averaging the two separate relative frequencies instead, (24/40 + 33/60) ÷ 2 = (0.6 + 0.55) ÷ 2 = 0.575 = 23/40, treats the two trials as equally weighted even though Ben made more flips, which is not correct. Using only Leah's data gives 24/40 = 3/5. Using only Ben's data gives 33/60 = 11/20.
- (d) 12/25 — The group holds 40 of the 250 tickets, so for any one prize the probability the group wins it is 40/250 = 4/25. There are 3 prizes and the group has the same chance at each one, so the expected number won is 3 × 4/25 = 12/25. Writing 4/25 is wrong because it is the chance of winning just ONE prize, without multiplying by the 3 prizes available. Writing 4/75 is wrong because it divides by the 3 prizes instead of multiplying (4/25 ÷ 3 = 4/75), which would mean the group did worse the more prizes were on offer. Writing 64/15625 is wrong because it multiplies the single-prize probability by itself three times, (4/25)³, as though all three prizes had to be won together, instead of adding up the expected number across the three separate prizes. The expected number of prizes won by the group is 12/25.
- (a) 1276 — Method: an unbiased relative frequency tends towards the theoretical probability as the number of trials increases, so use the record resting on the most trials, then multiply by the number of new trials. Working: the three records rest on 50, 200 and 1000 drops, so the most reliable is the one after 1000 drops, namely 0.638, and the run is indeed settling as the trials increase. The expected number of point up landings in 2000 further drops is 2000 × 0.638 = 1276. Answer: about 1276 times. The distractors: 1440 uses the earliest record, which rests on only 50 drops, giving 2000 × 0.720 = 1440; 1330 uses the middle record, treating 200 drops as a safe compromise when 1000 drops is better still, giving 2000 × 0.665 = 1330; 1348 comes from averaging the three records, since 0.720 + 0.665 + 0.638 = 2.023 and 2.023 ÷ 3 = 0.674, then 2000 × 0.674 = 1348, which gives the 50 drop record the same weight as the 1000 drop record.
- (a) 6.75% — Finishing, retiring and being disqualified are exhaustive, so the three percentages sum to 100%: 100% − 68.5% − 24.75% = 6.75%. Adding the two given percentages instead of subtracting them from 100% gives 68.5% + 24.75% = 93.25%, the combined probability of finishing or retiring, not of being disqualified. Subtracting only the retiring percentage from 100% and forgetting the finishing percentage gives 100% − 24.75% = 75.25%. Subtracting only the finishing percentage and forgetting the retiring percentage gives 100% − 68.5% = 31.50%.
- (a) 20 — Dice A is fair, so its expected number of sixes is 150 × 1/6 = 25. Dice B has P(6) = 0.3, so its expected number of sixes is 150 × 0.3 = 45. The difference is 45 − 25 = 20. Adding the two expected values instead of subtracting them gives 25 + 45 = 70. Reporting Dice B's expected sixes on their own, without comparing to Dice A, gives 45. Using the fair probability 1/6 for Dice B as well as Dice A ignores the bias altogether, giving 150 × 1/6 = 25 for both dice and a difference of 0.
- (a) 0.45, different from 0.4 for all the households — Method: work out the probability inside the restricted group of garden owners, then work out the probability across the whole survey, and compare the two. Working: 54 of the 120 households with a garden own a dog, so the conditional probability is 54 divided by 120, which is 0.45. Across the whole survey 80 of the 200 households own a dog, which is 0.4. Since 0.45 is not 0.4, having a garden changes the chance of owning a dog and the two events are not independent. Answer: 0.45, different from 0.4 for all the households. The distractors: 0.27 is 54/200, dividing the households with both by the whole survey instead of by the 120 with a garden; 0.675 is 54/80, the probability that a household has a garden given that it owns a dog, which is the condition and the event the wrong way round; 0.4 is 80/200, the probability of owning a dog with the garden information never used, which is why that route also reports no difference.
- (b) 50% — The total number of customers who bought a cake is 54 + 21 = 75, combining both hot-drink and non-hot-drink customers. As a percentage of all 150 customers, this is (75 ÷ 150) × 100 = 50%. Choosing 36% comes from only counting the hot-drink customers who bought a cake, (54 ÷ 150) × 100 = 36%, and forgetting the 21 non-hot-drink customers who also bought a cake. Choosing 14% comes from only counting the non-hot-drink customers who bought a cake, (21 ÷ 150) × 100 = 14%, and forgetting the 54 hot-drink customers who also bought a cake. Choosing 60% comes from dividing by the hot-drink total of 90 instead of the grand total of 150, (54 ÷ 90) × 100 = 60%.
- (c) 9/28 — Method: two linked steps. Total everyone whose test is positive, since the person picked is known to be one of them, then divide the positive tests that belong to people with the condition by that total. Working: 45 positive tests come from people who have the condition and 95 come from people who do not, so 140 tests are positive. The people with the condition give 45/140, and dividing the numerator and the denominator by 5 gives 9/28. Answer: the probability is 9/28. The distractors: 9/10 is 45/50, the probability of a positive test given that the person has the condition, which is the condition and the event the wrong way round and is the figure a candidate quotes when the two are confused; 9/200 is 45/1000, dividing by everyone tested rather than by the 140 who tested positive; 1/20 is 50/1000, the probability that a person has the condition before the test result is used at all.
- (d) 56 — Method: count the trials over the whole period first, then multiply the number of trials by the probability. Working: 4 weeks is 4 × 7 = 28 days, and at 25 trains a day that is 25 × 28 = 700 trains. The expected number of late trains is 700 × 0.08 = 56. Answer: about 56 late trains over the 4 weeks. The distractors: 2 is the expected number for a single day, 25 × 0.08 = 2, with the 28 days never brought in; 14 uses one week instead of four, 25 × 7 × 0.08 = 14; 644 is 700 − 56 and counts the trains expected to be on time.
- (c) 1/10 — On the adult branch, 210 − 189 = 21 appointments were missed. There are 300 − 210 = 90 child appointments, and 90 − 81 = 9 of those were missed. In total, 21 + 9 = 30 appointments were missed, out of 300: 30/300 = 1/10. Writing 7/100 is wrong because 21/300 simplifies to 7/100, and 21 only counts the adult branch, leaving out the 9 missed child appointments. Writing 3/100 is wrong because 9/300 simplifies to 3/100, and 9 only counts the child branch, leaving out the 21 missed adult appointments. Writing 1/9 is wrong because it divides the 30 missed appointments by the 270 that were attended (300 − 30) instead of by the whole 300 booked. The probability is 1/10.
Build your own mix at the worksheet builder.