Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Answer key: Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- (d) No, because 3/8 + 5/12 + 1/6 = 23/24 — Using a common denominator of 24: 3/8 = 9/24, 5/12 = 10/24 and 1/6 = 4/24. Adding these numerators gives 9 + 10 + 4 = 23, so the three probabilities sum to 23/24, which is less than 1 — Zara is not correct. Adding the original numerators (3 + 5 + 1 = 9) over a denominator of 12 instead of converting each fraction properly gives 9/12 = 3/4, still less than 1 but the wrong fraction. Converting 1/6 to 5/24 instead of 4/24 (using the wrong scaling) makes the total 9/24 + 10/24 + 5/24 = 24/24 = 1, wrongly suggesting the probabilities are valid. Judging validity from the fact that each individual fraction lies between 0 and 1 ignores that an exhaustive set must sum to exactly 1, not merely contain valid individual values.
- (c) 325 — Method: first find the relative frequency of NOT landing on red from the 200 spins, then scale that up to 500 spins. Working: non-red results = 50 + 80 = 130, out of 200 spins, so P(not red) = 130 ÷ 200 = 0.65. Expected non-red results in 500 spins = 500 × 0.65 = 325. Answer: 325. Watch out: writing down 175 finds the expected number of RED results instead, 70 ÷ 200 × 500 = 175, answering the opposite of what was asked. Writing down 250 assumes landing red or not landing red must be a fair 50-50 split, but the spinner is biased and the actual results do not split evenly. And writing down 130 stops after finding how many of the 200 spins were non-red and forgets to scale that figure up to the 500 spins asked for.
- (a) £150 — Each game, the expected payout is 0.1 × £20 = £2, so the fête's expected profit per game is the £3 charged minus the £2 expected payout, £1. Over 150 games, that is 150 × £1 = £150. Writing £300 is wrong because 150 × £2 = £300 is the total expected PAYOUT, not the profit — it has not been subtracted from the entry fees. Writing £450 is wrong because 150 × £3 = £450 is the total money taken in entry fees, without accounting for what is expected to be paid out in prizes. Writing £1 is wrong because that is only the expected profit for ONE game — it has not been scaled up to all 150 games. The fête's expected profit is £150.
- (d) 56 — Method: count the trials over the whole period first, then multiply the number of trials by the probability. Working: 4 weeks is 4 × 7 = 28 days, and at 25 trains a day that is 25 × 28 = 700 trains. The expected number of late trains is 700 × 0.08 = 56. Answer: about 56 late trains over the 4 weeks. The distractors: 2 is the expected number for a single day, 25 × 0.08 = 2, with the 28 days never brought in; 14 uses one week instead of four, 25 × 7 × 0.08 = 14; 644 is 700 − 56 and counts the trains expected to be on time.
- (b) 0.225 — The relative frequency of rain is the number of rainy days out of all days recorded: 9 ÷ 40 = 0.225, which is noticeably less than the forecaster's claimed 0.3. Using the number of dry days, 40 − 9 = 31, as the denominator instead of the total of 40 gives 9 ÷ 31 = 0.29 (2 d.p.). Simply reporting the forecaster's claimed value, 0.3, without calculating anything from the data at all, ignores the recorded results completely. Misplacing the decimal point, treating 9 out of 40 as 9%, gives 0.09 instead of 0.225.
- (c) Red — Theoretical probability is 1/3 ≈ 0.333 for each colour. Red's relative frequency is 38/90 ≈ 0.422, above 1/3, so red is over-represented. Blue's relative frequency is 26/90 ≈ 0.289, below 1/3, so blue is under-represented, not over. Green's relative frequency is also 26/90 ≈ 0.289, below 1/3 for the same reason. Since red's relative frequency clearly exceeds 1/3, it is not true that none of the colours are over-represented.
- (c) 80 — Since 180 calls are 0.75 of all the technical support calls, the technical support total is 180 ÷ 0.75 = 240. The billing calls make up the rest of the 320 calls, so 320 − 240 = 80. Choosing 240 comes from stopping after finding the technical support total and forgetting the question asks for the billing calls, which are the rest. Choosing 185 comes from multiplying 180 × 0.75 = 135 instead of dividing, then working out 320 − 135 = 185. Choosing 140 comes from using 180 directly as the whole technical support total, ignoring the probability altogether, then working out 320 − 180 = 140.
- (a) 952 — 68% = 0.68. The relative frequency from the survey applies to the new group of 1400 shoppers, so the expected number is 0.68 × 1400 = 952. Working out 1 − 0.68 = 0.32 and applying that instead, 0.32 × 1400 = 448, finds the number who have NOT used a self-checkout, not the number who have. Applying 68% to the original sample size of 250 instead of the new total of 1400 gives 0.68 × 250 = 170. Applying the complement percentage to the original sample size, 0.32 × 250 = 80, compounds both mistakes.
- (a) 0.45, different from 0.4 for all the households — Method: work out the probability inside the restricted group of garden owners, then work out the probability across the whole survey, and compare the two. Working: 54 of the 120 households with a garden own a dog, so the conditional probability is 54 divided by 120, which is 0.45. Across the whole survey 80 of the 200 households own a dog, which is 0.4. Since 0.45 is not 0.4, having a garden changes the chance of owning a dog and the two events are not independent. Answer: 0.45, different from 0.4 for all the households. The distractors: 0.27 is 54/200, dividing the households with both by the whole survey instead of by the 120 with a garden; 0.675 is 54/80, the probability that a household has a garden given that it owns a dog, which is the condition and the event the wrong way round; 0.4 is 80/200, the probability of owning a dog with the garden information never used, which is why that route also reports no difference.
- (b) 24/125 — On the fiction branch, 160 − 112 = 48 books were returned late. None of the non-fiction books were late, so the total number of late books is 48, out of 250 books altogether: 48/250 = 24/125. Writing 3/10 is wrong because it divides the 48 late fiction books by the fiction total (160) instead of the whole library total (250). Writing 56/125 is wrong because 112/250 simplifies to 56/125, and 112 is the number of fiction books returned ON TIME, not late. Writing 9/25 is wrong because 90/250 simplifies to 9/25, and 90 is simply the number of non-fiction books, which has nothing to do with late returns. The probability is 24/125.
- (b) 500 — Method: when a dice is known to be fair, the theoretical probability is the best thing to work from, and the more trials there are the closer the results tend to it. Working: for a fair dice the probability of a six is 1/6, so the expected number of sixes in 3000 rolls is 3000 × 1 ÷ 6 = 500. The class experiment gave a relative frequency of 14/60, but 60 trials is far too few to overturn a known theoretical value, and the school's 3000 rolls will tend towards 1/6 in any case. Answer: about 500 sixes. The distractors: 700 comes from using the class relative frequency instead of the theory, 3000 × 14 ÷ 60 = 700; 600 comes from splitting the difference between the two, since 1/6 is about 0.167 and 14/60 is about 0.233, whose mean is 0.2, and 3000 × 0.2 = 600; 2500 uses 5/6 instead of 1/6 and counts the rolls expected not to be a six.
- (a) 20 — Dice A is fair, so its expected number of sixes is 150 × 1/6 = 25. Dice B has P(6) = 0.3, so its expected number of sixes is 150 × 0.3 = 45. The difference is 45 − 25 = 20. Adding the two expected values instead of subtracting them gives 25 + 45 = 70. Reporting Dice B's expected sixes on their own, without comparing to Dice A, gives 45. Using the fair probability 1/6 for Dice B as well as Dice A ignores the bias altogether, giving 150 × 1/6 = 25 for both dice and a difference of 0.
- (c) Noah — 11/32 = 0.34375, above 34%. — Converting 11/32 to a decimal gives 11 ÷ 32 = 0.34375, which is greater than 34% (0.34), so Noah has the better save rate: 'Noah — 11/32 = 0.34375, above 34%.' Comparing the raw numbers 34 and 11 directly, without converting the fraction to the same form, gives 'Elin — 34 is bigger than 11.' Treating a larger denominator as meaning a bigger value, rather than smaller equal shares, gives 'Noah — 32 is a bigger denominator.' Rounding 34.375% to 34% to the nearest whole percent hides the difference and gives 'Equal — both round to 34% to the nearest percent.'
- (a) 1200 — Method: take the estimate from the larger sample, because an unbiased relative frequency tends towards the true probability as the sample grows, then multiply by the number of bulbs made in a week. Working: Inspector B tested 500 bulbs, far more than Inspector A's 40, so use B's relative frequency: 30 ÷ 500 = 0.06. A week's production is 4000 × 5 = 20000 bulbs. The expected number of faulty bulbs is 20000 × 0.06 = 1200. Answer: about 1200 faulty bulbs a week. The distractors: 2000 uses Inspector A's estimate, 4 ÷ 40 = 0.1, giving 20000 × 0.1 = 2000, and so rests on a sample of only 40 bulbs; 1600 comes from averaging the two estimates of 0.1 and 0.06 to get 0.08, and 20000 × 0.08 = 1600, which gives the small sample equal weight with the large one; 240 uses the right estimate but stops at a single day, 4000 × 0.06 = 240.
- (b) 38% — Method: the two swimming percentages are quoted inside different age groups, so weight each one by the size of its group and add the two results. Working: the under 18s are 45% of the members and 60% of them swim, giving 0.45 × 60 = 27% of all the members. The members aged 18 or over are 55% of the members and 20% of them swim, giving 0.55 × 20 = 11% of all the members. Adding these gives 38%. Answer: 38% of the members swim each week. The distractors: 80% comes from adding 60% and 20% straight off, treating two rates quoted inside different groups as though they could be added; 40% is the mean of 60% and 20%, which would be right only if the two age groups were the same size, and they are not; 42% comes from pairing each swimming rate with the wrong age group, working out 0.45 × 20 added to 0.55 × 60.
Build your own mix at the worksheet builder.