Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Answer key: Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- (b) £13.50 — The total cost of Nadia's 25 tickets is 25 × £1.50 = £37.50. The expected number of winning tickets is 25 × 0.12 = 3, so the expected prize money is 3 × £8 = £24.00. Nadia's expected loss is the cost minus the expected prize money: £37.50 − £24.00 = £13.50. A candidate who answers £24.00 has given the expected prize money and mistaken it for the loss. A candidate who answers £37.50 has given the total cost of the tickets, forgetting to subtract the expected prize money. A candidate who answers £34.50 has subtracted the expected number of wins, 3, from the cost instead of first converting it to prize money by multiplying by £8.
- (d) 1/12 — Method: list the full possibility space of sandwich-and-drink pairs, then divide the one matching pair by the size of the whole space. Working: there are 3 × 4 = 12 equally likely sandwich-and-drink pairs, and exactly one of them is egg and water. Answer: 1/12. Watch out: writing down 1/7 comes from adding the two counts, 3 + 4 = 7, instead of multiplying them to build the possibility space. Writing down 1/3 uses only the chance of choosing egg out of 3 sandwiches and ignores the drink altogether. And writing down 1/4 uses only the chance of choosing water out of 4 drinks and ignores the sandwich altogether.
- (b) 37/75 — There are 150 − 84 = 66 men. 84 − 50 = 34 women prefer weight training, and 40 men prefer weight training, so 34 + 40 = 74 people in total prefer weight training, out of 150: 74/150 = 37/75. Writing 4/15 is wrong because it only counts the men who prefer weight training (40/150, simplified), leaving out the 34 women. Writing 17/75 is wrong because it only counts the women who prefer weight training (34/150, simplified), leaving out the 40 men. Writing 37/42 is wrong because it uses the number of women (84) as the denominator instead of the whole gym (150) — 74/84 simplifies to 37/42, but that is not a probability out of the whole group. The probability is 37/75.
- (b) 4 — With 150 rolls and probability 1/6 for each number, the expected count is 150 ÷ 6 = 25. Comparing each actual count with 25: 1 is 22 (3 below), 2 is 27 (2 above), 3 is 24 (1 below), 4 is 34 (9 above), 5 is 21 (4 below) and 6 is 22 (3 below). Number 4 is furthest above its expected count, so it is the most over-represented. Number 2 is also above its expected count, but by only 2, far less than 4's 9. Number 3's count of 24 is below the expected 25, so it is under-represented, not over. Number 6's count of 22 is also below the expected 25, so it too is under-represented.
- (a) 50 — To find the number of shots needed for an expected 12 hits, divide the number of hits wanted by the probability of a hit: 12 ÷ 0.24 = 50. Multiplying the number of hits by the probability instead of dividing gives 12 × 0.24 = 2.88, which rounds to 3 shots. Rounding 0.24 to 0.25 before dividing gives 12 ÷ 0.25 = 48. Using the probability of missing, 1 − 0.24 = 0.76, instead of the probability of hitting, gives 12 ÷ 0.76 = 15.79, which rounds to 16.
- (b) 500 — Method: when a dice is known to be fair, the theoretical probability is the best thing to work from, and the more trials there are the closer the results tend to it. Working: for a fair dice the probability of a six is 1/6, so the expected number of sixes in 3000 rolls is 3000 × 1 ÷ 6 = 500. The class experiment gave a relative frequency of 14/60, but 60 trials is far too few to overturn a known theoretical value, and the school's 3000 rolls will tend towards 1/6 in any case. Answer: about 500 sixes. The distractors: 700 comes from using the class relative frequency instead of the theory, 3000 × 14 ÷ 60 = 700; 600 comes from splitting the difference between the two, since 1/6 is about 0.167 and 14/60 is about 0.233, whose mean is 0.2, and 3000 × 0.2 = 600; 2500 uses 5/6 instead of 1/6 and counts the rolls expected not to be a six.
- (d) 3400 — Combining all three greenhouses gives 200 + 150 + 250 = 600 seeds planted in total, and 172 + 126 + 212 = 510 germinated, so the combined estimate of the germination probability is 510/600 = 0.85. Out of a new batch of 4000 seeds, the expected number to germinate is 4000 × 0.85 = 3400. Writing 3440 is wrong because it uses only Greenhouse 1's rate, 172/200 = 0.86, instead of the combined rate from all three: 4000 × 0.86 = 3440. Writing 3360 is wrong because it uses only Greenhouse 2's rate, 126/150 = 0.84: 4000 × 0.84 = 3360. Writing 510 is wrong because that is the total number that germinated in the ORIGINAL trial, not scaled up to the new batch of 4000 seeds at all. The best estimate is 3400 seeds.
- (b) 22.8% — Relative frequency as a percentage is the faulty count divided by the total, then multiplied by 100: 33 ÷ 145 × 100 = 22.76, which rounds to 22.8%. Giving 33.0% as the answer uses the frequency, 33, directly as a percentage without dividing by the total 145 at all. Rounding 22.76 down to 22.7% instead of up applies the wrong rounding direction at the first decimal place. Finding the relative frequency of the bulbs that were NOT faulty first: 145 − 33 = 112, and 112 ÷ 145 × 100 = 77.24, answers the opposite question and rounds to 77.2%.
- (a) 1200 — Method: take the estimate from the larger sample, because an unbiased relative frequency tends towards the true probability as the sample grows, then multiply by the number of bulbs made in a week. Working: Inspector B tested 500 bulbs, far more than Inspector A's 40, so use B's relative frequency: 30 ÷ 500 = 0.06. A week's production is 4000 × 5 = 20000 bulbs. The expected number of faulty bulbs is 20000 × 0.06 = 1200. Answer: about 1200 faulty bulbs a week. The distractors: 2000 uses Inspector A's estimate, 4 ÷ 40 = 0.1, giving 20000 × 0.1 = 2000, and so rests on a sample of only 40 bulbs; 1600 comes from averaging the two estimates of 0.1 and 0.06 to get 0.08, and 20000 × 0.08 = 1600, which gives the small sample equal weight with the large one; 240 uses the right estimate but stops at a single day, 4000 × 0.06 = 240.
- (a) 0.45, different from 0.4 for all the households — Method: work out the probability inside the restricted group of garden owners, then work out the probability across the whole survey, and compare the two. Working: 54 of the 120 households with a garden own a dog, so the conditional probability is 54 divided by 120, which is 0.45. Across the whole survey 80 of the 200 households own a dog, which is 0.4. Since 0.45 is not 0.4, having a garden changes the chance of owning a dog and the two events are not independent. Answer: 0.45, different from 0.4 for all the households. The distractors: 0.27 is 54/200, dividing the households with both by the whole survey instead of by the 120 with a garden; 0.675 is 54/80, the probability that a household has a garden given that it owns a dog, which is the condition and the event the wrong way round; 0.4 is 80/200, the probability of owning a dog with the garden information never used, which is why that route also reports no difference.
- (b) 3/10 — The number who use at least one app is 90 − 20 = 70. Since 55 + 42 double-counts the overlap, n(X ∩ Y) = 55 + 42 − 70 = 27, so P(both) = 27/90 = 3/10. Forgetting to subtract the 20 who use neither, and using the full 90 as the union, gives 55 + 42 − 90 = 7, so 7/90. Reporting the probability of using X or Y (or both), 70/90 = 7/9, answers a different question about the union, not the overlap. Reporting the probability of using neither app, 20/90 = 2/9, is the complement of the union, not the intersection.
- (b) 3/5 — The question asks about the girls only, so use the girls' total of 70 as the denominator: 42 out of 70 girls walk, giving 42/70 = 3/5. Choosing 21/65 comes from using the whole survey of 130 pupils as the denominator instead of just the 70 girls, 42/130 = 21/65. Choosing 33/70 comes from using the boys' walking count, 33, over the girls' total of 70, mixing up the two rows of the table. Choosing 2/5 comes from using the number of girls who CYCLE, 70 − 42 = 28, instead of the number who walk, giving 28/70 = 2/5.
- (a) 6.75% — Finishing, retiring and being disqualified are exhaustive, so the three percentages sum to 100%: 100% − 68.5% − 24.75% = 6.75%. Adding the two given percentages instead of subtracting them from 100% gives 68.5% + 24.75% = 93.25%, the combined probability of finishing or retiring, not of being disqualified. Subtracting only the retiring percentage from 100% and forgetting the finishing percentage gives 100% − 24.75% = 75.25%. Subtracting only the finishing percentage and forgetting the retiring percentage gives 100% − 68.5% = 31.50%.
- (d) 0.80 — Evens is exactly halfway along the scale, at 0.5, and 3/10 written as a decimal is 0.3. A probability 3/10 higher than evens is 0.5 + 0.3 = 0.80. Giving 0.30 as the answer converts the increase but never adds it to the value of evens. Adding the increase to 1, the 'certain' end of the scale, instead of to evens gives 1 + 0.3 = 1.30, which cannot be a probability. Writing 3/10 as 0.03 instead of 0.3, a place-value slip, gives 0.5 + 0.03 = 0.53.
- (b) 39 — Method: put the counts into a two-way table and fill each missing cell by subtracting along a row or down a column. Working: the number of female members is 150 − 80 = 70. The pool column holds 66 members and 35 of them are male, so the number of female pool users is 66 − 35 = 31. Subtracting along the female row, 70 − 31 = 39 female members use the gym. Answer: 39 female members use the gym. The distractors: 45 comes from subtracting along the male row instead, 80 − 35 = 45, which counts male gym users; 31 is the female pool cell, written down one step before the gym cell; 84 is 150 − 66 and counts every gym user, male and female together.
Build your own mix at the worksheet builder.