Printable · GCSE Foundation · ages 14-16
Probability worksheet — GCSE Foundation
Fifteen questions across the probability statements at Foundation tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Answer key: Probability worksheet — GCSE Foundation
MathsUKwww.geekhero.co.uk
- (b) 57/100 — Pooling both trials: total heads = 24 + 33 = 57, total flips = 40 + 60 = 100, so the combined relative frequency is 57/100, which is already in its simplest form since 57 and 100 share no common factor. Averaging the two separate relative frequencies instead, (24/40 + 33/60) ÷ 2 = (0.6 + 0.55) ÷ 2 = 0.575 = 23/40, treats the two trials as equally weighted even though Ben made more flips, which is not correct. Using only Leah's data gives 24/40 = 3/5. Using only Ben's data gives 33/60 = 11/20.
- (a) 1200 — Method: take the estimate from the larger sample, because an unbiased relative frequency tends towards the true probability as the sample grows, then multiply by the number of bulbs made in a week. Working: Inspector B tested 500 bulbs, far more than Inspector A's 40, so use B's relative frequency: 30 ÷ 500 = 0.06. A week's production is 4000 × 5 = 20000 bulbs. The expected number of faulty bulbs is 20000 × 0.06 = 1200. Answer: about 1200 faulty bulbs a week. The distractors: 2000 uses Inspector A's estimate, 4 ÷ 40 = 0.1, giving 20000 × 0.1 = 2000, and so rests on a sample of only 40 bulbs; 1600 comes from averaging the two estimates of 0.1 and 0.06 to get 0.08, and 20000 × 0.08 = 1600, which gives the small sample equal weight with the large one; 240 uses the right estimate but stops at a single day, 4000 × 0.06 = 240.
- (b) 3/5 — The question asks about the girls only, so use the girls' total of 70 as the denominator: 42 out of 70 girls walk, giving 42/70 = 3/5. Choosing 21/65 comes from using the whole survey of 130 pupils as the denominator instead of just the 70 girls, 42/130 = 21/65. Choosing 33/70 comes from using the boys' walking count, 33, over the girls' total of 70, mixing up the two rows of the table. Choosing 2/5 comes from using the number of girls who CYCLE, 70 − 42 = 28, instead of the number who walk, giving 28/70 = 2/5.
- (c) No — 7/30 is the relative frequency; theory stays 1/6. — The theoretical probability of rolling a 6 on an ordinary dice is fixed at 1/6, worked out from the number of equally likely outcomes, and does not change however the dice is actually rolled. The relative frequency from this trial is 7/30, found from what happened in these particular 30 rolls. Since 7/30 and 1/6 are different numbers, the correct statement is 'No — 7/30 is the relative frequency; theory stays 1/6.' Assuming the two values must always match because they describe the same event gives 'Yes — relative frequency always equals theory.' Believing that an observed result redefines the theoretical probability gives 'Yes — the theoretical probability has now become 7/30.' Refusing to work out either value at all gives 'Neither can be found — 30 rolls is too few to tell', which ignores that both numbers CAN be calculated from the information given.
- (c) 80 — Since 180 calls are 0.75 of all the technical support calls, the technical support total is 180 ÷ 0.75 = 240. The billing calls make up the rest of the 320 calls, so 320 − 240 = 80. Choosing 240 comes from stopping after finding the technical support total and forgetting the question asks for the billing calls, which are the rest. Choosing 185 comes from multiplying 180 × 0.75 = 135 instead of dividing, then working out 320 − 135 = 185. Choosing 140 comes from using 180 directly as the whole technical support total, ignoring the probability altogether, then working out 320 − 180 = 140.
- (b) 40 — n(P ∪ Q) = n(P) + n(Q) − n(P ∩ Q) = 34 + 27 − 11 = 50. The complement is everyone outside both sets: n((P ∪ Q)′) = 90 − 50 = 40. Adding P and Q without subtracting the overlap gives 34 + 27 = 61, so 90 − 61 = 29 double-subtracts the 11 who are in both. Reporting n(P ∪ Q) itself, 50, forgets to take the complement at all. Subtracting only n(P) from the universal set, 90 − 34 = 56, ignores set Q altogether.
- (d) 3400 — Combining all three greenhouses gives 200 + 150 + 250 = 600 seeds planted in total, and 172 + 126 + 212 = 510 germinated, so the combined estimate of the germination probability is 510/600 = 0.85. Out of a new batch of 4000 seeds, the expected number to germinate is 4000 × 0.85 = 3400. Writing 3440 is wrong because it uses only Greenhouse 1's rate, 172/200 = 0.86, instead of the combined rate from all three: 4000 × 0.86 = 3440. Writing 3360 is wrong because it uses only Greenhouse 2's rate, 126/150 = 0.84: 4000 × 0.84 = 3360. Writing 510 is wrong because that is the total number that germinated in the ORIGINAL trial, not scaled up to the new batch of 4000 seeds at all. The best estimate is 3400 seeds.
- (b) 100 — The sample shows a proportion of 8/60 = 2/15 faulty. Apply that proportion to the new batch of 750: 750 × 2/15 = 100. Flipping the ratio, calculating 8/750 × 60 instead of 8/60 × 750, gives 0.64, which rounds to about 1. Assuming the same number of faulty items applies to the new batch, without scaling for its larger size, just repeats the sample's count of 8. Rounding the proportion 8/60 = 0.1333... down to 0.1 before multiplying gives 750 × 0.1 = 75.
- (b) 24/125 — On the fiction branch, 160 − 112 = 48 books were returned late. None of the non-fiction books were late, so the total number of late books is 48, out of 250 books altogether: 48/250 = 24/125. Writing 3/10 is wrong because it divides the 48 late fiction books by the fiction total (160) instead of the whole library total (250). Writing 56/125 is wrong because 112/250 simplifies to 56/125, and 112 is the number of fiction books returned ON TIME, not late. Writing 9/25 is wrong because 90/250 simplifies to 9/25, and 90 is simply the number of non-fiction books, which has nothing to do with late returns. The probability is 24/125.
- (b) 3/10 — The number who use at least one app is 90 − 20 = 70. Since 55 + 42 double-counts the overlap, n(X ∩ Y) = 55 + 42 − 70 = 27, so P(both) = 27/90 = 3/10. Forgetting to subtract the 20 who use neither, and using the full 90 as the union, gives 55 + 42 − 90 = 7, so 7/90. Reporting the probability of using X or Y (or both), 70/90 = 7/9, answers a different question about the union, not the overlap. Reporting the probability of using neither app, 20/90 = 2/9, is the complement of the union, not the intersection.
- (b) £100 — Method: find the expected number of wins, turn that into the expected pay out, then compare it with what the games cost. Working: the expected number of wins is 200 × 0.15 = 30. Each win pays £10, so the expected pay out is 30 × 10 = 300 pounds. Playing 200 times at £2 a go costs 200 × 2 = 400 pounds. The expected loss is 400 − 300 = 100 pounds. Answer: Amir should expect to be about £100 down. The distractors: £300 is the expected winnings on their own, with the cost of playing never taken off; £400 is the total cost of playing, with the winnings never taken off; £700 comes from adding the two totals, 400 + 300 = 700, instead of subtracting one from the other.
- (a) Game B, by £8 — Game A's expected total is 40 × 0.2 × £5 = £40. Game B's expected total is 40 × 0.1 × £12 = £48. Game B is higher, by £48 − £40 = £8. Writing 'Game A, by £8' is wrong because it has the right difference but the wrong game — Game A's total (£40) is actually LOWER than Game B's, not higher. Writing 'Game B, by £48' is wrong because £48 is Game B's whole expected total, not the DIFFERENCE between the two games. Writing 'Game A, by £40' is wrong in the same way, using Game A's whole total as if it were the margin, and naming the wrong game as the winner. Game B gives the higher expected total, by £8.
- (b) likely — On the probability scale, 'certain' is reserved for a probability of exactly 1, and 'evens' describes a probability of exactly 0.5. A probability of 0.9 is high but not equal to 1, so the correct word is 'likely'. Choosing 'certain' treats a probability close to 1 as if it were exactly 1, which it is not. Choosing 'evens' misjudges 0.9 as being close to the midpoint of the scale, when it is close to the 'certain' end instead. Choosing 'unlikely' reads the scale the wrong way round, as if a high probability meant a low chance of happening.
- (a) 160 — Expected number = probability × number of trials = 0.08 × 2,000 = 160. Moving the decimal point one place too far, using 0.008 instead of 0.08, gives 2,000 × 0.008 = 16. Working out the expected number of customers who do NOT buy a bag, using the complement 1 − 0.08 = 0.92, gives 2,000 × 0.92 = 1,840. Rounding 0.08 up to 0.1 before multiplying gives 2,000 × 0.1 = 200.
- (d) 136 — 15% of 160 = 0.15 × 160 = 24 patients reported side effects, so 160 − 24 = 136 did not. Stopping after finding the number who reported side effects, 24, answers the wrong question — it is not the number who did NOT report them. Misreading '15%' as a raw count of 15 patients, rather than a percentage, gives 160 − 15 = 145. Subtracting 15% of 160 twice, 160 − 24 − 24 = 112, double-counts the side-effect group.
Build your own mix at the worksheet builder.