Printable · GCSE Foundation · ages 14-16
Probability worksheet — GCSE Foundation
Fifteen questions across the probability statements at Foundation tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Foundation
MathsUKwww.geekhero.co.uk
- 1.Two pupils each flip the same coin to estimate the probability of heads. Leah flips it 40 times and gets 24 heads. Ben flips it 60 times and gets 33 heads. By combining both pupils' results, work out the relative frequency of heads, giving your answer as a fraction in its simplest form.
- 2.A machine makes 4000 light bulbs a day and runs 5 days a week. Two inspectors test bulbs from this machine. Inspector A tests 40 bulbs and finds 4 faulty. Inspector B tests 500 bulbs and finds 30 faulty. Using the better of the two estimates, work out how many faulty bulbs the machine is expected to make in one week.
- 3.A two-way table records how 130 pupils travel to school. 70 of the pupils are girls and the rest are boys. 42 of the girls walk to school and the rest of the girls cycle. 33 of the boys walk to school. Work out what fraction of the girls walk to school.
- 4.An ordinary six-sided dice, numbered 1 to 6, is rolled 30 times and lands on a 6 seven times. Ravi says the theoretical probability of rolling a 6 and the relative frequency of rolling a 6 in this trial are the same number. Is Ravi right?
- 5.A frequency tree records 320 calls to a helpline. The calls first split into calls about billing and calls about technical support. Of the technical support calls, the probability that a call was resolved on the first contact is 0.75. If 180 technical support calls were resolved on the first contact, work out how many calls were about billing.
- 6.A Venn diagram shows two sets, P and Q, inside a universal set. n(P) = 34, n(Q) = 27, n(P ∩ Q) = 11, and n(ξ) = 90, where ξ is the universal set. Work out n((P ∪ Q)′), the number of elements in neither P nor Q.
- 7.A seed company tests germination using results from three greenhouses. Greenhouse 1 plants 200 seeds and 172 germinate. Greenhouse 2 plants 150 seeds and 126 germinate. Greenhouse 3 plants 250 seeds and 212 germinate. Using the combined results from all three greenhouses, work out the best estimate of the number of seeds, out of a new batch of 4000 seeds, that would be expected to germinate.
- 8.A quality inspector examines a sample of 60 items from a production line and finds that 8 are faulty. Using this sample's proportion, work out how many faulty items would be expected in a new batch of 750 items.
- 9.A library recorded loans for 250 books over a week, using a frequency tree. The first branch splits the books into 160 fiction and 90 non-fiction. Of the fiction books, 112 were returned on time and the rest were returned late. All the non-fiction books were returned on time. Work out the probability that a book, chosen at random from the 250, was returned late. Give your answer as a fraction in its simplest form.
- 10.A survey of 90 people records which of two apps, X and Y, they use. 55 people use app X, 42 people use app Y, and 20 people use neither app. Work out the probability that a randomly chosen person uses both apps, giving your answer as a fraction in its simplest form.
- 11.At a fair, a game costs £2 to play. The probability of winning the game is 0.15, and each win pays out £10. Amir plays the game 200 times. Work out how much money Amir should expect to lose in total.
- 12.Two games are offered at a school fair, and every one of the 40 pupils in a class plays each game once. In Game A, the probability of winning a prize is 0.2 and the prize is worth £5. In Game B, the probability of winning a prize is 0.1 and the prize is worth £12. Work out which game gives a higher expected total value of prizes for the class, and by how much.
- 13.On the probability scale from 0 to 1, which word best describes an event with probability 0.9?
- 14.On a Saturday, 2,000 customers go through the tills at a supermarket. The probability that a customer buys a reusable bag is 0.08. Work out how many customers you would expect to buy a reusable bag.
- 15.A frequency tree records the results of 160 patients who took a new medicine. It splits them into those who reported side effects and those who did not. 15% of the patients reported side effects. Work out how many of the 160 patients did not report side effects.
Answer key
- (b) 57/100 — Pooling both trials: total heads = 24 + 33 = 57, total flips = 40 + 60 = 100, so the combined relative frequency is 57/100, which is already in its simplest form since 57 and 100 share no common factor. Averaging the two separate relative frequencies instead, (24/40 + 33/60) ÷ 2 = (0.6 + 0.55) ÷ 2 = 0.575 = 23/40, treats the two trials as equally weighted even though Ben made more flips, which is not correct. Using only Leah's data gives 24/40 = 3/5. Using only Ben's data gives 33/60 = 11/20.
- (a) 1200 — Method: take the estimate from the larger sample, because an unbiased relative frequency tends towards the true probability as the sample grows, then multiply by the number of bulbs made in a week. Working: Inspector B tested 500 bulbs, far more than Inspector A's 40, so use B's relative frequency: 30 ÷ 500 = 0.06. A week's production is 4000 × 5 = 20000 bulbs. The expected number of faulty bulbs is 20000 × 0.06 = 1200. Answer: about 1200 faulty bulbs a week. The distractors: 2000 uses Inspector A's estimate, 4 ÷ 40 = 0.1, giving 20000 × 0.1 = 2000, and so rests on a sample of only 40 bulbs; 1600 comes from averaging the two estimates of 0.1 and 0.06 to get 0.08, and 20000 × 0.08 = 1600, which gives the small sample equal weight with the large one; 240 uses the right estimate but stops at a single day, 4000 × 0.06 = 240.
- (b) 3/5 — The question asks about the girls only, so use the girls' total of 70 as the denominator: 42 out of 70 girls walk, giving 42/70 = 3/5. Choosing 21/65 comes from using the whole survey of 130 pupils as the denominator instead of just the 70 girls, 42/130 = 21/65. Choosing 33/70 comes from using the boys' walking count, 33, over the girls' total of 70, mixing up the two rows of the table. Choosing 2/5 comes from using the number of girls who CYCLE, 70 − 42 = 28, instead of the number who walk, giving 28/70 = 2/5.
- (c) No — 7/30 is the relative frequency; theory stays 1/6. — The theoretical probability of rolling a 6 on an ordinary dice is fixed at 1/6, worked out from the number of equally likely outcomes, and does not change however the dice is actually rolled. The relative frequency from this trial is 7/30, found from what happened in these particular 30 rolls. Since 7/30 and 1/6 are different numbers, the correct statement is 'No — 7/30 is the relative frequency; theory stays 1/6.' Assuming the two values must always match because they describe the same event gives 'Yes — relative frequency always equals theory.' Believing that an observed result redefines the theoretical probability gives 'Yes — the theoretical probability has now become 7/30.' Refusing to work out either value at all gives 'Neither can be found — 30 rolls is too few to tell', which ignores that both numbers CAN be calculated from the information given.
- (c) 80 — Since 180 calls are 0.75 of all the technical support calls, the technical support total is 180 ÷ 0.75 = 240. The billing calls make up the rest of the 320 calls, so 320 − 240 = 80. Choosing 240 comes from stopping after finding the technical support total and forgetting the question asks for the billing calls, which are the rest. Choosing 185 comes from multiplying 180 × 0.75 = 135 instead of dividing, then working out 320 − 135 = 185. Choosing 140 comes from using 180 directly as the whole technical support total, ignoring the probability altogether, then working out 320 − 180 = 140.
- (b) 40 — n(P ∪ Q) = n(P) + n(Q) − n(P ∩ Q) = 34 + 27 − 11 = 50. The complement is everyone outside both sets: n((P ∪ Q)′) = 90 − 50 = 40. Adding P and Q without subtracting the overlap gives 34 + 27 = 61, so 90 − 61 = 29 double-subtracts the 11 who are in both. Reporting n(P ∪ Q) itself, 50, forgets to take the complement at all. Subtracting only n(P) from the universal set, 90 − 34 = 56, ignores set Q altogether.
- (d) 3400 — Combining all three greenhouses gives 200 + 150 + 250 = 600 seeds planted in total, and 172 + 126 + 212 = 510 germinated, so the combined estimate of the germination probability is 510/600 = 0.85. Out of a new batch of 4000 seeds, the expected number to germinate is 4000 × 0.85 = 3400. Writing 3440 is wrong because it uses only Greenhouse 1's rate, 172/200 = 0.86, instead of the combined rate from all three: 4000 × 0.86 = 3440. Writing 3360 is wrong because it uses only Greenhouse 2's rate, 126/150 = 0.84: 4000 × 0.84 = 3360. Writing 510 is wrong because that is the total number that germinated in the ORIGINAL trial, not scaled up to the new batch of 4000 seeds at all. The best estimate is 3400 seeds.
- (b) 100 — The sample shows a proportion of 8/60 = 2/15 faulty. Apply that proportion to the new batch of 750: 750 × 2/15 = 100. Flipping the ratio, calculating 8/750 × 60 instead of 8/60 × 750, gives 0.64, which rounds to about 1. Assuming the same number of faulty items applies to the new batch, without scaling for its larger size, just repeats the sample's count of 8. Rounding the proportion 8/60 = 0.1333... down to 0.1 before multiplying gives 750 × 0.1 = 75.
- (b) 24/125 — On the fiction branch, 160 − 112 = 48 books were returned late. None of the non-fiction books were late, so the total number of late books is 48, out of 250 books altogether: 48/250 = 24/125. Writing 3/10 is wrong because it divides the 48 late fiction books by the fiction total (160) instead of the whole library total (250). Writing 56/125 is wrong because 112/250 simplifies to 56/125, and 112 is the number of fiction books returned ON TIME, not late. Writing 9/25 is wrong because 90/250 simplifies to 9/25, and 90 is simply the number of non-fiction books, which has nothing to do with late returns. The probability is 24/125.
- (b) 3/10 — The number who use at least one app is 90 − 20 = 70. Since 55 + 42 double-counts the overlap, n(X ∩ Y) = 55 + 42 − 70 = 27, so P(both) = 27/90 = 3/10. Forgetting to subtract the 20 who use neither, and using the full 90 as the union, gives 55 + 42 − 90 = 7, so 7/90. Reporting the probability of using X or Y (or both), 70/90 = 7/9, answers a different question about the union, not the overlap. Reporting the probability of using neither app, 20/90 = 2/9, is the complement of the union, not the intersection.
- (b) £100 — Method: find the expected number of wins, turn that into the expected pay out, then compare it with what the games cost. Working: the expected number of wins is 200 × 0.15 = 30. Each win pays £10, so the expected pay out is 30 × 10 = 300 pounds. Playing 200 times at £2 a go costs 200 × 2 = 400 pounds. The expected loss is 400 − 300 = 100 pounds. Answer: Amir should expect to be about £100 down. The distractors: £300 is the expected winnings on their own, with the cost of playing never taken off; £400 is the total cost of playing, with the winnings never taken off; £700 comes from adding the two totals, 400 + 300 = 700, instead of subtracting one from the other.
- (a) Game B, by £8 — Game A's expected total is 40 × 0.2 × £5 = £40. Game B's expected total is 40 × 0.1 × £12 = £48. Game B is higher, by £48 − £40 = £8. Writing 'Game A, by £8' is wrong because it has the right difference but the wrong game — Game A's total (£40) is actually LOWER than Game B's, not higher. Writing 'Game B, by £48' is wrong because £48 is Game B's whole expected total, not the DIFFERENCE between the two games. Writing 'Game A, by £40' is wrong in the same way, using Game A's whole total as if it were the margin, and naming the wrong game as the winner. Game B gives the higher expected total, by £8.
- (b) likely — On the probability scale, 'certain' is reserved for a probability of exactly 1, and 'evens' describes a probability of exactly 0.5. A probability of 0.9 is high but not equal to 1, so the correct word is 'likely'. Choosing 'certain' treats a probability close to 1 as if it were exactly 1, which it is not. Choosing 'evens' misjudges 0.9 as being close to the midpoint of the scale, when it is close to the 'certain' end instead. Choosing 'unlikely' reads the scale the wrong way round, as if a high probability meant a low chance of happening.
- (a) 160 — Expected number = probability × number of trials = 0.08 × 2,000 = 160. Moving the decimal point one place too far, using 0.008 instead of 0.08, gives 2,000 × 0.008 = 16. Working out the expected number of customers who do NOT buy a bag, using the complement 1 − 0.08 = 0.92, gives 2,000 × 0.92 = 1,840. Rounding 0.08 up to 0.1 before multiplying gives 2,000 × 0.1 = 200.
- (d) 136 — 15% of 160 = 0.15 × 160 = 24 patients reported side effects, so 160 − 24 = 136 did not. Stopping after finding the number who reported side effects, 24, answers the wrong question — it is not the number who did NOT report them. Misreading '15%' as a raw count of 15 patients, rather than a percentage, gives 160 − 15 = 145. Subtracting 15% of 160 twice, 160 − 24 − 24 = 112, double-counts the side-effect group.
Build your own mix at the worksheet builder.