Printable · GCSE Foundation · ages 14-16
Probability worksheet — GCSE Foundation
Fifteen questions across the probability statements at Foundation tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Foundation
MathsUKwww.geekhero.co.uk
- 1.A survey of 90 people records which of two apps, X and Y, they use. 55 people use app X, 42 people use app Y, and 20 people use neither app. Work out the probability that a randomly chosen person uses both apps, giving your answer as a fraction in its simplest form.
- 2.150 people at a gym were asked whether they prefer weight training or cardio; each person chose exactly one. 84 of the people are women. 50 of the women prefer cardio. 40 of the men prefer weight training. Work out the probability that a person, chosen at random from the 150, prefers weight training. Give your answer as a fraction in its simplest form.
- 3.A leisure centre has 150 members. 80 of the members are male and the rest are female. Every member uses either the pool or the gym, but not both. 66 members use the pool, and 35 of those pool users are male. Work out how many female members use the gym.
- 4.A machine makes 4000 light bulbs a day and runs 5 days a week. Two inspectors test bulbs from this machine. Inspector A tests 40 bulbs and finds 4 faulty. Inspector B tests 500 bulbs and finds 30 faulty. Using the better of the two estimates, work out how many faulty bulbs the machine is expected to make in one week.
- 5.A school trip took 200 pupils to a museum, travelling either by coach or by minibus. A frequency tree records the trip: 150 pupils travelled by coach, split into 90 in Year 8 and the rest in Year 9. All the pupils who travelled by minibus were in Year 8. Work out the total number of Year 8 pupils on the trip.
- 6.A basketball player takes two free throws, and the throws are independent. The probability of scoring on each throw is 0.6, and the probability of missing is 0.4. Work out the probability that she scores exactly one of the two throws.
- 7.A fair coin is flipped again and again. After the first 10 flips there have been 7 heads. After 1000 flips there have been 528 heads. Which statement best describes what these results show?
- 8.A survey found that, of 250 shoppers questioned, 68% said they had used a self-checkout in the past month. A different store expects 1400 shoppers this week. Using this relative frequency, work out how many of the 1400 shoppers would be expected to have used a self-checkout in the past month.
- 9.A survey of 160 employees at a company recorded whether they cycle to work, using a frequency tree. The first branch splits them into 90 who work full-time and 70 who work part-time. Of the full-time employees, 27 cycle to work. Of the part-time employees, 14 cycle to work. Work out the probability that an employee, chosen at random from the 160, cycles to work. Give your answer as a fraction in its simplest form.
- 10.A bag is known to contain red, blue and green counters in equal numbers, so the theoretical probability of taking each colour is 1/3. In 90 trials with replacement, red was taken 38 times, blue was taken 26 times and green was taken 26 times. Which colour is over-represented compared with its theoretical probability?
- 11.At a fun run, each runner finishes, retires or is disqualified, and cannot do more than one of these. The probability that a runner finishes is 68.5% and the probability that a runner retires is 24.75%. Work out the probability, as a percentage, that a runner is disqualified.
- 12.At a fun run, a raffle stall charges £1.50 per ticket. The probability that any one ticket wins a prize worth £8 is 0.12, and a losing ticket wins nothing. Nadia buys 25 tickets. Work out how much money Nadia should expect to lose in total.
- 13.Ellie drops a bottle top and records whether it lands open end up. In her first 20 drops it lands open end up 13 times. She carries on, and after 200 drops in total it has landed open end up 84 times. Work out the best estimate of the probability that the bottle top lands open end up.
- 14.A fair six-sided dice is rolled 150 times. The table shows how many times each number came up: 1 came up 22 times, 2 came up 27 times, 3 came up 24 times, 4 came up 34 times, 5 came up 21 times and 6 came up 22 times. The theoretical probability of each number is 1/6. Which number is most over-represented compared with its theoretical probability?
- 15.A bag contains 20 counters. 8 of the counters are red and 12 are blue. A counter is taken at random, its colour is recorded, and it is put back in the bag. This is done 150 times. Work out how many red counters you would expect to be recorded.
Answer key
- (b) 3/10 — The number who use at least one app is 90 − 20 = 70. Since 55 + 42 double-counts the overlap, n(X ∩ Y) = 55 + 42 − 70 = 27, so P(both) = 27/90 = 3/10. Forgetting to subtract the 20 who use neither, and using the full 90 as the union, gives 55 + 42 − 90 = 7, so 7/90. Reporting the probability of using X or Y (or both), 70/90 = 7/9, answers a different question about the union, not the overlap. Reporting the probability of using neither app, 20/90 = 2/9, is the complement of the union, not the intersection.
- (b) 37/75 — There are 150 − 84 = 66 men. 84 − 50 = 34 women prefer weight training, and 40 men prefer weight training, so 34 + 40 = 74 people in total prefer weight training, out of 150: 74/150 = 37/75. Writing 4/15 is wrong because it only counts the men who prefer weight training (40/150, simplified), leaving out the 34 women. Writing 17/75 is wrong because it only counts the women who prefer weight training (34/150, simplified), leaving out the 40 men. Writing 37/42 is wrong because it uses the number of women (84) as the denominator instead of the whole gym (150) — 74/84 simplifies to 37/42, but that is not a probability out of the whole group. The probability is 37/75.
- (b) 39 — Method: put the counts into a two-way table and fill each missing cell by subtracting along a row or down a column. Working: the number of female members is 150 − 80 = 70. The pool column holds 66 members and 35 of them are male, so the number of female pool users is 66 − 35 = 31. Subtracting along the female row, 70 − 31 = 39 female members use the gym. Answer: 39 female members use the gym. The distractors: 45 comes from subtracting along the male row instead, 80 − 35 = 45, which counts male gym users; 31 is the female pool cell, written down one step before the gym cell; 84 is 150 − 66 and counts every gym user, male and female together.
- (a) 1200 — Method: take the estimate from the larger sample, because an unbiased relative frequency tends towards the true probability as the sample grows, then multiply by the number of bulbs made in a week. Working: Inspector B tested 500 bulbs, far more than Inspector A's 40, so use B's relative frequency: 30 ÷ 500 = 0.06. A week's production is 4000 × 5 = 20000 bulbs. The expected number of faulty bulbs is 20000 × 0.06 = 1200. Answer: about 1200 faulty bulbs a week. The distractors: 2000 uses Inspector A's estimate, 4 ÷ 40 = 0.1, giving 20000 × 0.1 = 2000, and so rests on a sample of only 40 bulbs; 1600 comes from averaging the two estimates of 0.1 and 0.06 to get 0.08, and 20000 × 0.08 = 1600, which gives the small sample equal weight with the large one; 240 uses the right estimate but stops at a single day, 4000 × 0.06 = 240.
- (d) 140 — 90 pupils were in Year 8 on the coach branch. The minibus branch has 200 − 150 = 50 pupils, and all of them are Year 8 too, so the total number of Year 8 pupils is 90 + 50 = 140. Writing 90 alone is wrong because it only counts the coach's Year 8 pupils and misses the minibus ones. Writing 60 is wrong because that is the number of Year 9 pupils on the coach (150 − 90 = 60), not Year 8 at all. Writing 50 alone is wrong because it only counts the minibus pupils and misses the coach's Year 8 pupils. The total is 140.
- (a) 0.48 — There are two ways to score exactly one throw: scoring on the first and missing the second, 0.6 × 0.4 = 0.24, or missing the first and scoring the second, 0.4 × 0.6 = 0.24. Adding these gives 0.24 + 0.24 = 0.48. Choosing 0.24 comes from working out only one of the two paths and forgetting the other one also gives exactly one score. Choosing 0.36 comes from working out the probability of scoring BOTH throws, 0.6 × 0.6 = 0.36, instead of exactly one. Choosing 0.84 comes from working out the probability of scoring AT LEAST one throw, 1 − 0.4 × 0.4 = 0.84, instead of exactly one.
- (b) The relative frequency is settling near 0.5 — Method: turn each result into a relative frequency before comparing them, because it is the relative frequency, and not the difference between the two counts, that tends towards the theoretical probability. Working: after 10 flips the relative frequency of a head is 7 ÷ 10 = 0.7, which is a long way from 0.5. After 1000 flips it is 528 ÷ 1000 = 0.528, which is much closer to 0.5. Meanwhile the gap between the two counts has grown rather than shrunk: it was 7 − 3 = 4 after 10 flips and is 528 − 472 = 56 after 1000 flips. Answer: the relative frequency is settling near 0.5, which is what an unbiased experiment does as the sample grows. The distractors: saying the counts are levelling out is the usual form of this idea and the figures contradict it, since the gap went from 4 to 56; saying the coin is biased treats 28 extra heads in 1000 flips as proof, when 0.528 sits close to 0.5 and a fair coin gives results like this often; saying the next flip is more likely to be a tail is the gambler's fallacy, since each flip stays at 1/2 whatever came before.
- (a) 952 — 68% = 0.68. The relative frequency from the survey applies to the new group of 1400 shoppers, so the expected number is 0.68 × 1400 = 952. Working out 1 − 0.68 = 0.32 and applying that instead, 0.32 × 1400 = 448, finds the number who have NOT used a self-checkout, not the number who have. Applying 68% to the original sample size of 250 instead of the new total of 1400 gives 0.68 × 250 = 170. Applying the complement percentage to the original sample size, 0.32 × 250 = 80, compounds both mistakes.
- (b) 41/160 — In total, 27 + 14 = 41 of the 160 employees cycle to work, so the probability is 41/160 (41 and 160 share no common factor, so this is already in its simplest form). Writing 27/160 is wrong because it only counts the full-time cyclists and leaves out the 14 part-time cyclists. Writing 41/90 is wrong because it uses the full-time total (90) as the denominator instead of the whole survey (160). Writing 1/5 is wrong because it only uses the part-time branch, simplifying 14/70 to 1/5 and ignoring the full-time cyclists completely. The probability is 41/160.
- (c) Red — Theoretical probability is 1/3 ≈ 0.333 for each colour. Red's relative frequency is 38/90 ≈ 0.422, above 1/3, so red is over-represented. Blue's relative frequency is 26/90 ≈ 0.289, below 1/3, so blue is under-represented, not over. Green's relative frequency is also 26/90 ≈ 0.289, below 1/3 for the same reason. Since red's relative frequency clearly exceeds 1/3, it is not true that none of the colours are over-represented.
- (a) 6.75% — Finishing, retiring and being disqualified are exhaustive, so the three percentages sum to 100%: 100% − 68.5% − 24.75% = 6.75%. Adding the two given percentages instead of subtracting them from 100% gives 68.5% + 24.75% = 93.25%, the combined probability of finishing or retiring, not of being disqualified. Subtracting only the retiring percentage from 100% and forgetting the finishing percentage gives 100% − 24.75% = 75.25%. Subtracting only the finishing percentage and forgetting the retiring percentage gives 100% − 68.5% = 31.50%.
- (b) £13.50 — The total cost of Nadia's 25 tickets is 25 × £1.50 = £37.50. The expected number of winning tickets is 25 × 0.12 = 3, so the expected prize money is 3 × £8 = £24.00. Nadia's expected loss is the cost minus the expected prize money: £37.50 − £24.00 = £13.50. A candidate who answers £24.00 has given the expected prize money and mistaken it for the loss. A candidate who answers £37.50 has given the total cost of the tickets, forgetting to subtract the expected prize money. A candidate who answers £34.50 has subtracted the expected number of wins, 3, from the cost instead of first converting it to prize money by multiplying by £8.
- (d) 0.420 — Method: use the relative frequency worked out from the larger number of trials as the best estimate of the probability, since a bigger sample tends to sit closer to the true probability. Working: 84 out of 200 drops land open end up, so the relative frequency from the total is 84 ÷ 200 = 0.420. Answer: 0.420. Watch out: writing down 0.650 comes from 13 ÷ 20, using only the first, much smaller sample instead of the total. Writing down 0.535 comes from averaging 0.650 and 0.420, treating the 20-drop run and the 200-drop run as equally reliable instead of using the larger sample on its own. And writing down 0.580 comes from 1 − 0.420, working out the probability that the bottle top lands the other way up instead of open end up.
- (b) 4 — With 150 rolls and probability 1/6 for each number, the expected count is 150 ÷ 6 = 25. Comparing each actual count with 25: 1 is 22 (3 below), 2 is 27 (2 above), 3 is 24 (1 below), 4 is 34 (9 above), 5 is 21 (4 below) and 6 is 22 (3 below). Number 4 is furthest above its expected count, so it is the most over-represented. Number 2 is also above its expected count, but by only 2, far less than 4's 9. Number 3's count of 24 is below the expected 25, so it is under-represented, not over. Number 6's count of 22 is also below the expected 25, so it too is under-represented.
- (a) 60 — Method: because the counter goes back each time, every draw is the same experiment, so the expected number of reds is the number of draws multiplied by P(red). Working: there are 8 red counters out of 20, so P(red) = 8 ÷ 20 = 0.4. Over 150 draws the expected number of reds is 150 × 0.4 = 60. Answer: about 60 red counters would be expected. The distractors: 90 uses P(blue) by mistake, 150 × 12 ÷ 20 = 90; 100 comes from writing the probability from the red to blue ratio of 8 to 12, giving 150 × 8 ÷ 12 = 100; 75 comes from treating red and blue as equally likely because there are two colours, which gives 150 ÷ 2 = 75.
Build your own mix at the worksheet builder.