Printable · GCSE Foundation · ages 14-16
Probability worksheet — GCSE Foundation
Fifteen questions across the probability statements at Foundation tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Foundation
MathsUKwww.geekhero.co.uk
- 1.A factory checked 400 items. A frequency tree splits them into 250 items made by machine A and 150 items made by machine B. On machine A's branch, 15 of the items were faulty. On machine B's branch, 5 of the items were faulty. One of the 400 items is picked at random. Work out the probability that it is faulty. Give your answer as a fraction in its simplest form.
- 2.A bag is known to contain red, blue and green counters in equal numbers, so the theoretical probability of taking each colour is 1/3. In 90 trials with replacement, red was taken 38 times, blue was taken 26 times and green was taken 26 times. Which colour is over-represented compared with its theoretical probability?
- 3.A quality inspector tests 145 light bulbs and finds that 33 are faulty. Work out the relative frequency of a bulb being faulty, as a percentage correct to 1 decimal place.
- 4.A four-colour spinner (red, blue, green, yellow) is spun repeatedly, and the relative frequency of landing on green is recorded as the number of spins increases: after 20 spins it is 0.350; after 200 spins it is 0.290; after 2000 spins it is 0.251. Using the result from 2000 spins as the best estimate of the probability, work out the number of times the spinner would be expected to land on green in a further 3000 spins.
- 5.A survey of 160 employees at a company recorded whether they cycle to work, using a frequency tree. The first branch splits them into 90 who work full-time and 70 who work part-time. Of the full-time employees, 27 cycle to work. Of the part-time employees, 14 cycle to work. Work out the probability that an employee, chosen at random from the 160, cycles to work. Give your answer as a fraction in its simplest form.
- 6.A factory finds that the probability a randomly chosen light bulb is defective is 0.035. In a batch of 4,000 bulbs, work out how many bulbs you would expect to work correctly.
- 7.In a game, Zara says the probability of scoring is 3/8, the probability of missing is 5/12 and the probability of a rebound is 1/6, and that these are the only three outcomes. Is Zara correct that her three probabilities are valid?
- 8.A two-way table records 200 members of a gym. 80 of the members are women and the rest are men. 60 of the women attend yoga classes. Work out the percentage of the women who attend yoga classes.
- 9.A drawing pin is dropped many times and lands either point up or point down. The relative frequency of landing point up is recorded as the experiment goes on: after 50 drops it is 0.720, after 200 drops it is 0.665, and after 1000 drops it is 0.638. The pin is to be dropped a further 2000 times. Work out the best estimate of the number of times it will land point up.
- 10.A library recorded loans for 250 books over a week, using a frequency tree. The first branch splits the books into 160 fiction and 90 non-fiction. Of the fiction books, 112 were returned on time and the rest were returned late. All the non-fiction books were returned on time. Work out the probability that a book, chosen at random from the 250, was returned late. Give your answer as a fraction in its simplest form.
- 11.A leisure centre has 150 members. 80 of the members are male and the rest are female. Every member uses either the pool or the gym, but not both. 66 members use the pool, and 35 of those pool users are male. Work out how many female members use the gym.
- 12.A survey found that, of 250 shoppers questioned, 68% said they had used a self-checkout in the past month. A different store expects 1400 shoppers this week. Using this relative frequency, work out how many of the 1400 shoppers would be expected to have used a self-checkout in the past month.
- 13.A fair six-sided dice is rolled once. Work out the probability that the score is an odd number. Give your answer as a decimal.
- 14.A two-way table records 150 customers at a café. 90 of the customers bought a hot drink and the rest did not. 54 of the hot-drink customers also bought a cake. 21 of the customers who did not buy a hot drink bought a cake. Work out the percentage of all 150 customers who bought a cake.
- 15.A survey of 90 people records which of two apps, X and Y, they use. 55 people use app X, 42 people use app Y, and 20 people use neither app. Work out the probability that a randomly chosen person uses both apps, giving your answer as a fraction in its simplest form.
Answer key
- (d) 1/20 — Method: add the counts on the faulty end branches, divide by the total number of items in the experiment, then cancel. Working: the faulty items number 15 + 5 = 20, and 400 items were checked, so the probability is 20/400. Dividing the top and the bottom by 20 gives 1/20. Answer: the probability is 1/20. The distractors: 3/50 is 15/250 and comes from dividing machine A's faults by machine A's output, which is that machine's own fault rate rather than the probability for the whole batch; 1/30 is 5/150 and does the same on machine B's branch; 19/20 is 380/400 and gives the probability that the item picked is not faulty.
- (c) Red — Theoretical probability is 1/3 ≈ 0.333 for each colour. Red's relative frequency is 38/90 ≈ 0.422, above 1/3, so red is over-represented. Blue's relative frequency is 26/90 ≈ 0.289, below 1/3, so blue is under-represented, not over. Green's relative frequency is also 26/90 ≈ 0.289, below 1/3 for the same reason. Since red's relative frequency clearly exceeds 1/3, it is not true that none of the colours are over-represented.
- (b) 22.8% — Relative frequency as a percentage is the faulty count divided by the total, then multiplied by 100: 33 ÷ 145 × 100 = 22.76, which rounds to 22.8%. Giving 33.0% as the answer uses the frequency, 33, directly as a percentage without dividing by the total 145 at all. Rounding 22.76 down to 22.7% instead of up applies the wrong rounding direction at the first decimal place. Finding the relative frequency of the bulbs that were NOT faulty first: 145 − 33 = 112, and 112 ÷ 145 × 100 = 77.24, answers the opposite question and rounds to 77.2%.
- (a) 753 — The estimate from 2000 spins is the most reliable, since it comes from the largest sample size, so the best estimate of the probability is 0.251. Over a further 3000 spins, the expected number landing on green is 3000 × 0.251 = 753. Writing 1050 is wrong because 3000 × 0.350 = 1050 uses the estimate from only 20 spins, the LEAST reliable of the three. Writing 870 is wrong because 3000 × 0.290 = 870 uses the estimate from 200 spins rather than the more reliable 2000-spin estimate. Writing 750 is wrong because 3000 × 0.25 = 750 ignores the recorded data completely and simply assumes each of the 4 colours is equally likely. The best estimate is 753 expected green spins.
- (b) 41/160 — In total, 27 + 14 = 41 of the 160 employees cycle to work, so the probability is 41/160 (41 and 160 share no common factor, so this is already in its simplest form). Writing 27/160 is wrong because it only counts the full-time cyclists and leaves out the 14 part-time cyclists. Writing 41/90 is wrong because it uses the full-time total (90) as the denominator instead of the whole survey (160). Writing 1/5 is wrong because it only uses the part-time branch, simplifying 14/70 to 1/5 and ignoring the full-time cyclists completely. The probability is 41/160.
- (a) 3,860 — The probability a bulb works correctly is the complement of being defective: 1 − 0.035 = 0.965. Expected number working correctly = 0.965 × 4,000 = 3,860. Using the probability of being defective instead of its complement gives 4,000 × 0.035 = 140, the expected number of DEFECTIVE bulbs, not working ones. Shifting the decimal point in the complement, using 0.0965 instead of 0.965, gives 4,000 × 0.0965 = 386. Assuming every bulb works, ignoring the 0.035 probability altogether, gives the full batch of 4,000.
- (d) No, because 3/8 + 5/12 + 1/6 = 23/24 — Using a common denominator of 24: 3/8 = 9/24, 5/12 = 10/24 and 1/6 = 4/24. Adding these numerators gives 9 + 10 + 4 = 23, so the three probabilities sum to 23/24, which is less than 1 — Zara is not correct. Adding the original numerators (3 + 5 + 1 = 9) over a denominator of 12 instead of converting each fraction properly gives 9/12 = 3/4, still less than 1 but the wrong fraction. Converting 1/6 to 5/24 instead of 4/24 (using the wrong scaling) makes the total 9/24 + 10/24 + 5/24 = 24/24 = 1, wrongly suggesting the probabilities are valid. Judging validity from the fact that each individual fraction lies between 0 and 1 ignores that an exhaustive set must sum to exactly 1, not merely contain valid individual values.
- (a) 75% — 'Percentage of the women' restricts the group to the 80 women, of whom 60 attend yoga: 60/80 = 0.75 = 75%. Dividing by the number of men (200 − 80 = 120) instead of the number of women gives 60/120 = 0.5 = 50%. Dividing by all 200 members instead of just the 80 women gives 60/200 = 0.3 = 30%. Using the 20 women who do NOT attend yoga (80 − 60) as the numerator instead of the 60 who do gives 20/80 = 0.25 = 25%.
- (a) 1276 — Method: an unbiased relative frequency tends towards the theoretical probability as the number of trials increases, so use the record resting on the most trials, then multiply by the number of new trials. Working: the three records rest on 50, 200 and 1000 drops, so the most reliable is the one after 1000 drops, namely 0.638, and the run is indeed settling as the trials increase. The expected number of point up landings in 2000 further drops is 2000 × 0.638 = 1276. Answer: about 1276 times. The distractors: 1440 uses the earliest record, which rests on only 50 drops, giving 2000 × 0.720 = 1440; 1330 uses the middle record, treating 200 drops as a safe compromise when 1000 drops is better still, giving 2000 × 0.665 = 1330; 1348 comes from averaging the three records, since 0.720 + 0.665 + 0.638 = 2.023 and 2.023 ÷ 3 = 0.674, then 2000 × 0.674 = 1348, which gives the 50 drop record the same weight as the 1000 drop record.
- (b) 24/125 — On the fiction branch, 160 − 112 = 48 books were returned late. None of the non-fiction books were late, so the total number of late books is 48, out of 250 books altogether: 48/250 = 24/125. Writing 3/10 is wrong because it divides the 48 late fiction books by the fiction total (160) instead of the whole library total (250). Writing 56/125 is wrong because 112/250 simplifies to 56/125, and 112 is the number of fiction books returned ON TIME, not late. Writing 9/25 is wrong because 90/250 simplifies to 9/25, and 90 is simply the number of non-fiction books, which has nothing to do with late returns. The probability is 24/125.
- (b) 39 — Method: put the counts into a two-way table and fill each missing cell by subtracting along a row or down a column. Working: the number of female members is 150 − 80 = 70. The pool column holds 66 members and 35 of them are male, so the number of female pool users is 66 − 35 = 31. Subtracting along the female row, 70 − 31 = 39 female members use the gym. Answer: 39 female members use the gym. The distractors: 45 comes from subtracting along the male row instead, 80 − 35 = 45, which counts male gym users; 31 is the female pool cell, written down one step before the gym cell; 84 is 150 − 66 and counts every gym user, male and female together.
- (a) 952 — 68% = 0.68. The relative frequency from the survey applies to the new group of 1400 shoppers, so the expected number is 0.68 × 1400 = 952. Working out 1 − 0.68 = 0.32 and applying that instead, 0.32 × 1400 = 448, finds the number who have NOT used a self-checkout, not the number who have. Applying 68% to the original sample size of 250 instead of the new total of 1400 gives 0.68 × 250 = 170. Applying the complement percentage to the original sample size, 0.32 × 250 = 80, compounds both mistakes.
- (c) 0.5 — Method: count the favourable outcomes, write them over the total number of equally likely outcomes and then divide to turn the fraction into a decimal. Working: the odd scores are 1, 3 and 5, which is 3 of the 6 equally likely scores, so the probability is 3/6, and 3 ÷ 6 = 0.5. Answer: 0.5, the middle of the 0 to 1 probability scale. The distractors: 0.33 comes from listing only 3 and 5 as odd and working out 2 ÷ 6; 0.17 comes from giving the probability of one particular odd score, 1 ÷ 6; 0.3 comes from writing '3 out of 6' as 0.3, reading the 3 straight off as tenths instead of dividing.
- (b) 50% — The total number of customers who bought a cake is 54 + 21 = 75, combining both hot-drink and non-hot-drink customers. As a percentage of all 150 customers, this is (75 ÷ 150) × 100 = 50%. Choosing 36% comes from only counting the hot-drink customers who bought a cake, (54 ÷ 150) × 100 = 36%, and forgetting the 21 non-hot-drink customers who also bought a cake. Choosing 14% comes from only counting the non-hot-drink customers who bought a cake, (21 ÷ 150) × 100 = 14%, and forgetting the 54 hot-drink customers who also bought a cake. Choosing 60% comes from dividing by the hot-drink total of 90 instead of the grand total of 150, (54 ÷ 90) × 100 = 60%.
- (b) 3/10 — The number who use at least one app is 90 − 20 = 70. Since 55 + 42 double-counts the overlap, n(X ∩ Y) = 55 + 42 − 70 = 27, so P(both) = 27/90 = 3/10. Forgetting to subtract the 20 who use neither, and using the full 90 as the union, gives 55 + 42 − 90 = 7, so 7/90. Reporting the probability of using X or Y (or both), 70/90 = 7/9, answers a different question about the union, not the overlap. Reporting the probability of using neither app, 20/90 = 2/9, is the complement of the union, not the intersection.
Build your own mix at the worksheet builder.