Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- 1.A quality inspector tests 145 light bulbs and finds that 33 are faulty. Work out the relative frequency of a bulb being faulty, as a percentage correct to 1 decimal place.
- 2.A weather station records whether it rains each day for 40 days: it rains on 9 of the days and does not rain on the rest. A local forecaster claims that the probability of rain on any day is 0.3. Work out the relative frequency of rain from the recorded data.
- 3.Dice A is a fair six-sided dice. Dice B is biased so that P(6) = 0.3. Dice A is rolled 150 times and Dice B is rolled 150 times. Work out how many more sixes you would expect from Dice B than from Dice A.
- 4.A biased spinner is spun 200 times. It lands on red 70 times, on blue 50 times, and on green 80 times. Using these results, work out the expected number of times the spinner does NOT land on red, in 500 spins of the same spinner.
- 5.A ferry company finds that on 20% of days the sea is rough. If the sea is rough, the probability that a crossing is delayed is 0.75. If the sea is calm, the probability that a crossing is delayed is 0.1. Given that a crossing was delayed, work out the probability that the sea was rough that day.
- 6.A test for a medical condition is given to 1000 people. 50 of the people have the condition and 950 do not. The test is positive for 45 of the 50 people who have the condition, and it is also positive for 95 of the 950 people who do not have the condition. One of the people whose test is positive is picked at random. Work out the probability that this person has the condition.
- 7.A factory checked 400 items. A frequency tree splits them into 250 items made by machine A and 150 items made by machine B. On machine A's branch, 15 of the items were faulty. On machine B's branch, 5 of the items were faulty. One of the 400 items is picked at random. Work out the probability that it is faulty. Give your answer as a fraction in its simplest form.
- 8.A fair six-sided dice is rolled 90 times. Work out how many more times you would expect it to land on a number less than 4 than on a 6.
- 9.A fair six-sided dice is rolled 300 times. Work out how many more times you would expect it to land on an even number than on a six.
- 10.A spinner can land on red, blue, green or yellow. The probability that it lands on red is 0.24 and the probability that it lands on yellow is 0.16. The probabilities of landing on blue and on green are equal, and each is called x. Work out the value of x.
- 11.A factory finds that the probability a randomly chosen light bulb is defective is 0.035. In a batch of 4,000 bulbs, work out how many bulbs you would expect to work correctly.
- 12.A game uses a fair spinner with 5 equal sections numbered 1 to 5. A player wins £12 if the spinner lands on 5, and wins nothing otherwise. It costs £2 to play. The game is played 250 times. Work out the expected profit for the players, in total, over the 250 games.
- 13.A train company runs 25 trains a day, every day. The probability that any one train is late is 0.08. Work out how many late trains the company should expect over a period of 4 weeks.
- 14.A four-colour spinner (red, blue, green, yellow) is spun repeatedly, and the relative frequency of landing on green is recorded as the number of spins increases: after 20 spins it is 0.350; after 200 spins it is 0.290; after 2000 spins it is 0.251. Using the result from 2000 spins as the best estimate of the probability, work out the number of times the spinner would be expected to land on green in a further 3000 spins.
- 15.Two goalkeepers face penalty kicks. Elin saves 34% of the penalties she faces. Noah saves 11/32 of the penalties he faces. Which statement correctly compares them?
Answer key
- (b) 22.8% — Relative frequency as a percentage is the faulty count divided by the total, then multiplied by 100: 33 ÷ 145 × 100 = 22.76, which rounds to 22.8%. Giving 33.0% as the answer uses the frequency, 33, directly as a percentage without dividing by the total 145 at all. Rounding 22.76 down to 22.7% instead of up applies the wrong rounding direction at the first decimal place. Finding the relative frequency of the bulbs that were NOT faulty first: 145 − 33 = 112, and 112 ÷ 145 × 100 = 77.24, answers the opposite question and rounds to 77.2%.
- (b) 0.225 — The relative frequency of rain is the number of rainy days out of all days recorded: 9 ÷ 40 = 0.225, which is noticeably less than the forecaster's claimed 0.3. Using the number of dry days, 40 − 9 = 31, as the denominator instead of the total of 40 gives 9 ÷ 31 = 0.29 (2 d.p.). Simply reporting the forecaster's claimed value, 0.3, without calculating anything from the data at all, ignores the recorded results completely. Misplacing the decimal point, treating 9 out of 40 as 9%, gives 0.09 instead of 0.225.
- (a) 20 — Dice A is fair, so its expected number of sixes is 150 × 1/6 = 25. Dice B has P(6) = 0.3, so its expected number of sixes is 150 × 0.3 = 45. The difference is 45 − 25 = 20. Adding the two expected values instead of subtracting them gives 25 + 45 = 70. Reporting Dice B's expected sixes on their own, without comparing to Dice A, gives 45. Using the fair probability 1/6 for Dice B as well as Dice A ignores the bias altogether, giving 150 × 1/6 = 25 for both dice and a difference of 0.
- (c) 325 — Method: first find the relative frequency of NOT landing on red from the 200 spins, then scale that up to 500 spins. Working: non-red results = 50 + 80 = 130, out of 200 spins, so P(not red) = 130 ÷ 200 = 0.65. Expected non-red results in 500 spins = 500 × 0.65 = 325. Answer: 325. Watch out: writing down 175 finds the expected number of RED results instead, 70 ÷ 200 × 500 = 175, answering the opposite of what was asked. Writing down 250 assumes landing red or not landing red must be a fair 50-50 split, but the spinner is biased and the actual results do not split evenly. And writing down 130 stops after finding how many of the 200 spins were non-red and forgets to scale that figure up to the 500 spins asked for.
- (c) 15/23 — Method: find P(rough and delayed) and the overall P(delayed) using the tree, then divide. Working: P(rough and delayed) = 0.2 × 0.75 = 0.15. P(calm and delayed) = 0.8 × 0.1 = 0.08. P(delayed) = 0.15 + 0.08 = 0.23. P(rough | delayed) = 0.15 ÷ 0.23 = 15/23. Answer: 15/23. Watch out: leaving the answer as 0.15 (3/20) gives P(rough and delayed) itself, without dividing by the overall probability that a crossing is delayed. Giving 0.75 (3/4) is the probability you were told to start with — that a crossing is delayed GIVEN the sea is rough — which is the reverse of what's being asked. And 0.2 (1/5) is just the original probability that the sea is rough, before you take the fact that the crossing was delayed into account.
- (c) 9/28 — Method: two linked steps. Total everyone whose test is positive, since the person picked is known to be one of them, then divide the positive tests that belong to people with the condition by that total. Working: 45 positive tests come from people who have the condition and 95 come from people who do not, so 140 tests are positive. The people with the condition give 45/140, and dividing the numerator and the denominator by 5 gives 9/28. Answer: the probability is 9/28. The distractors: 9/10 is 45/50, the probability of a positive test given that the person has the condition, which is the condition and the event the wrong way round and is the figure a candidate quotes when the two are confused; 9/200 is 45/1000, dividing by everyone tested rather than by the 140 who tested positive; 1/20 is 50/1000, the probability that a person has the condition before the test result is used at all.
- (d) 1/20 — Method: add the counts on the faulty end branches, divide by the total number of items in the experiment, then cancel. Working: the faulty items number 15 + 5 = 20, and 400 items were checked, so the probability is 20/400. Dividing the top and the bottom by 20 gives 1/20. Answer: the probability is 1/20. The distractors: 3/50 is 15/250 and comes from dividing machine A's faults by machine A's output, which is that machine's own fault rate rather than the probability for the whole batch; 1/30 is 5/150 and does the same on machine B's branch; 19/20 is 380/400 and gives the probability that the item picked is not faulty.
- (c) 30 — The numbers less than 4 are 1, 2 and 3, so the probability of that event is 3/6, and the expected count in 90 rolls is 90 × 3/6 = 45. The probability of rolling a 6 is 1/6, and the expected count is 90 × 1/6 = 15. The difference between the two expected counts is 45 − 15 = 30. A candidate who answers 45 has given the expected count for 'less than 4' only, forgetting to subtract the other expected count. A candidate who answers 15 has given the expected count for '6' only. A candidate who answers 36 has used a dice with 5 possible numbers instead of 6, giving 90 × 3/5 = 54 and 90 × 1/5 = 18, a difference of 36.
- (c) 100 — There are 3 even numbers on a fair dice (2, 4 and 6), so the probability of landing on an even number is 3/6 = 1/2, and 300 × 1/2 = 150. The probability of landing on a six is 1/6, so 300 × 1/6 = 50. The dice is expected to land on an even number 150 − 50 = 100 more times than on a six. Writing 50 is wrong because that is just the expected number of sixes on its own, without comparing it to the expected number of evens. Writing 150 is wrong because that is just the expected number of evens on its own, without subtracting the sixes. Writing 200 is wrong because it adds the two expected frequencies together (150 + 50 = 200) instead of finding the difference between them. The dice is expected to land on an even number 100 more times than on a six.
- (a) 0.30 — Red, blue, green and yellow are exhaustive, so all four probabilities sum to 1: 0.24 + 0.16 + x + x = 1, so 2x + 0.40 = 1, giving 2x = 0.60 and x = 0.30. Stopping at 2x = 0.60 without dividing by 2 leaves 0.60, the combined probability of both blue and green together, not the value of x on its own. Sharing the 0.60 across all four colours instead of just the two unknown ones gives 0.60 ÷ 4 = 0.15. Leaving out the 0.16 for yellow gives 2x + 0.24 = 1, so 2x = 0.76 and x = 0.38.
- (a) 3,860 — The probability a bulb works correctly is the complement of being defective: 1 − 0.035 = 0.965. Expected number working correctly = 0.965 × 4,000 = 3,860. Using the probability of being defective instead of its complement gives 4,000 × 0.035 = 140, the expected number of DEFECTIVE bulbs, not working ones. Shifting the decimal point in the complement, using 0.0965 instead of 0.965, gives 4,000 × 0.0965 = 386. Assuming every bulb works, ignoring the 0.035 probability altogether, gives the full batch of 4,000.
- (b) £100 — Over 250 games, the expected total winnings are 250 × (1/5) × £12 = £600, since a player wins on 1 of the 5 equally likely sections. The total cost of playing is 250 × £2 = £500. The players' expected profit is the winnings minus the cost: £600 − £500 = £100. Writing £500 is wrong because that is only the total cost of playing, without any winnings included. Writing £600 is wrong because that is only the total expected winnings, without subtracting what was paid to play. Writing £2,500 is wrong because it assumes a win on every single game (250 × £12 = £3,000) instead of using the 1-in-5 probability, then subtracts the cost: £3,000 − £500 = £2,500. The players' expected profit over the 250 games is £100.
- (d) 56 — Method: count the trials over the whole period first, then multiply the number of trials by the probability. Working: 4 weeks is 4 × 7 = 28 days, and at 25 trains a day that is 25 × 28 = 700 trains. The expected number of late trains is 700 × 0.08 = 56. Answer: about 56 late trains over the 4 weeks. The distractors: 2 is the expected number for a single day, 25 × 0.08 = 2, with the 28 days never brought in; 14 uses one week instead of four, 25 × 7 × 0.08 = 14; 644 is 700 − 56 and counts the trains expected to be on time.
- (a) 753 — The estimate from 2000 spins is the most reliable, since it comes from the largest sample size, so the best estimate of the probability is 0.251. Over a further 3000 spins, the expected number landing on green is 3000 × 0.251 = 753. Writing 1050 is wrong because 3000 × 0.350 = 1050 uses the estimate from only 20 spins, the LEAST reliable of the three. Writing 870 is wrong because 3000 × 0.290 = 870 uses the estimate from 200 spins rather than the more reliable 2000-spin estimate. Writing 750 is wrong because 3000 × 0.25 = 750 ignores the recorded data completely and simply assumes each of the 4 colours is equally likely. The best estimate is 753 expected green spins.
- (c) Noah — 11/32 = 0.34375, above 34%. — Converting 11/32 to a decimal gives 11 ÷ 32 = 0.34375, which is greater than 34% (0.34), so Noah has the better save rate: 'Noah — 11/32 = 0.34375, above 34%.' Comparing the raw numbers 34 and 11 directly, without converting the fraction to the same form, gives 'Elin — 34 is bigger than 11.' Treating a larger denominator as meaning a bigger value, rather than smaller equal shares, gives 'Noah — 32 is a bigger denominator.' Rounding 34.375% to 34% to the nearest whole percent hides the difference and gives 'Equal — both round to 34% to the nearest percent.'
Build your own mix at the worksheet builder.