Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- 1.A drawing pin is dropped many times and lands either point up or point down. The relative frequency of landing point up is recorded as the experiment goes on: after 50 drops it is 0.720, after 200 drops it is 0.665, and after 1000 drops it is 0.638. The pin is to be dropped a further 2000 times. Work out the best estimate of the number of times it will land point up.
- 2.At a fun run, each runner finishes, retires or is disqualified, and cannot do more than one of these. The probability that a runner finishes is 68.5% and the probability that a runner retires is 24.75%. Work out the probability, as a percentage, that a runner is disqualified.
- 3.A survey found that, of 250 shoppers questioned, 68% said they had used a self-checkout in the past month. A different store expects 1400 shoppers this week. Using this relative frequency, work out how many of the 1400 shoppers would be expected to have used a self-checkout in the past month.
- 4.A Venn diagram shows two sets, P and Q, inside a universal set. n(P) = 34, n(Q) = 27, n(P ∩ Q) = 11, and n(ξ) = 90, where ξ is the universal set. Work out n((P ∪ Q)′), the number of elements in neither P nor Q.
- 5.A bag is known to contain red, blue and green counters in equal numbers, so the theoretical probability of taking each colour is 1/3. In 90 trials with replacement, red was taken 38 times, blue was taken 26 times and green was taken 26 times. Which colour is over-represented compared with its theoretical probability?
- 6.A weather station records whether it rains each day for 40 days: it rains on 9 of the days and does not rain on the rest. A local forecaster claims that the probability of rain on any day is 0.3. Work out the relative frequency of rain from the recorded data.
- 7.A phone network sends automatic text alerts to customers. On average, 1,500 alerts are sent each day, and the probability that a customer replies 'STOP' to an alert is 0.18. Work out how many replies of 'STOP' the network should expect over a 30-day month.
- 8.A factory checked 400 items. A frequency tree splits them into 250 items made by machine A and 150 items made by machine B. On machine A's branch, 15 of the items were faulty. On machine B's branch, 5 of the items were faulty. One of the 400 items is picked at random. Work out the probability that it is faulty. Give your answer as a fraction in its simplest form.
- 9.A biased spinner is spun 200 times. It lands on red 70 times, on blue 50 times, and on green 80 times. Using these results, work out the expected number of times the spinner does NOT land on red, in 500 spins of the same spinner.
- 10.A survey of 160 employees at a company recorded whether they cycle to work, using a frequency tree. The first branch splits them into 90 who work full-time and 70 who work part-time. Of the full-time employees, 27 cycle to work. Of the part-time employees, 14 cycle to work. Work out the probability that an employee, chosen at random from the 160, cycles to work. Give your answer as a fraction in its simplest form.
- 11.Two goalkeepers face penalty kicks. Elin saves 34% of the penalties she faces. Noah saves 11/32 of the penalties he faces. Which statement correctly compares them?
- 12.A fair six-sided dice is rolled 90 times. Work out how many more times you would expect it to land on a number less than 4 than on a 6.
- 13.A council surveyed 140 households about recycling. 85 of the households are in Zone A. 52 of the Zone A households recycle glass. 21 of the Zone B households do not recycle glass. Work out the probability that a household, chosen at random from the 140, recycles glass. Give your answer as a fraction in its simplest form.
- 14.A machine makes 4000 light bulbs a day and runs 5 days a week. Two inspectors test bulbs from this machine. Inspector A tests 40 bulbs and finds 4 faulty. Inspector B tests 500 bulbs and finds 30 faulty. Using the better of the two estimates, work out how many faulty bulbs the machine is expected to make in one week.
- 15.Two pupils each flip the same coin to estimate the probability of heads. Leah flips it 40 times and gets 24 heads. Ben flips it 60 times and gets 33 heads. By combining both pupils' results, work out the relative frequency of heads, giving your answer as a fraction in its simplest form.
Answer key
- (a) 1276 — Method: an unbiased relative frequency tends towards the theoretical probability as the number of trials increases, so use the record resting on the most trials, then multiply by the number of new trials. Working: the three records rest on 50, 200 and 1000 drops, so the most reliable is the one after 1000 drops, namely 0.638, and the run is indeed settling as the trials increase. The expected number of point up landings in 2000 further drops is 2000 × 0.638 = 1276. Answer: about 1276 times. The distractors: 1440 uses the earliest record, which rests on only 50 drops, giving 2000 × 0.720 = 1440; 1330 uses the middle record, treating 200 drops as a safe compromise when 1000 drops is better still, giving 2000 × 0.665 = 1330; 1348 comes from averaging the three records, since 0.720 + 0.665 + 0.638 = 2.023 and 2.023 ÷ 3 = 0.674, then 2000 × 0.674 = 1348, which gives the 50 drop record the same weight as the 1000 drop record.
- (a) 6.75% — Finishing, retiring and being disqualified are exhaustive, so the three percentages sum to 100%: 100% − 68.5% − 24.75% = 6.75%. Adding the two given percentages instead of subtracting them from 100% gives 68.5% + 24.75% = 93.25%, the combined probability of finishing or retiring, not of being disqualified. Subtracting only the retiring percentage from 100% and forgetting the finishing percentage gives 100% − 24.75% = 75.25%. Subtracting only the finishing percentage and forgetting the retiring percentage gives 100% − 68.5% = 31.50%.
- (a) 952 — 68% = 0.68. The relative frequency from the survey applies to the new group of 1400 shoppers, so the expected number is 0.68 × 1400 = 952. Working out 1 − 0.68 = 0.32 and applying that instead, 0.32 × 1400 = 448, finds the number who have NOT used a self-checkout, not the number who have. Applying 68% to the original sample size of 250 instead of the new total of 1400 gives 0.68 × 250 = 170. Applying the complement percentage to the original sample size, 0.32 × 250 = 80, compounds both mistakes.
- (b) 40 — n(P ∪ Q) = n(P) + n(Q) − n(P ∩ Q) = 34 + 27 − 11 = 50. The complement is everyone outside both sets: n((P ∪ Q)′) = 90 − 50 = 40. Adding P and Q without subtracting the overlap gives 34 + 27 = 61, so 90 − 61 = 29 double-subtracts the 11 who are in both. Reporting n(P ∪ Q) itself, 50, forgets to take the complement at all. Subtracting only n(P) from the universal set, 90 − 34 = 56, ignores set Q altogether.
- (c) Red — Theoretical probability is 1/3 ≈ 0.333 for each colour. Red's relative frequency is 38/90 ≈ 0.422, above 1/3, so red is over-represented. Blue's relative frequency is 26/90 ≈ 0.289, below 1/3, so blue is under-represented, not over. Green's relative frequency is also 26/90 ≈ 0.289, below 1/3 for the same reason. Since red's relative frequency clearly exceeds 1/3, it is not true that none of the colours are over-represented.
- (b) 0.225 — The relative frequency of rain is the number of rainy days out of all days recorded: 9 ÷ 40 = 0.225, which is noticeably less than the forecaster's claimed 0.3. Using the number of dry days, 40 − 9 = 31, as the denominator instead of the total of 40 gives 9 ÷ 31 = 0.29 (2 d.p.). Simply reporting the forecaster's claimed value, 0.3, without calculating anything from the data at all, ignores the recorded results completely. Misplacing the decimal point, treating 9 out of 40 as 9%, gives 0.09 instead of 0.225.
- (b) 8,100 — First find the total number of alerts sent in the month: 1,500 × 30 = 45,000. Then apply the probability of a 'STOP' reply: 45,000 × 0.18 = 8,100. Stopping after finding only one day's expected replies, 1,500 × 0.18 = 270, forgets to scale up to the whole month. Multiplying the number of days by the probability instead of by the daily total of alerts gives 30 × 0.18 = 5.4, which rounds to 5. Shifting the decimal point in the probability, using 0.018 instead of 0.18, gives 45,000 × 0.018 = 810.
- (d) 1/20 — Method: add the counts on the faulty end branches, divide by the total number of items in the experiment, then cancel. Working: the faulty items number 15 + 5 = 20, and 400 items were checked, so the probability is 20/400. Dividing the top and the bottom by 20 gives 1/20. Answer: the probability is 1/20. The distractors: 3/50 is 15/250 and comes from dividing machine A's faults by machine A's output, which is that machine's own fault rate rather than the probability for the whole batch; 1/30 is 5/150 and does the same on machine B's branch; 19/20 is 380/400 and gives the probability that the item picked is not faulty.
- (c) 325 — Method: first find the relative frequency of NOT landing on red from the 200 spins, then scale that up to 500 spins. Working: non-red results = 50 + 80 = 130, out of 200 spins, so P(not red) = 130 ÷ 200 = 0.65. Expected non-red results in 500 spins = 500 × 0.65 = 325. Answer: 325. Watch out: writing down 175 finds the expected number of RED results instead, 70 ÷ 200 × 500 = 175, answering the opposite of what was asked. Writing down 250 assumes landing red or not landing red must be a fair 50-50 split, but the spinner is biased and the actual results do not split evenly. And writing down 130 stops after finding how many of the 200 spins were non-red and forgets to scale that figure up to the 500 spins asked for.
- (b) 41/160 — In total, 27 + 14 = 41 of the 160 employees cycle to work, so the probability is 41/160 (41 and 160 share no common factor, so this is already in its simplest form). Writing 27/160 is wrong because it only counts the full-time cyclists and leaves out the 14 part-time cyclists. Writing 41/90 is wrong because it uses the full-time total (90) as the denominator instead of the whole survey (160). Writing 1/5 is wrong because it only uses the part-time branch, simplifying 14/70 to 1/5 and ignoring the full-time cyclists completely. The probability is 41/160.
- (c) Noah — 11/32 = 0.34375, above 34%. — Converting 11/32 to a decimal gives 11 ÷ 32 = 0.34375, which is greater than 34% (0.34), so Noah has the better save rate: 'Noah — 11/32 = 0.34375, above 34%.' Comparing the raw numbers 34 and 11 directly, without converting the fraction to the same form, gives 'Elin — 34 is bigger than 11.' Treating a larger denominator as meaning a bigger value, rather than smaller equal shares, gives 'Noah — 32 is a bigger denominator.' Rounding 34.375% to 34% to the nearest whole percent hides the difference and gives 'Equal — both round to 34% to the nearest percent.'
- (c) 30 — The numbers less than 4 are 1, 2 and 3, so the probability of that event is 3/6, and the expected count in 90 rolls is 90 × 3/6 = 45. The probability of rolling a 6 is 1/6, and the expected count is 90 × 1/6 = 15. The difference between the two expected counts is 45 − 15 = 30. A candidate who answers 45 has given the expected count for 'less than 4' only, forgetting to subtract the other expected count. A candidate who answers 15 has given the expected count for '6' only. A candidate who answers 36 has used a dice with 5 possible numbers instead of 6, giving 90 × 3/5 = 54 and 90 × 1/5 = 18, a difference of 36.
- (a) 43/70 — There are 140 − 85 = 55 Zone B households, and 55 − 21 = 34 of them recycle glass. In total, 52 + 34 = 86 households recycle glass, out of 140: 86/140 = 43/70. Writing 13/35 is wrong because 52/140 simplifies to 13/35, and 52 only counts Zone A, leaving out the 34 Zone B recyclers. Writing 17/70 is wrong because 34/140 simplifies to 17/70, and 34 only counts Zone B, leaving out the 52 Zone A recyclers. Writing 73/140 is wrong because it adds the 52 Zone A recyclers to the 21 Zone B households that do NOT recycle, mixing up two different groups instead of adding the two recycling groups. The probability is 43/70.
- (a) 1200 — Method: take the estimate from the larger sample, because an unbiased relative frequency tends towards the true probability as the sample grows, then multiply by the number of bulbs made in a week. Working: Inspector B tested 500 bulbs, far more than Inspector A's 40, so use B's relative frequency: 30 ÷ 500 = 0.06. A week's production is 4000 × 5 = 20000 bulbs. The expected number of faulty bulbs is 20000 × 0.06 = 1200. Answer: about 1200 faulty bulbs a week. The distractors: 2000 uses Inspector A's estimate, 4 ÷ 40 = 0.1, giving 20000 × 0.1 = 2000, and so rests on a sample of only 40 bulbs; 1600 comes from averaging the two estimates of 0.1 and 0.06 to get 0.08, and 20000 × 0.08 = 1600, which gives the small sample equal weight with the large one; 240 uses the right estimate but stops at a single day, 4000 × 0.06 = 240.
- (b) 57/100 — Pooling both trials: total heads = 24 + 33 = 57, total flips = 40 + 60 = 100, so the combined relative frequency is 57/100, which is already in its simplest form since 57 and 100 share no common factor. Averaging the two separate relative frequencies instead, (24/40 + 33/60) ÷ 2 = (0.6 + 0.55) ÷ 2 = 0.575 = 23/40, treats the two trials as equally weighted even though Ben made more flips, which is not correct. Using only Leah's data gives 24/40 = 3/5. Using only Ben's data gives 33/60 = 11/20.
Build your own mix at the worksheet builder.