Printable · GCSE Higher · ages 14-16
Probability worksheet — GCSE Higher
Fifteen questions across the probability statements at Higher tier. Choose the non-calculator filter to rehearse Paper 1, which counts for a third of the marks.
Calculator
Probability worksheet — GCSE Higher
MathsUKwww.geekhero.co.uk
- 1.A game uses a fair spinner with 5 equal sections numbered 1 to 5. A player wins £12 if the spinner lands on 5, and wins nothing otherwise. It costs £2 to play. The game is played 250 times. Work out the expected profit for the players, in total, over the 250 games.
- 2.A survey found that, of 250 shoppers questioned, 68% said they had used a self-checkout in the past month. A different store expects 1400 shoppers this week. Using this relative frequency, work out how many of the 1400 shoppers would be expected to have used a self-checkout in the past month.
- 3.A weather station records whether it rains each day for 40 days: it rains on 9 of the days and does not rain on the rest. A local forecaster claims that the probability of rain on any day is 0.3. Work out the relative frequency of rain from the recorded data.
- 4.Dice A is a fair six-sided dice. Dice B is biased so that P(6) = 0.3. Dice A is rolled 150 times and Dice B is rolled 150 times. Work out how many more sixes you would expect from Dice B than from Dice A.
- 5.A basketball player takes two free throws, and the throws are independent. The probability of scoring on each throw is 0.6, and the probability of missing is 0.4. Work out the probability that she scores exactly one of the two throws.
- 6.A fair six-sided dice is rolled 90 times. Work out how many more times you would expect it to land on a number less than 4 than on a 6.
- 7.A two-way table records 150 customers at a café. 90 of the customers bought a hot drink and the rest did not. 54 of the hot-drink customers also bought a cake. 21 of the customers who did not buy a hot drink bought a cake. Work out the percentage of all 150 customers who bought a cake.
- 8.At a fair, a game costs £2 to play. The probability of winning the game is 0.15, and each win pays out £10. Amir plays the game 200 times. Work out how much money Amir should expect to lose in total.
- 9.An ordinary six-sided dice, numbered 1 to 6, is rolled 30 times and lands on a 6 seven times. Ravi says the theoretical probability of rolling a 6 and the relative frequency of rolling a 6 in this trial are the same number. Is Ravi right?
- 10.A bag is known to contain red, blue and green counters in equal numbers, so the theoretical probability of taking each colour is 1/3. In 90 trials with replacement, red was taken 38 times, blue was taken 26 times and green was taken 26 times. Which colour is over-represented compared with its theoretical probability?
- 11.A café offers 3 types of sandwich, cheese, ham and egg, and 4 types of drink, tea, coffee, juice and water. A customer chooses one sandwich and one drink at random. Work out the probability that the customer chooses egg and water.
- 12.In a survey of 200 households, 120 have a garden and 80 own a dog. 54 of the households have a garden and own a dog. Work out the probability that a household owns a dog given that it has a garden, and compare it with the probability that a household picked from the whole survey owns a dog.
- 13.A train company runs 25 trains a day, every day. The probability that any one train is late is 0.08. Work out how many late trains the company should expect over a period of 4 weeks.
- 14.150 people at a gym were asked whether they prefer weight training or cardio; each person chose exactly one. 84 of the people are women. 50 of the women prefer cardio. 40 of the men prefer weight training. Work out the probability that a person, chosen at random from the 150, prefers weight training. Give your answer as a fraction in its simplest form.
- 15.A biased spinner is spun 40 times and lands on red 16 times. It is then spun a further 60 times and lands on red 21 times. Work out the best estimate of the probability that the spinner lands on red, using the results of all 100 spins together.
Answer key
- (b) £100 — Over 250 games, the expected total winnings are 250 × (1/5) × £12 = £600, since a player wins on 1 of the 5 equally likely sections. The total cost of playing is 250 × £2 = £500. The players' expected profit is the winnings minus the cost: £600 − £500 = £100. Writing £500 is wrong because that is only the total cost of playing, without any winnings included. Writing £600 is wrong because that is only the total expected winnings, without subtracting what was paid to play. Writing £2,500 is wrong because it assumes a win on every single game (250 × £12 = £3,000) instead of using the 1-in-5 probability, then subtracts the cost: £3,000 − £500 = £2,500. The players' expected profit over the 250 games is £100.
- (a) 952 — 68% = 0.68. The relative frequency from the survey applies to the new group of 1400 shoppers, so the expected number is 0.68 × 1400 = 952. Working out 1 − 0.68 = 0.32 and applying that instead, 0.32 × 1400 = 448, finds the number who have NOT used a self-checkout, not the number who have. Applying 68% to the original sample size of 250 instead of the new total of 1400 gives 0.68 × 250 = 170. Applying the complement percentage to the original sample size, 0.32 × 250 = 80, compounds both mistakes.
- (b) 0.225 — The relative frequency of rain is the number of rainy days out of all days recorded: 9 ÷ 40 = 0.225, which is noticeably less than the forecaster's claimed 0.3. Using the number of dry days, 40 − 9 = 31, as the denominator instead of the total of 40 gives 9 ÷ 31 = 0.29 (2 d.p.). Simply reporting the forecaster's claimed value, 0.3, without calculating anything from the data at all, ignores the recorded results completely. Misplacing the decimal point, treating 9 out of 40 as 9%, gives 0.09 instead of 0.225.
- (a) 20 — Dice A is fair, so its expected number of sixes is 150 × 1/6 = 25. Dice B has P(6) = 0.3, so its expected number of sixes is 150 × 0.3 = 45. The difference is 45 − 25 = 20. Adding the two expected values instead of subtracting them gives 25 + 45 = 70. Reporting Dice B's expected sixes on their own, without comparing to Dice A, gives 45. Using the fair probability 1/6 for Dice B as well as Dice A ignores the bias altogether, giving 150 × 1/6 = 25 for both dice and a difference of 0.
- (a) 0.48 — There are two ways to score exactly one throw: scoring on the first and missing the second, 0.6 × 0.4 = 0.24, or missing the first and scoring the second, 0.4 × 0.6 = 0.24. Adding these gives 0.24 + 0.24 = 0.48. Choosing 0.24 comes from working out only one of the two paths and forgetting the other one also gives exactly one score. Choosing 0.36 comes from working out the probability of scoring BOTH throws, 0.6 × 0.6 = 0.36, instead of exactly one. Choosing 0.84 comes from working out the probability of scoring AT LEAST one throw, 1 − 0.4 × 0.4 = 0.84, instead of exactly one.
- (c) 30 — The numbers less than 4 are 1, 2 and 3, so the probability of that event is 3/6, and the expected count in 90 rolls is 90 × 3/6 = 45. The probability of rolling a 6 is 1/6, and the expected count is 90 × 1/6 = 15. The difference between the two expected counts is 45 − 15 = 30. A candidate who answers 45 has given the expected count for 'less than 4' only, forgetting to subtract the other expected count. A candidate who answers 15 has given the expected count for '6' only. A candidate who answers 36 has used a dice with 5 possible numbers instead of 6, giving 90 × 3/5 = 54 and 90 × 1/5 = 18, a difference of 36.
- (b) 50% — The total number of customers who bought a cake is 54 + 21 = 75, combining both hot-drink and non-hot-drink customers. As a percentage of all 150 customers, this is (75 ÷ 150) × 100 = 50%. Choosing 36% comes from only counting the hot-drink customers who bought a cake, (54 ÷ 150) × 100 = 36%, and forgetting the 21 non-hot-drink customers who also bought a cake. Choosing 14% comes from only counting the non-hot-drink customers who bought a cake, (21 ÷ 150) × 100 = 14%, and forgetting the 54 hot-drink customers who also bought a cake. Choosing 60% comes from dividing by the hot-drink total of 90 instead of the grand total of 150, (54 ÷ 90) × 100 = 60%.
- (b) £100 — Method: find the expected number of wins, turn that into the expected pay out, then compare it with what the games cost. Working: the expected number of wins is 200 × 0.15 = 30. Each win pays £10, so the expected pay out is 30 × 10 = 300 pounds. Playing 200 times at £2 a go costs 200 × 2 = 400 pounds. The expected loss is 400 − 300 = 100 pounds. Answer: Amir should expect to be about £100 down. The distractors: £300 is the expected winnings on their own, with the cost of playing never taken off; £400 is the total cost of playing, with the winnings never taken off; £700 comes from adding the two totals, 400 + 300 = 700, instead of subtracting one from the other.
- (c) No — 7/30 is the relative frequency; theory stays 1/6. — The theoretical probability of rolling a 6 on an ordinary dice is fixed at 1/6, worked out from the number of equally likely outcomes, and does not change however the dice is actually rolled. The relative frequency from this trial is 7/30, found from what happened in these particular 30 rolls. Since 7/30 and 1/6 are different numbers, the correct statement is 'No — 7/30 is the relative frequency; theory stays 1/6.' Assuming the two values must always match because they describe the same event gives 'Yes — relative frequency always equals theory.' Believing that an observed result redefines the theoretical probability gives 'Yes — the theoretical probability has now become 7/30.' Refusing to work out either value at all gives 'Neither can be found — 30 rolls is too few to tell', which ignores that both numbers CAN be calculated from the information given.
- (c) Red — Theoretical probability is 1/3 ≈ 0.333 for each colour. Red's relative frequency is 38/90 ≈ 0.422, above 1/3, so red is over-represented. Blue's relative frequency is 26/90 ≈ 0.289, below 1/3, so blue is under-represented, not over. Green's relative frequency is also 26/90 ≈ 0.289, below 1/3 for the same reason. Since red's relative frequency clearly exceeds 1/3, it is not true that none of the colours are over-represented.
- (d) 1/12 — Method: list the full possibility space of sandwich-and-drink pairs, then divide the one matching pair by the size of the whole space. Working: there are 3 × 4 = 12 equally likely sandwich-and-drink pairs, and exactly one of them is egg and water. Answer: 1/12. Watch out: writing down 1/7 comes from adding the two counts, 3 + 4 = 7, instead of multiplying them to build the possibility space. Writing down 1/3 uses only the chance of choosing egg out of 3 sandwiches and ignores the drink altogether. And writing down 1/4 uses only the chance of choosing water out of 4 drinks and ignores the sandwich altogether.
- (a) 0.45, different from 0.4 for all the households — Method: work out the probability inside the restricted group of garden owners, then work out the probability across the whole survey, and compare the two. Working: 54 of the 120 households with a garden own a dog, so the conditional probability is 54 divided by 120, which is 0.45. Across the whole survey 80 of the 200 households own a dog, which is 0.4. Since 0.45 is not 0.4, having a garden changes the chance of owning a dog and the two events are not independent. Answer: 0.45, different from 0.4 for all the households. The distractors: 0.27 is 54/200, dividing the households with both by the whole survey instead of by the 120 with a garden; 0.675 is 54/80, the probability that a household has a garden given that it owns a dog, which is the condition and the event the wrong way round; 0.4 is 80/200, the probability of owning a dog with the garden information never used, which is why that route also reports no difference.
- (d) 56 — Method: count the trials over the whole period first, then multiply the number of trials by the probability. Working: 4 weeks is 4 × 7 = 28 days, and at 25 trains a day that is 25 × 28 = 700 trains. The expected number of late trains is 700 × 0.08 = 56. Answer: about 56 late trains over the 4 weeks. The distractors: 2 is the expected number for a single day, 25 × 0.08 = 2, with the 28 days never brought in; 14 uses one week instead of four, 25 × 7 × 0.08 = 14; 644 is 700 − 56 and counts the trains expected to be on time.
- (b) 37/75 — There are 150 − 84 = 66 men. 84 − 50 = 34 women prefer weight training, and 40 men prefer weight training, so 34 + 40 = 74 people in total prefer weight training, out of 150: 74/150 = 37/75. Writing 4/15 is wrong because it only counts the men who prefer weight training (40/150, simplified), leaving out the 34 women. Writing 17/75 is wrong because it only counts the women who prefer weight training (34/150, simplified), leaving out the 40 men. Writing 37/42 is wrong because it uses the number of women (84) as the denominator instead of the whole gym (150) — 74/84 simplifies to 37/42, but that is not a probability out of the whole group. The probability is 37/75.
- (c) 0.37 — Method: pool the two runs into one combined set of results, then find the relative frequency of red across all of the spins together. Working: total reds = 16 + 21 = 37. Total spins = 40 + 60 = 100. Relative frequency = 37 ÷ 100 = 0.37. Answer: 0.37. Watch out: writing down 0.40 uses only the first run, 16 ÷ 40, and throws away the extra evidence from the second 60 spins. Writing down 0.35 uses only the second run, 21 ÷ 60, and throws away the first run instead. And writing down 0.375 averages the two runs' separate rates, (0.40 + 0.35) ÷ 2, which treats a run of 40 spins and a run of 60 spins as equally weighted, when pooling the actual counts gives the larger run its fair share of influence.
Build your own mix at the worksheet builder.