Probability · Basic probability
1 / 18
Relative frequency & expected outcomes
Estimating a probability from data — relative frequency is frequency divided by the number of trials — why more trials sharpen the estimate, how to test a dice or spinner for bias, and why the expected number is probability times trials.
Probability · Basic probability
Relative frequency & expected outcomes
Estimating a probability from data — relative frequency is frequency divided by the number of trials — why more trials sharpen the estimate, how to test a dice or spinner for bias, and why the expected number is probability times trials.
Why it works
Some probabilities you can work out by pure thought. A fair six-sided dice has six equally likely faces, so — no experiment needed. But now bend the dice, or drop a drawing pin, or ask "what is the probability this bus is late?". There is no symmetry to argue from, so there is nothing to calculate. The only way in is to watch what actually happens.That is what relative frequency is:
Drop a drawing pin times, see it land point-up times, and the relative frequency is . Because it is a part over a whole it always lands between and , exactly like a probability — and the relative frequencies of all the possible outcomes add to , because every trial produced exactly one of them.
Why does measuring tell you anything about probability? Because a probability is a long-run proportion. Saying a coin has is a claim about what happens when you flip it many, many times: the proportion of heads settles down near . So recording the proportion is not a second-best substitute for the "real" probability — it is a direct measurement of the very thing the probability describes. That is why relative frequency is also called experimental probability, and why it is only ever an estimate: you measured a finite stretch of the long run, not the whole of it.
Why more trials give a better estimate. Short runs are lumpy. Flip a fair coin times and getting or more heads happens about one time in six — a relative frequency of , nowhere near , out of a perfectly fair coin. Flip the same coin times and the number of heads is almost always between and , so the relative frequency is pinned between and . Nothing has changed about the coin. What changed is that a run of luck of a given size becomes a smaller and smaller proportion of the total as the total grows: five extra heads is half of ten flips but only a two-hundredth of a thousand. So the wobble shrinks, and the relative frequency closes in on the true probability.*The running relative frequency of heads for a fair coin, against the flat line . The luck does not go away; it just counts for less and less.*
Testing for bias. That gives you the whole method for deciding whether a dice or a spinner is fair. Work out the fair (theoretical) probability, work out the relative frequency your data actually gives, and compare — but weigh the comparison against the number of trials. Three sixes in rolls (relative frequency ) proves nothing; that gap sits well inside the ordinary lumpiness of a short run. Seventy-two sixes in rolls (relative frequency , against a fair value of ) is a different matter: over that many rolls a fair dice would give about sixes, and is a long way clear of . Equally, a relative frequency that is not exactly is not evidence of bias — you should never expect an exact match.
Expected number probability number of trials. If the proportion of successes settles near , then in trials the count of successes settles near . That is the entire derivation: a proportion of a total is that fraction times the total. So if , then in drops you expect about point-ups.
Two warnings about the word "expected". First, it does not have to be a whole number. The expected number of sixes in rolls of a fair dice is , and you obviously cannot roll a third of a six. The is the average over many sets of rolls, and an average of whole numbers need not itself be whole — in the same way that families average children. Do not round it unless the question tells you to. Second, it is a prediction, not a promise: you might get sixes this time, or .
The trap, made concrete: never average the probabilities. Ann spins a spinner times and gets green times. Ben spins the same spinner times and gets green times. Their separate estimates are and , and it is very tempting to split the difference:
Averaging the two fractions treats Ann's spins as carrying the same weight as Ben's — it quietly promotes each of her spins to three times the importance of each of his. Go back to the definition instead and pool the raw counts. Between them they spun times and saw green times, so the best estimate is
and now every one of the spins counts exactly once. It is also the more trustworthy estimate for the reason above: it rests on trials rather than or . (Averaging happens to give the right answer only when the two sets of trials are the same size — which is exactly when it saves you nothing.)