The same arithmetic in a rating system
The mark and the rating service are the same idea in two sports: keep a number, compare it with the opponent, move both by a share of the surprise. The share is a choice, and it decides how fast the number moves.
- Sample system
- a points rating
- Update weight
- 24
- Prior gap
- 100 points
- Move in one match
- 15.4 each
Sample R is one invented match. Team one is rated 1500 and team two 1600. The gap is turned into an expectation, the underdog wins, and both ratings move by the same amount in opposite directions.
Fifteen point four points from one match is a large move, and it is a direct consequence of the update weight of 24. Halve the weight and the same result moves each rating by 7.7, so a season of results changes the ranking half as quickly. The weight is not a detail of the system; in any practical sense, it is the system.
Why two systems disagree
Four decisions separate two rating systems that both work. The share of the surprise that is applied (the update weight), the period over which older results count (some systems decay results, some reset at the start of a season, some count everything ever played), whether a home or venue advantage is folded into the expectation before the update, and the population the number is calibrated against. Two systems can agree on the method and disagree on every number, which is exactly the position a reader of a speed figure is in when the standard time and the allowance are not published.
What the two systems share
Both are iterative: the published number is the previous number plus the last result's effect. Both are relative: they order a population rather than measure an absolute quantity. Both are stable by construction: the update weight is what prevents a single result from rewriting the number, and it is also what makes the number slow to notice a real change. And in both cases the numbers on the inside - the update weight for a rating, the standard time and the allowance for a figure - are the ones a reader cannot see.
- Ask what the gap is turned into before you ask what the numbers are.
- Find the update weight or its equivalent: it sets how quickly everything else moves.
- Check the rating period; a rating that counts everything ever is answering a different question from one that counts a season.
- Do not compare two systems' numbers directly. Compare their orderings, if anything.
- Treat the period of stability as evidence about the weight, not about the competitors.