International football rankings determine seeding, which determines draws, which affects who qualifies. When the formula could be exploited by fixture choice, the exploitation was rational.
Averaging created the wrong incentive
A ranking that averages points earned over a period rewards playing fewer matches if the expected value of an additional match is below a team's current average.
Under such a system, a highly ranked team can protect its position by playing less, and by choosing opponents whose defeat carries a high weighting.
Associations noticed. Scheduling decisions began to reflect ranking arithmetic alongside preparation needs, which is not what the ranking was supposed to measure.
Why the friendly became a problem
Friendlies carried weight in the formula, so the choice of opponent was a ranking decision as well as a sporting one.
Teams could improve their position by arranging matches against opposition with favorable weighting rather than against opposition that would prepare them.
Since seeding flows from rankings, the incentive fed back into competitive outcomes. A scheduling choice could change a World Cup group.
What a rating system does differently
A rating approach assigns each team a number and updates it after every match by the difference between the result and what the ratings predicted.
Because the update depends on expectation, beating a weaker team gains almost nothing and losing to one costs a great deal. Playing more matches is neither rewarded nor punished by itself.
The incentive to avoid matches disappears, which was the specific defect being addressed. The system is descended from chess rating methods and behaves similarly.
What the rewrite did not fix
Confederations play largely within themselves, so ratings across confederations are connected by relatively few matches and are correspondingly less certain.
Any rating system inherits this. Comparing teams that rarely meet is an extrapolation, and the number's precision overstates how much is actually known.
Competition weightings remain a judgment: how much more a tournament match counts than a friendly is set by administrators, not derived from data.
Why the transition was awkward
Switching formulas means the ranking series breaks. Positions before and after the change are not measuring the same thing, even though they are printed identically.
Governing bodies handled this by initializing the new system from the old ratings, which preserves continuity of appearance while the underlying method differs.
Historical ranking comparisons therefore cross a methodological seam. It is a familiar problem in sports record-keeping: the number survives the redefinition and the caveat does not.

