A player's serve and return statistics from two tournaments look like the same measurements. They are frequently produced by different systems working to different rules.
Instrumentation is not uniform
Ball tracking, radar and automated line calling are installed on some courts and not others, often varying between show courts and outside courts at the same event.
A match played on an uninstrumented court yields only manually collected statistics. The published line for that match is thinner and produced by a different method.
Aggregating a player's season therefore combines measurements from several instrumentation regimes without any indication of which match came from which.
Radar placement changes serve speeds
Serve speed depends heavily on where the measurement is taken, since the ball decelerates continuously. Equipment sited differently produces systematically different readings.
Calibration practice also varies between events. Two tournaments can each be internally consistent while disagreeing with each other by a meaningful margin.
Fastest-serve lists compiled across events consequently mix incompatible readings. The leaderboard partly ranks tournaments rather than players.
Surface changes what the numbers mean
Rally length, first-serve effectiveness and net approach frequency all shift with surface. A player's aggregate season figures are a weighted average of the surfaces played.
A season heavier in clay events produces different numbers from a season heavier in hard courts for the same player playing the same way.
Careful analysis splits by surface for this reason. Headline career statistics almost never do, which makes them a poor basis for comparison between players with different schedules.
Definitions differ between providers
Categories such as forced errors, net points and return depth are coded to house definitions. Different events contract different providers, so the codebook changes between tournaments.
Even a mechanical statistic like an ace has edge cases, such as a serve touched but not controlled, and those are resolved differently by different operations.
The disagreements are individually small and cumulatively real. They set a floor on how precisely any cross-tournament comparison can be stated.
What comparison remains valid
Within-match comparison is sound, because both players were measured by the same system under the same conditions on the same day.
Within-tournament comparison is usually sound for matches on instrumented courts, which is why event-level analysis tends to be more defensible than season-level analysis.
Across a career, the honest framing is directional rather than precise. The numbers show shape and trend; they do not support the decimal places they are printed with.

