Comparing the strength of two domestic leagues sounds straightforward and is not. The clubs involved rarely play each other, which removes the basis for direct comparison.
Leagues are closed systems
Within a league, results establish a reliable internal order because every club plays every other repeatedly across a season.
Across leagues there is no such structure. Two clubs in different countries may never meet, and the only bridge is a small number of continental fixtures each year.
That handful of matches has to carry the entire weight of the comparison, which is far more than they can reliably support.
Continental results are a biased sample
Only the strongest clubs from each league qualify, so continental results compare the top of one league with the top of another rather than the leagues as wholes.
A league with a very strong leading club and a weak middle can outperform a deeper league in continental competition while being the weaker division overall.
Because qualification numbers differ by country, the sample is also uneven, with some leagues represented by several clubs and others by one.
Transfer flows carry information
Which direction players move, and at what fees and wages, reveals how clubs themselves value the standard of each league.
A consistent flow of players moving in one direction, with the receiving clubs paying more, is evidence about relative standard that does not depend on match results.
It is imperfect because wages reflect a league's commercial revenue as much as its playing quality, and those two are not the same thing.
Style differences confound direct comparison
Leagues differ in pace, physical contact, refereeing interpretation and tactical convention, and a club optimised for one environment may underperform in another.
This is why individual results between clubs from different leagues carry less information than they appear to, particularly in a single knockout tie.
Players moving between leagues face the same adjustment, which is why a transfer that looks like a clear upgrade on paper often takes a season to look like one on the pitch.
Rating models attempt the bridge indirectly
Statistical models rate every club on a common scale and use the few cross-league matches, plus player movement, to calibrate the leagues against each other.
The uncertainty in those calibrations is large and is usually reported alongside the estimate, though it rarely survives into public discussion of the results.
What the models do establish clearly is the ordering at the extremes. The difficulty is in the middle, where several leagues are close enough that the available evidence cannot separate them.

