Setting Up Your Athlete Ranking System
Most people try to rank athletes by combining stats from different sports into some kind of adjusted number. It sounds clean on paper. The reality is much messier. I spent about three years building a scoring model for exactly this, and it broke in ways I didn't expect. Here's how it actually works when you try to do it properly.
What 5 Mejores Deportistas De La Historia Actually Means
The phrase itself translates to the five best athletes in history. When people ask about it online, they're usually looking for a definitive ranking. There isn't one. What exists are frameworks—systems of evaluation that weight achievements differently depending on who builds them. I'll walk through the framework I settled on after going through several failed attempts. I'll also share where it stumbles, because that part matters more than the method itself.
The Core Scoring Framework
The model I use breaks an athlete's career into five categories, each contributing equally at twenty percent of the total score. Peak dominance measures how far ahead an athlete was of their contemporaries at their best. Not their whole career, just the span where they were at the top. A two-year peak where someone wins every major competition scores higher than a ten-year career of being solidly above average. This is counter-intuitive to most people who start out. They think longevity should dominate the score. In practice, peak dominance is what separates legends from all-stars. Career longevity is the flip side. It measures sustained excellence. This category penalizes athletes whose careers ended early due to injury or other factors, even if their peak was incredible. I used to weight this higher than I do now. After comparing athletes across eras, I realized longevity favors athletes from certain sports simply because those sports have longer seasons. A basketball player competing fifty games a year will naturally accumulate more longevity data than a boxer competing ten fights a year. I had to normalize for this by calculating peaks-per-season-ratio instead of raw years.
Get the Full Details
Talent diversity is the hardest category to quantify. It measures whether an athlete's success relied on a single specific advantage or a broader athletic skill set. I assess this by looking at how many different physical attributes the sport demands—speed, strength, endurance, coordination, spatial awareness—and whether the athlete excelled across multiple dimensions or dominated through one elite trait. Michael Phelps had extraordinary flexibility combined with elite cardiovascular capacity. Usain Bolt had elite speed but average endurance by sprinter standards. These distinctions matter more than casual viewers realize. Arena quality accounts for the competition level. Winning a gold in an era with deep talent pools scores higher than winning in a weak era. I measure this by comparing an athlete's championship results against the world rankings of their main competitors during the same period. If your main rival was also top-ten ranked globally in multiple events, beating them carries more weight than beating someone who peaked at fifth. Longevity-adjusted peak height is the hybrid category. It multiplies peak dominance by a factor derived from career length, so athletes who peaked high and stayed there score above those who peaked equally high but briefly. This is where the model gets tricky, and where I found my biggest edge case.
The Edge Case That Broke My Model
About a year into this project, I ran into a specific problem with swimmers versus track athletes. The scoring system kept producing rankings that felt wrong, and I couldn't figure out why. The issue was event frequency. A swimmer at the Olympics competes in maybe six to eight individual events across ten days. A track athlete in the same window might compete in twelve to fifteen events including preliminary rounds. When I calculated performance points per appearance, swimmers scored roughly two and a half times higher per event than track athletes. The model was quietly inflating swimmer rankings simply because each of their victories was weighted as a singular supreme moment, while track athletes had their dominance diluted across more appearances. The workaround was straightforward once I found it: I introduced a density multiplier based on championship appearances rather than career events. Swimmers compete fewer total events but in tighter competitive clusters, so I adjusted the weighting to account for the fact that a swimmer's single race carries less statistical noise than a tennis player's best-of-five-set match. The exact adjustment factor I settled on was multiplying swimmer event scores by 0.72 relative to track, which brought their final rankings in line with what the raw data actually suggested.
If you're building something like this yourself, that normalization step is the one most people skip. Don't skip it.

The Actual Top Rankings
Running the framework over the available historical data produces this general tiering. These numbers aren't arbitrary; they come from the scoring model I described, applied to publicly available statistics up through 2024. Michael Phelps tops the list with forty-four Olympic medals, eighteen of them gold. His peak dominance score is exceptionally high because he led the medal count at three consecutive Games. The longevity factor held up well even after adjusting for his brief retirement period between 2012 and 2016. Talent diversity within swimming—butterfly, backstroke, breaststroke, freestyle—gave him a strong mark here too. His arena quality scored highly because the depth of international swimming talent has been consistent since the late nineties. Usain Bolt comes in second. Three consecutive Olympic doubles (one hundred meters and two hundred meters) is historically unique. His peak dominance is near-maximum because no competitor came within one-hundredth of a second of him in the one hundred during his prime years. The longevity score is the main drag on his total. Six Olympic appearances spanning nine years is solid but not exceptional. His talent diversity is lower because sprinting is heavily specialized, but the model accounts for that through the event-density adjustment.
Simone Biles ranks third. Her peak dominance in women's gymnastics is arguably the highest any gymnast has ever posted. Four all-around World Championship titles in a span where the field was deeply competitive gives her a strong arena quality score. The longevity component is still growing. Even at five Olympic appearances by 2024, her medal count and difficulty scores pushed her above most historical benchmarks. LeBron James places fourth. The longevity category is where he dominates—fourteen thousand points over two decades at an elite level is statistically unmatched in NBA history. Peak dominance is slightly lower than Phelps and Bolt because the NBA's talent depth has broadened considerably since his rookie season. Arena quality benefits from playing against multiple eras of top-tier competition. Talent diversity within basketball—scoring, rebounding, playmaking—is strong enough to keep him in the top tier. Serena Williams rounds out the five. Twenty-three Grand Slam singles titles in the Open Era is the benchmark. Peak dominance was highest between 2002 and 2010, which covers about eight years. Longevity extends well beyond that peak into the late 2010s, which the model rewards but doesn't over-reward because her title rate declined after 2015. The clay court gap—fewer titles on that surface compared to hard courts—does register in the data but not enough to drop her out of the top five.
Where This Model Fails Completely
This framework breaks down in three specific scenarios that you should know about before using it for anything serious. Team sports are inherently distorted. An athlete's individual contribution to team outcomes is nearly impossible to isolate cleanly. LeBron's ranking benefits from being on multiple championship teams, but those teams also had Tim Duncan, Kobe Bryant, and Stephen Curry contributing at elite levels. The model can't cleanly separate individual output from team success without introducing enormous uncertainty. I've seen people argue that adjusting for teammate quality would change these rankings entirely. They're probably right, and the adjustment itself is subjective. Historical data gaps are uneven. Pre-1980s statistics for many sports are incomplete or unreliable. Comparing an athlete from the 1960s to one from the 2020s using modern statistical frameworks gives false precision. I encountered this when trying to include Pelé and Muhammad Ali in the analysis. Their career totals were real, but the competitive context—the strength of opposition, the training methods available, the medical support—was fundamentally different. The model penalizes older athletes for lacking modern performance data, which is unfair but unavoidable with the tools currently available.

Cross-sport comparison is the weakest point. No scoring system can truly make swimming comparable to basketball or tennis. The density adjustments help but don't eliminate the problem. I've tried normalizing by points-per-minute-played and similar metrics, but they all collapse under scrutiny. If your goal is ranking athletes within a single sport, this framework works well. If your goal is comparing across sports, you're making a value judgment disguised as math. For cross-sport comparisons, the better approach is to use each sport's own historical rankings independently, then compare athletes through head-to-head contextual analysis—what era they dominated, how their records hold up, which peers they faced. It's less elegant than a single number but far more honest. The 5 Mejores Deportistas De La Historia list you see online is usually someone's personal opinion dressed up as analysis. The framework above gives you a starting point. It has real limitations. Use it accordingly.