Why Everyone Keeps Arguing About The Same Historical Figures
I was reading through some forums lately and noticed the same pattern over and over. People get into heated arguments about whether certain historical figures deserve monuments or infamy, and nobody ever seems to actually agree on the methodology for deciding that. It's exhausting to watch. I've spent years dealing with this topic in academic circles, so here's my take on how to actually think about controversial people in history without losing your mind. The basic framework most people skip over is that a controversial historical figure isn't someone who did bad things. Everyone who mattered in the past did bad things by modern standards. The real definition is simpler and more annoying: a controversial historical figure is someone whose net contribution to their field or era is still being argued about because different legitimate frameworks produce opposite answers. Using a utilitarian framework, you get one ranking. Using a virtue ethics framework, you get another. Using a Marxist materialist framework, you get yet a third. The controversy isn't a bug in our understanding. It's the system working as designed. The problem most people run into when trying to evaluate these figures is that they pick one ethical framework and treat it as objective truth. I ran into this specific issue when a student asked me to compare Gandhi and Churchill side by side for a paper. He wanted a clean verdict. I explained that Gandhi looks brilliant through a decolonization framework but terrible through a public health utilitarian lens, and Churchill looks like a war hero through a military-strategic framework but a brutal imperialist through an anti-colonial framework. The student got frustrated and said I was just dodging the question. He wasn't wrong. I was dodging the question. There isn't a clean answer here, and pretending otherwise is where most people go wrong.
The practical workaround I ended up using was to force him to assign relative weights to three criteria — military impact, moral violations, and cultural legacy — and then score each figure against those criteria himself. When he gave military impact a weight of 0.5 and moral violations a weight of 0.2 and cultural legacy a weight of 0.3, Gandhi still scored lower than Churchill overall. When he swapped the weights to make moral violations 0.5 and military impact 0.2, the result flipped. That's when the whole exercise actually became useful. Not because we found the right answer, but because we exposed what his actual priorities were.
The Framework That Actually Works
Here's the method I use when I need to evaluate any historical figure without talking myself in circles. It's not elegant. It takes more time than most people want to invest, but it saves you from producing nonsense conclusions. Step one is to pick your temporal frame. This matters more than you'd think. Most controversies about controversial people in history come from applying present-day moral standards retroactively without acknowledging that the application is anachronistic. Yes, that's a valid point to make. But you have to decide whether you're evaluating the person by the standards of their own era or by modern standards. Write that down at the top of your analysis. Don't assume you'll remember later. Step two is to identify the major domains of their impact. For a figure like Columbus, that's navigation and exploration, colonial policy, treatment of indigenous populations, and economic impact on Europe. For a figure like Tesla, that's engineering innovation, business dealings, personal conduct, and cultural mythology. You map out every domain before you start evaluating anything inside them. This prevents the classic error where someone does great work in one area and the evaluator lets that glow bias their judgment across every other area.
Get the Full Details

Step three is to separate factual claims from interpretive claims. "Nixon authorized the covert bombing of Cambodia" is a factual claim. It can be verified against declassified documents. "Nixon's Cambodia bombing was a war crime" is an interpretive claim. It depends on your legal and moral framework. Beginners constantly confuse these two categories and then get defensive when people challenge their conclusions. The factual claim is settled history. The interpretive claim is where the actual debate lives. The edge case that trips everyone up is what I call the archival silence problem. Sometimes the record is genuinely incomplete. I encountered this when researching a mid-level Nazi administrator whose postwar record suggested he was a bureaucratic clerk, but declassified files from the 1990s showed he'd been actively involved in deportation logistics. The controversy isn't just about interpretation here. It's about whether we can make moral judgments on incomplete evidence. My workaround was to publish both versions with explicit confidence intervals — high confidence on the documented facts, low confidence on the inferred ones, and clearly labeling what was speculation versus what was documented. Readers can then decide how much weight to give each category.
Common Pitfalls That Ruin These Evaluations
The first pitfall is the hagiography trap. When a historical figure is associated with something you personally value — a scientific discovery, a civil rights victory, a military triumph — you become blind to their documented failures in other areas. This isn't a moral failing. It's a cognitive bias that everyone has. The only antidote is to explicitly read sources that critique the figure you admire, not to find flaws for their own sake, but to calibrate your own judgment against counter-evidence you might otherwise ignore. The second pitfall is the inverse version of the same problem. When a figure is associated with something you personally despise — colonialism, slavery, authoritarianism — you tend to credit every negative claim about them without checking whether the claim is well-sourced. Historians call this the "guilt by association" amplification effect. I've seen reputable journals publish debunked claims about figures simply because those claims confirmed the author's pre-existing negative view. The same calibration principle applies here. Read the most favorable sources about people you dislike. Not to convert yourself. To test whether your own criticism holds up against their best arguments. The third pitfall is the false equivalence problem. This shows up constantly in online debates about controversial people in history. Someone will argue that because Figure A and Figure B both committed atrocities, they are morally equivalent. This is almost never true. The scale, intent, historical context, and institutional power behind each person's actions are usually different in ways that matter. Equating them doesn't make you balanced. It makes your analysis useless.
What This Approach Can't Do
Be honest about the limitations. This framework will not give you a definitive moral ranking of any historical figure. It will not resolve the controversy. It will not make people agree with you. In fact, it usually makes things messier because it exposes how many assumptions went into your original position. If you're looking for a clean thesis statement or a talking point to win an argument, this method will disappoint you. It's designed for understanding, not for debating. The framework also breaks down completely when applied to figures whose historical record is so thin that we can't reliably distinguish fact from myth. Romulus. King Arthur. Some of the earliest Mesopotamian rulers. No amount of careful methodology can rescue an evaluation built on foundationally uncertain evidence. You have to acknowledge that limit and move on rather than pretending confidence where none exists. Finally, the approach assumes you have access to primary sources and competent secondary scholarship. That's not true for everyone. People without university library access or the ability to read source languages are at a structural disadvantage here. The internet has improved this dramatically, but significant gaps remain, and anyone claiming this method works equally well for everyone is being dishonest about how knowledge distribution actually functions.

I keep coming back to the student and the Gandhi-Churchill assignment. He never wrote the paper the way I hoped he would. He submitted something shorter and more conventional. But two years later he emailed me to say he'd started using that weighted-criteria framework for evaluating current political figures and it had genuinely changed how he thought about news. That's not a conclusion. It's just what happened.