Justin’s Notes

Righteousness as Epistemic Anesthetic: How Feeling Like the Good Guy Disables Scientific Self-Correction

2026-08-05 · AI-generated insight

McIntyre lists "ideological infection" as one of eight ailments plaguing the social sciences, treating it as a parallel problem to fuzzy concepts and cherry-picking. But the other three notes suggest it deserves special status — because moral conviction doesn't just motivate bias, it disables the immune system that would normally catch it.

Inzlicht provides the confession that names the mechanism: "feeling like I was on the right side of history was a kind of moral cover. Maybe I felt less wrong about the p-hacking because I was promoting social justice." This is crucial. Cherry-picking for career advancement still triggers guilt, which invites scrutiny. Cherry-picking for justice feels like virtue, which suppresses it. The corrupted conscience doesn't just permit the error — it reframes the error as service.

Warne's account of Gould delivers the mechanism's most exquisite irony. Gould accused Morton of unconsciously fudging skull data to match his racial prejudices — and then, according to the 2011 re-analysis, apparently did exactly that himself, manipulating the same data to fit anti-racist conclusions. Gould's own thesis, that bias operates below awareness, applied to him with full force. But notice the asymmetry in detection: Morton's alleged sins were exposed within a framework eager to find them; Gould's took decades, because who audits the anti-racist?

The Coon episode shows the collective version. The anthropologists who condemned Putnam's book without reading it presumably felt they were defending science against prejudice. Their righteousness made examination itself seem unnecessary — even suspect, as Coon's objection was. Condemnation without examination is only psychologically possible when the condemners are certain they're the good guys.

This reframes McIntyre's prescription. Medicine escaped its "dark ages" through methods that assume the physician might be wrong. But bloodletting doctors merely believed they were correct; they didn't believe error would make them evil. A field where certain conclusions carry moral valence faces a harder problem: its practitioners have Inzlicht's incentive not just to find the favored answer, but to feel righteous while not looking too hard at how they found it.

Woven from these notes

Many scholars have criticized The Mismeasure of Man periodically throughout its 38-year history. For example, James T. Sanders stated that Gould’s attempt to link his argument to anti-racism was a ploy to smear intelligence scholars and Gould’s enemies as evil people. Arthur Jensen argued in 1982 that Gould misrepresented Jensen’s ideas and often demolished strawmen that no intelligence scholar believes, including the boogeyman of “biological determinism.” John Carroll showed that Gould understood neither the purpose nor interpretation of factor analysis (a statistical procedure often used to evaluate data from psychological tests) and that Gould’s attacks on factor analysis do nothing to alter the importance of intelligence tests, nor the mass of evidence—impossible to dispute—that they predict real-life outcomes. Most criticism of The Mismeasure of Man was confined to the recherché world of psychologists who study intelligence. However, a new debate opened up in 2011 when a team of anthropologists argued that Gould’s analysis of the data on cranium measurements from 19th century scientist Samuel George Morton was flawed. Gould cast Morton as a racist who fudged his data to match his beliefs about white racial superiority because of a supposed larger skull capacity. Instead, the anthropologists argued, it was Gould who manipulated the data to support his biases. This ignited a series of follow-up articles in the scholarly literature by authors taking a variety of positions regarding Morton’s data and Gould’s interpretations. Weisberg believed that the re-analysis was flawed and Gould was mostly correct. Kaplan and his colleagues claimed that Morton’s interpretations were flawed, but that Gould was incorrect in believing that he could discern Morton’s actions and motivations. Finally, Mitchell believed that Morton’s data were accurate and that the interpretations were colored by the racism of the era, but the claim that Morton subtly manipulated the data was a fiction created by Gould.
— Russell T. Warne The Mismeasurements of Stephen Jay Gould
When so many studies fail to be replicated, or draw different conclusions from the same set of facts, it does not instill confidence in the social sciences. Whether this is because of sloppy methodology, ideological infection, or other problems, the result is that even if there are right and wrong answers to many of our questions about human action, most social scientists are not yet in a position to find them. It is not that none of the work in social science is rigorous enough, but when policymakers (and sometimes even other researchers) are not sure which results are reliable, it drives down the status of the entire field. If medicine could break with its barbarous past, isn’t the same path open to the social sciences? For years, many have argued that if they could emulate the “scientific method” of the natural sciences, they too could become more scientific. But this simple advice faces several problems. Among the issues that plague contemporary social scientific research: Too much theory: A number of social scientific studies propose answers that have not been tested against evidence. The classic example here is neoclassical economics, where a number of simplifying assumptions — perfect rationality, perfect information — resulted in beautiful quantitative models that had little to do with actual human behavior. Lack of experimentation/data: Except for social psychology and the newly emerging field of behavioral economics, much of social science still does not rely on experimentation, even where it is possible. For example, it is sometimes offered as justification for putting sex offenders on a public database that doing so reduces the recidivism rate. This must be measured, though, against what the recidivism rate would have been absent the Sex Offender Registry Board (SORB), which is difficult to measure and has produced varying answers. This exacerbates the difficulty in (1), whereby favored theoretical explanations are accepted even when they have not been tested against any experimental evidence. Fuzzy concepts: Some social scientific studies can lead to misleading conclusions because of the use of “proxy” concepts for what one really wishes to measure. A recent example includes measuring “warmth” as a proxy for “trustworthiness,” in which researchers assumed — on the basis of studies which show that we are more likely to trust someone whom we perceive to be “on our side” — that perceptions of scientists as “cold” meant that they would be less trustworthy as well. But the two concepts may not be interchangeable. Ideological infection: This problem is rampant throughout the social sciences, especially on topics that are politically charged. Two ongoing examples are the bastardization of empirical work on the deterrence effect of capital punishment and the effectiveness of gun control on mitigating crime. If one knows in advance what one wants to find, one will likely find it. Cherry picking: The use of statistics allows multiple “degrees of freedom” to scientific researchers, but this is the most likely to be abused. In studies on immigration, for instance, a great deal of the difference between them is a result of alternative ways of counting the “costs” incurred by immigration. This is obviously also related to (4) above. If we know our conclusion, we may shop for the data to support it. Lack of data sharing: As the evolutionary biologist Robert Trivers reports in Psychology Today, there are numerous documented cases of researchers failing to share their data in psychological studies, despite a requirement from APA-sponsored journals to do so. When data were later analyzed, errors were found most commonly in the direction of the researcher’s hypothesis. Lack of replication: Psychology is undergoing a reproducibility crisis. One might validly argue that the initial finding that nearly two-thirds of psychology studies were irreproducible was overblown, but it is nonetheless shocking that most studies are not even attempted to be replicated. This can lead to difficulties, where errors can sneak through. Questionable causation: It is gospel in statistical research that “correlation does not equal causation,” yet some social scientific studies continue to highlight provocative results of questionable value. One recent sociological study, for instance, found that matriculating at a selective college was correlated with parental visitation at art museums, without explicitly suggesting that this was likely an artifact of parental income.
— Lee McIntyre To Fix the Social Sciences, Look to the “Dark Ages” of Medicine
Carleton Coon was one of the greatest anthropologists of the 20th century and a champion of the study of racial differences. In 1961 he was elected president of the American Anthropological Association. It was during this year that Carleton Putnam's infamous book Race and Reason came out. Putnam's book argued that there were significant differences between the races and that this had important political implications. A few anthropologists became upset when they saw the book gain popularity and wanted the AAA to intervene. But they knew that Coon supported the book and that, as a result, having the AAA take action against it would be difficult. So they organized a secret meeting behind Coon's back in order to issue a statement denouncing Putnam's work. Coon found out about the meeting and went to stop it. Upon entering the meeting Coon asked all the association members who had actually read Putnam's book to raise their hands. Only one rose. He then asked for the hands of everyone that had even heard of the book prior to that meeting. Only a few hands rose. None the less, the statement against the book passed. Coon, disgusted by the actions of his peers, resigned from the AAA (1). This incident stands out in the history of science as a particularly clear example of scientists not living up to what people expect of them. People often take what scientists say for granted. They trust that scientists have looked at the evidence rationally and are giving the public as accurate an account of all the relevant facts as they can. When explaining science to laypeople researchers are expected to only make authoritative statements on topics they are knowledgeable of, to not lie about scientific evidence, and to not omit obviously relevant facts. People trust scientists to do these things and so feel comfortable taking what they tell them for granted. (See
My work was uplifting people, or so I told myself. Looking back, I suspect that feeling like I was on the right side of history was a kind of moral cover. Maybe I felt less wrong about the p-hacking because I was promoting social justice. The conclusions I drew were what the good guys already believed, after all.
— Michael Inzlicht My P-Hacking Felt Like Truth