The Splintered Mind: AI Slop and Evidence about Evidence: Why Philosophy Journals Should Reject AI-Written Prose
skip to main |<br>skip to sidebar
The Splintered Mind
reflections in philosophy of psychology, broadly construed
Wednesday, July 15, 2026
AI Slop and Evidence about Evidence: Why Philosophy Journals Should Reject AI-Written Prose
I don't want to read your AI-generated email. Lots of reasons why, but here's the core of it: I want the text to reflect your thoughts. I want to know that you, the human, actually had those ideas and had them authentically enough to express them in exactly the form I see on the page. For personal emails, I want this for personal reasons. For philosophically substantive emails, I want this for evidential reasons. Philosophy journals should also want human-generated rather than AI-generated text for the same evidential reasons.
There Probably Are Good Reasons You Phrased It the Way You Did
The evidential reason: Human experts think differently and better than LLMs. Their word choices, even subtle ones, reflect sensitivities that they might not themselves be aware of. Typically, an expert's prose will be more sensitive to the matters on which they are expert than the output of a language model. When I receive a philosophical email from you -- and more so when I read a journal article -- I want your expert word choices, not blurry LLM approximations.
You might object as follows: Of course I read the LLM outputs before sending, and I wouldn't send the email, much less submit the article, unless I endorsed every word! So, the objection continues, you did think the thoughts expressed. The text reflects your expert best judgment -- maybe even something better than your expert best judgment: your expert best judgment combined with the expertise of an LLM.
I reply: There's a huge cognitive difference between nodding along while reading something and actually productively generating a text. Two reasons: First, once the text is on the page, it's easy to passively let the approximate word suffice, rather than thinking about word choice in the same effortful, active way we do when generating prose de novo. Second, as I suggested above, I doubt that human beings, even experts, have a good sense of all the factors that shape word choice -- everything they're being sensitive to. You would have phrased it slightly differently, and even if you don't know that, or why, a different signal is sent and received.
Evidence about Evidence
I'm talking about evidence about evidence: meta-epistemology. Your email or your article presents evidence for a particular philosophical view (alternatively, evidence that you support a particular philosophical view). In a simple world, I could evaluate this evidence entirely on its face: How good is the proposed view? But in the actual, complex world, it helps to have evidence about the quality of the evidence. The fact that you, an expert human, generated the text is evidence that the view is worth thinking about -- more so than if the text were generated by an LLM. This holds even if the text is exactly the same, which of course it wouldn't be.
An increasingly large part of the function of journals is to provide evidence about evidence -- the value of their imprimatur. The fact that an article appears in Nous or Ethics is evidence that it has been through rigorous review and was judged worthy by several expert humans applying unusually demanding standards of quality and importance. Its appearance in those journals is thus evidence (imperfect of course!) that the reasoning is of high quality and the arguments worth taking seriously.
Similarly, if I know that an email or an article was written by a respected colleague, reflecting their positive creative exertion in trying to choose the right words, guided by their intuitive expertise in how to phrase things, I have better reason to take it seriously than if I know that it was generated by an LLM and reflects only their passive after-the-fact assent.
Philosophers sometimes suggest that we shouldn't care if an argument was human-generated or AI-generated -- that insisting on human-generated prose is fetishizing personal human interaction rather than facts and argument quality. In a way, that's true: A sound argument is a sound argument. Similarly, we shouldn't care if an article was written by David Chalmers and published in Philosophical Review or whether it was written by someone with no institutional affiliation and published on an obscure blog. If the argument is good, it's good -- of course, of course!
But at the same time, we have limited attention, limited time, limited ability to understand the nuances when matters drift even a little from our tightest foci of expertise, and in these cases it's helpful to have meta-evidence. What should I read? How far should I trust the author has the details right, versus how much should I pause critically and chase down independent sources? How much...