A Meta-Epistemological Reason for Rejecting AI-Written Philosophy


“The fact that you, an expert human, generated the text is evidence that the view is worth thinking about—more so than if the text were generated by an LLM.”

That’s Eric Schwitzgebel (UCR), writing at The Splintered Mind about why he doesn’t want to receive AI-written emails about philosophy, and why philosophy journals should reject AI-written submissions.

Here’s more context:

I’m talking about evidence about evidence: meta-epistemology. Your email or your article presents evidence for a particular philosophical view (alternatively, evidence that you support a particular philosophical view). In a simple world, I could evaluate this evidence entirely on its face: How good is the proposed view? But in the actual, complex world, it helps to have evidence about the quality of the evidence. The fact that you, an expert human, generated the text is evidence that the view is worth thinking about—more so than if the text were generated by an LLM. This holds even if the text is exactly the same, which of course it wouldn’t be.

An increasingly large part of the function of journals is to provide evidence about evidence—the value of their imprimatur. The fact that an article appears in Nous or Ethics is evidence that it has been through rigorous review and was judged worthy by several expert humans applying unusually demanding standards of quality and importance. Its appearance in those journals is thus evidence (imperfect of course!) that the reasoning is of high quality and the arguments worth taking seriously.

Similarly, if I know that an email or an article was written by a respected colleague, reflecting their positive creative exertion in trying to choose the right words, guided by their intuitive expertise in how to phrase things, I have better reason to take it seriously than if I know that it was generated by an LLM and reflects only their passive after-the-fact assent.

Philosophers sometimes suggest that we shouldn’t care if an argument was human-generated or AI-generated—that insisting on human-generated prose is fetishizing personal human interaction rather than facts and argument quality. In a way, that’s true: A sound argument is a sound argument. Similarly, we shouldn’t care if an article was written by David Chalmers and published in Philosophical Review or whether it was written by someone with no institutional affiliation and published on an obscure blog. If the argument is good, it’s good—of course, of course!

But at the same time, we have limited attention, limited time, limited ability to understand the nuances when matters drift even a little from our tightest foci of expertise, and in these cases it’s helpful to have meta-evidence. What should I read? How far should I trust the author has the details right, versus how much should I pause critically and chase down independent sources? How much should I let their way of phrasing things, their habitual patterns of thinking, the presuppositions hidden in their word choices and sentence structures, slip gently into my brain, silently strengthening my own associations and predilections?

Why does “knowing this text was written by a human philosopher” provide some evidence that the text is worth a certain degree of attention?

Human experts think differently and better than LLMs. Their word choices, even subtle ones, reflect sensitivities that they might not themselves be aware of. Typically, an expert’s prose will be more sensitive to the matters on which they are expert than the output of a language model. 

And checking an LLM-written text to make sure it reflects your own thoughts is risky:

There’s a huge cognitive difference between nodding along while reading something and actually productively generating a text. Two reasons: First, once the text is on the page, it’s easy to passively let the approximate word suffice, rather than thinking about word choice in the same effortful, active way we do when generating prose de novo. Second… I doubt that human beings, even experts, have a good sense of all the factors that shape word choice—everything they’re being sensitive to. You would have phrased it slightly differently, and even if you don’t know that, or why, a different signal is sent and received.

See his post for a good example of what he’s getting at, along with an explanation for why his argument “is not a call for blanket rejection of the use of LLMs in philosophical (or other) writing.”


Related:
Ethics Announces AI Policy
The Ethics of Using AI in Philosophical Research

guest

52 Comments
Oldest
Newest Most Voted
Inline Feedbacks
View all comments
Matt
Matt
20 days ago

Eric Schwitzgebel makes valuable points about AI-written papers ‘nodded-through’ by their human instigator; the value of one’s de novo, even clumsy, own wording; and (in the underlying post) about careful use of AI. However, there is a slightly concerning note of anthropocentric self-satisfaction. The assertion, ‘Human experts think differently and better than LLMs’, appears open to challenge on both counts. It seems that human brains may predict continuations subconsciously in a similar manner to LLMs (https://www.sciencedaily.com/releases/2026/06/260624025514.htm), while (from personal observation) higher-version LLMs, at least, seem rather good at grasping what we wanted to say but didn’t quite manage to. Humans, at least for the time being, have advantages, e.g. the passion to press deeper into an idea, the ability to say what matters. But a little more circumspection might be in order.

John
John
20 days ago

I wonder what everyone’s opinion is on AI4Science, AI4math, and vibe coding.

Kenny Easwaran
Reply to  John
15 days ago

Some of these things are likely really useful!

I haven’t yet read through the text of any AI-generated mathematical proofs to figure out if the AI-generated text is itself fine as a section of the mathematical publication, or if you want to rewrite it and refine it. I’ve heard claims that they’re not as good as humans at generating re-usable structures and ideas in the process of coming up with the proof, so that they might burn through our stock of conjectures without actually growing the mathematical idea pool – but again, I don’t know if that’s exactly accurate (or if it might change in a few months).

Vibe coding has been great for little side projects I’ve been interested in doing outside of academia, and also for putting together little activities for students to do in Zoom breakout rooms to illustrate some ideas in online classes. Now that I understand that I can generate code to do things systematically, I’m trying to think if there are things that would be nice to have for bits of my work (beyond TikZ code for diagrams in LaTeX).

Eric Steinhart
Reply to  Kenny Easwaran
15 days ago

At this point the AIs don’t seem to be generating many deep new insights. They’re good at combining old ideas or grinding through a bunch of lemmas to get to improved versions of old results.

The recent claim of the cycle double cover theorem (if true) might mean AIs have shifted into a higher level of mathematical intelligence. But that proof-candidate is currently undergoing human review.

The most interesting thing recently to come out of the math work is the use of multi-agent generator-verifier systems. One found a flaw in a hypothesis about the Benjamini-Hochberg procedure in statistics.

Eli Alshanetsky
20 days ago

100% with Eric on the epistemic value of expert word choice. My worry is the move from “human word choice is valuable as a guidance signal” to “journals should require it.” Without a public verification method, making human authorship a formal requirement for acceptance collapses into a mix of honor-system pledges and subjective sniff tests by editors and reviewers. To get published, either your word already needs to be trusted, or you need to know the private handshake a given group will read as human. Both take peer review from something that can at least pretend to be anonymous to something that can’t. Plagiarism norms work because the source is traceable, which is not the case here. I wrote about this issue and the threat it poses to public reasoning here.

Jessie Ewesmont
Jessie Ewesmont
Reply to  Eli Alshanetsky
18 days ago

There’s room for debating where the line should be in edge cases. Maybe some benefit of the doubt can be extended when it’s not totally clear. But when it is, rejecting the paper for that reason makes sense. (And, presumably, publicly stating you don’t accept AI submissions discourages some people from submitting in the first place.)

Dale E. Miller
20 days ago

When I’m trying to decide whether to read a random unpublished paper that I come across on the Internet, the fact that it was written by a competent human philosopher does seem to be useful meta-evidence that it’s worth my attention. However, if it’s been published in Ethics, then that seems to be such strong evidence that it’s worth reading that whether a human wrote it is no longer so important. So the implications of this meta-evidence point for journal practice are at best unclear.

Grumpy Logician
Grumpy Logician
Reply to  Dale E. Miller
18 days ago

That’s a good point, but I don’t think the fact that the paper is published comes close to screening off author identity. In my experience published papers vary dramatically in quality, even at the top journals (at least some of them). In the end, usually the person who by far has thought most deeply about the topic is the author themselves, so their judgment has significant evidential value over and above that of reviewers and editors. For example, I’d be much more interested to read a paper by a philosopher whom I respect than a random paper on the same topic at the same journal.

Dale E. Miller
Reply to  Grumpy Logician
16 days ago

I would be, too. But if the comparison is between two random papers, one written by a human whose work I know nothing about and one an AI, then it’s not clear to me how more inclined to read the human-written paper I should be.

Patrick Lin
20 days ago

Makes sense. And here’s the additional rationale I offered in the last AI post on Daily Nous. I wouldn’t call it “meta-evidence” but only a practical way to triage your reading load, esp. in discussion forums:

I’m not necessarily saying that your AI-assisted post was bad or wrong, just that it’s “tainted” with AI in the view of some or many readers. Here’s why that matters:

For many of us, we already have way too much to read and can’t possibly engage with everything worthwhile. So, one reasonable way to triage our workload is to avoid reading or engaging with anything that used AI.

One reason for doing this is to avoid the Gish gallop, where AI-assisted interlocutors can “flood the zone.” With AI, they can engage in argumentum verbosum to overwhelm the conversation with rapid-fire, low-effort, instant arguments that no human can respond to fast enough. If you’ve never been on the receiving end of that, you might not know.

There can be other reasons, such as feeling disrespected by those who write with AI. That becomes clearer in more personal communications, such as emails or eulogies.

But more generally, separating the message from the messenger has been a problem for literally thousands of years, so I don’t think anyone should be surprised much. And it’s not always wrong; sometimes, who the messenger is matters in a non-ad hominem way…

Meme
Meme
20 days ago

I’ve tinkered with Claude by feeding it one of my half-written papers, just to see if what it produces will come close to what I had imagined. (Don’t worry, these aren’t for publication.) I occasionally find it difficult to understand a bit of reasoning in some of the generated text. The weird part is that, in these cases, I have no confidence that what Claude wrote actually *does* make sense, and that it’s therefore worth trying to understand. It made me realize that, before AI, I tended to be highly deferential to authors; I always just assumed (with exceptions for obscurantists and the like) that anything hard to follow would still be coherent, and thus that struggling through it would be rewarded. I find this experience odd and interesting, and maybe sort of related loosely to Schwitzgebel’s discussion. Have others experienced it, or thought about it, or whatever?

Alice
Alice
Reply to  Meme
20 days ago

Not so long ago, LLMs still hallucinated a lot. And they also have a much higher rate of confidently asserting things they are not sure about than experts at comparable levels. The latter remains true. So I am kinda not surprised by your experience. We were *a lot more* skeptical not so long ago, no?

Meme
Meme
Reply to  Alice
19 days ago

That’s fair, I should have been a bit more careful. I realize that hallucination is probably *why* I don’t trust a difficult AI-written passage. What I found interesting was (a) the strangeness of the experience (of reading a chain of reasoning and not knowing if it would actually reward the effort of trying to understand it) and (b) the realization of how much I relied on basic trust when reading human work. Maybe it reminded me of the claim that future superintelligence will—if it ever exists—so vastly outpace us that we won’t even understand how it verifies its factual claims. The claims themselves might be tractable, but we’d have to take the method which got to them on faith(?)

Maybe that’s not interesting, or I don’t know what I find interesting about it; for some reason I found it worth mentioning.

Alice
Alice
Reply to  Meme
19 days ago

Fair!

Jessie Ewesmont
Jessie Ewesmont
Reply to  Meme
18 days ago

Yes, I’ve had this experience too. In the past, when I read a piece of difficult text, I tried to find a way to make sense of it, because I knew the author must’ve meant something by it. But nowadays, there’s some chance I later discover the text was just some garbled nonsense spat out by an LLM and feel annoyed that my time is wasted.

Meme
Meme
Reply to  Jessie Ewesmont
17 days ago

I find it very jarring, and sort of philosophically interesting, too. Not sure why. Glad I’m not the only one.

Alice
Alice
19 days ago

The thing is, we don’t know if an article is written with substantial help from AI. As the latest models become genuinely more capable in certain subfields, I am starting to suspect a lot of articles submitted are like that. As a referee, I do not know what to think of it (obviously I am not talking about AI slops, in which case there is no question). It feels like I *should* be using substantial help from AI so as to write reports more efficiently. AI is coming for us, I think, whether or not we are using it or not.

I fear that OP and readers are not fearing enough.

John
John
Reply to  Alice
19 days ago

The paradoxical question is that when Large Language Models write well enough and are examined meticulously, it’s impossible for humans to know how the author accomplished it. In fact, this involves a question: Is the hostility towards AIGC out of the fear of losing the anthropocentric status or due to concerns about responsibility? I’ve noticed more of this hostility in the humanities. Currently, Internet companies have widely adopted AI technologies, including vibe coding, to generate content. As far as I know, in fields such as physics, biology, chemistry, computer science, and artificial intelligence, the concept of AI4XXX has emerged and been quickly recognized. However, in fields like literature, philosophy, and art, there are still some people who strongly resist. In fact, I don’t think humans have any unique advantages in these fields. In other words, I don’t think human agency endows purely human – made products with any mysterious properties.

ERW
ERW
Reply to  John
19 days ago

You don’t have to speculate why people are critical of AI, and attribute to them a “fear of losing the anthropocentric status” or whatever, you can just read their actual criticisms (e.g., unreliability, encouraging bad epistemic habits in people, etc.).

There is an annoying tendency of (certain) AI boosters to just ignore the stated criticisms of critics of AI and conspiratorially speculate on hidden motives. No, I dislike it for the reasons I said.

Both Alice and John also just assume that AI writing is undetectable, which is a bit funny after coming from that other Daily Nous article, where the comment section immediately noticed the author used AI without it being disclosed. So I reject that premise.

John
John
Reply to  ERW
19 days ago

In fact, I have of course read relevant criticisms, but regrettably, the vast majority of the arguments in these criticisms focus on the differences between human writing style and that of Large Language Models. And this difference can be addressed through deep engagement. This has led to a situation where people constantly speculate about whether others’ papers are completed by AI, causing some previously normal language usage to be regarded as incorrect behavior. This is an overextension, and it is limited to the humanities and social sciences. In disciplines such as physics, biology, computer science, and artificial intelligence, linguistic expression only needs to be accurate, with more attention paid to the accuracy of experimental data and the standardization of procedures.

We can draw a conclusion that the main factor in determining whether an article is worth reading is the accuracy of its viewpoints and arguments, rather than the identity of the author. If an AI proves a certain mathematical conjecture (and such a case already exists), should we refuse to read it and directly dismiss it simply because it was generated by AI? This is in fact another form of identity discrimination. Sometimes we can observe a phenomenon where, in single-blind journals, articles written by renowned scholars are more likely to pass review. Correspondingly, some authors, concerned that scholars holding your viewpoint might discriminate against their articles, choose not to disclose AI assistance in their papers (He, Y., & Bu, Y. (2025). Academic journals’ AI policies fail to curb the surge in AI-assisted academic writing. Proceedings of the National Academy of Sciences of the United States of America, 123 9, e2526734123). For example, in the article I have cited, after major publishers announced their AI policies, among 75,000 papers in 2023, only 76 chose to disclose AI assistance. In fact, AI assistance certainly far exceeds 76, but the other AI-assisted articles all underwent peer review and were published. This shows that if a paper involves AI participation and has been revised, even professionals would find it difficult to detect. Moreover, it is precisely because of discrimination against such assistance that relevant policies are difficult to implement.

I believe that the importance of an article depends on its viewpoints and arguments, not on the author. As long as the quality of the paper is good enough, whether the author is a renowned scholar, an ordinary person, someone without a diploma, or with AI assistance, or even entirely completed by AI, it deserves to be taken seriously.

Jessie Ewesmont
Jessie Ewesmont
Reply to  John
18 days ago

It’s patently absurd to call it identity discrimination. You can’t discriminate against AI, because it isn’t a person.

Kenny Easwaran
Reply to  Jessie Ewesmont
15 days ago

To “discriminate” just means to tell the difference between two classes of things and treat them differently. We in fact often think of discrimination as a *good* thing when it’s not about people – we talk about a “discriminating taste” or a “discriminating palate”.

Even when it’s about people, it’s sometimes good – math professors want to write hard tests so they can discriminate the students with strong skills from the students without them, rather than indiscriminately giving everyone A’s, or discriminating only on the basis of minor slip-ups.

Of course one can discriminate against AI – and probably sometimes one should, but sometimes one should be discriminating on some basis that is more relevant for what is actually important.

Greg Guy
Greg Guy
Reply to  ERW
16 days ago

No, the question was why these criticisms are worse in the humanities than in other subjects. Especially as AI keeps improving.

Matthew Braham
Matthew Braham
Reply to  John
18 days ago
Sam Duncan
Sam Duncan
Reply to  Matthew Braham
17 days ago

This is pretty biased toward the claim people can’t spot AI in that hotel reviews are incredibly formulaic — especially the positive ones which these all are—and each text is very short. A better test would be longer text in something where more creativity or expertise is expected. Having said that I got 75% on this and honestly only skimmed each once while half paying attention to Peanuts cartoon my kids are watching.

Matthew Braham
Matthew Braham
Reply to  Sam Duncan
17 days ago

From my colleagues, the success rates so far range from 53% to 66%. You are more successful. The point is that you are still getting 25% wrong. As AI improves, that will drive to 50%.

The point of the test is to test our epistemic modesty. How confident are we as humans at AI detection.

I totally agree that the longer and more complex the text, the easier it is to spot. But again, this will get more difficult with time.

Marketeer
Marketeer
19 days ago

If an AI-generated or AI-assisted article passes peer review at a prestigious journal, which goes on the publish it, how is the fact that the journal published it not still weighty higher-order evidence? In general, it seems that use of AI in the writing process doesn’t compromise the higher-order evidential value of *publication*. It does negatively impact the higher-order value of “so-and-so sent me a draft,” of course.

Michel
Reply to  Marketeer
18 days ago

What, like _Social Text_?

As far as I recall, the consensus there weighed in the other direction…

ikj
ikj
Reply to  Michel
18 days ago

just a reminder that sokal’s paper wasn’t peer reviewed in a crucial sense. only the distinctly non-physicist editorial board reviewed that paper.

Matthew Braham
Matthew Braham
18 days ago

Step outside philosophy for a moment. Suppose an AI solves the Riemann Hypothesis. Does that mean its not worth reading? By analogy, suppose that an AI provides a convincing answer to the Hard Problem of Consciousness. Also not worth reading.

Humans also produce intellectual slop. Consider this experience:

I did a test with Claude’s lates model yesterday. I put in short answer (30 words) responses to my ethics exam. I asked Claude to write model answers from my lecture scripts where each answer is. And then create a grading scheme for the 3 points fro the question (with 0.5 steps). Across 49 scripts for one question, it out performed me. I missed stuff and was inconsistent.

What I discovered was this: I tired very quickly from the handwriting and tended to be lenient with clean and elegant script and harsh with the “crackle”; then I tended to have runs of leniency and runs of harshness. I tended to drift off in frustration (why am I still doing this?) … you name it.

So why would human error be preferable to silicon accuracy? “I failed, but at least a human failed me, even though I passed by AI”. Clearly we can’t go “full AI”. That is, just accept automated grading for these kinds of scripts. It has to be some kind of combination. Is it a sampling? A double control?

Michel
Reply to  Matthew Braham
17 days ago

Consider this experience: I asked Claude to find a verse couplet from a famous epic poem for a translation. It attributed it to a number of sources, but not the right one. I pointed Claude to the right text. It then made up locations (books and line numbers.) After I corrected it, it decided a better way to proceed would be to identify the first known quotation, and see where it claimed to have found the couplet. I agreed, so Claude made up a source whose (fictional) date was ten years later than the text I was translating.

Anecdotes =/= data.

Matthew Braham
Matthew Braham
Reply to  Michel
17 days ago

Michel,

The two cases are not comparable. Yours is one of retrieval; mine is pattern-matching against a fixed text. This is a different task, so the anecdotes don’t match up.

There is also a common misunderstanding. An LLM is a probabilistic prediction machine, which means we have to constrain it — just as we do with ourselves. Thus we need to bind every claim to a provided source such as a DOI or the actual text and also require license abstention, ie. return nothing if uncertain. While you can’t get to a literal zero error rate you drive it near-zero. Your couplet task did none of this. Instead it left the model ungrounded and asked it to source a fact it didn’t hold, which is precisely the setup that produces confabulation. My grading task supplied the ground truth and asked only for consistent application. And “anecdotes =/= data” cuts your way harder than mine: yours was n=1, mine was 49 scripts measured against my own marking. What was striking was that the AI caught me confabulating and hallucinating … due to very natural attention deficits.

Michel
Reply to  Matthew Braham
17 days ago

I’m sorry, you’re saying it was ungrounded and unconstrained _when I gave it the 700-page PDF of the right epic poem_?

You’re right, mine was a different–more basic–task. That it could not complete it is disturbing.

Matthew Braham
Matthew Braham
Reply to  Michel
17 days ago

I did not have this information. However, I have done similar things. But I constrain it. I give Claude very precise instructions: just as I would do a research assistant. The reason is because I know that for efficiency reasons it will estimate and guestimate — just as humans do. My sense is we are expecting too much from LLMs. They are just machines. Nothing more and nothing less. And they do better or worse jobs based on the competence we have as users and for the task at hand. Its not enough to give it a text. You have to give it exact instructions and limit its error behaviour.

Michel
Reply to  Matthew Braham
16 days ago

I did say I pointed it in the right direction.

I’m not sure how much more precise I can be than asking for the book and line number, or even just the page number, where a piece of text appears. If you have a suggestion, I’ll try it.

Matthew Braham
Matthew Braham
Reply to  Michel
16 days ago

See below. But if you need detailed support just email me. That would be easiest. I use Claude Code as research team. Sometimes it writes me scripts to do certain jobs.

Matthew Braham
Matthew Braham
Reply to  Michel
17 days ago

A further point. I asked Claude why you would have such a problem. Its reply (verbatim):

“A PDF in the context window is not grounding in the sense that matters. Supplying a document and binding a claim to a located, verifiable position in it are different things. If you say “find the couplet” over 700 pages, the model still predicts a plausible answer — including plausible book and line numbers — because you gave it a haystack and asked for a needle without requiring it to show the needle. Grounding isn’t “the source is present.” It’s “every claim must carry its location and abstain if it can’t.”

Maybe that works better.

Michel
Reply to  Matthew Braham
16 days ago

A “plausible” answer in this case is any old book and line number. E.g. book 7, line 332. So basically any answer returned is ‘plausible;’ it’s practically analytic.

But sure, I can add the further restriction not to hallucinate if it has no success. Though having to say that every time you interact seems non-ideal. Especially since this is a keyword search. Google used to be able to do this before Alphabet deliberately sabotaged it. And, of course, ctrl+f does it, too.

Matthew Braham
Matthew Braham
Reply to  Michel
16 days ago

Have you tried using Claude to write a Skill for this? Also, are you just using Claude Chat or Claude Code or Cowork?

The advantage of CC in VS Code is that you can fine tune your work, develop special analytical tools, and read the deep thinking output.

Just think of Claude Code as a tireless and smart research team that needs instructions. These instructions can be written up and automated. You can even get CC to write you a log of your chat, and summarize results. My logs tell me what it is certain about, and what it is not certain about.

Kenny Easwaran
Reply to  Michel
15 days ago

Not any old book and line number are equally plausible! If I heard someone guess what act and line number “double double, toil and trouble” appears at in Macbeth, I would find guesses like “line 40 of Act I” plausible and guesses like “line 700 of Act V” implausible. There’s more to plausibility than just getting an Act and line number that exist – having a sense of the shape of the plot is relevant!

But of course, it’s not that useful to be able to generate plausible guesses of this sort.

In any case, finding precise locations of text is one of the things that LLMs are particularly bad at, while (as you note) many other computer systems (like simple ctrl-f) are really good at.

I’d recommend trying some other tasks, like “find passages where there is a metaphor involving birds” or “find places where a conclusion is asserted without an argument”. This sort of request tends to trigger it to engage more closely with the text in ways that produce better results.

It’s still bad if it’s dealing with the full 700 pages. What will work much better is if you write a script to take your 700-page text, break it into paragraphs, and then iterate through the paragraphs, repeatedly calling an LLM with your prompt for that paragraph. The script should gather the results and then present them to you with the original paragraph.

With this structure, you get the advantage in thoroughness of a traditional computer program, with the advantage of text interpretation from an LLM. You’re extremely unlikely to miss any examples or get any false positives this way.

And if, like me, you aren’t skilled enough at any scripting language to actually write this script yourself, most modern LLMs will be able to write it for you. (Though you might need to set up an “API key” to let your script call the LLM for you, and you might end up getting charged a few cents or dollars for using this API key enough times to process 700 pages. The LLM can usually explain how to do this too.)

Kenny Easwaran
Reply to  Michel
15 days ago

It’s a mistake to think that someone who doesn’t know your favorite musician must therefore not know anything serious about music, because knowing an individual musician is “more basic” than knowing serious things about music. And if you’ve worked at all with AI, you’ll understand how weird their outputs can be, having some truly effective and powerful capacities and also some major incapacities, often on tasks that for humans seem very similar.

If you’re not familiar with using AI, it might be useful to treat all its outputs equally negatively. But if you get some practice with it, you’ll get a better sense of what sorts of things it can do well and what sorts of things it does badly at. It’s just a rookie mistake to assume that if it does badly at one thing it must therefore do worse at everything that seems harder to you.

Matthew Braham
Matthew Braham
Reply to  Kenny Easwaran
14 days ago

On this point, Kenny is absolutely spot on. I have been testing grading of short answers and constrained Claude’s behaviour. Reading its “thinking transcript” it gave me this

I need to flag a transcription-reliability issue before compiling: three scripts in the second batch (student numbers redacted here) came back with near-identical A3 wording in my reading, which is almost certainly my conflation, not three students writing the same sentence. Rather than propagate that, I’ll re-read those three cleanly — the verify-don’t-fabricate discipline matters most on real grades.”

It also constrained its task management:

Continuing on Opus. One judgment call I’ll flag rather than bury: for this task a parallel workflow would hurt, not help. The scarce resource is single-grader calibration — the anchor rules, and especially A3’s F-anchor (expressivism/prescriptivism confusion = wahrheitswertlos statt falsch, max 1,0). Fragmenting grading across subagents reintroduces exactly the inter-rater variance your human comparison is meant to measure. So I’ll finish A3 solo to hold calibration, then the natural place for a workflow is an adversarial second-pass verifying the borderline/F cases — which I’ll offer once the grades are down, since it partly overlaps your planned human check.
Reading the final 15 scripts”.

So here Claude is managing its own processes to minimize hallucinations and confabulations much like humans would do.

So it is the usual GIGO adage. We, as users have to manage that.

Kenny Easwaran
Reply to  Matthew Braham
13 days ago

> Suppose an AI solves the Riemann Hypothesis. Does that mean its not worth reading?

I recently had a conversation with someone who was working at OpenAI on AI proofs. He said he had probably generated about 1000 mathematical papers using AI over the previous few weeks, and that several dozen of them had produced what appear to be correct proofs of previously-unsolved problems that people cared about, but only about 10 of them were actually worth reading – and he emphasized the implicature that this meant that a good number of those papers with correct proofs of previously-unsolved problems were still so badly-written that they didn’t feel worth reading.

At the moment, it’s hard to imagine a paper with a correct proof of the Riemann Hypothesis that is so badly written that it’s not worth reading, but AI has been doing a lot of things lately that were very hard to imagine until very recently.

Matthew Braham
Matthew Braham
Reply to  Kenny Easwaran
13 days ago

There are two issues here: one epistemic (correct proof), the other aesthetic (worth reading). My guess is that if an AI generated an “ugly” but correct proof, it would engage mathematicians immediately. And what would follow would attempts to improve the readability or beauty of the proof both with and without AI and to determine if an AI can generate a beautiful proof, and if not, why not. That is, AI proofing itself becomes an object of study.

Similarly, I think if an AI produced a readable and original response to the Hard Problem of Consciousness, I am sure it would occupy philosophers in the same way. Obviously some would dismiss it for it being AI. I, for one, would not. I would have an open mind as to this possibility.

Kris Rhodes
Kris Rhodes
Reply to  Matthew Braham
8 days ago

If an AI solves the Riemann Hypothesis, that’s worth reading.

But that’s not relevant to the point made by this post. The relevant question is, “If a computer claims to have solved the Riemnann Hypothesis, is that worth reading?”

The answer to that is still a hard no.

But it’s true that in a few years we may be at a point where if Chatgpt 11.0 says it solved the RH, that might be worth checking out. That’s still importantly irrelevant to the issues posed by this post, because when the RH is proved, it will be something that can be formally checked, where “word choice” and other aspects of human introspection and interaction don’t have to be important.

Unlike philosophy.

Matthew Braham
Matthew Braham
10 days ago

And so it has happened. The Riemann Hypothesis has not been proven, but on the same list of problems, the 87-year-old Jacobian conjecture was recently disproved when mathematician Levent Alpöge used Anthropic’s Claude Fable 5 model to find an explicit three-variable counterexample.

https://www.digitalapplied.com/blog/claude-fable-5-jacobian-conjecture-disproof-ai-mathematics

So what did mathematicians do? Exactly what would be expected: they set to work to check it. Follow the discussion on the web site given. Its very enlightening.

So why should philosophy be any different?

Patrick Lin
Reply to  Matthew Braham
10 days ago

Because philosophy isn’t mathematics. Usually, there’s no clear, verifiable answer in philosophy, beyond logic; otherwise, it wouldn’t be philosophy.

For the same reason, philosophy also isn’t like finding a cure for cancer, etc. as other commentators have compared it to on Daily Nous…

So, there’s a real danger of wasting your time engaging with philosophers who rely on AI, even if AI can be more useful in math. See my comment above about triaging one’s work to avoid the Gish gallop and argumentum verbosum, as just one of many reasons.

If anyone has time to kill and doesn’t care about the broader reasons and ethics, they’re free to read all the AI papers they like. Or more likely, they’d use AI to read those AI papers, which seems absurd.

Matthew Braham
Matthew Braham
Reply to  Patrick Lin
10 days ago

Are you now claiming that philosophical statements are not truth apt?

Patrick Lin
Reply to  Matthew Braham
10 days ago

That’s not at all implied by what I said. What I’m suggesting is only that the truth of most philosophical claims is unknowable by humans, currently. That is, we can debate them but can’t “check the work” in the way you had suggested re: math proofs to get a clear consensus on the answer.

Or if you think we can, please provide some examples.

If it ever becomes possible to empirically confirm the claims — e.g., is the world composed of atoms? is aether the fifth element? — then they would no longer be philosophy questions but science questions or something else.

I know that there’s no clear consensus on the super-basic question of what philosophy even is, which sort of proves my point…and I’m not really interested in rehashing that debate here…

Patrick Lin
Reply to  Matthew Braham
9 days ago

Also, Matthew, if you’re following AI in mathematics, then I assume you’ve familiar with the Leiden Declaration on AI and Mathematics from last month.

The declaration explains why AI-generated work is different from human-created work. E.g., here’s the first threat from AI they identified (which gestures at the Gish gallop, i.e., “fast-moving developments”):

Current automated techniques can produce plausible but unreliable (or even incorrect) arguments which are difficult to distinguish from correct mathematical proofs. This applies not only to informal arguments, but also to formalizations, where the difficulty lies in the translation between computer-encoded and human presentations of concepts. These fast-moving developments put our present system of review under increasing pressure, jeopardizing our ability to implement traditional standards for the correctness, transparency, and independent verifiability of proof.

And note that the declaration’s very first recommendation is to disclose AI use, as relevant to the other recent DN thread about journal policies.

So, sure, philosophers can learn something about how mathematicians are approaching AI, but I don’t think it’s what you think it is…

Matthew Braham
Matthew Braham
Reply to  Patrick Lin
9 days ago

Hi Patrick,

Not all philosophical statements are metaphysical. Hume on induction, the underivability of ought from is, Kant’s claim that his formulations of the CI are equivalent: each is checkable, and the third, I think, fails. I would say that these are loosely “philosophical facts”.

You’ll might, of course, say that’s still logic, which you already cordoned off. But the exception comes at a cost. Impossibility results, equivalence claims, showing a position collapses into the one it was built to avoid — that is what a lot of substantive philosophy does. If all of it counts as “just logic,” what’s left of philosophy is unverifiable claims. I’m not sure you mean that.

On the Alpöge case: mathematicians had no consensus the night the map appeared. But they had a procedure. That is all I am claiming for philosophy, and we are exercising it on this comment right now.

Also, I made no claim about relying on AI, or about disclosure. Only that some results are worth checking regardless of who or what generated them—and that counts for philosophy.

To clarify the record: I happen to take philosophy to be a valuable human activity. But, what I am unprepared to do is to assert that it is uniquely human. Of course the standard objection now hits: an AI only recombines what humans produced. But here recombination is not the opposite of philosophical originality and interest. On Hume’s own account, that is what imagination does: it compounds and transpose materials we already have. In fact, artists of all types do this all the time. The history of art is a history of recombination. If that suffices for human novelty, it cannot be disqualifying elsewhere.

humantooth
humantooth
Reply to  Patrick Lin
6 days ago

Thank you for sharing this, I just finished reading a piece from one of the working group authors of the Leiden Declaration: Knowledge Collapse – Michael Harris (Boston Review) June 10, 2026
In it, Harris shares some of the impetus for its creation and provides an excellent overview of how AI is affecting the field of mathematics, along with speculating on what makes the Riemann Hypothesis in particular so enticing to AI firms- which ties back to some of the other comment threads under this post. He throws some shade at philosophers, but its still a great read

 In the last two years alone, Yacin Hamami and Rebecca Lea Morris, Jeremy Avigad, Jessica Carter, and Michael Friedman and Kati Kish-Bar-On have all made efforts to pin down mathematical understanding and label its parts. And since they are philosophers, no two of them understand understanding, to say nothing of its mathematical subspecies, in the same way.