I recently read this William Zinsser quote that inspired a nickname for this: Clotted Claude [1].
> Nobody has made the point better than George Orwell in his translation into modern bureaucratic fuzz of this famous verse from Ecclesiastes:
> > I returned and saw under the sun, that the race is not to the swift, nor the battle to the strong, neither yet bread to the wise, nor yet riches to men of understanding, nor yet favor to men of skill; but time and chance happeneth to them all.
> Orwell's version goes:
> > Objective consideration of contemporary phenomena compels the conclusion that success or failure in competitive activities exhibits no tendency to be commensurate with innate capacity, but that a considerable element of the unpredictable must invariably be taken into account.
> First notice how the two passages look. The first one at the top invites us to read it. The words are short and have air around them; they convey the rhythms of human speech. The second one is clotted with long words. It tells us instantly that a ponderous mind is at work. We don't want to go anywhere with a mind that expresses itself in such suffocating language. We don't even start to read.
> The Ecclasiast looked under the sun, but there was something he didn't understand. Something that wasn't right. Something that was not as it was supposed to be. And here is what the Ecclesiast didn't understand. Here is what nobody understood. Not then. Not in the years that followed. Not now. It was not the swift who won the race. Not the strong who won the battle. Not the wise who earned the bread. Not the men of understanding who gained the riches. Not the men of skill who gained the favor. And here is what I found: to any story of success, there is an element of unpredictability and chance.
(There are really just two possible outcomes: either the article is right or in, say, two years, we will be all writing and talking like this, as in humans learning from mediamatically reinforced human feedback.)
Shoot me now.
I am a bit of a luddite in this domain and have so far managed to resist the lure of using the generator to expand my thoughts, and I still catch myself writing "it's not just an X it's a Y" and other generator type tells. If it infecting my patterns it is totally entering the wider subconscious as "How to write" (Sighs)
Despite its bizarre look, the sentence is evocative and eloquent. It does make me get a clear mental image from the very first word. It leaves little room for roaming and guessing, as it firmly nails elements one by one, and, by the time I reach the end of the sentence, I get the full meaning almost immediately.
This sentence is not randomly written; this is crafted with intention. TBH, it would take me hours, if not days, to write a sentence this much condensed and easy to understand. I seriously like it.
Perhaps, this is more about context -- which style to use in which situation. I'm only guessing here, but, since Orwell is offering an interpretation, he probably chose to be more clinical. He probably had a point to make and didn't want to risk vagueness up-front.
It's like having to sit through a party with acclaimed academics: every single one is so full of themselves, they will constantly one-up each other by belittling everyone in their workplace s.a. to make you feel how great of an intellect they possess and how much more they would accomplish, had they not been surrounded by all these bumbling idiots.
It feels great to use, direct your machine minion to fill out your thoughts for you, but holy hell does it suck to be on the receiving end. Least of all is the disrespect, they don't care enough to even talk to you but worse is having to try and reason through that big incoherent blob.
Probably to only reasonable thing to do is to try and get your own mechanical agents to produce summaries. Inventing the lossy expansion algorithm(like compression but things get bigger on the wire), And we wept.
Now I am all depressed because it is probably inevitable, apparently thinking is hard and in general people are all to happy to outsource it to the machines.
Apparently, even though they want to spew AI prose everywhere, they want it read by humans, not by other bots, so when a few holdout places are insisting that prose be human authored they fight very hard against the rule.
A 5m search got me the following:
https://news.ycombinator.com/item?id=49410941
https://news.ycombinator.com/item?id=49411042
https://news.ycombinator.com/item?id=49059571
This thread, in particular, stands out - reader makes the claim that Pangram found that the US constitution was 100% AI generated, when others tried they found 0% (or close to it) https://news.ycombinator.com/item?id=48378191
I feel that if one doesn't want to send that message, they shouldn't be attempting to convince others that rejection of AI prose must stop.
I feel that if you want humans to read your stuff, those humans insisting that you write your own stuff is not an unreasonable position to take.
https://www.orwellfoundation.com/the-orwell-foundation/orwel...
(I'm guessing Zinsser's comments are from "On Writing Well", which you can also find online even though it is still under copyright.)
It's a bit of a pet peeve when people include quotes on a blog post without linking or otherwise references their source.
>iv. Never use the passive where you can use the active.
Orwell himself routinely ignores it, even in the first sentence of the essay:
>Most people who bother with the matter at all would admit that the English language is in a bad way, but it is generally assumed that we cannot by conscious action do anything about it.
The second clause could be rewritten in active voice by changing it to "but people generally assume". But this would make the writing worse, and Orwell, as a good writer, probably didn't even consider the option of making it worse, and therefore didn't notice the passive voice.
Passive voice is an essential tool for all good writers of English. I always give the example of the opening of Pride and Prejudice [0]:
>It is a truth universally acknowledged, that a single man in possession of a good fortune, must be in want of a wife.
The joke doesn't work in active voice. If you attribute this acknowledgement to some specific group of people then it's simply false, not a comedic exaggeration.
For example, nobody would seriously suggest avoiding the passive participle in a sentence like "Put the broken plate in the bin". ("Put the plate that someone broke into the bin"?)
You’re peeved with good reason. It’s the blog equivalent of posting a screenshot of an article to social media. People, please post your sources! In the age of misinformation, that’s more important than ever.
Is there something about tuning for desirable qualities that forces LLMs to have this voice?
Firstly, some parts of the RLHF involve human graders on the LLM's performance. I suspect their general bias towards a punchy, persuasive writing style could come from what biases the graders towards preferring that response, especially in shorter segments and when the grader is not focused on writing style
Secondly, later parts of the finetuning involve reinforcement learning on achieving certain tasks which are automatically graded: stuff like coding tasks. I think this can create a kind of feedback loop where the style drifts further, and you get the kind of LLM tics which are even more extreme (it might be that they incidentally help somehow with the actual tasks, or it might be a drift that comes from the grader also now being an LLM or some of this finetuning happening on output from other models). The more recent claude models seem to suffer from this a lot, moreso than earlier ones.
I also have to say that the first strikes me as being written by someone that might be smarter than I am, the second as being written by someone significantly less intelligent than I, yet somehow placed by society in a position of authority over me.
> the first strikes me as being written by someone that might be smarter than I am
This is why Joseph Smith tried to imitate the language of the King James Bible in the Book of Mormon, albeit not very successfully.
I also don't have a problem with large words as long as I'm well familiar with the words. The length of a word has nothing to do with the complexity of its meaning. We just have a limit to the number of pronounceable combinations of 5 letters.
Does that make sense?
The second transcribed considerable information bandwidth through intentionally structured word choice for maximal density.
I think the second requires deeper concentration, but is still quite readable compared to the kind of low-content engagement / SEO stuff one read on the internet even before LLMs
For example, the Lexham English Bible:
> I looked again and saw under the sun that the race does not belong to the swift, the battle does not belong to the mighty, food does not belong to the wise, wealth does not belong to the intelligent, and success does not belong to the skillful, for time and chance befalls all of them.
IMHO a big problem with Pangram in particular is that they market it as a reliable tool that can be used to catch students cheating. This can obviously have disastrous effects on young lives, because it is not as reliable as they suggest.
Per their own benchmarks, they do not achieve 100% accuracy even on text that is published on the Internet, and which is likely encoded into the models themselves.
There is validity to their goals, but that is overshadowed by the irresponsible way in which it is marketed.
(All of this, swirling in a context where students are being told that they absolutely must become proficient at using LLMs to do exactly this kind of work by the highest levels of state and federal governments, faculty leadership, as well as the leaders of the workforce into which they hope to graduate. The message to youth is extremely muddled at best.)
Always a pleasure reading Bryan's writing; it's like Bryan is sitting there with you and saying the words (hard to convey the feeling).
Having dealt with Enterprise Hardware(TM) in a previous job, it’s refreshing simply to see someone look at that pile of crap and go “it doesn’t have to be that way” and then actually set out to prove it.
I tried searching for good opensource / reasonably priced alternatives, pangram themselves even have some of their older architecture and training data on github/hugging face, but i never got that working reliably enough.
Glad those words proved prophetic!
[0] https://rustfoundation.org/media/how-the-rust-standard-libra...
AAUGH IT BURNS
Yes, quantity famously has a quality all its own, but perhaps not where correctness checks for something this central is concerned.
I wonder if that idea could be modernized now for this.
Unfortunately, it's not a browser extension and doesn't seem to have an API. I'd make a browser extension for this myself if it didn't involve paying for expensive Pangram usage.
I'm not entirely against using AI to help content creators improve their narrative, like finding common storytelling mistakes. But that's very different than using yourself as merely an avatar for LLM content.
To me, the glaring question is: What are we doing? The act of writing exists to 1) externalize and organize one's own thoughts for the purpose of considering and revising those thoughts; and 2) share one's own thoughts with other minds.
When we hand writing to a machine, we hand thinking to a machine, denying both our humanity and our role in the conversation.
No. I have good anecdata: readers cannot reliably distinguish my own prose from LLM-written one apart from cases where LLMs use odd metaphors or one of their specific patterns. I've been specifically experimenting with that.
It's driving me nuts, I constantly have to prompt it to "explain in plain, simple English"
Schwitzgebel, Strasser, and Crosby fine-tuned GPT-3 on Dennett's corpus and asked whether readers could pick Dennett's real answers to ten philosophical questions from four machine-generated alternatives, with no cherry-picking beyond mechanical length filters. Even Dennett experts averaged only 5.1 out of 10 (well below the 80% the authors predicted), blog readers got 4.8, and lay participants barely beat chance — though experts did rate Dennett's answers as more Dennett-like overall. Schwitzgebel stresses this isn't a Turing test (one-shot text is far easier to fake than extended interaction), but argues it foreshadows a future where machine outputs are humanlike enough that their moral status becomes genuinely uncertain, motivating his "Design Policy of the Excluded Middle": build machines that clearly lack moral status or clearly have it, not ambiguous ones in between.
My own take is : don't focus on the symbols on paper. focus on the facts about the world it is talking about. Isn't objectivity all about the facts? In future AI will have all the memory about what I have already read and it will just furnish the delta new information in the blog/writing so that I don't spend time on refreshing what I already know.
If a person reads AI generated text and does not notice, they by definition will not know about it.
There have been numerous cases of people accessing human created content as being AI.
There are instances where it seems relatively uncontroversial that it is AI generated, but without knowing both the amount of AI content people are exposed toand the amount that they register I don't think you can draw a conclusion of the overall state.
I do think the sensitivity to it can vary a lot: it depends a lot on how much and how closely you read the text, and how much exposure you have to LLM writing. Certainly it seems like a lot of people just don't really notice, or at least don't care much.
This is just "em dash redux." Except now we've moved on to accusing anyone who does "It's not X. It's Y." of being AI. In six months, it'll be "use of the word 'petrichor'" or something.
(TBH I think the biggest likelihood for false positives comes from heavy LLM users picking up their tics: it's a natural tendency and I've already seen a few cases where it seems like that has happened).
I'm also not reading pumpkin spice murder mysteries for a similar reason. I'm also not reading stories where everybody clapped. Actually, I'm already familiar with petrichor, so unless someone has surrounded the word "petrichor" with non-cliché prose, I'm also not going to read all that.
https://github.com/ucsandman/declick/blob/main/README.md
Please let me know if you
(a) believe this is human prose
(b) enjoy reading this prose
(c) would enjoy reading 100 READMEs like this.
As for invoking petitio principii and questioning other commenters' logical coherence [0], can you politely shove the argumentum ad Latinum up your ass?https://github.com/prathish-ks/isthmus/blob/main/README.md
https://github.com/prathish-ks/isthmus/blob/main/docs/thesis...
https://github.com/prathish-ks/isthmus/blob/main/docs/host-d...
https://github.com/prathish-ks/isthmus/blob/main/docs/threat...
https://github.com/prathish-ks/isthmus/blob/main/docs/threat...
https://github.com/prathish-ks/isthmus/blob/main/docs/baseli...
But Wait, There's More!
https://github.com/prathish-ks/isthmus/tree/main/docs
What—do—you—think————is this human?
LLM writing is verbose and meandering, but people are making a much bigger deal over this stuff than necessary for virtue signalling purposes. You don't want to read someone else's LLM writing? Get a summary of the page from yours. No time wasted, no pretentious posturing, and you don't make the error of assuming because the piece was written by an LLM that there was no thought put into the subject or there's no value in what is being communicated.
This statement does not logically cohere. "We can spot it because so many people make it easy to spot." You don't see how this is just petitio principii in action?
Maybe some people stop reading LLM slop purely because it violates their moral principles or whatever but most people bail out because slop is mentally painful to read. If you are a human and you write like today’s AI find a different writing style, not because reads like AI, but because it reads like shit.
There are lots of people who belong to the above groups, sure, but at least here in Germany Nazi is now applied to basically anyone who doesn't vote green, it's ridiculous.
It's common enough that it's training me to recognize and recoil from AI tics through sheer classical conditioning.
What about answers that an LLM gave to a question that we ourselves asked? Should we “labor” to understand that answer?
I think the argument, as presented in this and other similar pieces of critique, is too simplistic.
I do understand the criticism, but I think it should be framed in a different manner. The problem, when we read a long form piece by an author, is that we imagine that there’s another “mind” at the other side. We imagine that we are following the reasoning within the mind of a fellow human being, the writer. There’s an implied sort of “intimacy” to it. And the breach is when we are fooled into thinking that we are engaged in human communication, only to discover that there is a machine on the other side.
When we ask questions to an AI, this problem does not exist, because we are fully aware that the entity on the other side is not a human being.
Yet there is no doubt that the reply from an AI can contain information that is very much worthy of our time, and of our “labor” and effort to understand it.
So I think this ultimately will be about disclosure. As long as we are being made aware of the percentage of AI use in a text, explicitly or implicitly, I think we will actually grow to accept it.
I (am kinda forced to) use LLM to generate maybe 40% of the code at work, that is after my review and modifications. But I pretty much wrote all of the comments by myself. I can get into the flow by writing comments.
I liked your piece, and agree with almost all of it, but I'm surprised by your faith in the accuracy of Pangram at detecting AI writing. Is your faith based on testing it with lots of writing of known origins, or are you just saying that it reaches the same conclusion that you do as a talented human?
In particular, I wondered if you have tried running all of your own writings through it to verify that it thinks you are human. I was struck by Freddie deBoer's recent piece where he did this and said it often failed: https://freddiedeboer.substack.com/p/i-wouldnt-say-pangram-i...
What percentage of false positive rejections would you find acceptable? Would you accept this even if it forced you to change the way you write?
As for my own writing, I didn't do this experiment, but one of my co-workers did -- and over 176 posts spanning 22 years, all 176 (well, 177 now with my latest) are 100% human. This is not hugely surprising in that (in addition to me having actually written them!) my voice is very... distinctive. What would be more entertaining would be to try to get an LLM to write like me and fool Pangram that way. I still think that this would be difficult based on the experiences that I've heard, but it wouldn't surprise me if you could pull it off (and I would assuredly find the result entertaining!).
In the dimensions that we use Pangram in the most actionable sense (namely, to audit our own public writing), I am unconcerned about false positives, and leave it to Oxide authors to rework/recast as needed. (Though it sounds like Freddie didn't even need to do that -- he just needed to provide a longer sample.)
False negatives are mentioned, but the false positive is what could hurt people.
To human writing. Thank you!
https://www.atomic14.com/2026/08/18/detecting-claude-with-le...
It’s very hard to make reliable though. Different models have different characteristics and you can prompt your way out of being detected.
Thanks a lot for sharing!
Feeding it samples of a long-going conversation with Gemini 3.1 Pro is interesting. The first message seems to get flagged instantly, but later ones sometimes pass as human. Or at least more human-ish.
If I read the blogpost correctly, you've only "trained" on prompt<->response and not interactive sessions?
[0] https://bcantrill.dtrace.org/2008/11/03/concurrencys-shyster...
It’s now quite hard to get non AI training data…
Where I kind of disagree is that I don't think most readers will revolt. I think the mountain of LLM slop has actually changed people's behavior in more ways than one. Some are already relying on LLMs to summarize articles: then it doesn't matter to them who wrote it, they're just consuming machine-condensed content with no way to tell if a human or an LLM wrote the original piece. Or if their summarizer hallucinated.
Only saying "LLM writing" is honestly lazy writing. Specifically what?
I get the glaring cases, I get the idea that if the prose is generated then maybe also the idea, I get the feeling when reading a complete LLM authored piece.
But that doesn't help the piece, because - beside those glaring cases - most writing today is a mix between authors ideas and LLM prose.
I guess there is a kind of participatory element to the discourse where, if you want an audience, there is an editing process. Whereas in other cases, we wrote these as progress notes on an unknown journey, breadcrumbs or upturned stones to mark a path to the horizon.
Maybe it's the difference between writing as a mode of discovery, retreading the mental arc of a solution, and writing something honed to leave a mark.
The chief grief appears to be phoning in the whole process.
And in that case, no one needs to read it ai or not.
I wish that were true, but I fear it may not be.
https://arstechnica.com/ai/2026/07/canadian-legislator-reads...
Another great day where Google only gives me 2 results on the first page.
Pretty much sums up the issue re: workplace lazy AI dumping on folks as well.
> Why do people have this reaction? Beyond having to endure aggravating stylistic tics, when reading a piece that has had substantial LLM assistance, we — the readers — don’t know what is real and what isn’t.
This is well said. But, here too, I would pause and reflect on what it means to (think you) know what is real and what isn't in a pre-LLM setting. For example, authority bias predates LLMs, and can have disastrous consequences.
* If I need to learn something before I write about it, I rely on LLMs heavily to answer questions that I have about other source materials, e.g. to clear up ambiguities.
* I've recently started prompting it to find grammatical and spelling errors.
* And I've prompted it to find technical errors, places where I'm just wrong.
For all the prompting, I additionally tell it to not rewrite anything or offer any prose suggestions. It can keep all that to itself, thank you.
And I verify what it gives back for correctness.
(I'd encourage non-native speakers to use LLMs in much the same way. Don't sacrifice your human voice by letting the AI rewrite your words. Personally, I'd very much rather hear it from you, blemishes and all, than hear it from an AI.)
But if I could step back for a minute:
Why write anything?
If your writing goal is to flood the zone and make as much money as humanly possible from ads, then hell yeah, paperclip the everliving shit out of that.
But if your writing goal is to learn material or share material, then put that LLM on the back burner and don't use it to directly generate your text. It's bad for you, and the results are subpar.
When I'm learning something, I can go through reams of tokens and then, once I understand it, I digest that to single a paragraph about the topic. The paragraph is as concise and as helpful as I can make it. Now, I could just share the prompts that I went through with those pages of back-and-forth with the LLM... but wouldn't you rather just read the concise paragraph that gets the point across?
It's not hard to be better than an AI at writing for humans, so the minimum low bar to aim for is "better than an AI". And we can all get there with a small amount of practice. The real goal is to greatly exceed the LLMs' capabilities for sharing information.
Finally, I think everyone should write a lot. Blogs, morning pages, fiction, technical books, letters, whatever. Especially when it comes to technical content, nothing makes you do your research like putting your ass out in the ether to get flamed by 5 billion people. And teachers the world over know the best way to learn something is to teach it. Pick a topic, research, and write it up more clearly and concisely than anyone else ever has. You'll learn so much, and your readers will, as well. Writing fires up your brain. Don't give that up to an LLM.
Yes! Well said. If I already know someone, reading their own words, technical or businesss or personal, is meaningful to me. Warts and all. And if I don’t yet know the author then I definitely want to read their own words so I can get to know them.
Either way, taking the time to think and then write is a gift and I respect that.
this goes 10x for all the slide decks and google docs and wikislop everyone's trying to pass off as an accomplishment lately
Even then, I would say that using an LLM is robbing you of the process of writing, a process that is crucial to developing and understanding your own ideas.
Think about the last time you wrote something for consumption and the sentence to sentence thought processes you’re going through. I bet a lot of that was “is that right?” Or “does that make sense?” Or “am I communicating this at the level of my reader?”.
All of that is fundamental to your readers understanding, but more importantly, its fundamental to YOUR understanding.
[0] https://oxide-and-friends.transistor.fm/episodes/ai-detectio...
>This argument can probably be leveled at vanilla raw output from an LLM, but even the slightest attempt at obfuscation bears solid fruit
Whoops, disproven by bcantrill's comment:
https://news.ycombinator.com/item?id=49582629
Let's talk about the detection ability of corporate normies instead:
>pretty much everything “product” in corporate America is now LLM generated with some marginal oversight. It passes muster for the most part.
Goalposts: moved.
They just don't care to put the slightest attempt because they have a blindness to the problem. They are not doing it to intentionally mislead people.
You can still influence their writing style in a broad manner that might look correct at a glance, but the repetitive little patterns that give it away will always be there - if it was that easy to get rid of them, don't you think the AI labs themselves would've done it before releasing the models?
I dunno, man, according to Hardcover, I've read 76 fiction books this year, and I can't tell. All the "AI tells" fail the vibe check. I'm a writer and I get flagged by many of them.
And according to PhD linguists with expertise in the field, most AI tells are just the equivalent of old wives' tales. https://www.youtube.com/watch?v=ORgKY9AlybA
I vaguely recall that researchers were able to train people to tell, but only for a minority language that AIs likely aren't particularly good at mimicking, and after training.
This whole thing reminds me of how "you can recognize a vegan because they'll tell you." There, you have a ton of false negatives (i.e., since you aren't polling people to find out if they're vegan, you're only flagging the obvious vegans and missing all the regular people who happen to be vegan).
Except here, it's a bunch of false positives and negatives I bet. You don't really have a way of knowing, so you're accusing some people (without complete accuracy) and missing some people (without complete accuracy). But you have no way of knowing, so you're just like "hell yeah, my vibes tell me I'm right."
Research and experts disagree.
Is obvious AI-assisted writing better or worse than an obvious PR quid pro quo and/or cross-promotion?
That's pretty normal, but the point is that a blog post which is 35% Pangram promotion may not actually be less annoying than the use of AI to help write blog posts.
If you want to read something good, read a good book.
I'm much more interested in the content itself than the author that wrote it.
I don’t think this is remotely true. Sure, they’re legally obliged to let you unsubscribe, and sure, it’s not dick pills, but every US company will immediately send you a newsletter when you purchase something, review requests and, if they/you use Shop for checkout, expect an abandoned cart reminder.
PR pieces and software companies don’t write tutorials to be helpful, they are advertising to you. If the LLM can do it for cheap, they really don’t care.
Readers think they don't like LLM-authored text because they only recognize bad LLM-authored text as LLM-authored.
Blind trials have actually shown that readers generally prefer LLM authored books to human-authored ones on the same subject.
https://www.cambridge.org/core/journals/judgment-and-decisio...