Our Community Often Overestimates Pangram's Accuracy

By SoniaAlbrecht 🔸 @ 2026-09-27T15:54 (+38)

The EA forum measures the level of AI usage in forum posts using the product Pangram. But how much does Pangram's "AI-Generated" label really indicate the degree to which an author has outsourced their thinking?

When they tested their 4.0 product, Pangram found that, by their definition, the proportion of AI-Assisted documents it classified as AI-Generated was 0.01%, 4%, or 7%, depending on the experiment. Then they omitted the experiments that found 4% and 7% false positive rates (FPRs) on their website outside their technical report, while advertising that the product detects AI-Assisted writing. Before I contacted them about this issue on September 17th, their claim on their website was more misleading—"99.9%+ Accuracy" was displayed directly adjacent to the phrase "Detects AI Assistance" on the text detection input box that many people don't read past.[1] Similar claims remain repeated elsewhere on the main page instead of by the text box itself. I do not know if my message was the cause of the change.

Their experiment that found the 0.01% FPR might be more appropriate for identifying human-written text rather than AI-Assisted writing. For example, the prompt they gave to Claude for this experiment was "Fix spelling, punctuation, and clear grammar errors only." In my experience, human editors normally provide conceptual feedback as well. Since we generally don't cite human editors, shouldn't the label AI-Assisted indicate more assistance from AI than would be provided by a human editor?

In my experience, our community also heavily relied on Pangram before their 4.0 release on July 29th. The previous version had FPRs for the AI-Assisted vs. AI-Generated labels for these experiments of 0.2%, 15%, and 22% respectively.

At least an order of magnitude is a big difference! But it may be much worse than that.

It’s likely that the false positives will be concentrated in the work of writers who write in a way that particularly confuses Pangram — the most elite writers not seeing false positives for their writing does not establish that their experience is typical. Sometimes I write in a style that Pangram defines as AI-Assisted writing in their experiments that found 4% and 7% FPRs. I spend so much time on these pieces that I am able to match almost every sentence to Pangram's definitions. They believe they can identify the AI involvement in every sentence, so every sentence gets an AI-Generated, AI-Assisted, or Human-Written label. I have run about 5,000 words of my writing that uses this process through Pangram 4.0, and it labels about 30% of this text AI-Generated, often with high confidence. Pangram claims to have much lower FPRs than false negative rates, so it is implied that this is an underestimate. The vast majority of my instances of false positives would not have been labeled false positives by Pangram’s paper—the 4% and 7% FPRs only counted instances where the entire excerpt was mislabeled AI-Generated. In general my excerpts are labeled Mixed. If my experience is common, the problem is much worse than a 4% or 7% FPR would indicate.

Sometimes I keep so much of the rewording from the LLM that reasonable people can disagree about whether that part is AI-Generated or AI-Assisted. However, these parts seem to be no more likely to be labeled AI-Generated than my distinct phrases or sentences that have zero AI input are. I have included a footnote with examples of these sentences with zero AI input that are objectively mislabeled. [2] 

Pangram's definitions of AI-Assisted rely on more objective measures than the subjective task of comparing prompts. However, to give a sense of the writing style that so confuses the product, Claude's prompt for the experiment that found a 4% FPR was "Substantially rewrite for polished academic style while preserving meaning." My process is similar. First, I provide an LLM many thousands of words of my writing and ask it to emulate my style in its suggestions. After I've produced writing I would have called finished a year ago, I give an LLM a prompt like "Please make the wording of the blander parts of this excerpt more poetic and visceral, to match the parts of it that are most in that style. Do not remove core concepts or add additional ones." After receiving a draft produced by such a prompt, I spend roughly six more hours per 1,000 words editing it. Overall, I reject the majority of suggestions the LLM gives me. I have included a footnote with a representative example of my process. [3] 

I believe the AI-Assisted label describes my process well. I incorporate the LLMs’ input more often than I do with human editors, so—especially given that they could be sentient soon—I want to list them as co-authors.  But many people besides me think there is a large difference between this writing and having an LLM entirely generate a piece on its own from a brief prompt. I have never met someone in our community who values a product that can distinguish between the two and was previously aware of how often Pangram struggles with this task. Some people do not value a product that can distinguish between the two, and only care about whether any AI input was used. That debate is beyond the scope of this post.

Like others, I have noticed that when I run my work by Pangram in chunks of a few hundred words instead of the whole piece, the FPR increases dramatically despite the writing being identical. Pangram acknowledges that its results are more uncertain for shorter passages. We should remember that Pangram's headline accuracy figures are unlikely to apply to the X posts and emails many of us are now using the product for.

Because of my high FPRs with Pangram, I only use AI editing in the way Pangram defines as AI-Assisted in fiction. It is better than having my more important work dismissed as “AI slop”. Because I come from a relatively disadvantaged background for our community, this makes it hard for me to compete with writers who have enough education to not require as much editing assistance. I am particularly concerned about how overreliance on Pangram might affect people like a friend of mine. He is a brilliant scientist whose native language isn't English. He once relied on me to help him sound fluent in his papers. I never had enough time for him, and he was overjoyed when AI editing could finally give him the English help he needed. He needed to use AI editing in the way that Pangram found 4% and 7% FPRs for. Indeed, one of the prompts Pangram used in the experiment that found a 4% FPR was "Sound fluent." Overreliance on Pangram could be depriving us of the insights of people like him, as well as native English speakers who are simply too busy doing interesting research to waste time specializing in writing instead. 

If AI editing typically made human writing worse, these points would be weak. However, in my experience, AI editing is very good now. I've experimented with almost every new Claude and ChatGPT update, and in 2026 their wording suggestions have gone from almost useless to usually better than the best I can produce. AI-assisted work frequently pops up on bestseller lists now. I invite people basing their opinions of the quality of AI assisted writing on examples from prior to the last six months to experiment with Fable and Astra and notice the difference for themselves.

Many people find Pangram’s results plausible because they believe they can notice whether something is AI-generated themselves as well, using AI writing tells. But this is a difficult task with such rapidly changing models, and many people who dislike AI editing tells describe a visceral feeling of stress when they see them. This is automation of a skill many of us have put a lot of love into developing—it is reasonable to feel some discomfort. But basing decisions on something as subjective as noticing AI tells while experiencing such stress seems likely to be error-prone. Someone posted a real Monet painting with the claim that it was AI and asked for an explanation of why it was inferior to the real thing. People responded with a flood of scathing critiques. Why wouldn’t a similar cognitive bias be happening with writing? Perhaps Pangram gives people who would normally be more careful a reason to abstain from deeper reflection. 

I actually think Pangram is a very impressive product. I've looked at about a dozen papers evaluating them, and believe their paper was perhaps the best. For example, they did an unusually good job of producing an AI-Assisted sample that is similar to how AI editing is really used. They are attempting a very difficult technical problem and doing it well. But I think being aware of the limitations of the product would be helpful when we evaluate the degree to which writers have outsourced their thinking, and which writers to trust.

  1. ^

    Pangram's text box prior to September 17th. The 99.9%+ accuracy claim has since been removed from the text box itself, but similar claims remain repeated prominently elsewhere on the main page. 

  2. ^

    Examples of my sentences that have zero AI input that Pangram labels 100% AI-Generated when part of their whole piece:

    "It had not been enough."

    "It is a celebration of some of the most essential things that make me who I am."

    "But this was no speech in front of Baghdad’s greatest scholars or the courage it had taken in her youth to carve out a name for herself among men."

    "Not the brilliant woman he once respected."

    "This is what I am now. This is what they made me."

    "Something new had entered the world, and he could feel it."

    "He was the god of the people with the gleaming spears. But he had grown weary of blind worship."

    "We entered the pool and emerged changed, touched by the divine in a way we would indeed carry for all eternity."

  3. ^

    The following is an example of my writing process that resulted in a passage that Pangram 4.0 called 100% AI-generated, part of a 3,000 word submission that was overall labeled Mixed. For context, all changes in meaning from the original draft are references to other parts of the piece. I selected the first passage Pangram mislabeled to avoid accidentally providing a biased sample. My editor warned me that the phrase I kept from Claude would be misunderstood by many people. However, I kept it because it was a particularly perceptive emulation of my writing style and reference to the themes of the book, and I originally didn’t plan to share my work publicly.

    Pre-AI draft: Erhan’s spies told him this was where Nasreen was now working, and he climbed the steps to her office hoping to find some documents left behind in a hurried evacuation. If his scholars could decipher some of her work, perhaps it could keep famine away for a while. 

    Claude's draft: Erhan’s spies said this was where Nasreen worked now, and he climbed the stairs to her office with a thief’s hope in his throat: papers left behind, a ledger, a packet of notes—anything his scholars could worry into meaning before winter made a mouth of the countryside.

    My final Draft: Erhan’s spies had told him this was where Nasreen was now working, and he climbed the steps to her office with a kind of hope that felt like thirst: for papers left behind, a ledger, a jar of precious seeds—anything his scholars could worry into meaning before another harsh winter came to collect what war had spared.


Austin @ 2026-09-29T16:41 (+6)

Thank you very much for writing this! We've integrated Pangram into Manifund, as a flag and visibility filter on projects; on net it's been very helpful due to the volume of low-quality AI submissions we get through the site. However, I am sympathetic eg to your example of a non-native speaker using AI to get their points across. Recently, I had a back and forth with a Chinese-speaking student on Manifund who started with an AI writing proposal which put me off, but I eventually decided to give him a small grant after more dialogue.

I appreciate your headline statement that Pangram's 99.9% accuracy is overstated as well, that was a useful update.

(Disclaimer, I've also recently made a minor ($10k) personal investment into Pangram.)

NickLaing @ 2026-09-28T11:16 (+4)

Thanks I think your overall point is fair.

My problem is that AI editing still brings about a regression to the mean, and still means we lose some of the human's voice. I agree it could make writing technically "better" in one sense, but that's not my problem with AI writing. Whether AI makes writing better or not isn't important to me when it comes to this question of pre-labelling writing. My problem is that voice-authenticity and writing diversity is lost when AI is involved.

IMO AI makes writing sound more sam-ey and for me at least less interesting. Others might be more excited to read AI edited/written work. That's great!

I'm OK with AI detectors overcalling it somewhat, if someone actually is using a lot of AI in the editing process. This doesn't mean they have done anything "wrong", its just tells me in advance that AI is heavily involved.

If someone is using AI in the writing process, that's fine it's no sin. I just think everyone should be able to know and then we can choose in advance how or whether we should read it. Pangram has been a game-changer and I appreciate it a lot!

Ryan Baker @ 2026-10-01T00:40 (+5)

There are many other human practices that bring about a regression to the mean in writing. In fact, most all education steers toward this, rather than away. Of course, as writing is for communication, a high degree of regression is inherently necessary for function. If language was random, rather than common, it would be non-functional.

Now, we still do value creativity between these norms, but is this argument that AI editing regresses to a mean more a reflection of a regression toward a shared communication mechanism or a regression of creative elements, like ideas or high level concepts?

I would find it very interesting if someone had asked that question before this phrase of "regression to the mean" became globally accepted as a description of a result of AI writing. The phrase may be mathematically correct, but the connotation attached to it is the second, while logically the first is probably predominant. The first is simpler and more basic. I'd expect the simpler and more basic factor to be more rapidly regressed. The second is more complex. To regress the second, the training process would first have to form representations of the concepts and ideas to regress toward. It should come later.

Also, keep in mind that "regression to the mean" in mathematics accepts a connotational conflict with "regress" more generally in English. In general use, there is progress and regress, and regress carries the negative connotation. "Regression to the mean" swallows both the progress and regress as one movement. You still might dislike a regression to the mean if you value the loss of the peaks more than the filling of the valleys, but you might wonder, if the peaks are simply held out, and not shaved off, is there loss at all?

I don't want to fully dismiss such concerns, I consider them unresolved. But, I do want to disrupt the consensus that has formed by a lazy acceptance of false premises.

Of course the best mechanism would be a test, but as far as I know, no one using that phrase has even considered such a test. So I imagine it will fall upon those who do not support it's use to do that work.

Examples of human practices that regress writing style, just to hit the tip of the iceberg:

  • Active voice
  • Confident writing
  • Study writers you like

Honestly, in many domains, we should just leave it up to the author to decide and not care. If they think the output is better than the input, and we can identify an idea or concept that's valuable inside, we shouldn't care what any tool says.

This isn't to excuse many ways of dishonestly seeking status or wasting people's time with meaningless drivel (two practices that commonly coincide). And yes, the easiest of all possible ways to do both of those is "Write a 2,000 word post arguing for my opinion being right / my product being the best", but it's merely at the top of a long list.

NickLaing @ 2026-10-01T07:07 (+2)

I agree many human practises do this, and we need it to be functional.

I just hate the extreme of it we get with AI assisted writing. I don't have a perfect answer to... "this argument that AI editing regresses to a mean more a reflection of a regression toward a shared communication mechanism or a regression of creative elements"

It just makes everything sound more the-same and less human to me. This is so clear to me and so many others I don't think it needs any kind of specific test, although tests have been done.

I agree we can leave it to the author to decide how to write, but the reader/consumer needs agency to decide what and how to read as well. I want to know in advance before I read something how much AI is involved so I can decide whether to read it slowly and enjoy, skim for content or just not read at all. I think labelling systems like for food safety or "made in XXXX country" usually bring more benefit than harm and AI assisted writing is no exception.

Ryan Baker @ 2026-10-01T15:37 (+1)

I'll reply in two parts, one dealing with the suitability of the phrase, a second dealing with the wider perspective.

I think you are not getting a proper sample set. I do not see how your uncontrolled observations could possibly bring back a clear result. I do see how you could easily make the mistake of thinking you did.

First, a few acknowledgments. Each particular model does have a "style". That style will have some cohesiveness. This will be moderately true for everything done, but especially true for prompt-generated content without any steering. That's the type of content we both agree shouldn't exist, or at least, shouldn't comingle with other writing.

I'd challenge this alone makes your clarity dubious. Much of that clarity will come from experience with that type of work. That's the type of work that will show up as 100% AI written in Pangram, even in long form.

Now, I'll skip toward the other end of the spectrum, and avoid describing every in between example. If an a person writes a full first draft, and then asks for critique, acknowledges some or most of the critique and requests a rewrite, some of the common style will enter, but less. Many of those common styles will be the ones a human editor would drive you toward. 

You'd also see more of them if a human editor actually rewrites part. Human editors almost never do this, probably for a host of reasons. Time of course is one. Convention, is a second. Conventions are influenced by schooling where the pedagogical value of rewriting is far less than critique that a writer has to puzzle through.

But I somewhat digress, my point here is that this content will likely show as AI assisted, or maybe more. You'll also probably see a tic or two show that triggers your AI radar. But I think you should be very careful about assuming that the effect of the assistance in the search of clarity has a lower ratio than the normal editing process. The normal editing process drives out many variations. The editing process that professional publications demand does more than others.

If the mean is defined as "professional writing", they are all regressing toward that mean, intentionally. If the mean is "the average of all writing", they may be progressing away from that mean.

Apologies for the long explanation, but you didn't accept my simpler initial explanation, deciding to substitute your unchecked impressions, so it seems necessary to spell that out more. 

You also did not consider the connotational aspects of the statement. I'd consider those in deciding whether it's really appropriate to use that phrase as a justification.

Now, I said I'd also deal with the wider perspective. The phrase is just one leg of a stool that holds up assumptions there. The wider perspective considers what people do.  For this I might want to point you to some prior writing of my own (https://substack.norabble.com/p/the-perpetrator-is-not-the-tool).

My point here is that for the problem you want to solve, Pangram is a pretty incomplete solution. If we're complacent enough to never build something better, I guess it's your best option. If you use it responsibly you'd filter the 100% AI only. That is not what people are doing overall. I myself have been bullied out of using AI for any editing. I want my writing to have a chance of being read, and it's clear the reach is less with even the smallest hint of AI influence. It's of course called dishonest, even though one of the first things I wrote when I restarted my public writing described my process (https://substack.norabble.com/p/writing-with-ai).

For about 6-months now my interaction is critique only to ensure my writing doesn't run into that type of reaction. I made a decision long before that about what I felt was going to help vs. harm my writing, and I did tune that over time. But this decision wasn't made with that in mind. This decision was made as a result of what I see as bullying. For you, I'd take some awareness of the wider group that your phrasing is shared with. I can tell your views are much better considered, but it doesn't change that you're under a banner in terms of common phrases.