Hacker News new | past | comments | ask | show | jobs | submit

Ask HN: Add flag for AI-generated articles

We don't allow genai text on HN itself - see https://news.ycombinator.com/newsguidelines.html#generated and https://news.ycombinator.com/item?id=47340079. How to enforce it is a separate question, of course, but the rule exists.

We don't have a similar rule yet about article content but my sense is that the community mostly doesn't want to read it—or, to put it more conservatively, discounts it. This is why we see so many "just show me the prompt" responses, along with others like this: https://news.ycombinator.com/genai-pushback. I built that list so I have something to send to users who email about why their genai articles got flagged.

It's a fascinating arms race right now: the AIs are training on the humans but the human hivemind is also training on the AIs. Readers are developing allergic sensitivities to language that sounds like an LLM produced it. The AIs will adapt to this, but the humans will adapt in turn. Where it ends up is anyone's guess.

For the present, there is an emerging class distinction between writing (and writers) that use genai vs. writing that does not. As soon as the "this sounds like an LLM" allergy kicks in, the writing instantly gets relegated to a low-status bucket in the reader's mind. That doesn't mean it won't still get looked at - but it is now under a stigma.

(I was rather pleased with the originality of this until I remembered pg had come up with "writes and write-nots" in https://paulgraham.com/writes.html. Oh well, it's the point that matters.)

This has the happy flipside that anyone who would like readers to classify their article as high-status rather than low-status can apply the judo move of simply writing it themselves.

Now I need to add the disclaimer that none of this is a dismissal of LLM technology per se. We rely on it heavily, and there's no question that it's useful. The question is how to use it (pg again: https://x.com/paulg/status/2058871512451412457) and whether one should use it on writing that one publishes to other humans.

To turn to OP's questions:

> Should HN add the ability to flag articles as AI-generated? [...] it could just show up as an indicator

Flagging-as-just-an-indicator would be tagging, which we've always resisted adding to HN, but I wouldn't rule it out.

What I do think we'll (finally) add is a "please give a reason why you flagged this post" step, and "because I think it's genai" will be one choice among several (spam, offtopic, mean, etc.)

> Why is the regular voting system not enough?

The regular voting system is never enough. https://hn.algolia.com/?dateRange=all&page=0&prefix=false&so...

> Should HN change in response to the gen AI era?

To this I am tempted to reply with https://news.ycombinator.com/item?id=48887149 in homage to https://news.ycombinator.com/item?id=3742902.

I think classifying it as an allergy or a status thing is a little too glib. I’ve read and reviewed, conservatively, hundreds of AI generated documents for work, and “written”/commissioned a bunch too. My biggest issue is that it’s impossible to engage with and give feedback on an AI written document, because it’s impossible to know whether misconceptions or gaps are because the author doesn’t understand the material deeply enough or the author does but the AI doesn’t and the author’s not proofreading carefully enough. Or if a surprising idea is raised — is it the authors insight, can they elaborate on it, where did it come from, etc?

Hackernews isn’t work, obviously, but “it’s impossible to engage deeper with the material because the author doesn’t really exist” is sort of a problem for a discussion site. If the human coauthor puts in enough work, they can make sure the doc really reflects their views and their understanding, but in my experience that’s much less common.

loading story #48889965
loading story #48893629
loading story #48888731
loading story #48890338
loading story #48900616
loading story #49005418
loading story #48888456
> The current picture is that there is an emerging class distinction between writing (and writers) that use genai vs. writing that does not. As soon as the "this sounds like an LLM" allergy kicks in, the writing instantly gets relegated to a low-status bucket in the reader's mind.

It's not a purity test, it's as the author is communicating they don't care whether the reader has any signals of what is accurate vs inaccurate information, which puts the burden of investigating how much is accurate on the reader at every step when there's some minimum expectation that should be an author's role (outside of topics where there is some expectation of ulterior motives/biases and one would naturally engage more critically minded).

When people complain here it's more often than not when an article has no disclaimer about AI use or what has been human-reviewed, so the burden again falls on the reader who is now even more skeptical. That is more to ask of a reader than when it's coming from say a known expert and the reader is receptive to engage and learn.

That's the reason tired cliches and turns of phrase (overused by LLMs) have become a heuristic for whether to pay attention, because it's a sign that there's some unknown quantity of of the article that hasn't had human review and it's easier to put in the bucket of 'maybe worthwhile but would need a fully human analysis of this' or just outright rejection (as we've seen from comments).

Edit: I see a sibling comment has raised the same observation.

loading story #48888567
> What I do think we'll (finally) add is a "please give a reason why you flagged this post" step, and "because I think it's genai" will be one choice among several

Does that mean that an article being AI-generated is a flaggable offense? Should we be flagging suspected AI-generated articles already, or should we wait for the flagging system to support reasons first?

loading story #48887907
I hope HN doesn't get into moderating the politics of articles.

I can see a grim future (present?) where "AI generated" turns into a slur, warranted or otherwise, in a world where the difference between human trained to talk like an AI and AI masquerading as human becomes increasingly difficult to discern, and some hidden cabal passes judgement.

That is wholly different from taking a stance on HN being a place for humans to comment on articles.

loading story #48889269
loading story #48887334
loading story #48888108
loading story #48888686
loading story #48887552
loading story #48888091
loading story #48887457
loading story #48887391
loading story #48887594
I'd prefer a tag to the mounds of "this looks like it is AI generated, I can tell from the pixels and from having seen quite a few AIs in my time" comments. That way the people who reject AI content can filter it out rather than having to argue about it in the comments.
If AI-generated comments are disallowed, why are AI-generated articles allowed? Seems like they have the same issues.
loading story #48888684
loading story #48888205
> This has the happy flipside that anyone who would like readers to classify their article as high-status rather than low-status can apply the judo move of simply writing it themselves.

I really really hope more people take up pen and paper! My last blog post [0] came with proof-of-work attached.

[0] https://abner.page/post/are-we-harold-bloom/

loading story #48892356
loading story #48888391
What I do think we'll (finally) add is a "please give a reason why you flagged this post" step

Do you believe adding friction to flagging will reduce the quantity of low quality articles?

Or is the flagging of high quality articles a bigger and more pressing problem?

Or is the problem simply too-many-damn-flags?

Just curious.

loading story #48888717
> The AIs will adapt to this

I don't think this is true, at least not right now, and in a way I'm actually thankful for it.

The frantic rush to chase the only potentially profitable use case for LLMs found so far (writing code) and the resulting focus on coding RLHF means models are actively becoming worse at sounding like humans.

This is my favorite example, and it's already relatively outdated: https://progress.openai.com/?prompt=10

I think of LLM tells like grammatical issues. If you read an essay full of grammatical mistakes you’d immediately start thinking less of the author, even if the essay isn’t about grammar. You wonder if someone who doesn’t pay close enough attention to catch a mistake “their” from “they’re” took attention to the rest of their work. This isn’t necessarily fair because the content of the essay might still be good. But on the internet I don’t have the time to evaluate the quality of every piece of writing I come across. It is very much the burden of the author to, as fast as possible, prove to me that the rest of the article will not waste my time. There is already so much content to read, and in some sense the amount of time to evaluate if an article is well-founded can be unbounded (imagine how long it would take to tell if an article about why a programming language is thoughtful without going out and also learning that language).

I find LLM-isms to be exactly the same as grammatical errors, but worse. At least when writing before you had to take the effort to type every word, so there was a minimum amount of effort you’d need to expend. If you aren’t catching obvious things like “the honest part” then that likely says bad things about your attention to detail elsewhere.

Community generated tags seems like the obvious solution to all of this. You could easily give a setting to turn them off, breaks no existing systems, and allows for a broad emergent taxonomy. Only surface tags above X community upvotes except to superusers who are allowed to propose tags.

But then again, there’s always reddit :)

loading story #48887436
The tricky thing about tags is that we get a tag for genai this year, what about next year’s thing and the year after that? We’d end up with a list of tags attached haphazardly all over the place. Flagging with a box to add a reason sounds like an excellent idea.
loading story #48887595
Sensible policy. I think with mainstream news publications now obviously using LLMs in their day to day workflows, it's going to be hard to take a purist stance here. Some do this more responsibly than others. But I don't think it's necessarily a bad thing if it is done responsibly. It's only a problem if it is done poorly.

Poor writing is not a new thing, of course. Most of the moderation mechanisms that it uses were perfected a quarter century ago when sites like Slashdot were popular as a defense mechanism against bad user behavior. Bad user behavior impacts commenting, article submissions, and moderating itself. While bad users now have AI to abuse, the problem of a large volume of low quality content is is largely the same. The moderation mechanisms end up targeting the problem at the source: identifying good and bad users. So, the same amount of bad users generating a lot more garbage isn't that big of a deal. Getting good karma still is a lot of work and it makes identifying all the garbage created by users without that relatively straightforward.

Using LLMs to tag, flag and filter content might not be a bad thing to experiment with. There are a lot of low quality AI generated opinion pieces that somehow make it to the front page. Same for political and controversial stuff, which of course is against the HN guidelines for content. Auto flagging things that obviously violate guidelines should not be that hard. It's just a matter of having good guard rails. @dang might actually already be doing that. I know I would be if staying on top of piles of generated garbage was part of my job description. It might also be done to give good content a little boost.

The new articles section has a very low signal to noise ratio currently and the window for good content to make it past that is very short. Often articles on the front page will have many duplicate submissions that never made it past that. IMHO duplicate submissions should just count as upvotes on the original. Auto de-duplicating based on canonical URL should not be that hard.

loading story #48896359
I would love to see a 'flag as AI' coupled with a profile setting to hide posts that have had a certain number of AI flags. Allow people who don't want to see likely AI content to filter it silently.
Great response. Thanks for weighing in and all the moderation you do!
>We don't have a similar rule yet about article content but my sense is that the community mostly doesn't want to read it—or, to put it more conservatively, discounts it.

It's definitely not universal. I've seen articles that seem clearly AI-generated, but still get upvoted because the community likes the title/thesis.

The quality of HN articles has degraded rapidly in the last year. It seems a meaningful fraction of articles posted here, especially most blog posts, are now AI generated. (Of course, this is the case for the rest of the internet too, but HN has always been a haven from the rest of the internet.)
The times I've pointed point out pretty blatant AI comments, I get nuked with downvotes. So often that I've stopped pointing them out.

Before you say I'm just falsely calling them out, it's typical ChatGPT style of either very amicable or Nobel Laureate tone, lots of formatting, with a couple of paragraphs and then a clever one-line punchline at the end. If you look at those commenters their history, it's all like that. Either generated or assisted. For older accounts you can see the steep increase of it around 2025ish.

Seems like the HN crowd absolutely adores AI comments and the rule banning them is (sadly) unnecessary. Or at least not what 'the people' want.

Out of curiosity I looked through your post history to find an example of a time you got downvoted for calling out AI comments. The first one I could find was 3 months ago (https://news.ycombinator.com/item?id=47493096), where you got downvoted for calling out an AI comment...to a comment that had zero common signs of AI writing.

The OP then replied:

> Not AI. Not sure how I feel getting my writing style called out like that though :D

loading story #48887645
OP wrote in the parent:

".. Fault-tolerant and highly available hardware must facilitate low-latency, single-threaded communication with high semantic density in order to achieve multi-dimensional consensus in a safety-critical, heterogeneous, adversarial environment. .."

I am not sure why you think someone saying "not AI trust me bro" carries any merit.

At any rate, like I said, I've given up the war. People enjoy reading that stuff, I'll just be the old man no longer yelling at the clouds.

loading story #48887497
loading story #48887703
loading story #48887499
to be honest, I hope you'll also recognize that non-native English speakers have limited vocabulary and phrasing when participating in HN.
IMO posting "This article is AI" does not add anything to the conversation.

The HN guidelines[1] include:

    Please don't complain about tangential annoyances—e.g. article or website formats, name collisions, or back-button breakage.
and

    Please don't post insinuations about astroturfing, shilling, brigading, foreign agents, and the like. It degrades discussion and is usually mistaken.

I'd argue pointing out that you think an article is AI is very similar in value to pointing out any of the above. None of us like AI slop. But I wouldn't be surprised if, by the end of 2026, 90-95% of articles posted online are AI slop. Pointing it out is useless. As useless as pointing out that the article breaks the scrollbar (which happens often) or that the article is formatted badly or has poor text contrast, or that an article is Chinese propaganda. Probably true, but posting about it adds nothing to the discussion, and is not allowed on HN.

All we really need is to add "Don't complain that an article is AI" to the guidelines.

1: https://news.ycombinator.com/newsguidelines.html

loading story #48888323
I don't think I agree with that. Complaining about AI written articles is more about the quality of the writing. it's on par with a piece of writing that wasn't proof read, well researched or some stream of consciousness rant.

I think the criticism also signals to the submitter or idk co-author? That the article isn't valued

loading story #48887407
loading story #48892057
loading story #48888521
loading story #48887365
loading story #48887408
> my sense is that the community mostly doesn't want to read it

I can confirm. Most LLM-written content is low effort, low value. This is somewhat by construction. You get the blandest takes in the blandest language.

A fixed dropdown list of flag reasons would be a very good change, I think, because it would somewhat counteract strategic flagging of stories as a downvote. I think you'd want to keep the list pretty tight, because it seems like a huge source of meta drama.
loading story #48888180
I think the problem space needs to be divided into two - AI generated articles (which the OP's question is about) and AI generated content/comments. The articles - honestly, normal flagging just works. If the content and the information in it is worthwhile, I don't think people would flag it. Isn't this the whole point of flagging?

However, the second part of the problem - the AI comments, that's really what kills the discussion and eventually the community. If I have trust issues that I'm talking to a bot but not a real user, I am not going to engage further. I might or might not flag the account, but at some point, I'm going to be tired of reporting if everywhere around me it's just bots pretending to be humans. I think this is the more serious problem that needs focus.

In my experience, sloppy AI content almost always, sometime instantly even gets flagged out.

My only thought is that "good" genai authors/articles will easily get through the filter while "generic" ones will fail. So the outcome won't be all genai articles getting flagged (unless people self-report), just the low quality ones. I'm guessing that's okay and I'm one of those people who discounts genai writing the second I read one of the tells, so the flag would save me time.

Ultimately though this is the same debate as "should we allow genai code in codebases". High quality code lands naturally while slop is slop. Not much value in banning AI outright--the desire is predominantly to ban the slop.

Maybe the tag should be [slop] rather than [genai]...

loading story #48888732
I find PG's article so dismissive.

I know several people that have difficulty writing but are still f*cking smart. So what, they should refrain from using AI to help them because the AI police says so?

But yeah, let's continue classifying people based on their outer qualities and habits... History showed us were this leads us to.

And here an em dash -- to freak out the AI police.

loading story #48888955
loading story #48889323
loading story #48891291
loading story #48888863
loading story #48894489
loading story #48889397
Let's make flags public, too. Flagged_by link.

Congratulations on your recursive ascension, too

loading story #48900214
Some of us use genAI as an accessibility tool. It enables people to write and publish work that otherwise wouldn't exist.

Some people already dismiss genuinely useful content solely based on the use of AI to assist in writing it - i am not sure what flagging would do other than to reinforce that prejudice.

I posted a couple of my articles here, and the one that got traction was generally well received (and also received some constructive feedback from those who acknowledged that is was AI assisted) - but it is evident across HN there is a vocal minority who outright dismiss content solely because it was "AI generated" completely disregarding the content itself. I appreciate this is personal taste, or LLM fatigue, or whatever, but its not really constructive.

If what you want to do is target the slop while not targeting the quality content, then that is what the voting mechanism already does. If people don't like something they can downvote it. Flagging content as AI generated is just a dogwhistle to those who want to downvote AI generated content. If anything, id rather see a rule that stops people commenting on stuff just to dismiss it as "LLM slop".

Its already trivial to avoid detection with fine-tuned humanisers [i] built on non-instruction-tuned models. That makes the flag mostly useless - or worse - a way of penalising any content you disagree with. I'd rather not hide what I am doing and have something that I feel reads well than hide it and sacrifice the message to satisfy a vocal minority.

[i] https://arxiv.org/abs/2605.19516

EDIT: downvotes, as expected. hope you see this anyway @dang. Downvotes kind of make my point for me.

loading story #48893821
loading story #48899843
loading story #48892167
We should try replacing forum mods with AI

just for a while :)

loading story #48913171
This is a real conundrum.

For example, if I quote a GenAI response –even in criticism– (See what I did, there?), it can get flagged, and result in a shadowban (has happened to me -lesson learned).

But a good use for LLMs, is as a copyeditor. They do a great job. Some unedited stuff is so bad, I'd rather read slop, any day.

The problem is, what's the threshold? If they just fix a few typos and misspellings, that's fine, but what if they offer more substantial changes? How much text must change, before we can legit dismiss as "slop"?

Also, what if there's a significant GenAI component, but the article really is something that we want on the HN frontpage, because of its content?

>This is why we see so many "just show me the prompt" responses, along with others like this: https://news.ycombinator.com/genai-pushback.

This is totally tangential to your point, but what is that page? It's not your usual link to an algolia search. Is this already part of some sort of manual tagging system? Clicking on the first one, these comments don't seem to be moderated. Are you using these complaints to help detect AI generated content? I think the existence of that page just leaves me confused on whether you actually want people to comment like this or not.

loading story #48888756
loading story #48888555
I'm going to 'detach this comment and move it to' the second level because I'm late to the party and the chance you'll see it will drop from slim to almost none (unless you have a reply detector?).

https://news.ycombinator.com/item?id=48887942

> They can't know for sure whether what they're saying is true or false, and we can't know for sure how we should moderate it.

(Regarding people identifying text as genai).

You could moderate like another widespread but hard to prove problem: Propaganda and dis/misinformation campaigns. It's easy to see lots of people repeating well-crafted talking points (too well-crafted for an ordinary person), shouting down those who disagree. Yet users are forbidden from identifying it as such in comments; we're told to email hn@ycombinator.com. You could do that for genai suspicions too (assuming you have unlimited time to review such things).

It seems like a witch hunt to me; I have little reason or evidence to believe people know - because people speaking with no evidence is never reliable, because the angry mob is prone to witch hunts, because genai's are trained on human writing making it harder to distinguish, because some factors people target (e.g., em dashes) have been widely used, and because a highly disruptive technology at this stage of adoption is highly prone to strong emotional reaction and misunderstanding.

I'm not sure it matters: As a rule, we should hold the human who puts their name on it fully responsible. If their assistant, their genai software, or their dog writes it, it doesn't matter. Their name is on it. Speakers don't blame their speechwriters - the speaker said it.

I'm not sure it matters because right now it seems like a big reactionary over-response. I doubt we'll care much in a few years.

The automated spam potential, in posts and in comments, is a problem. One solution is somehow raising the standard of posts and comments: if it's that easy to generate middling crap, it makes better content more valuable and available and it may raise the standard naturally.

Thanks!

i love the allergy hall of fame and i was expecting to find many comments of mine
> This is why we see so many "just show me the prompt" responses, along with others like this

Many are quite a bit more subtle, like this: https://news.ycombinator.com/item?id=48844062

The more subversive undercurrent is interesting to me. People intentionally fucking with someone's bot, burning tokens for the lulz.

loading story #48888437
Why resist tagging?
loading story #48899885
We already have tags for PDF and Video so I could definitely see one for [LLM] working!

Generally I'd appreciate a top level comment from the submitter saying this is LLM generated but I read it and found it interesting because x,y,z because I'd rather read slop someone vouches for than slop someone hasn't

Can we just make it official in the guidelines that "Articles written by genai" should generally not be submitted to HN? (And by extension, that they are okay to flag?)

(I already flag submissions I think were written by AI.)

lots of things happening in this post

1st: the presumption that AI generated text is actually unsuccessful, rather than proliferating broadly unchallenged today

2nd: the disposition that negative attitudes towards AI text are unjustified discrirmination, rather than working as a strong latent predictor of low-effort content

3rd: the assumption that human writing is reliably doled benefits, rather than some poor proxy of it (winning the social contest for claiming authenticity)

Operationally, only a very small minority of humans actually successfully identify AI generated text at rates ≥ Pangram. People discriminate against the label of "AI", but mostly fail to vote accurately. It's not uncommon to see bots abusing this gap for their own success -- accusing humans, sympathizing with generated profiles... FUD environment where people routinely get away with dismissing true accusations.

For someone who is mediocre at detection, this would structurally feel like an unhinged, unjustified bias: look at all these good posts, these honest people, getting undermined by discrimination...

loading story #48887722
> I was rather pleased with the originality of this until I remembered pg had come up with "writes and write-nots" in (…)

Is Paul arguing that there will be people who can’t write because of AI? People have been crap at writing before AI, and much of Gen Z (and Alpha) literally don’t know how to write (not just how to write well) at ages where previous generations could.

That’s not his prediction, and not a prediction about technology, as claimed at the top of the post. School teachers could have told him that years ago.

It’s frankly dangerous that so many people lap up Paul’s words, when his world view is so distorted and out of touch with reality and devoid of understanding of regular people. He’s been rich and of high status for too long for his own intellectual good.

I don't rely on LLMs and I don't find them useful, hence I don't use LLMs and everything about them deserves to be questioned.
loading story #48888447
loading story #48887196
loading story #48886992
loading story #48887127
loading story #48886918
loading story #48888339
loading story #48889899
loading story #48887882
loading story #48889327
loading story #48887075
loading story #48888404
loading story #48887870
loading story #48892848
loading story #48889591
loading story #48886964
loading story #48968028
loading story #48888651
loading story #48891890
loading story #48889490
loading story #48888719
loading story #48887265
loading story #48886961
loading story #48890703
loading story #48886898
loading story #48888805
loading story #48889555
loading story #48888842
loading story #48904324
loading story #48890202
loading story #48888165
loading story #48887074
loading story #48891564
loading story #48888596
loading story #48888665
loading story #48891522
loading story #48909851
loading story #48889410
loading story #48887716
loading story #48891105
loading story #48886976
loading story #48887053
loading story #48889117
loading story #48887192
loading story #48898220
loading story #48889677
loading story #48887306
loading story #48887569
loading story #48889345
loading story #48888410
loading story #48898533
loading story #48903946
This makes sense if AI articles are bad or low quality, but what if one day, the AI generated content is actually good? As good or even better than what any human creates?

Is it purely just a "human supremacist" desire that fuels the motivation to ban or block such articles?

loading story #48890007
loading story #48889330
loading story #48887403
loading story #48888818
loading story #48904152
loading story #48889535
loading story #48887061
loading story #48899217
loading story #48888723
loading story #48965968
Not too long until next gen LLMs to produce text indistinguishable from human's. Voting system is usually enough to filter out low quality submissions.
loading story #48889276
loading story #48891286
loading story #48891004
loading story #48923862
loading story #48886883
loading story #48893916
loading story #48886991
loading story #48893133
loading story #48886923
loading story #48888578
loading story #48893391
loading story #48895069
loading story #48891176
loading story #48887999
loading story #48887396
loading story #48898728
+1, I would love to stop reading AI slop.
loading story #48890773
loading story #48930845
loading story #48892625
loading story #48888147
loading story #48888416
loading story #48897187
loading story #48886993
loading story #48888052
loading story #48892159
loading story #48889745
loading story #48887266
loading story #48888309
loading story #48887798
loading story #48887145