Ask HN: Add flag for AI-generated articles

1102 points by levkk ↗ HN
Should HN add the ability to flag articles as AI-generated? This doesn't have to act as a regular flag, i.e., it won't de-rank the article; it could just show up as an indicator, allowing others (like myself) who don't like reading AI-generated text, to skip it.

Open questions:

1. Why is the regular voting system not enough?

2. Should HN change in response to the gen AI era? It has been successful not changing fundamentals.

117 comments

[ 1.9 ms ] story [ 121 ms ] thread
Maybe just adding down-vote to submissions would do?
Considering YC invests in AI I doubt you’ll get anything of the sort. Too many people here also think you just have to give in and accept (abuser mentality IMO).
> why is the regular voting system not enough

Voting systems can be gamed and as HN becomes bigger and bigger it'll start to attract unsavory audiences who have an agenda.

A problem I see is that what someone may consider to be AI-generated actually isn't. And the AI checkers aren't reliable enough to definitely enough say something is AI-generated.
I'm of the deepest conviction AI-generated text should not show up at all. Proving that however can be difficult (obvious LLM tells aside). Requiring evidence of authentic human authorship is also difficult, though increasingly I lean towards communities where that is a given for any legitimate shares.
The recent rule addition to the Guidelines says this: "Don't post generated text or AI-edited text. HN is for conversation between humans." And I think that covers comments, but I'd be happy to see it also cover articles that are blatantly and primarily if not exclusively AI-generated. But how much AI is permitted? For instance: I'm writing a blog post now. It's all mine. If I include an AI-generated cartoon at the end, just to illustrate something, but not to be the whole or primary point of the article, is that AI-generated? Would the rule be conservative in nature to the extent that mostly human but clearly also AI-enhanced might get flagged but it's in the discretion of the moderators? How would you propose enforcing as to articles (versus comments which are usually quite obvious and thankfully have pretty much stopped being AI-generated since the rule was implemented, for the most part)?
That will only further increase the stigma surrounding LLMs. On Lobsters it actually got to the point where I no longer felt welcome on the site, even though I don't use LLMs to generate articles. The constant "this is AI slop" commentary is noisy and tiresome as well.
Regarding 1, I think a) a sizeable fraction of voters are not able to recognize AI-generated text b) many who notice don't care, or are willing to overlook it if the premise is interesting enough. (The latter is true for me, on occasion)

Maybe we need a two-dimensional voting system: good/bad, ai/human. I think the second axis could cut down on meta-discussions over how much of the article was AI-generated.

Sounds like a good job for AI. Why should humans have to waste their time on it? Accounts that post any should just get banned and deleted.
Great so I can use a CSS rule to hide anything with the flag.
The regular voting system is not enough because posts can't be downvoted and for some reason many people are not bothered by the notion of reading something no one bothered to write.

The issue is complicated by the fact that there can be substantial effort invested in a process outside of the writing itself - and so AI written does not guarantee that the content will not valuable. But I'm inclined to punish it anyway to establish a norm of valuing genuine human communication. I think this norm has always been present but we didn't know until we'd really explored the alternatives.

I spend a LOT of time reading AI generated content because I use AI a lot for various purposes - maybe I'm more sensitive to its voice than some. AI voice always bothers me and its been getting more annoying the more I notice it, but there is a huge difference in reading responses to my own prompts and in reading the response to a prompt I haven't seen, when I don't know how many revisions there were, when I don't know if a human mind reviewed it at all before clicking send.

It becomes an unacceptable distraction because I don't know if I'm investing more time in the content than the author did, when in normal written communication the author would be putting in at least 5x the work.

Might be more appropriate to add a "not AI" flag at this point.
+1, I would love to stop reading AI slop.
This makes sense if AI articles are bad or low quality, but what if one day, the AI generated content is actually good? As good or even better than what any human creates?

Is it purely just a "human supremacist" desire that fuels the motivation to ban or block such articles?

This is something that works better on paper in practice. Namely, there are a hell of a lot of false positives of AI use which frequently causes shitstorms on social media where someone says "AI?" in bad faith and now the OP has to defend themselves and in the case of writing a blog post there aren't as concrete ways to defend yourself. (no, demanding the edit history of the post is not reasonable)

Hacker News adopting such a feature would likely do more harm than good.

Most parsimonious explanation IMV: site staff can't see most AI slop. Reasons unimportant, but moderation systems are guaranteed to break down when the moderators themselves have poor classification ability.

A simple beneficial step that would lead to modest improvements and little downside: partner with Pangram. Either adding it as an automated spam filter, or by simply attaching the detection % to all posts.

We don't allow genai text on HN itself - see https://news.ycombinator.com/newsguidelines.html#generated and https://news.ycombinator.com/item?id=47340079. How to enforce it is a separate question, of course, but the rule exists.

We don't have a similar rule yet about article content but my sense is that the community mostly doesn't want to read it—or, to put it more conservatively, discounts it. This is why we see so many "just show me the prompt" responses, along with others like this: https://news.ycombinator.com/genai-pushback. I built that list so I have something to send to users who email about why their genai articles got flagged.

It's a fascinating arms race right now: the AIs are training on the humans but the human hivemind is also training on the AIs. Readers are developing allergic sensitivities to language that sounds like an LLM produced it. The AIs will adapt to this, but the humans will adapt in turn. Where it ends up is anyone's guess.

For the present, there is an emerging class distinction between writing (and writers) that use genai vs. writing that does not. As soon as the "this sounds like an LLM" allergy kicks in, the writing instantly gets relegated to a low-status bucket in the reader's mind. That doesn't mean it won't still get looked at - but it is now under a stigma.

(I was rather pleased with the originality of this until I remembered pg had come up with "writes and write-nots" in https://paulgraham.com/writes.html. Oh well, it's the point that matters.)

This has the happy flipside that anyone who would like readers to classify their article as high-status rather than low-status can apply the judo move of simply writing it themselves.

Now I need to add the disclaimer that none of this is a dismissal of LLM technology per se. We rely on it heavily, and there's no question that it's useful. The question is how to use it (pg again: https://x.com/paulg/status/2058871512451412457) and whether one should use it on writing that one publishes to other humans.

To turn to OP's questions:

> Should HN add the ability to flag articles as AI-generated? [...] it could just show up as an indicator

Flagging-as-just-an-indicator would be tagging, which we've always resisted adding to HN, but I wouldn't rule it out.

What I do think we'll (finally) add is a "please give a reason why you flagged this post" step, and "because I think it's genai" will be one choice among several (spam, offtopic, mean, etc.)

> Why is the regular voting system not enough?

The regular voting system is never enough. https://hn.algolia.com/?dateRange=all&page=0&prefix=false&so...

> Should HN change in response to the gen AI era?

To this I am tempted to reply with https://news.ycombinator.com/item?id=48887149 in homage to https://news.ycombinator.com/item?id=3742902.

This is a real conundrum.

For example, if I quote a GenAI response –even in criticism– (See what I did, there?), it can get flagged, and result in a shadowban (has happened to me -lesson learned).

But a good use for LLMs, is as a copyeditor. They do a great job. Some unedited stuff is so bad, I'd rather read slop, any day.

The problem is, what's the threshold? If they just fix a few typos and misspellings, that's fine, but what if they offer more substantial changes? How much text must change, before we can legit dismiss as "slop"?

Also, what if there's a significant GenAI component, but the article really is something that we want on the HN frontpage, because of its content?

We should try replacing forum mods with AI

just for a while :)

I wish!
I'm not volunteering, and I know you aren't asking, but why don't you have more moderators than you seem to? Do what other forums do and have regular mod drives, take applications from the community, take on more people during high traffic, etc. They don't all need to be employees of YC with backend experience do they?
Some of us use genAI as an accessibility tool. It enables people to write and publish work that otherwise wouldn't exist.

Some people already dismiss genuinely useful content solely based on the use of AI to assist in writing it - i am not sure what flagging would do other than to reinforce that prejudice.

I posted a couple of my articles here, and the one that got traction was generally well received (and also received some constructive feedback from those who acknowledged that is was AI assisted) - but it is evident across HN there is a vocal minority who outright dismiss content solely because it was "AI generated" completely disregarding the content itself. I appreciate this is personal taste, or LLM fatigue, or whatever, but its not really constructive.

If what you want to do is target the slop while not targeting the quality content, then that is what the voting mechanism already does. If people don't like something they can downvote it. Flagging content as AI generated is just a dogwhistle to those who want to downvote AI generated content. If anything, id rather see a rule that stops people commenting on stuff just to dismiss it as "LLM slop".

Its already trivial to avoid detection with fine-tuned humanisers [i] built on non-instruction-tuned models. That makes the flag mostly useless - or worse - a way of penalising any content you disagree with. I'd rather not hide what I am doing and have something that I feel reads well than hide it and sacrifice the message to satisfy a vocal minority.

[i] https://arxiv.org/abs/2605.19516

EDIT: downvotes, as expected. hope you see this anyway @dang. Downvotes kind of make my point for me.

> my sense is that the community mostly doesn't want to read it

I can confirm. Most LLM-written content is low effort, low value. This is somewhat by construction. You get the blandest takes in the blandest language.

>We don't have a similar rule yet about article content but my sense is that the community mostly doesn't want to read it—or, to put it more conservatively, discounts it.

It's definitely not universal. I've seen articles that seem clearly AI-generated, but still get upvoted because the community likes the title/thesis.

The quality of HN articles has degraded rapidly in the last year. It seems a meaningful fraction of articles posted here, especially most blog posts, are now AI generated. (Of course, this is the case for the rest of the internet too, but HN has always been a haven from the rest of the internet.)
I think of LLM tells like grammatical issues. If you read an essay full of grammatical mistakes you’d immediately start thinking less of the author, even if the essay isn’t about grammar. You wonder if someone who doesn’t pay close enough attention to catch a mistake “their” from “they’re” took attention to the rest of their work. This isn’t necessarily fair because the content of the essay might still be good. But on the internet I don’t have the time to evaluate the quality of every piece of writing I come across. It is very much the burden of the author to, as fast as possible, prove to me that the rest of the article will not waste my time. There is already so much content to read, and in some sense the amount of time to evaluate if an article is well-founded can be unbounded (imagine how long it would take to tell if an article about why a programming language is thoughtful without going out and also learning that language).

I find LLM-isms to be exactly the same as grammatical errors, but worse. At least when writing before you had to take the effort to type every word, so there was a minimum amount of effort you’d need to expend. If you aren’t catching obvious things like “the honest part” then that likely says bad things about your attention to detail elsewhere.

i love the allergy hall of fame and i was expecting to find many comments of mine
Why resist tagging?
I'm going to 'detach this comment and move it to' the second level because I'm late to the party and the chance you'll see it will drop from slim to almost none (unless you have a reply detector?).

https://news.ycombinator.com/item?id=48887942

> They can't know for sure whether what they're saying is true or false, and we can't know for sure how we should moderate it.

(Regarding people identifying text as genai).

You could moderate like another widespread but hard to prove problem: Propaganda and dis/misinformation campaigns. It's easy to see lots of people repeating well-crafted talking points (too well-crafted for an ordinary person), shouting down those who disagree. Yet users are forbidden from identifying it as such in comments; we're told to email hn@ycombinator.com. You could do that for genai suspicions too (assuming you have unlimited time to review such things).

It seems like a witch hunt to me; I have little reason or evidence to believe people know - because people speaking with no evidence is never reliable, because the angry mob is prone to witch hunts, because genai's are trained on human writing making it harder to distinguish, because some factors people target (e.g., em dashes) have been widely used, and because a highly disruptive technology at this stage of adoption is highly prone to strong emotional reaction and misunderstanding.

I'm not sure it matters: As a rule, we should hold the human who puts their name on it fully responsible. If their assistant, their genai software, or their dog writes it, it doesn't matter. Their name is on it. Speakers don't blame their speechwriters - the speaker said it.

I'm not sure it matters because right now it seems like a big reactionary over-response. I doubt we'll care much in a few years.

The automated spam potential, in posts and in comments, is a problem. One solution is somehow raising the standard of posts and comments: if it's that easy to generate middling crap, it makes better content more valuable and available and it may raise the standard naturally.

Thanks!

> I was rather pleased with the originality of this until I remembered pg had come up with "writes and write-nots" in (…)

Is Paul arguing that there will be people who can’t write because of AI? People have been crap at writing before AI, and much of Gen Z (and Alpha) literally don’t know how to write (not just how to write well) at ages where previous generations could.

That’s not his prediction, and not a prediction about technology, as claimed at the top of the post. School teachers could have told him that years ago.

It’s frankly dangerous that so many people lap up Paul’s words, when his world view is so distorted and out of touch with reality and devoid of understanding of regular people. He’s been rich and of high status for too long for his own intellectual good.

how are we detecting AI gen text?

Humans? We're not particularly effective at this as a whole...

AI service ? We'd probably have to pay for that AI to detect that AI and well.. Its also not particularly effective

Effectiveness is important, because we dont want real human produced data to be accidentally removed from view, just as much if not more so than having AI gen data being left on the site.

Nobody wants to label their stuff as AI generated because they removed credibility. Communities can flag posts as AI generated based on speculation and telltales but it won’t be 100% and will take extra work.

I think the era of the blog is simply dead now and that’s mostly ok. Blogspam and corporate blogs had killed quality bogs ages ago even before AI was a thing. The real question is what replaces it.

Oh and of course the $64k question is this: if an AI generated article is indistinguishable from a human written article and it is accurate and interesting, do you care who wrote it? We want to avoid low quality, not AI generation, right?

> Blogspam and corporate blogs had killed quality bogs ages ago

One of the main reasons that I (and I assume others) am here is because I can discover interesting content. It is true that there is a lot of spam in the internet but if I wanted that I would be in x, linkedin or sth. My problem with AI right now is that I consider machine generated content low quality one, and I would like to be able to decide if I want to read such an article without having to waste time before I realise it is ai generated.

> if an AI generated article is indistinguishable from a human written article and it is accurate and interesting, do you care who wrote it

Ideally I would like to know regardless. Practically in such a hypothetical future scenario it may be impossible, but I think that not doing sth right now because of some hypothetical future that may or may not happen is not a very good argument. Right now the AI content is pretty distinguishable, and if one took the time and effort to make it not seem like AI then at least that text contains some more human effort.

> do you care who wrote it?

I do.

> We want to avoid low quality, not AI generation, right?

We want to avoid further dehumanization of the already semi-dehumanized humankind.

I think the era of blogging is pretty alive. I find most of the signal on agentic coding in articles (Armin Ronacher, Simon Willison, etc.). It's just harder to find because of all the noise.
> I think the era of the blog is simply dead now and that’s mostly ok. Blogspam and corporate blogs had killed quality bogs ages ago

I simply instruct my RSS Reader to fetch articles only from blogs which I believe to be high quality.

> if an AI generated article is indistinguishable from a human written article and it is accurate and interesting, do you care who wrote it?

Well, yes, because of that "accurate" part. Humans, in general, actually attempt to validate what they publish. I don't feel the need to treat literally everything as a possible hallucination when I read an article published by a human. Less cognitive overhead makes for a more pleasant reading experience. In other words, I trust a human to be accurate most of the time, and I trust an AI to be accurate only some of the time, so there's much more to verify.

> if an AI generated article is indistinguishable from a human written article and it is accurate and interesting, do you care who wrote it?

I care about the fact that I’m interacting with a human. If an LLM generates content that reads like a human and it isn’t disclosed it is deceptive by nature. It doesn’t matter to me that a human typed the words on their keyboard, they could use text to speech or whatever. But I care that a human crafted the content. A LLM doesn’t just type the words, it takes style, tone, voice decisions, in addition to the choice of framing. That’s everything that matters about human to human communication

Even if you did, how would you even enforce it? Say it was a pure text article, do you count the number of em dashes? Even AI detection scanners purpose built for this are extremely faulty.
The voting system could be enough if downvoting was added.

AI writing is not the problem - low effort is the problem. Low effort AI articles are full of tics which are obvious, if you've done a lot of AI writing. To write well with AI you need to spend a good deal of time editing.

If you submit something that's low effort but has a clickbait headline that appeals to HN, you may well make the front page even if the article is lightweight (it does happen!) This is true both for AI and for human written articles.

On the flip side, somebody could spend an enormous amount of effort creating a masterpiece with AI. Penalizing that because of the tool that was used is arbitrary.

(comment deleted)