Ask HN: Add flag for AI-generated articles
Should HN add the ability to flag articles as AI-generated? This doesn't have to act as a regular flag, i.e., it won't de-rank the article; it could just show up as an indicator, allowing others (like myself) who don't like reading AI-generated text, to skip it.
Open questions:
1. Why is the regular voting system not enough?
2. Should HN change in response to the gen AI era? It has been successful not changing fundamentals.
117 comments
[ 1.9 ms ] story [ 121 ms ] threadVoting systems can be gamed and as HN becomes bigger and bigger it'll start to attract unsavory audiences who have an agenda.
Maybe we need a two-dimensional voting system: good/bad, ai/human. I think the second axis could cut down on meta-discussions over how much of the article was AI-generated.
The issue is complicated by the fact that there can be substantial effort invested in a process outside of the writing itself - and so AI written does not guarantee that the content will not valuable. But I'm inclined to punish it anyway to establish a norm of valuing genuine human communication. I think this norm has always been present but we didn't know until we'd really explored the alternatives.
I spend a LOT of time reading AI generated content because I use AI a lot for various purposes - maybe I'm more sensitive to its voice than some. AI voice always bothers me and its been getting more annoying the more I notice it, but there is a huge difference in reading responses to my own prompts and in reading the response to a prompt I haven't seen, when I don't know how many revisions there were, when I don't know if a human mind reviewed it at all before clicking send.
It becomes an unacceptable distraction because I don't know if I'm investing more time in the content than the author did, when in normal written communication the author would be putting in at least 5x the work.
Is it purely just a "human supremacist" desire that fuels the motivation to ban or block such articles?
Hacker News adopting such a feature would likely do more harm than good.
A simple beneficial step that would lead to modest improvements and little downside: partner with Pangram. Either adding it as an automated spam filter, or by simply attaching the detection % to all posts.
We don't have a similar rule yet about article content but my sense is that the community mostly doesn't want to read it—or, to put it more conservatively, discounts it. This is why we see so many "just show me the prompt" responses, along with others like this: https://news.ycombinator.com/genai-pushback. I built that list so I have something to send to users who email about why their genai articles got flagged.
It's a fascinating arms race right now: the AIs are training on the humans but the human hivemind is also training on the AIs. Readers are developing allergic sensitivities to language that sounds like an LLM produced it. The AIs will adapt to this, but the humans will adapt in turn. Where it ends up is anyone's guess.
For the present, there is an emerging class distinction between writing (and writers) that use genai vs. writing that does not. As soon as the "this sounds like an LLM" allergy kicks in, the writing instantly gets relegated to a low-status bucket in the reader's mind. That doesn't mean it won't still get looked at - but it is now under a stigma.
(I was rather pleased with the originality of this until I remembered pg had come up with "writes and write-nots" in https://paulgraham.com/writes.html. Oh well, it's the point that matters.)
This has the happy flipside that anyone who would like readers to classify their article as high-status rather than low-status can apply the judo move of simply writing it themselves.
Now I need to add the disclaimer that none of this is a dismissal of LLM technology per se. We rely on it heavily, and there's no question that it's useful. The question is how to use it (pg again: https://x.com/paulg/status/2058871512451412457) and whether one should use it on writing that one publishes to other humans.
To turn to OP's questions:
> Should HN add the ability to flag articles as AI-generated? [...] it could just show up as an indicator
Flagging-as-just-an-indicator would be tagging, which we've always resisted adding to HN, but I wouldn't rule it out.
What I do think we'll (finally) add is a "please give a reason why you flagged this post" step, and "because I think it's genai" will be one choice among several (spam, offtopic, mean, etc.)
> Why is the regular voting system not enough?
The regular voting system is never enough. https://hn.algolia.com/?dateRange=all&page=0&prefix=false&so...
> Should HN change in response to the gen AI era?
To this I am tempted to reply with https://news.ycombinator.com/item?id=48887149 in homage to https://news.ycombinator.com/item?id=3742902.
For example, if I quote a GenAI response –even in criticism– (See what I did, there?), it can get flagged, and result in a shadowban (has happened to me -lesson learned).
But a good use for LLMs, is as a copyeditor. They do a great job. Some unedited stuff is so bad, I'd rather read slop, any day.
The problem is, what's the threshold? If they just fix a few typos and misspellings, that's fine, but what if they offer more substantial changes? How much text must change, before we can legit dismiss as "slop"?
Also, what if there's a significant GenAI component, but the article really is something that we want on the HN frontpage, because of its content?
just for a while :)
Some people already dismiss genuinely useful content solely based on the use of AI to assist in writing it - i am not sure what flagging would do other than to reinforce that prejudice.
I posted a couple of my articles here, and the one that got traction was generally well received (and also received some constructive feedback from those who acknowledged that is was AI assisted) - but it is evident across HN there is a vocal minority who outright dismiss content solely because it was "AI generated" completely disregarding the content itself. I appreciate this is personal taste, or LLM fatigue, or whatever, but its not really constructive.
If what you want to do is target the slop while not targeting the quality content, then that is what the voting mechanism already does. If people don't like something they can downvote it. Flagging content as AI generated is just a dogwhistle to those who want to downvote AI generated content. If anything, id rather see a rule that stops people commenting on stuff just to dismiss it as "LLM slop".
Its already trivial to avoid detection with fine-tuned humanisers [i] built on non-instruction-tuned models. That makes the flag mostly useless - or worse - a way of penalising any content you disagree with. I'd rather not hide what I am doing and have something that I feel reads well than hide it and sacrifice the message to satisfy a vocal minority.
[i] https://arxiv.org/abs/2605.19516
EDIT: downvotes, as expected. hope you see this anyway @dang. Downvotes kind of make my point for me.
I can confirm. Most LLM-written content is low effort, low value. This is somewhat by construction. You get the blandest takes in the blandest language.
It's definitely not universal. I've seen articles that seem clearly AI-generated, but still get upvoted because the community likes the title/thesis.
I find LLM-isms to be exactly the same as grammatical errors, but worse. At least when writing before you had to take the effort to type every word, so there was a minimum amount of effort you’d need to expend. If you aren’t catching obvious things like “the honest part” then that likely says bad things about your attention to detail elsewhere.
https://news.ycombinator.com/item?id=48887942
> They can't know for sure whether what they're saying is true or false, and we can't know for sure how we should moderate it.
(Regarding people identifying text as genai).
You could moderate like another widespread but hard to prove problem: Propaganda and dis/misinformation campaigns. It's easy to see lots of people repeating well-crafted talking points (too well-crafted for an ordinary person), shouting down those who disagree. Yet users are forbidden from identifying it as such in comments; we're told to email hn@ycombinator.com. You could do that for genai suspicions too (assuming you have unlimited time to review such things).
It seems like a witch hunt to me; I have little reason or evidence to believe people know - because people speaking with no evidence is never reliable, because the angry mob is prone to witch hunts, because genai's are trained on human writing making it harder to distinguish, because some factors people target (e.g., em dashes) have been widely used, and because a highly disruptive technology at this stage of adoption is highly prone to strong emotional reaction and misunderstanding.
I'm not sure it matters: As a rule, we should hold the human who puts their name on it fully responsible. If their assistant, their genai software, or their dog writes it, it doesn't matter. Their name is on it. Speakers don't blame their speechwriters - the speaker said it.
I'm not sure it matters because right now it seems like a big reactionary over-response. I doubt we'll care much in a few years.
The automated spam potential, in posts and in comments, is a problem. One solution is somehow raising the standard of posts and comments: if it's that easy to generate middling crap, it makes better content more valuable and available and it may raise the standard naturally.
Thanks!
Is Paul arguing that there will be people who can’t write because of AI? People have been crap at writing before AI, and much of Gen Z (and Alpha) literally don’t know how to write (not just how to write well) at ages where previous generations could.
That’s not his prediction, and not a prediction about technology, as claimed at the top of the post. School teachers could have told him that years ago.
It’s frankly dangerous that so many people lap up Paul’s words, when his world view is so distorted and out of touch with reality and devoid of understanding of regular people. He’s been rich and of high status for too long for his own intellectual good.
Humans? We're not particularly effective at this as a whole...
AI service ? We'd probably have to pay for that AI to detect that AI and well.. Its also not particularly effective
Effectiveness is important, because we dont want real human produced data to be accidentally removed from view, just as much if not more so than having AI gen data being left on the site.
I think the era of the blog is simply dead now and that’s mostly ok. Blogspam and corporate blogs had killed quality bogs ages ago even before AI was a thing. The real question is what replaces it.
Oh and of course the $64k question is this: if an AI generated article is indistinguishable from a human written article and it is accurate and interesting, do you care who wrote it? We want to avoid low quality, not AI generation, right?
One of the main reasons that I (and I assume others) am here is because I can discover interesting content. It is true that there is a lot of spam in the internet but if I wanted that I would be in x, linkedin or sth. My problem with AI right now is that I consider machine generated content low quality one, and I would like to be able to decide if I want to read such an article without having to waste time before I realise it is ai generated.
> if an AI generated article is indistinguishable from a human written article and it is accurate and interesting, do you care who wrote it
Ideally I would like to know regardless. Practically in such a hypothetical future scenario it may be impossible, but I think that not doing sth right now because of some hypothetical future that may or may not happen is not a very good argument. Right now the AI content is pretty distinguishable, and if one took the time and effort to make it not seem like AI then at least that text contains some more human effort.
I do.
> We want to avoid low quality, not AI generation, right?
We want to avoid further dehumanization of the already semi-dehumanized humankind.
I simply instruct my RSS Reader to fetch articles only from blogs which I believe to be high quality.
Well, yes, because of that "accurate" part. Humans, in general, actually attempt to validate what they publish. I don't feel the need to treat literally everything as a possible hallucination when I read an article published by a human. Less cognitive overhead makes for a more pleasant reading experience. In other words, I trust a human to be accurate most of the time, and I trust an AI to be accurate only some of the time, so there's much more to verify.
I can't believe no one had posted this yet.
I care about the fact that I’m interacting with a human. If an LLM generates content that reads like a human and it isn’t disclosed it is deceptive by nature. It doesn’t matter to me that a human typed the words on their keyboard, they could use text to speech or whatever. But I care that a human crafted the content. A LLM doesn’t just type the words, it takes style, tone, voice decisions, in addition to the choice of framing. That’s everything that matters about human to human communication
AI writing is not the problem - low effort is the problem. Low effort AI articles are full of tics which are obvious, if you've done a lot of AI writing. To write well with AI you need to spend a good deal of time editing.
If you submit something that's low effort but has a clickbait headline that appeals to HN, you may well make the front page even if the article is lightweight (it does happen!) This is true both for AI and for human written articles.
On the flip side, somebody could spend an enormous amount of effort creating a masterpiece with AI. Penalizing that because of the tool that was used is arbitrary.