HACKER Q&A
📣 levkk

Add flag for AI-generated articles


Should HN add the ability to flag articles as AI-generated? This doesn't have to act as a regular flag, i.e., it won't de-rank the article; it could just show up as an indicator, allowing others (like myself) who don't like reading AI-generated text, to skip it.

Open questions:

1. Why is the regular voting system not enough?

2. Should HN change in response to the gen AI era? It has been successful not changing fundamentals.


  👤 simonreiff Accepted Answer ✓
The recent rule addition to the Guidelines says this: "Don't post generated text or AI-edited text. HN is for conversation between humans." And I think that covers comments, but I'd be happy to see it also cover articles that are blatantly and primarily if not exclusively AI-generated. But how much AI is permitted? For instance: I'm writing a blog post now. It's all mine. If I include an AI-generated cartoon at the end, just to illustrate something, but not to be the whole or primary point of the article, is that AI-generated? Would the rule be conservative in nature to the extent that mostly human but clearly also AI-enhanced might get flagged but it's in the discretion of the moderators? How would you propose enforcing as to articles (versus comments which are usually quite obvious and thankfully have pretty much stopped being AI-generated since the rule was implemented, for the most part)?

👤 dawnerd
Considering YC invests in AI I doubt you’ll get anything of the sort. Too many people here also think you just have to give in and accept (abuser mentality IMO).

👤 jaredcwhite
I'm of the deepest conviction AI-generated text should not show up at all. Proving that however can be difficult (obvious LLM tells aside). Requiring evidence of authentic human authorship is also difficult, though increasingly I lean towards communities where that is a given for any legitimate shares.

👤 CqtGLRGcukpy
A problem I see is that what someone may consider to be AI-generated actually isn't. And the AI checkers aren't reliable enough to definitely enough say something is AI-generated.

👤 edoceo
Maybe just adding down-vote to submissions would do?

👤 JimsonYang
> why is the regular voting system not enough

Voting systems can be gamed and as HN becomes bigger and bigger it'll start to attract unsavory audiences who have an agenda.


👤 ranger_danger

👤 dang
We don't allow genai text on HN itself - see https://news.ycombinator.com/newsguidelines.html#generated and https://news.ycombinator.com/item?id=47340079. How to enforce this is a separate question, of course, but the rule exists.

We don't have a similar rule yet about article content but my sense is that the community mostly doesn't want to read it—or, to put it a bit more conservatively, sharply discounts it. This is why we see so many "just show me the prompt" responses, along with others like this: https://news.ycombinator.com/genai-pushback. I built that list so I have something to send to users who email about why their genai articles got flagged.

A fascinating arms race is happening: the AIs are training on the humans but the human hivemind is also training on the AIs. Readers are developing allergic sensitivities to language that sounds like an LLM produced it. The AIs will adapt to this, but the humans will adapt in turn. Where it ends up is anyone's guess. I have an optimistic view, but I've already been wrong about this so many times that I have low confidence in it.

The near-term picture is that there is a class distinction between writing (and writers) that use genai vs. writing that does not. That is, as soon as the allergic "this sounds like an LLM" reaction kicks in, the writing immediately gets relegated to a low-status bucket in the reader's mind. That doesn't mean it won't still get looked at - but it is now under a stigma.

(I was rather pleased with the originality of this observation until I remembered pg had written about "writes and write-nots" in https://paulgraham.com/writes.html. Oh well, it's the point that matters.)

This has the happy flipside that anyone who would like readers to classify their article in a high-status bucket, rather than a low-status one, can apply the judo move of simply writing it themselves.

None of what I'm saying is a dismissal of LLM technology per se. We rely on it heavily, and there's no question that it's useful. The question is how best to use it (pg again: https://x.com/paulg/status/2058871512451412457) and whether one should use it to generate or edit (<-- that bit is important) writing which one publishes to human readers.

To turn to OP's questions:

> Should HN add the ability to flag articles as AI-generated? [...] it could just show up as an indicator

Flagging as "just an indicator" would be tagging, which we've always resisted adding to HN, but I wouldn't rule it out.

> Why is the regular voting system not enough?

The regular voting system is never enough. (https://hn.algolia.com/?dateRange=all&page=0&prefix=false&so...)

[editing - bear with me... [you guys are adding replies faster than I can finish my own damn comment...]]


👤 Retr0id
Regarding 1, I think a) a sizeable fraction of voters are not able to recognize AI-generated text b) many who notice don't care, or are willing to overlook it if the premise is interesting enough. (The latter is true for me, on occasion)

Maybe we need a two-dimensional voting system: good/bad, ai/human. I think the second axis could cut down on meta-discussions over how much of the article was AI-generated.


👤 minimaxir
This is something that works better on paper in practice. Namely, there are a hell of a lot of false positives of AI use which frequently causes shitstorms on social media where someone says "AI?" in bad faith and now the OP has to defend themselves and in the case of writing a blog post there aren't as concrete ways to defend yourself. (no, demanding the edit history of the post is not reasonable)

Hacker News adopting such a feature would likely do more harm than good.


👤 IgorPartola
Nobody wants to label their stuff as AI generated because they removed credibility. Communities can flag posts as AI generated based on speculation and telltales but it won’t be 100% and will take extra work.

I think the era of the blog is simply dead now and that’s mostly ok. Blogspam and corporate blogs had killed quality bogs ages ago even before AI was a thing. The real question is what replaces it.

Oh and of course the $64k question is this: if an AI generated article is indistinguishable from a human written article and it is accurate and interesting, do you care who wrote it? We want to avoid low quality, not AI generation, right?


👤 nickandbro
Even if you did, how would you even enforce it? Say it was a pure text article, do you count the number of em dashes? Even AI detection scanners purpose built for this are extremely faulty.

👤 smallerfish
The voting system could be enough if downvoting was added.

AI writing is not the problem - low effort is the problem. Low effort AI articles are full of tics which are obvious, if you've done a lot of AI writing. To write well with AI you need to spend a good deal of time editing.

If you submit something that's low effort but has a clickbait headline that appeals to HN, you may well make the front page even if the article is lightweight (it does happen!) This is true both for AI and for human written articles.

On the flip side, somebody could spend an enormous amount of effort creating a masterpiece with AI. Penalizing that because of the tool that was used is arbitrary.


👤 mattas
Might be more appropriate to add a "not AI" flag at this point.

👤 jeremyjh
The regular voting system is not enough because posts can't be downvoted and for some reason many people are not bothered by the notion of reading something no one bothered to write.

The issue is complicated by the fact that there can be substantial effort invested in a process outside of the writing itself - and so AI written does not guarantee that the content will not valuable. But I'm inclined to punish it anyway to establish a norm of valuing genuine human communication. I think this norm has always been present but we didn't know until we'd really explored the alternatives.

I spend a LOT of time reading AI generated content because I use AI a lot for various purposes - maybe I'm more sensitive to its voice than some. AI voice always bothers me and its been getting more annoying the more I notice it, but there is a huge difference in reading responses to my own prompts and in reading the response to a prompt I haven't seen, when I don't know how many revisions there were, when I don't know if a human mind reviewed it at all before clicking send.

It becomes an unacceptable distraction because I don't know if I'm investing more time in the content than the author did, when in normal written communication the author would be putting in at least 5x the work.


👤 152334H
Most parsimonious explanation IMV: site staff can't see most AI slop. Reasons unimportant, but moderation systems are guaranteed to break down when the moderators themselves have poor classification ability.

A simple beneficial step that would lead to modest improvements and little downside: partner with Pangram. Either adding it as an automated spam filter, or by simply attaching the detection % to all posts.


👤 deadbabe
This makes sense if AI articles are bad or low quality, but what if one day, the AI generated content is actually good? As good or even better than what any human creates?

Is it purely just a "human supremacist" desire that fuels the motivation to ban or block such articles?


👤 matheusmoreira
That will only further increase the stigma surrounding LLMs. On Lobsters it actually got to the point where I no longer felt welcome on the site, even though I don't use LLMs to generate articles. The constant "this is AI slop" commentary is noisy and tiresome as well.

👤 user3939382
Great so I can use a CSS rule to hide anything with the flag.

👤 kgwxd
Sounds like a good job for AI. Why should humans have to waste their time on it? Accounts that post any should just get banned and deleted.

👤 senectus1
how are we detecting AI gen text?

Humans? We're not particularly effective at this as a whole...

AI service ? We'd probably have to pay for that AI to detect that AI and well.. Its also not particularly effective

Effectiveness is important, because we dont want real human produced data to be accidentally removed from view, just as much if not more so than having AI gen data being left on the site.


👤 wxw
+1, I would love to stop reading AI slop.

👤 sjs382
It seems like we get a post like this every other day or so...

👤 jvwww
Personally I just Pangram's chrome extension, which is amazing for spotting AI generated text.

👤 postalcoder
This is coming to hcker.news soon

👤 rddbs
I care more about hiding AI related articles than I do about AI-authored content in articles. Feels like most of the content on HN is now AI related and it’s just exhausting. Anyone have a way they’re correcting this balance of content, like a browser extension maybe?

👤 lazyasciiart
What about a title marker similar to [1997]? It’s content that people may or may not be interested in but will react differently if they know this about it’s origin.

👤 danieltanfh95
1. because your belief that AI-generated text is worse than human written text is flat out illogical and wrong.

2. judge content not by its cover and think.


👤 feverzsj
Just ban anything AI.