example.com/path/to/article
000 points · username · 0 hours ago
example.com1102 points · 459 comments · 2 months ago · levkk
Open questions:
1. Why is the regular voting system not enough?
2. Should HN change in response to the gen AI era? It has been successful not changing fundamentals.
dang
IgorPartola
I think the era of the blog is simply dead now and that’s mostly ok. Blogspam and corporate blogs had killed quality bogs ages ago even before AI was a thing. The real question is what replaces it.
Oh and of course the $64k question is this: if an AI generated article is indistinguishable from a human written article and it is accurate and interesting, do you care who wrote it? We want to avoid low quality, not AI generation, right?
Retr0id
Maybe we need a two-dimensional voting system: good/bad, ai/human. I think the second axis could cut down on meta-discussions over how much of the article was AI-generated.
minimaxir
Hacker News adopting such a feature would likely do more harm than good.
dawnerd
nunez
I'm glad the idea is picking up steam.
IMO, the post title should get "[AI Generated]" at the end if enough people flag the content as AI.
shantnutiwari
Quoting myself: https://smackernews.com/item/48063759 HN
> Every post that reaches the top of HN will have at least a few comments saying "This is LLM!"
It has become a proxy for "I don't like this article, so it must be a LLM"
To me, it feels like lazy karma farming, as these comments often do get a few upvotes.
And of course, accuse a 100 posts if being LLM, you are guaranteed to be right at least once, then like astrologers you can claim success.
Is there anything we can do to discourage this type of lazy and low effort posting?
lazyasciiart
mattas
Debugreality
Personally I recently published a white paper on an idea for democracy that has been bouncing around in my head for decades. AI helped flesh out the idea and write the pages of text and structure the whitepaper.
I'm not an academic so I probably would never have fleshed out this idea without AI but I think it's better to have the idea published so it can be seen than to have remained bouncing around in my head until I passed away.
I'd agree that an academic version of my idea written by hand would be better. But a mostly AI written version is better than the knowledge never being published and I still spent a good two days on it making sure I agreed with every paragraph and fixed every issue I could find.
Considering this in general I think the amount of human effort that goes into creating something is probably the right measure of it's quality over if AI was used to assist in it's creation.
DevKoala
rddbs
thinkingemote
If you can easily detect the pattern of LLM writing here's something for you to look at: I was reading some pre-consumer-LLM papers by AI researchers and founders and the way these are written are incredibly similar to existing LLM prose!
CM30
But I'm not sure there's a great way to handle it. Flagging works as AI generated is good in theory, but it's become a bit of a witch hunt online, with plenty of human created pieces getting wrongly flagged as AI generated due to using em dashes in text, having the hands drawn awkwardly in artwork, or using other stylistic traits that AI content overuses. I fear half the site could end up flagged as AI-generated, just because a lot of people are hyper-vigilant about such content and assume the worst for everything.
At the same time, there's not really much of a way to incentivise writers to flag their own articles here, since revealing that a work is created by AI is a great way to both kill your credibility and drive away about half the people who'd otherwise enjoy it. So, if it's up to the authors, the incentives are for them to lie through their teeth.
I also don't really trust many AI detection tools, since they usually use AI themselves and misflag a lot of content.
But I don't think the site should change that much. Maybe add an option for AI content in the flagging system, and assume good faith for submissions in general. I'd rather not see the site get too paranoid or restrictive over this stuff.
jaredcwhite
ramon156
Works on a personal level, unsure if this would work well in practice. Maybe just a tag is enough, so people can conclude for themselves.
I know there are extensions out there that are doing the democratic part right. Mostly YouTube-related extensions like DeArrow and SponsorBlock
gverrilla
shahzaibmushtaq
AI isn't a higher power than HI and it never shall.
It is the author's and writer's duty and sole responsibility to tell viewers/readers that their respective works are AI-generated or how much percentage of it.
KingOfCoders
I don't care if a human, an AI or a cat wrote it.
the__alchemist
nickandbro
CqtGLRGcukpy
edoceo
amelius
We can do it already if we ask the AI companies to use one of the special whitespace characters instead of ascii 0x20. It would also help them avoid the problem of feeding their training loop on generated data.
aryehof
warshinder
itsgrimetime
Xotic007
PaiDxng
[deleted]
jeremyjh
The issue is complicated by the fact that there can be substantial effort invested in a process outside of the writing itself - and so AI written does not guarantee that the content will not valuable. But I'm inclined to punish it anyway to establish a norm of valuing genuine human communication. I think this norm has always been present but we didn't know until we'd really explored the alternatives.
I spend a LOT of time reading AI generated content because I use AI a lot for various purposes - maybe I'm more sensitive to its voice than some. AI voice always bothers me and its been getting more annoying the more I notice it, but there is a huge difference in reading responses to my own prompts and in reading the response to a prompt I haven't seen, when I don't know how many revisions there were, when I don't know if a human mind reviewed it at all before clicking send.
It becomes an unacceptable distraction because I don't know if I'm investing more time in the content than the author did, when in normal written communication the author would be putting in at least 5x the work.
atomicthumbs
rdataguy
mleroy
deevus
tukunjil
sjs382
simonreiff
antfarm
Esophagus4
kgwxd
nyellin
senectus1
Humans? We're not particularly effective at this as a whole...
AI service ? We'd probably have to pay for that AI to detect that AI and well.. Its also not particularly effective
Effectiveness is important, because we dont want real human produced data to be accidentally removed from view, just as much if not more so than having AI gen data being left on the site.
brador
A good article is a good article. Doesn't matter who wrote it.
Obfuscation of facts is the major issue with AI articles that should be bopped.
cloudpilot
smallerfish
AI writing is not the problem - low effort is the problem. Low effort AI articles are full of tics which are obvious, if you've done a lot of AI writing. To write well with AI you need to spend a good deal of time editing.
If you submit something that's low effort but has a clickbait headline that appeals to HN, you may well make the front page even if the article is lightweight (it does happen!) This is true both for AI and for human written articles.
On the flip side, somebody could spend an enormous amount of effort creating a masterpiece with AI. Penalizing that because of the tool that was used is arbitrary.
jvwww
lukasbm
chickenuggies69
pritesh1908
deadbabe
Is it purely just a "human supremacist" desire that fuels the motivation to ban or block such articles?
teiferer
admiralrohan
Ma4etaSS
user3939382
teo_zero
The day you get your "AI generated" flag, I'll want my "lousy with mistakes", too.
hgs3
I used to love reading HN for the handwritten articles and handmade projects, but in the LLM era the quality has deteriorated significantly. I find myself flocking to other message boards where LLM content is flagged, discouraged, or banned.
ivanjermakov
rienbdj
armanckeser
wartywhoa23
feedweave
ranger_danger
saint-evan
matheusmoreira
ThierryRakt
ltbarcly3
2. Labeling the ones that are honest about being AI generated punishes honesty and rewards lying by boosting AI generated articles that say they are not AI generated.
JimsonYang
why is the regular voting system not enough
Voting systems can be gamed and as HN becomes bigger and bigger it'll start to attract unsavory audiences who have an agenda.
arjie
The problem with these texts to me is that the parts that are information-dense are often not real and the majority is not information-dense. It’s just filler text of a sort that’s pointless “35% ram. 3x throughout. No latency trade off. That’s the whole point”. Okay, what’s this random “that’s the whole point” added there. Useless.
I know it’s passé to say “HN is becoming like X” but this is pure LinkedIn slop. Someone publishes pure bullshit and their fan club posts a bunch of likes and “I’m so excited to see this. Great post”.
pocksuppet
spwa4
danieltanfh95
2. judge content not by its cover and think.
sean_pedersen
[deleted]
oleggromov
wxw
xeyownt
Would you skip articles because it's written with a text processor? You need to have it written bit by bit by fusing it directly in memory?
AI / LLM is the new word editor. Get over it.
What I find really annoying is all the comments that pretend to see / detect AI slop... with lot of false positives.
feverzsj
152334H
A simple beneficial step that would lead to modest improvements and little downside: partner with Pangram. Either adding it as an automated spam filter, or by simply attaching the detection % to all posts.
We don't have a similar rule yet about article content but my sense is that the community mostly doesn't want to read it—or, to put it more conservatively, discounts it. This is why we see so many "just show me the prompt" responses, along with others like this: https://news.ycombinator.com/genai-pushback. I built that list so I have something to send to users who email about why their genai articles got flagged.
It's a fascinating arms race right now: the AIs are training on the humans but the human hivemind is also training on the AIs. Readers are developing allergic sensitivities to language that sounds like an LLM produced it. The AIs will adapt to this, but the humans will adapt in turn. Where it ends up is anyone's guess.
For the present, there is an emerging class distinction between writing (and writers) that use genai vs. writing that does not. As soon as the "this sounds like an LLM" allergy kicks in, the writing instantly gets relegated to a low-status bucket in the reader's mind. That doesn't mean it won't still get looked at - but it is now under a stigma.
(I was rather pleased with the originality of this until I remembered pg had come up with "writes and write-nots" in https://paulgraham.com/writes.html. Oh well, it's the point that matters.)
This has the happy flipside that anyone who would like readers to classify their article as high-status rather than low-status can apply the judo move of simply writing it themselves.
Now I need to add the disclaimer that none of this is a dismissal of LLM technology per se. We rely on it heavily, and there's no question that it's useful. The question is how to use it (pg again: https://x.com/paulg/status/2058871512451412457) and whether one should use it on writing that one publishes to other humans.
To turn to OP's questions:
Should HN add the ability to flag articles as AI-generated? [...] it could just show up as an indicator
Flagging-as-just-an-indicator would be tagging, which we've always resisted adding to HN, but I wouldn't rule it out.
What I do think we'll (finally) add is a "please give a reason why you flagged this post" step, and "because I think it's genai" will be one choice among several (spam, offtopic, mean, etc.)
Why is the regular voting system not enough?
The regular voting system is never enough. https://hn.algolia.com/?dateRange=all&page=0&prefix=false&so...
Should HN change in response to the gen AI era?
To this I am tempted to reply with https://smackernews.com/item/48887149 HN in homage to https://smackernews.com/item/3742902 HN.