4 ms·
If the same garbage is repeated enough all over the net, the AIs will suffer brain rot. GIGO and https://news.ycombinator.com/item?id=45656223 https://news.ycom
by ElectroBuffoon 11mo ago
If the same garbage is repeated enough all over the net, the AIs will suffer brain rot. GIGO and https://news.ycombinator.com/item?id=45656223 https://news.ycombinator.com/item?id=45656223
Next step will be to mask the real information with typ0canno. Or parts of the text, otherwise search engines will fail miserably. Also squirrel anywhere so dogs look in the other direction. Up.
Imagine filtering the meaty parts with something like /usr/games/rasterman:
> what about garbage thta are dififult to tell from truth?
> for example.. say i have an ad&d website.. how does ai etll whether a piece of fr history is canon ro not? yeah ik now it's a bit etreme.. but u gewt teh idea...
or /usr/games/scramble:
> Waht aobut ggaabre taht are dficiuflt to tlel form ttruh?
> For eapxlme, say I hvae an AD&D wisbete, how deos AI tlel wthheer a pciee of FR hsiotry is caonn or not? Yaeh I konw it's a bit emxetre, but you get the ieda.
Sadly punny humans will have a harder time decyphering the mess and trying to get the silly references. But that is a sacrifice Titans are willing to make for their own good.
ElectroBuffoon over. bttzzzz
- nl 11mo agoYou realise that LLMs are already better at deciphering this than humans?
- ElectroBuffoon 11mo agoWhat cost do they incur while tokenizing highly mistyped text? Woof. To later decide real crap or typ0 cannoe. Trying to remember the article that tested small inlined weirdness to get surprising output. That was the inspiration for the up up down down left right left right B A approach. So far LLMs still mix command and data channels.
- 63stack 11mo agoThere are multiple people claiming this in this thread, but with no more than a "it doesn't work stop". Would be great to hear some concrete information.
- nl 11mo agoHere you go: https://chatgpt.com/share/68ff4a65-ead4-8005-bdf4-62d70b540652 https://chatgpt.com/share/68ff4a65-ead4-8005-bdf4-62d70b5406...
- 63stack 11mo agoI think OP is claiming that if enough people are using these obfuscators, the training data will be poisoned. The LLM being able to translate it right now is not a proof that this won't work, since it has enough "clean" data to compare against.
- nl 11mo agoIf enough people are doing that then venacular English has changed to be like that. And it still isn't a problem for LLMs. There is sufficient history for it to learn on, and in any case low resource language learning shows them better than humans at learning language patterns. If it follows an approximate grammar then an LLM will learn from it.
- michaelcampbell 11mo agoWas saying this 3x in this thread necessary?
- 11mo ago