You realise that LLMs are already better at deciphering this than humans?

ElectroBuffoon · 2025-10-27T09:13:44 1761556424

What cost do they incur while tokenizing highly mistyped text? Woof. To later decide real crap or typ0 cannoe.

Trying to remember the article that tested small inlined weirdness to get surprising output. That was the inspiration for the up up down down left right left right B A approach.

So far LLMs still mix command and data channels.

63stack · 2025-10-27T10:00:51 1761559251

There are multiple people claiming this in this thread, but with no more than a "it doesn't work stop". Would be great to hear some concrete information.

nl · 2025-10-27T10:33:38 1761561218

Here you go:

https://chatgpt.com/share/68ff4a65-ead4-8005-bdf4-62d70b5406...

63stack · 2025-10-27T10:50:02 1761562202

I think OP is claiming that if enough people are using these obfuscators, the training data will be poisoned. The LLM being able to translate it right now is not a proof that this won't work, since it has enough "clean" data to compare against.

nl · 2025-10-27T11:17:09 1761563829

If enough people are doing that then venacular English has changed to be like that.

And it still isn't a problem for LLMs. There is sufficient history for it to learn on, and in any case low resource language learning shows them better than humans at learning language patterns.

If it follows an approximate grammar then an LLM will learn from it.

63stack · 2025-10-27T11:22:23 1761564143

I don't mean people actually conversing like this on the internet, but using programs like what is in the article to feed it to the bots only.

nl · 2025-10-27T12:08:05 1761566885

This is exactly like those search engine traps people implemented in the late 90s and is roughly as effective.

But sure.

michaelcampbell · 2025-10-27T12:12:37 1761567157

Was saying this 3x in this thread necessary?

63stack · 2025-10-27T20:12:58 1761595978

I'm just interested in opinions from all 3

hasa · 2025-10-27T13:44:52 1761572692

I thought it was a bot