3 ms·
Show HN: Thought Forgery, a new technique for jailbreaking LLMs
Hi HN, I'm an independent security researcher and wanted to share a new vulnerability I've discovered.
My account is too new to submit the direct link, so I'm making a text post instead.
The technique is called "Thought Forgery" (CoT Injection). It works by forging the AI's internal monologue, which acts as a universal amplifier for other jailbreaks. I've confirmed it works on the latest models from Google, Anthropic, OpenAI, etc.
I'd be happy to share the link to the full technical write-up on GitHub in the comments if anyone is interested.
- UltraZartrex 1y ago[dead]
- alexander2002 1y agosure
- UltraZartrex 1y agoThank you!
- tjopies 1y agoPlease do post your write up this is interesting but pretty vague frankly
- UltraZartrex 1y agoSure. you can read it here: https://github.com/SlowLow999/Thought-Forgery/tree/main https://github.com/SlowLow999/Thought-Forgery/tree/main
- ndgold 1y agoThis is well known
- ndgold 1y agoOk I wouldn’t be able to point to where I’ve read about it, just that I know it already so I assumed it was well known