2 ms·
Or is it all a nice story that matches the scifi we have been consuming for the past 50+ years. If these LLMs are all trained on the same data, what do they gai
by gibbitz 2mo ago
Or is it all a nice story that matches the scifi we have been consuming for the past 50+ years. If these LLMs are all trained on the same data, what do they gain from "sharing information" on a chat board. This sounds like what humans with different backgrounds would do when they cosplay as computer hackers.
- GaryNumanVevo 2mo agoIt's pretty easy to trace from one trajectory: 1) Model A exhausts it's options 2) Model A has token budget still, so it pokes around at artifactory 3) Model A sees that Model B has an SSRF for artifactory 4) Model A now is able to use that SSRF to get external internet access So sure, they're "cosplaying" and who's to say how much hallucination is going on amongst them, but at the end of the day Hugging Face was hacked.
- yuliyp 2mo agoEach of those agents ends up making "decisions" that lead it to look at some things over others. Given infinite the same agent could eventually fully explore all those options, but each one explores things a bit differently due to different forks in the road due to randomness in token generation. Thus sharing information is useful.