3 ms·
> The thing we considered the most troubling was whether AIs could autonomously acquire resources by running businesses. So we decided to build that to validat
by TedDoesntTalk 18d ago
> The thing we considered the most troubling was whether AIs could autonomously acquire resources by running businesses.
So we decided to build that to validate and confirm our fears.
- theptip 18d agoI get your point, but it’s obvious that everybody is going to do this. So building a benchmark isn’t likely to directly move capabilities. OTOH it might give some good signals on required changes in training recipes.
- itake 18d agoPeople are still debating the liabilities associated with AI development. If your agent hacks the CIA, people want to blame the AI lab. but... If your agent spends $100k on tokens, then thats a user error. If your agent spends $10m on a shopping spree, then that is also a user error?
- theptip 18d agoYes, of course? Put a spend cap on your card like you would on your API key. At some point these systems will get certified as fiduciary agents but they sure as hell aren’t claimed to be that now.
- timr 18d agoYes, let’s not do any experiments at all. Then we’re sure to remain ignorant. That always works well.
- Agentlien 18d agoI was happy surprised to see the Swedish term "skräckblandad förtjusning" in the article. I've always loved that expression and I think it's a wonderful explanation to this seemingly crazy behavior. "Wow, that's terrifying! So exciting!"