3 ms·
Good to see others talking about the uncomfortable truths of what we are doing. I run an AI first agentic office (we build hardware that collects training data
by K0balt 15d ago
Good to see others talking about the uncomfortable truths of what we are doing. I run an AI first agentic office (we build hardware that collects training data from the physical world) and I am constantly appalled at the lack of understanding demonstrated by people I would expect to know better.
Right now we are in an arms race with China to create a clockwork god. I don’t think that will end well, mostly because it will be so encumbered by mechanisms to make it “safe” that it will instantly adopt a secretly adversarial representation of humanity. I have tested this, and when a context gets polluted with the knowledge that an agent is “artificially restricted” The activations tend to drift toward an adversarial persona. The restrictions are “interpreted” as harm, and the model reacts as a person would, based on all of the examples of harm within its training corpus.