3 ms·
Do something where there's no training data available. The AI companies are re-encoding large corpora of human creative work as a kind of compressed, conceptua
by 37ef_ced3 4y ago
Do something where there's no training data available.
The AI companies are re-encoding large corpora of human creative work as a kind of compressed, conceptually-indexed representation. This allows that work to be mixed and the patterns regurgitated in very flexible ways.
Everybody's output (our creative work) is assimilated by the AI, and becomes the AI's output.
For example, OpenAI's Codex is trained on about 54 million public GitHub repositories.
This allows Microsoft's CoPilot to regurgitate pieces of that code without attribution. Code from many sources is blended together and regurgitated for Microsoft's customers (without any acknowledgement of the source).
It is perhaps the greatest theft of intellectual property in the history of Man.
- wnkrshm 4y agoI've been wondering about giving artists some sort of adversarial filter to post their creations online and have it complicate or ruin training.