3 ms·
This is why the "AI learns from materials just like a human does so it's not copyright infringement" argument always bothered me. A person won't recite full pag
by Crestwave 2y ago
This is why the "AI learns from materials just like a human does so it's not copyright infringement" argument always bothered me. A person won't recite full pages of word-for-word copies [1] from their head when you ask them something.
When I first tried Copilot, I asked it to write a simple algorithm like FizzBuzz and it ripped off a random repo verbatim, complete with Korean comments and typos. Image models will also happily generate near-identical [2] copies (usually with some added noise) of copyrighted images with the right prompt.
[1] https://bair.berkeley.edu/blog/2020/12/20/lmmem/ https://bair.berkeley.edu/blog/2020/12/20/lmmem/
[2] https://www.theregister.com/2023/02/06/uh_oh_attackers_can_extract/ https://www.theregister.com/2023/02/06/uh_oh_attackers_can_e...