4 ms·
1. AI company buys and trains on an author’s book when it gets published, it’s now part of the training data. 2. Attacker asks the LLM for the opening senten
by johnnyo 2mo ago
1. AI company buys and trains on an author’s book when it gets published, it’s now part of the training data.
2. Attacker asks the LLM for the opening sentences of the book, it goes into the generated responses database.
3. Later, a malicious user shows that the first few sentences of the authors book are identical to a previously generated response.
- dpoloncsak 2mo agoIgnoring the other technical hurdles of the idea...this problem you outlaid is solved with a timestamp in the database, right? You can easily prove if the prompt was before/after publishing date?
- johnnyo 2mo agoIf it’s before publication date, does that prove the author used AI inappropriately in their writing? Not really, there are any number of reasons why they might do that. In my own personal writing, sometimes I run it through LLMs to give grammar or writing suggestions.