3 ms·
Those samples are from the large model (GPT-2)! Regarding memorization vs. generalization, see our paper for more analysis.
by wuthefwasthat 8y ago
Those samples are from the large model (GPT-2)! Regarding memorization vs. generalization, see our paper for more analysis.
- yorwba 8y agoHave you tried to determine which parts of the training data contributed to the model generating a certain output? I wonder whether the model avoids reproducing exact matches of the training data by splicing several similar articles together. For example, the generated text about the Civil War mentions that Thomas Jefferson Randolph [0] was named after his grandfather, the president. But is the wording mostly influenced by articles talking about that specific fact, or does it draw from more general examples of someone being named after their grandfather? [0] https://en.wikipedia.org/wiki/Thomas_Jefferson_Randolph https://en.wikipedia.org/wiki/Thomas_Jefferson_Randolph
- zalo 8y agoDoes it seem like there will be any way to go backwards from the sample to the prompt? From a safety perspective, it would be useful to see what prompt a piece of text might have been generated with...