4 ms·
The reason is that actual human-written text generally doesn't follow the most probable next word rule, but rather there are occasional 'decisions' made which m
by f_devd 4y ago
The reason is that actual human-written text generally doesn't follow the most probable next word rule, but rather there are occasional 'decisions' made which make the text unique and therefore more interesting/coherent. There are other sampling methods which try to avoid this issue like "Locally Typical" and "Nucleus" sampling.