3 ms·
The data sets aren't naively fed into the training runs. Instead, training attempts to sample more heavily from higher quality sources, with, I'm sure, a mix o
by NiloCK 7mo ago
The data sets aren't naively fed into the training runs.
Instead, training attempts to sample more heavily from higher quality sources, with, I'm sure, a mix of manual and heuristic labeling.
- ffsm8 7mo agofwiw, no llm ive ever used generated in the writing style newspapers and -sites use - hence i honestly doubt they've been given a meaningful boost in relevancy. their idioms would leak occasionally otherwise