2 ms·
Interesting exploratory comparison, but I be cautious about treating it as a model benchmark With only three runs per model, the results are highly sensitive to
by plumb_samji 2mo ago
Interesting exploratory comparison, but I be cautious about treating it as a model benchmark
With only three runs per model, the results are highly sensitive to randomness