3 ms·
> In April we published a paper on a new training approach for better & faster LLMs using multi-token prediction. To enable further exploration by researchers,
by skilled 2y ago
> In April we published a paper on a new training approach for better & faster LLMs using multi-token prediction. To enable further exploration by researchers, we’ve released pre-trained models for code completion using this approach.
The HN discussion on the paper,
Better and Faster Large Language Models via Multi-Token Prediction - https://news.ycombinator.com/item?id=40220851 https://news.ycombinator.com/item?id=40220851 - May 2024 (128 comments)