6 ms·
It's not just that it's accessible, it's also significantly higher in quality than previous local runnable causal LMs. I suspect people saying it's not good ar
by Vetch 4y ago
It's not just that it's accessible, it's also significantly higher in quality than previous local runnable causal LMs.
I suspect people saying it's not good are prompting it like ChatGPT, not realizing how much trickier a raw model is to prompt. Getting the hyperparameters for good sampling is another stumbling block. The models are very good if you do everything properly.
- reasonabl_human 4y agoInteresting, where can I learn more about prompting, and tuning a raw model?
- simonw 4y agoThere are a few initial tips here in the LLaMAA FAQ: https://github.com/facebookresearch/llama/blob/main/FAQ.md#2-generations-are-bad https://github.com/facebookresearch/llama/blob/main/FAQ.md#2...