4 ms·
The article shows (fine tuned) Mistral 7B outperforming GPT-4, never mind GPT-3.5.
by coder543 3y ago
The article shows (fine tuned) Mistral 7B outperforming GPT-4, never mind GPT-3.5.
- m3kw9 3y agoThis model is not close to even 3.5 from when I used it. It first of all does not follow instructions properly and it just runs on and on
- deleted 3y ago[deleted]
- coder543 3y agoWhat you're describing is the behavior you get from any base model that has not been instruction-tuned. The article is clear that this model is not for "direct use". It needs tuning for a specific application.
- m3kw9 3y agohow does one fine tune it to follow instructions? I would have thought they have open source training set for these instruction-follow fine tunes?