3 ms·
> scored 18.9 on HumanEval (coding) where Llama2 7B scored 12.2 The article claims 18.9 for the base model, but also claims 20.7 for the fine tuned model.
by coder543 3y ago
> scored 18.9 on HumanEval (coding) where Llama2 7B scored 12.2
The article claims 18.9 for the base model, but also claims 20.7 for the fine tuned model.