Comment on Persimmon-8BparentComments−coder5433yscored 18.9 on HumanEval (coding) where Llama2 7B scored 12.2The article claims 18.9 for the base model, but also claims 20.7 for the fine tuned model.
Comments
The article claims 18.9 for the base model, but also claims 20.7 for the fine tuned model.