Comment on Understanding, using, and finetuning GemmaparentComments−neodymiumphish2yWhich still rounds to 9B and is 21.4% larger.−rasbtOP2yYes, it's definitely unfair to count it as a 7B model. In that case, we could call Llama 2, which is 6.6B parameters, a 6B (or even 5B) parameter model.−neodymiumphish2yExcept 6.6 rounds to 7. That’s completely reasonable. Arguing otherwise is pedantic.
Comments
Which still rounds to 9B and is 21.4% larger.
Yes, it's definitely unfair to count it as a 7B model. In that case, we could call Llama 2, which is 6.6B parameters, a 6B (or even 5B) parameter model.
Except 6.6 rounds to 7. That’s completely reasonable. Arguing otherwise is pedantic.