True, and makes sense that the logic is closing in but the breadth of the data is too narrow in 7GB's to ask questions about niche topics.
Mistral hasn't released their own official 13B/30B's yet, but i'm really looking forward to what they can do.
What is crazy is that Ultrafastbert, Speculative, Jacobi, or lookahead decoding could potentially speed up by up to 80x depending on size which could make GPT-4 like models feasible on entry level macs / Phones if similar wizardry is done memory wise.
..Yes im very optimistic after the insane progress over the last months with models like Mistral, Deepseek etc.
Comments
True, and makes sense that the logic is closing in but the breadth of the data is too narrow in 7GB's to ask questions about niche topics.
Mistral hasn't released their own official 13B/30B's yet, but i'm really looking forward to what they can do.
What is crazy is that Ultrafastbert, Speculative, Jacobi, or lookahead decoding could potentially speed up by up to 80x depending on size which could make GPT-4 like models feasible on entry level macs / Phones if similar wizardry is done memory wise.
..Yes im very optimistic after the insane progress over the last months with models like Mistral, Deepseek etc.