Comment on Cerebras launches inference for Llama 3.1; benchmarked at 1846 tokens/s on 8BparentComments−eth0up2yMuch appreciated. Thanks for this!
Comments
Much appreciated. Thanks for this!