Comment on Benchmarks and comparison of LLM AI models and API hosting providersparentComments−djsh2yIf you like that speed, you would love Mixtral running at >500 tokens/s @ Groq https://www.youtube.com/watch?v=5fJyOVtOk4YIn full disclosure, I have worked on getting this up @ Groq.PS: Experience the speed for yourself, LLama2-70B, at https://chat.groq.com/
Comments
If you like that speed, you would love Mixtral running at >500 tokens/s @ Groq https://www.youtube.com/watch?v=5fJyOVtOk4Y
In full disclosure, I have worked on getting this up @ Groq.
PS: Experience the speed for yourself, LLama2-70B, at https://chat.groq.com/