Comment on MTIA v1: Meta’s first-generation AI inference acceleratorComments−two_in_one3yI want one. This thing can run LLaMA 64b int8 easily.Meta is going to use it in datacenters, Much more efficient than NVidia generic GPUs. They are serious about putting AI everywhere.
Comments
I want one. This thing can run LLaMA 64b int8 easily.
Meta is going to use it in datacenters, Much more efficient than NVidia generic GPUs. They are serious about putting AI everywhere.