Skip to content

Comment on 4T transistors, one giant chip (Cerebras WSE-3) [video]parent

Comments

but if we oversimplify things and assume that every synapse can be modeled as a single parameter in a weight matrix

Which, it probably can't... but offsetting those simplifications and 4-20x difference is the massive difference in how quickly those synapses can be activated.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.