Skip to content

Comment on New attention mechanisms that outperform standard multi-head attentionparent

Comments

This is the paper, which appears to be cited in hundreds of others, some of which appear to be about efficiency gains. https://arxiv.org/abs/2105.03824

"Fnet: Mixing tokens with fourier transforms" (2021) https://arxiv.org/abs/2105.03824

https://scholar.google.com/scholar?cites=1423699627588508486...

Fourier Transform and convolution.

Shouldn't a deconvolvable NN be more explainable? #XAI

Deconvolution: https://en.wikipedia.org/wiki/Deconvolution

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.