Skip to content

Comment on A new link to an old model could crack the mystery of deep learningparent

Comments

VGG also has regularization (L2, dropout, maxpool) if I remember correctly so it's sort of odd to use that as an example of overfitting when it explicitly tunes against that. I'm guessing the epochs and learning rate were also tuned to lower overfitting. The same is true to every other popular network out there.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.