Skip to content

Comment on DyLoRA: Parameter Efficient Tuning of Pre-Trained Modelsparent

Comments

Highest posssible in which combination, though? If you’re fine tuning a model with N layers, then you could apply LoRA to any or all of them. Maybe it’s better to concentrate effort unevenly, in which case a uniform increase of adaptation rank (to compute budget) could still be subpar.

Right but the way that this paper proposes determining the best rank is by training a LoRA with the full rank.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.