Gsus4@mander.xyz to Actually Useful AI@programming.devEnglish · Aquileo | edit-27 months agoThe Lottery Ticket Hypothesis: finding sparse trainable NNs with 90% less params [2018]arxiv.orgAquileo | external-linkAquileo | message-square4Aquileo | linkAquileo | fedilinkAquileo | arrow-up17Aquileo | arrow-down10Aquileo | file-textcross-posted to: hackernews@lemmy.bestiver.se
Aquileo | arrow-up17Aquileo | arrow-down1Aquileo | external-linkThe Lottery Ticket Hypothesis: finding sparse trainable NNs with 90% less params [2018]arxiv.orgGsus4@mander.xyz to Actually Useful AI@programming.devEnglish · Aquileo | edit-27 months agoAquileo | message-square4Aquileo | linkAquileo | fedilinkAquileo | file-textcross-posted to: hackernews@lemmy.bestiver.se
Aquileo | minus-squarehowrarAquileo | linkAquileo | fedilinkAquileo | arrow-up5·7 months agoI don’t know about implementation, but a lot of theoretical work I’ve been seeing with regards to LLMs and other deep learning models appear to confirm the central claim of this paper. The most recent one I remember reading was this: https://arxiv.org/abs/2306.00978
I don’t know about implementation, but a lot of theoretical work I’ve been seeing with regards to LLMs and other deep learning models appear to confirm the central claim of this paper.
The most recent one I remember reading was this: https://arxiv.org/abs/2306.00978