7 ms·
Wait until you hear about frankenmodels. You rip parts of one model (often attention heads) and transplant them in another and somehow that produces coherent re
by keonix 3y ago
Wait until you hear about frankenmodels. You rip parts of one model (often attention heads) and transplant them in another and somehow that produces coherent results! Witchcraft
https://huggingface.co/chargoddard https://huggingface.co/chargoddard
- GaggiX 3y ago>somehow that produces coherent results with or without finetuning? Also is there a practical motivation for creating them?
- keonix 3y ago> with or without finetuning? With, but it's still bonkers that it works so well >Also is there a practical motivation for creating them? You could get in-between model sizes (like 20b instead of 13b or 34b). Before better quantization it was useful for inference (if you are unlucky with vram size), but now I see this being useful only for training because you can't train on quants
- ShamelessC 3y ago> With, but it's still bonkers that it works so well Ehhhh…