12 ms·
Pelican: https://gist.github.com/SerJaimeLannister/8fdef9c00175da0ca6cd97f158b8f14d?permalink_comment_id=6334222#gistcomment-6334222 https://gist.github.com/Ser
by Imustaskforhelp 21d ago
Pelican: https://gist.github.com/SerJaimeLannister/8fdef9c00175da0ca6cd97f158b8f14d?permalink_comment_id=6334222#gistcomment-6334222 https://gist.github.com/SerJaimeLannister/8fdef9c00175da0ca6...
Aside from the pelican, I am sort of impressed by the fact that things are going the way in terms of really impressive small models.
Also I love how this uses N-gram embedding. I think that Longcat was the first one who used it (I submitted that submission on hackernews because I really just loved the idea of it that I understood), I am certainly more interested in local LLM models and its interesting how they are utilizing new architectures to do some really impressive optimizations!
(Do note that I created it using a free rate limited end-point that I found on the huggingface space section: https://victor-chat-with-qwen3-8-flash-next.hf.space https://victor-chat-with-qwen3-8-flash-next.hf.space)
- stymaar 21d ago> I think that Longcat was the first one who used it Wasn't it introduced by Gemma?