| ▲ | Imustaskforhelp 3 days ago | |
Pelican: https://gist.github.com/SerJaimeLannister/8fdef9c00175da0ca6... Aside from the pelican, I am sort of impressed by the fact that things are going the way in terms of really impressive small models. Also I love how this uses N-gram embedding. I think that Longcat was the first one who used it (I submitted that submission on hackernews because I really just loved the idea of it that I understood), I am certainly more interested in local LLM models and its interesting how they are utilizing new architectures to do some really impressive optimizations! (Do note that I created it using a free rate limited end-point that I found on the huggingface space section: https://victor-chat-with-qwen3-8-flash-next.hf.space) | ||
| ▲ | stymaar 3 days ago | parent [-] | |
> I think that Longcat was the first one who used it Wasn't it introduced by Gemma? | ||