| ▲ | hxii 21 minutes ago | |
Interestingly, in my own benchmark and testing (in the hopes of finding a good-enough local model to run a personal assistant agent), Ornith-1.0-9B was worse than Qwen3.5-9B which according to their scores should've been reversed. I will definitely pass Ornith-1.5-9B through the gauntlet as well! | ||