| ▲ | docheinestages 7 hours ago | |
Labs are focusing on creating models, small or large, that perform well on various benchmarks, including general knowledge, domain-specific expertise, and agentic capabilities. Asking for such a model while wanting to be small and fast would align with what you're describing, which I believe is different from what I'm pointing to. The model I'm describing sacrifices domain knowledge and expertise for agentic reasoning and tool-calling capabilities at a reasonable speed. Think of Cactus Compute's Needle [1]. | ||
| ▲ | an hour ago | parent [-] | |
| [deleted] | ||