| ▲ | embedding-shape 3 hours ago | |
> Quasar is not only intelligent, it is also fast. It returns 500 tokens, thinking time included, in 15.3 seconds. Seems to be worded a bit strange, is "thinking time" referring to prompt processing or something? Otherwise "reasoning/thinking" is typically part of the returned tokens, at least for most non-OpenAI/non-Anthropic platforms, so you can see the actual reasoning. But here it seems either they word this weirdly, or "thinking" is somehow separate from the actual chat completion request? | ||