| ▲ | senordevnyc 2 hours ago | |
Yeah, I thought reasoning was literally just chain of thought in the output token stream, with the model itself adding delimiters to indicate what part of the output is internal reasoning, and what part is an answer to the user. Is that wrong? | ||
| ▲ | beering an hour ago | parent [-] | |
You are right, reasoning is unrelated to tokens per second. | ||