| ▲ | spijdar 2 hours ago | |
Yeah, I dunno. For me it does "talk normally" for the most part when used in an actual coding harness. One thing though, the actual prompt I used was pretty long (844 words), and ... generated by GPT-5.6 Sol (lol), with the intent of "benchmarking" model performance in being able to write stories where the model avoids explicitly stating every detail in the prompt. I wonder if the GPT-produced stream could steer the generation into GPT-think territory. That's all I've got, though. Then there's the actual geometry problem from the stolen thoughts paper: | ||