| ▲ | dannyw a day ago |
| This was tweeted about when it happened, with some explanation from Tibo here: https://x.com/thsottiaux/status/2076543065045795309 |
|
| ▲ | mkl a day ago | parent | next [-] |
| To see replies: https://xcancel.com/thsottiaux/status/2076543065045795309 The linked tweet is an unofficial reply to Tibo's official info and Tibo makes a correction in a reply. |
| |
|
| ▲ | causal a day ago | parent | prev | next [-] |
| Am I dumb or does this chart make no sense? Or why does the line only go up even with compaction? Or maybe "overall trajectory size" is hiding some meaning I don't understand? |
| |
| ▲ | chaos_emergent 16 hours ago | parent | next [-] | | The chart makes sense and is describing the cumulative cost of a trajectory. Cache tokens are created when a trajectory’s prefix is used more than once. A larger pre-compaction context window means that a greater number of cache tokens are used per turn, and a larger number of turns are completed before compaction runs. So you get a cumulative cost that grows quadratically until the compaction event. | |
| ▲ | crazylogger 11 hours ago | parent | prev | next [-] | | The Y axis is total cost in USD. For it to go down would mean OpenAI refunding you money. | |
| ▲ | ph4rsikal 20 hours ago | parent | prev [-] | | The blue line (200K context) is lower than the red line (300K context).
Indicating that it's cheaper to run blue rather than red over longer cycles. |
|
|
| ▲ | imgyuri a day ago | parent | prev [-] |
| How can the overall trajectory length be the same across reasoning efforts? I don't see how this is possible even if reasoning is not included in the trajectory length calculation. |
| |
| ▲ | chaos_emergent 16 hours ago | parent [-] | | I think Tibo was just keeping all else fixed and it’s an illustrative example rather than a perfect real-world trajectory. |
|