| ▲ | cromka 5 hours ago | ||||||||||||||||||||||
I was incredibly surprised how much better was Astra at rephrasing the docs that Fable came up for a project I am about to publish. The instructions are now actually what a human would expect, with clear logic flow diagrams, lists itemized, short paragraphs. It also actually correctly detected the train-of-thought leftovers, as well as stuff that never made it to final version of the project and removed them. Meanwhile Fable consistently ignores all my requests to write this exact way. I mean, the bare minimum I ask it for it to itemize lists and not write in single long passages using comas, semicolons and 'and's. Still ignores them. I honestly think it's time to call Astra the SOTA. It may not lead all the benchmarks but it genuinely feels much superior of a model. Not to mention the ¢20 Codex plan with frequent resets (https://codex-resets.com/) gives me roughly as much allowance as the ¢90 Claude plan, especially with recent limit cuts on Anthropic plans. | |||||||||||||||||||||||
| ▲ | hashstring 4 hours ago | parent | next [-] | ||||||||||||||||||||||
These resets also reset your weekly timer right. So it’s like, you may have 30% left for 1 days that you want to use. They reset it, and that means you your “new week” timer starts today. That sucks, because it doesn’t always work in your favour if you plan your weekly spend. I think a real reset shouldn’t also reset your week timer. | |||||||||||||||||||||||
| |||||||||||||||||||||||
| ▲ | _the_inflator 4 hours ago | parent | prev [-] | ||||||||||||||||||||||
Maybe check in the other direction as well: Astra to Fable. I frequently simply let one of the three review what something that looks like awesome output by one AI gets totally annihilated by the other. Finished outputs are easier to improve than bend a LLM to produce stuff like that in my observation. Same with Gemini. I yet have to find out how to handle this, whether I let agents check themselves and if on what process step. Tweaking is hard. I agree with your conclusion I am a huge ChatGPT and Codex fan, Gemini has to many infrequent quality changes when new models arrive ranging from great improvement to WTF. ChatGPT seems to get scaling well while Claude still feels unstable, unclear usage statistics. Really weird. Tough call I use all three. | |||||||||||||||||||||||
| |||||||||||||||||||||||