| ▲ | andxor 5 hours ago | ||||||||||||||||||||||
> Performance is significantly higher than Fable 5.1 That's not clear. Need to see independent benchmarks first. | |||||||||||||||||||||||
| ▲ | forgot-my-pw 5 hours ago | parent | next [-] | ||||||||||||||||||||||
We need them pelicans on bikes. | |||||||||||||||||||||||
| |||||||||||||||||||||||
| ▲ | andxor 5 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
Artificial Analysis just published their aggregate score (61). Still below Fable 5, let alone Fable 5.1. EDIT: This is suspiciously low. Calls the relevance of existing benchmarks into question. | |||||||||||||||||||||||
| |||||||||||||||||||||||
| ▲ | forgot-my-pw 5 hours ago | parent | prev [-] | ||||||||||||||||||||||
AA benchmark: https://artificialanalysis.ai/articles/benchmarking-gpt-6-as... TLDR: it's about the same intelligence level as Opus/Fable, but it's suppose to be 70% more token efficient than GPT 5.6 Sol. So it's currently the new leader for cost efficiency frontier. | |||||||||||||||||||||||