| ▲ | edude03 3 hours ago | |
I've been watching a bunch of bycloud on YouTube recently, and although he's done a great job reviewing papers from the big AI labs, I feel like I'm missing something - how have all the labs seemingly made a model that's cheaper, faster AND has better performance? Historically `flash` variants (like codex spark as well) have been faster but perform worse | ||
| ▲ | _3u10 3 hours ago | parent | next [-] | |
That’s how increasing performance works. You make a model 10x faster, then you make it think 2x as much. Its cost is now 1/10th per token, and 1/5th per task. Basically they have shitty hardware so they have to do a lot of optimization. Think of it like replacing an O(n) algorithm with O(log n). Anthropic / Open AI think the best path is the most intelligent models deepseek is more focused on tok/$ | ||
| ▲ | slickytail 3 hours ago | parent | prev [-] | |
[dead] | ||