| ▲ | nojs an hour ago | ||||||||||||||||||||||||||||
Can anyone comment on the economics and likely turnaround times of this process, when it’s more mature? Would it be realistic for a frontier lab to deploy this or would the turnaround time mean the model is always too out of date? Assuming the weights and architecture are eventually stable, how much cheaper would this end up being? | |||||||||||||||||||||||||||||
| ▲ | 2001zhaozhao an hour ago | parent | next [-] | ||||||||||||||||||||||||||||
There are always uses for outdated models. Claude Code is still using haiku 4.5 from ages ago for explore subagents for instance. Not to mention production uses like customer service that only need to be "good enough" | |||||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||
| ▲ | cogman10 an hour ago | parent | prev | next [-] | ||||||||||||||||||||||||||||
2 to 3 months optimistically assuming everything goes smoothly and is fully automated. 6 months or even a year if something goes wrong in the fabrication process and you need to update things. If they do more standard asic design, it could be a lot longer as the design needs to be validated on an FPGA cluster, which would necessarily need to be very big for something like a LLM. Easily up to 2 years. There's a reason chatjimmy isn't demonstrating newer models and why they only show of an 8B model. | |||||||||||||||||||||||||||||
| ▲ | shangofox an hour ago | parent | prev [-] | ||||||||||||||||||||||||||||
I mean even if it take a few months, it'll still be out of date. But there was a hypothetical when it came up in Feb, would you want Qwen 3.5 at like 10k tokens per second. At the time people were no doubt saying yes but now 3.8 is out, is that still desirable? | |||||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||