| ▲ | aroman 2 days ago | |
The LLM is reasoning about estimates from its training data... which is to say, from human engineering timescales. I suspect the labs could improve the models such that they are estimating these sorts of things but they don't prioritize doing so (or perhaps RLHF selects it away) because, as you say, it feels amazing to do a week's worth of work in an hour. | ||