| ▲ | pixl97 2 hours ago | |||||||||||||||||||||||||||||||||||||||||||
>suspect that most the real gains are actually taking place in the harness. Part of the reason harnesses work well is you can run a lot of agents in parallel. That doesn't slow down demand. | ||||||||||||||||||||||||||||||||||||||||||||
| ▲ | hypfer 2 hours ago | parent | next [-] | |||||||||||||||||||||||||||||||||||||||||||
That is true, but the eventual realization that more machines doing more coin flips in parallel does not mean "more work gets done" might. LLMs are amazing tech, but they're terrible without oversight. More agents faster just makes reality collapse on them quicker. But yeah, you're right, temporarily, this will still push demand. But the topic was about "diminishing returns" as in "tech getting better". Not as in "customer spending". | ||||||||||||||||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||||||||||||
| ▲ | bunderbunder 2 hours ago | parent | prev [-] | |||||||||||||||||||||||||||||||||||||||||||
I had actually been thinking more about all the non-LLM functionality that go into the harnesses. I'm not going to name names and I haven't done any rigorous testing, but my general impression is that choice of harness matters more than choice of model. In terms of basic task completion success specifically, not code aesthetics. | ||||||||||||||||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||||||||||||