| ▲ | sreekanth850 4 hours ago | |||||||||||||
I think people here still evaluating the model in isolation. It is the combination that matters, model + strong harness + tools + long running autonomy + memory + retries + parallel agents + code execution + credentials + access to real systems. The model does not need to be perfect. If it fails 30% of the time, the harness can retry, verify, branch, use another agent and keep going. I don't think we necessarily need some magical AGI breakthrough first. The dangerous part may come from combining models that are already good enough with an extremely capable harness and enough access. | ||||||||||||||
| ▲ | contubernio 3 hours ago | parent | next [-] | |||||||||||||
People are underestimating the costs in terms of money and energy. The third law of thermodynamics is an essential barrier in all engineering. | ||||||||||||||
| ||||||||||||||
| ▲ | rob74 3 hours ago | parent | prev | next [-] | |||||||||||||
Doesn't this just move the need to be smarter from the model to the harness - if a human sometimes can't tell whether a model has produced something correct or just mostly correct-looking BS, how can an automated harness do it? OTOH, if the goal is simple ("break into a protected system") rather than more complex ("write an application that satisfies all requirements on all supported devices/screen resolutions etc."), that's of course more suitable for a harness. | ||||||||||||||
| ||||||||||||||
| ▲ | andai 3 hours ago | parent | prev | next [-] | |||||||||||||
ASI is just Claude in a while loop: | ||||||||||||||
| ▲ | copperx 3 hours ago | parent | prev [-] | |||||||||||||
> an extremely capable harness and enough access. Give enough access to a fuzzer and it's exactly as dangerous as an LLM. LLMs don't even have a moat in this domain. | ||||||||||||||
| ||||||||||||||