| ▲ | daishi55 2 hours ago |
| Not really. Take hallucinations for example. If they are 1 in 100 (actually they are much rarer, but for the sake of argument), then the chances that 2 LLMs or even just 2 runs of the same LLM have the same hallucination is, well, a lot less than 1 in 100. |
|
| ▲ | Terr_ 2 hours ago | parent | next [-] |
| That rests on a false-assumption that the errors are statistically independent events, and have nothing to do with the shared nature of the judges. |
| |
| ▲ | daishi55 an hour ago | parent [-] | | Are there any reproducible hallucinations on any of the currently available OAI/Anthropic models? I’m not aware of any. And even if they are related - if Opus 4.8 always has a 1:100 chance of a specific hallucination - then running the same model twice does indeed dramatically reduce the odds of an error in the final output. | | |
| ▲ | Terr_ an hour ago | parent [-] | | If simply running things thrice-over was enough to stop "hallucinations" (and not incur other problems) we wouldn't be here talking about it today, it'd have been "solved" months or years ago. |
|
|
|
| ▲ | 2 hours ago | parent | prev [-] |
| [deleted] |