| ▲ | Terr_ 2 hours ago | |||||||
That rests on a false-assumption that the errors are statistically independent events, and have nothing to do with the shared nature of the judges. | ||||||||
| ▲ | daishi55 an hour ago | parent [-] | |||||||
Are there any reproducible hallucinations on any of the currently available OAI/Anthropic models? I’m not aware of any. And even if they are related - if Opus 4.8 always has a 1:100 chance of a specific hallucination - then running the same model twice does indeed dramatically reduce the odds of an error in the final output. | ||||||||
| ||||||||