Remix.run Logo
schappim 4 hours ago

> "frequently contains all too plausible nonsense"

This really isn’t the case with frontier models in 2026.

I’ve found (sadly) that every time I thought the model was hallucinating, I was in fact the one who was mistaken.

consp 4 hours ago | parent | next [-]

N=1, and biased towards the type of questions asked. N+1, If I use one of the search engine sloptools I get frequent inaccurate answers which is I guess what the majority of people do.

tempfile 4 hours ago | parent | prev [-]

I am prepared to admit that the model usually doesn't get things outright incorrect (unless you ask it to count letters). But the solution does contain a lot of nonsense. Not false nonsense, but meaningless or irrelevant sentences that make understanding the core of the fix much more difficult.