| ▲ | matt4711 2 days ago | |
> Besides the clear AI smell, this nonsensical claim also plainly contradicts the methodology's key evaluation claim that the quality of an engine's results should be measured against how much it overlaps with the reranked aggregate of the other engines. The benchmark thus seemingly values an engine's ability to "answer unanswerable questions" at zero. The engine itself is part of the reranked aggregate so if it finds something useful and everybody else does not it gets full credit. | ||