Always key to include the one bench where the smaller model inexplicably outperforms the larger model