| ▲ | mydreamof 4 hours ago | |
Seems like benchmaxing? For example for Terminal-Bench 4 it doesn't have great results. And why not show other benchmarks? | ||
| ▲ | harmonic18374 4 hours ago | parent [-] | |
Probably, FrontierCode is made by Cognition itself. The model also seems worse in every way than DeepSeek v4.1 Flash, launched today. Also the submitter's account is very new which makes me suspicious of self-promotion. | ||