| ▲ | LorenDB 9 hours ago | ||||||||||||||||
I've heard that DeepSeek v4 Flash 0731 has frequently assumed that it has vision capabilities and then resorts to inventing text-based image analysis tools when it finds that it actually can't see. In that case, this is a great upgrade for the model. Anecdotally, I had to tell 0731 to refrain from viewing screenshots since it kept breaking its sessions by trying to read images. | |||||||||||||||||
| ▲ | VulgarExigency 8 hours ago | parent | next [-] | ||||||||||||||||
It tried to recreate vision by analyzing pixels on 3 separate projects I had it working on. | |||||||||||||||||
| ▲ | trollbridge 6 hours ago | parent | prev | next [-] | ||||||||||||||||
I've mitigated this by giving it a "skill" that just means the harness using a different model. | |||||||||||||||||
| ▲ | mavamaarten 6 hours ago | parent | prev [-] | ||||||||||||||||
Yeah I've seen it a lot. It goes through the effort, unasked, of pulling screenshots off a connected device and then it's like... Oh shit yeah I can't see. | |||||||||||||||||
| |||||||||||||||||