|
| ▲ | Aurornis 6 hours ago | parent | next [-] |
| The Qwen models, especially when quantized, can be really bad about just trying things and seeing what works. If you’re not watching all the tool calls you may not see it, but it’s kind of scary to watch them just bump into wrong decisions and backtrack. They also have a bad habit of accidentally building URLs that hit Alibaba infrastructure, likely because their training environment had them use those URLs. If you haven’t watched the outgoing network requests you might be very surprised at what your Qwen agents do sometimes. |
| |
| ▲ | julianlam 4 hours ago | parent [-] | | > [Humans] can be really bad about just trying things and seeing what works. If you’re not watching all the tool calls you may not see it, but it’s kind of scary to watch them just bump into wrong decisions and backtrack. |
|
|
| ▲ | reactordev 7 hours ago | parent | prev | next [-] |
| I’m not talking coding agents. More so agent harnesses and using LLMs for decision making |
|
| ▲ | cowboylowrez 7 hours ago | parent | prev [-] |
| heh I got a completely stupid error from gemini about some apache configs, when I pointed it out, the llm apologised, said the line(s?) should look like this, then posted the same two lines uneditted haha |