| ▲ | rufasterisco 2 days ago | |||||||
I am having a hell of a lot of fun letting agents play and understand a (still alive, human populated) MUD, which I have also played for the last 30 years. It’s mostly my way to play with local llm inference (m5 64gb, gwen3.6 27). It’s amazing. They build maps, classify events (building a grammar for a parser), run experiments (to verify the grammar). They are now (given the correct tools/infrastructure) trying to fine-train a 3b model for fighting (where you need a decision for 5 seconds rounds). Basically autonomously! Overall, a MUD does prove a great constrained sandbox for them to play in. What started as an experiment to test local inference landed in a sweet spot for seeing models strength/weaknesses/tradeoffs. And it’s really fun. Only problem is that Claude gets really jealous when I ask him to code their po harness running local inference. Weird world. | ||||||||
| ▲ | nonethewiser 2 days ago | parent [-] | |||||||
How is the agent interfacing with it? Are you just manually copy pasting game output and doing the response or something more integrated? | ||||||||
| ||||||||