Remix.run Logo
▲ CuriouslyC 2 hours ago

GPT-5's auto routing was bad, you'd get off the cuff bad answers for hard questions and overthinking on the banal. Also, with a chatbot the idea is that it's a back and forth but 5 minutes of thinking kills that, whereas fire and forget is core to the idea of autonomous agents.

Agents are going to replace almost all the UI in computers. For a trivial example, things like a control panel or settings menu will cease to exist, you'll handle coarse things via conversation ("set volume to 20%"), and UIs will be presented only for things where chat is cumbersome ("the password for the printer is !-*Donkey_Loves_Shrek^^@"). The UI that remains will become heavily decoupled, like supercharged versions of the office ribbons, with the intent that the agent will compose menu fragments and capability widgets dynamically so you only ever need to see controls for the settings that are relevant and that you care about.

▲ch_sm an hour ago | parent [-]

> "set volume to 20%"

haha oh god no please don’t.

EDIT:

Sorry for being dismissive. What I mean is, volume is something people generally don’t think about in percentages, but by feel/ear. It’s also adjusted so often and impulsively that doing it via chat would be extremely slow and unergonomic. So yeah, I don’t think that should happen.

▲CuriouslyC 38 minutes ago | parent [-]

Not saying you won't be able to invoke the volume control widget on a screen, but when harnesses are always listening, models are lower latency and you're interacting with a device from the other side of the room in a lot of cases (even PCs, via 72" wall mounted monitors), it'll be a lot more ergonomic.

There was a study that showed users "felt" faster on a keyboard even while completing tasks faster on a mouse. Similarly, with low latency models, "chromium volume 20%, actually, 10%" will be faster than clicking the sound control icon, finding the chromium volume from the list and reducing it, not to mention you don't have to look away from what you're doing. Direct interaction will probably only be faster if you need extremely precise control, for example, you wouldn't do digital painting with an agent, which seems pretty intuitive.