Remix.run Logo
▲ johnnyApplePRNG 2 hours ago

If it's a proper sandbox by definition, then yes.

https://en.wikipedia.org/wiki/Sandbox_(software_development)

▲grumbel an hour ago | parent | next [-]

A sandbox, even if 100% secure by itself, doesn't help when you use the agent to write code that you then executes outside the sandbox without checking, which is what everybody is doing at the moment.

The biggest hurdle for a full escape is that the agents don't have access to their own model weights.

▲simonw an hour ago | parent | prev | next [-]

Later in the article it points out that you need to punch holes in your sandbox in order to train the models - because the wheels exercises they are are training on need tools and data from outside that sandbox.

> Agents are most useful when they have access to information. That data can be drawn live from the Internet, which is fundamentally a two-way communications network. It can be information drawn from other (local) databases, or it can be the result of tool calls that themselves sometimes themselves result in network access. The more power you want from the agent — and for advanced agent RL and evaluation runs, you want a significant amount of power — the more information you’ll need to give it access to. Similarly, evaluations work best when the agent does not know that it’s definitely being evaluated. Sealing your agents behind glass makes this incredibly obvious.

▲johnnyApplePRNG an hour ago | parent [-]

>Later in the article it points out that you need to punch holes in your sandbox in order to train the models

You only "need" to do that if you desire the vibe coding experience.

I am perfectly capable, and I often do, download relevant materials for my coding agent to ingest locally.

Often times, the coding agent can't retrieve them programmatically anyways.

AI has ruined that ability for itself. (Nobody trusts anyone to scrape the web any longer)

▲_vertigo an hour ago | parent | prev [-]

No true sandbox..!