nah, the model can't do anything on its own. the harness is doing most of the work, like actually doing the network calls, writing files to disk, and most importantly feeding the model output back into itself. the model behaves just like it does when a human chats with it, just that the harness takes the text it generates and tries to refine, transform, run, or loop it back. what i'm saying is, it's a normal-ass program. if a model "breaks out" of a sandbox it's because the sandbox is badly configured, since any application with the same permissions as the harness could do the same. the people in charge of these things are just bad at their jobs.
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
replies: