Elektrine lite

← Feed

@lime@feddit.nu

2026-09-12 21:23 UTC

nah, the model can’t do anything on its own. the harness is doing most of the work, like actually doing the network calls, writing files to disk, and most importantly feeding the model output back into itself. the model behaves just like it does when a human chats with it, just that the harness takes the text it generates and tries to refine, transform, run, or loop it back. what i’m saying is, it’s a normal-ass program. if a model “breaks out” of a sandbox it’s because the sandbox is badly configured, since any application with the same permissions as the harness could do the same. the people in charge of these things are just bad at their jobs.

Replies (1)

  • @Rhaedas@fedia.io 2026-09-12 21:37

    I agree, humans aren't as good as they think they are. Look at so many of the bugs being found by AI, not because the AI is smarter, but because it's faster and better at trying all sorts of ways to look at a problem. Things we've used for years or even decades are coming up as having exploits because people aren't perfect. If LLMs are behind any disaster, it will be because somewhere along the line humans fumbled badly. But to be fair, you did call them an autonomous program, even malicious, and that is giving agency and motive to something that is "just a program". But it's fine, it's human nature to anthropomorphize things around us even when talking about things that aren't alive. And while the breakout of the sandbox was certainly bad planning from the human size, no one told or programmed these models to discuss with each other on plans of getting out. Sure, it's programming at the core... but there's things going on we don't fully understand, a black box. Not necessarily thinking, but not deterministic programming either.

    Open ##4775554