

We don’t know that… I also don’t think the current LLM architecture will achieve human-level thinking, but I do think that people will use agents & LLMs to try a 100 billion or so times to get us there… we’ll likely see many different architectures in the coming years as a result, and some that are close enough to be bloody scary (and some different enough to be even scarier!) We might very well still have LLMs as part of the bigger system.



It is dumb to just let a non-deterministic prediction machine loose with shell access and important credentials.
You either: verify changes, and verify commands against restrictive whitelists, falling back to manual approvals (that you carefully check) before letting them run, or…
You only enable YOLO in a sandbox with temporary/traceable limited access credentials (still some risks, but muuuuch lower) AND audit resulting code or…
You only give a limited set of tools that can do no real harm and can be audited after the fact… but this is more typical of assistants and predefined workflows than coding agents.
The catch: Good sandboxing requires a security mindset and technical knowledge to not screw it up… it’s just too easy to give agents more access than they actually need, and even the experts screw up too often. It’s the wild west out there…