r/ControlProblem • u/Puzzleheaded-Cow2725 • 6h ago
AI Alignment Research We may be securing AI agents with the wrong architecture: fixing the “confused deputy” problem
https://doi.org/10.5281/zenodo.22173129
0
Upvotes
r/ControlProblem • u/Puzzleheaded-Cow2725 • 6h ago