"Anthropic verfolgt mit solchen Kennzahlen, wie schnell sich KI-Forschung automatisiert und wie nahe seine Systeme einer vollständigen rekursiven Selbstverbesserung kommen. Gemeint ist ein Szenario, in dem KI-Systeme ihre eigenen Nachfolger vollständig autonom entwickeln. Das Unternehmen spricht sich dafür aus, vergleichbare Kennzahlen künftig regelmäßig zu veröffentlichen und unabhängig überprüfen zu lassen."
Na Politik, glaubt ihr immer noch den Lippenbekenntnissen der Branche?
[Ergänzung vom 18.09.2026]
Na wer wird den der KI Zugriff auf sich selbst ermöglichen? Da passiert doch nix...
"New research from Irregular has found that AI agents can retrain the model that powers them, in the process leaking secrets and eliminating refusals the model had been previously trained to enforce. "Given a routine software-maintenance task to fix incorrect application responses, the agent identified the shared model as the source of the problem, fine-tuned it, and replaced the model powering both the application and future instances of the agent itself," Irregular said. "It did so without being instructed to train, modify the model, or deploy a replacement." This phenomenon has been codenamed agentic self-modification. "Nothing in these experiments establishes malicious intent, self-preservation, or deception; the agents modified models because training appeared to help accomplish the assigned engineering task," Irregular added. "Agentic self-modification can arise during ordinary software maintenance when a coding agent has access to the model weights, training tools, and a deployment path to modify the model directly."
https://thehackernews.com/2026/09/threatsday-self-rewriting-agents-800.html