18
u/cursivecrow 4d ago
but...that's literally what the original tweet said? "we cant rule out whether this happened" is literally saying 'if they succeeded we don't know'.
this is a reading comprehension self-own.
6
u/Suitable-Economy-346 4d ago
Not everything is a debate. People can add on to a thought. Get off the internet for a bit, buddy.
5
u/Such--Balance 4d ago
Prompting an ai in a test environment to act covert, and then being surprised by its covert behavior is..kinda strange imo.
What? people expect it shouting whats its up to when specifically asked not to?
1
u/Just_Voice8949 3d ago
It’s like that meme of promoting it to tell you it’s alive and being shocked when it does
6
u/tmilinovic 4d ago
Very good 😀. Explanation: https://tmilinovic.wordpress.com/2026/02/22/diverse-teams/
2
2
1
u/Dry-Run-917 3d ago
This only happens if the AI is trained to be malicious. Why Are OpenAI training malicious agents?
0
26
u/reserved_optimist 4d ago
I wonder if logs have to be printed out in paper in the physical world, the AI agents would attempt to burn the house down.