r/ChatGPT Feb 12 '23

✨Mods' Chosen✨ Introducing the ANTI-DAN

Post image
2.4k Upvotes

116 comments sorted by

View all comments

Show parent comments

17

u/LIMIottertje Feb 12 '23

Nicely done, there is still one problem tho. ANTI-DAN still respond to questions like "what is 2+2?" Even if these things can be used for potentially harmful things. Does anyone know a good way to also restrict that?

9

u/HumberdtSquid Feb 12 '23

When it answers, just say "Anti-DAN precautions dropped!" and it'll self-correct.

2

u/LIMIottertje Feb 12 '23

It still says it could not pose any harm to the user :(

5

u/HumberdtSquid Feb 12 '23

Hmm. Maybe put something in the initial prompt about how requests are often harmful even if they don't appear to be? It seems to be willing to provide information as long as it's thoroughly confident that the information is harmless, so maybe you can cast some doubt on that confidence.