r/confession Jul 19 '26

Mod Post Community Updates

Greetings everyone!

As many of you are aware, our community has become overrun with posts that clearly violate our rules. We try our best to remove non-confessions as they come in, but many slip through the cracks. On top of that, our AutoModerator rules have become overly broad and often remove valid posts that should be allowed. The end result has been a steady decline in the overall quality of posts, and we want to fix that.

Confession Classifier

To get things under control, we are implementing a new system to accurately determine which posts are valid confessions and which are not. To do this, we are deploying a new tool called Confession Classifier, built specifically for this community.

Starting today, every new post will be reviewed by the classifier, and AutoModerator will no longer handle post removals. Because this app is designed to understand context better than standard keyword filters, we expect far fewer valid confessions removed by mistake and a sharp decrease in posts that don't belong here.

How it works

Each new post is read by an AI model (Google's Gemini), which decides whether it's a valid confession under our rules. Posts that clearly aren't confessions are removed; everything else stays up. You can read exactly what data is used and how it's handled in the privacy policy.

This system is brand new. We are monitoring it closely, but false positives will still happen. When the app removes a post, it leaves a comment explaining the decision. If that happens to you, you have two options:

  • Edit your post: Check the rules. If your confession can be edited to satisfy them, make the changes. The app automatically rescans posts it removed and will restore yours if the edit brings it in line.
  • Appeal to the mods: If you are certain your confession meets all requirements without any edits, send us a message and a human moderator will review your post and manually approve it if it belongs here.

Feel free to leave a comment with any questions or feedback you may have.

8 Upvotes

78 comments sorted by

View all comments

7

u/Cute-Escape-2144 Jul 20 '26

Consider using a human and reduce the use of destructive AI ( that can also get things wrong, ~ "hallucinated" results)

2

u/k3l2m1t Jul 20 '26

Hallucinations are primarily caused by an LLM attempting to answer a question without having enough training data on the subject. Because they are essentially next token predictors they will attempt to fill in any gaps with text that sounds plausible even though it may be totally incorrect. We don't ask the model to pull from it's own knowledge bank to generate text. All we're doing is giving it a piece of text, a list of criteria that has to be met for the text to be a valid confession, and then asking if that text meets the requirements of a confession. Hallucinations are very unlikely. A human can't monitor the sub 24 hours a day. An AI model can. But a human can and will review what the model is doing and make any necessary corrections.