Reddit plans to rollout tools to allow mods to detect AI posts/comments

macdonjo3

Elite Member
Jr. VIP
Joined
Nov 8, 2009
Messages
8,887
Reaction score
9,192
I just answered a bunch of questions in their moderator survey. There were a few questions before this about detecting AI and then it tested my ability to identify AI.

They also had some questions about AI entirely taking over as moderator, asking what I would do as moderator if everything was done for me.

1731974305673.png
 
They seem to added it partly already as per their modnews announcement. To fully read these changes here is the link to their full modnews post

The relevant part of their post:

"
Comment Collapsing and Improved Spam Detection

We’ve all seen it: spam comments cluttering a thread, dragging the discussion down. Our latest update, rolling out over the next few weeks, adds automatic comment collapsing for messages likely to be spam or low-quality.

Mods will see these comments tagged as “Potential Spam” in their community, whereas users will see these comments collapsed automatically, helping to reduce disruption in your community without needing manual moderation. Early testing shows this tool is a powerful front-line defense, letting you keep an eye on what matters while spam fades into the background.

A big thank you to the communities who helped pilot this experience in r/PartnerCommunities which helped us collect valuable insight into how well the model operates.
"


As for my own experience:
I have noticed that any comments being made with chatgpt any openai gpt model results in instant account ban, except really old accounts or accounts that are already well established.
Comments that were written by me before adding them to my bots seems to be doing fine.

It was an inevitable step, I noticed reddit becoming very much full of AI low effort comments using perfect grammar and words that no human would ever use in non formal talks which makes it easy to be detected. It will be still possible to comment with AI of course just need a bit more creativity than simply entering a prompt and pasting whatever gets written.
 
Last edited:
Actually some AI comments are not bad but the tools should identify whether these comments are useful (with high upvote or interaction).
 
The only AI content that will be allowed on Reddit will be that produced by Reddit.
 
They seem to added it partly already as per their modnews announcement. To fully read these changes https://www.reddit.com/r/modnews/comments/1gqowid/streamlining_moderation_enhanced_safety_features/

The relevant part of their post:
I found this part interesting

  • Criteria modal: For those who don’t meet specific posting criteria (like karma or account age) within a community, a new criteria modal now appears. This page explains the rules in plain terms and points users to communities where they meet the requirements, keeping them active and engaged.
Have mods always had the ability the filter posts based on each subreddit karama?

I thought overall account karma was what mattered.
 
I'm wondering how they are able to identify AI content within 7-10 words phrases.

Wouldn't that rise a lot of false positives?
 
I'm wondering how they are able to identify AI content within 7-10 words phrases.

Wouldn't that rise a lot of false positives?
Probably, maybe they just compare all the comments on the account, to determined if the account itself is automated using AI, since it's possible that if a comment on an account is AI, then other comments made by the accounts would definitely be AI
 
I found this part interesting


Have mods always had the ability the filter posts based on each subreddit karama?

I thought overall account karma was what mattered.
They have what is called community positive or negative karma which basically means someone who is heavily downvoted in particular communitiy will automatically have their comments being removed or in modqueue, meanwhile positive karma in community will grant you to be able to instantly post comments.

That is why it is favorable to first engage in community where you plan onto promoting your links/products or doing anything.
 
I believe it's already live. A mod banned my account and asked for an explanation as to why I used AI when I challenged their decision.
There are likely apps for it, but I don't see anything in my standard set of tools offered by Reddit.
 
they cant stop whats coming. u cant prevent ai auto reply bots. we just change the prompts
 
They seem to added it partly already as per their modnews announcement. To fully read these changes here is the link to their full modnews post

The relevant part of their post:

"
Comment Collapsing and Improved Spam Detection

We’ve all seen it: spam comments cluttering a thread, dragging the discussion down. Our latest update, rolling out over the next few weeks, adds automatic comment collapsing for messages likely to be spam or low-quality.

Mods will see these comments tagged as “Potential Spam” in their community, whereas users will see these comments collapsed automatically, helping to reduce disruption in your community without needing manual moderation. Early testing shows this tool is a powerful front-line defense, letting you keep an eye on what matters while spam fades into the background.

A big thank you to the communities who helped pilot this experience in r/PartnerCommunities which helped us collect valuable insight into how well the model operates.
"


As for my own experience:
I have noticed that any comments being made with chatgpt any openai gpt model results in instant account ban, except really old accounts or accounts that are already well established.
Comments that were written by me before adding them to my bots seems to be doing fine.

It was an inevitable step, I noticed reddit becoming very much full of AI low effort comments using perfect grammar and words that no human would ever use in non formal talks which makes it easy to be detected. It will be still possible to comment with AI of course just need a bit more creativity than simply entering a prompt and pasting whatever gets written.
When these companies send out these surveys, they usually already have something in the works, so this isn't surprising.

Thanks for sharing this and confirming my suspicions.
 
I'm wondering how they are able to identify AI content within 7-10 words phrases.

Wouldn't that rise a lot of false positives?
They can't, it'll just lead to an increase in regular users being shadowbanned or having their comments constantly collapsed.

But that's not too bad for Reddit because people will just make new accounts and it'll improve their user numbers for their quarterly reports... Why do you think they don't outright ban most accounts and use this half-assed shadowban system?
 
Just change your prompts to ask can you make a few words misspelled such as 4 or 5 small words make them look human natural. Also put in 1 sentence where it needs maybe a command or restructuring. It'll pass ai detection.
 
How effective is their test? Because if google cant figure out AI content, how a shitty company like reddit could do it?

I am a reddit mod of various big subreddits.

Depending on the niche, some subreddits I just dont allow anyone to post/comment. because the spam/shitty content is too much.

On the ones I allow users, a single Automod rule with karma requeriments cuts down spam significantly.

And there is me at the end, I read every comment on new threads. Anyone using easy for me to spot AI, gets instant ban.

There is a whole in my system, old threads comments. I should just lock up threads after a day to avoid any spam there.

My point being, that an AI detection by reddit is wasted resources.

I would preffer stuff like Toolbox being added to the mod tools instead of punished by API restrictions.
 
I think you can ask AI to add some errors to make the text more human-like and you can get around all these restrictions
 
Today’s AI has a hard time understanding context, especially in complex or ambiguous situations
 
Today’s AI has a hard time understanding context, especially in complex or ambiguous situations
Just like probably 98% of this forum users don't understand this comment ;)
 
The AI comments when it types out the comment like the real user would do and does not follow the same pattern , it is possible.
 
Back
Top