{"id":576571,"date":"2026-08-06T13:36:09","date_gmt":"2026-08-06T13:36:09","guid":{"rendered":"https:\/\/www.newsbeep.com\/il\/576571\/"},"modified":"2026-08-06T13:36:09","modified_gmt":"2026-08-06T13:36:09","slug":"ai-isnt-enough-to-protect-social-media-communities-from-ai","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/il\/576571\/","title":{"rendered":"AI isn\u2019t enough to protect social media communities from AI"},"content":{"rendered":"<p>AI\u2019s biases<\/p>\n<p>Typical social media AI-based <a href=\"https:\/\/www.techtarget.com\/searchcontentmanagement\/tip\/Types-of-AI-content-moderation-and-how-they-work\" rel=\"nofollow noopener\" target=\"_blank\">moderating systems<\/a> use <a href=\"https:\/\/www.ibm.com\/think\/topics\/classification-machine-learning\" rel=\"nofollow noopener\" target=\"_blank\">machine learning classifiers<\/a> to analyze posts and identify and flag content that breaks platform rules. But it\u2019s difficult for a machine to understand the nuances of sarcasm, satire, and slang.<\/p>\n<p>Further, some research (examples <a href=\"https:\/\/pmc.ncbi.nlm.nih.gov\/articles\/PMC11420153\/\" rel=\"nofollow noopener\" target=\"_blank\">here<\/a>, <a href=\"https:\/\/aclanthology.org\/2023.trustnlp-1.10.pdf\" rel=\"nofollow noopener\" target=\"_blank\">here<\/a>, and <a href=\"https:\/\/www.researchgate.net\/publication\/342587849_Reading_Between_the_Demographic_Lines_Resolving_Sources_of_Bias_in_Toxicity_Classifiers\" rel=\"nofollow noopener\" target=\"_blank\">here<\/a>) suggests that marginalized groups can be disproportionately affected by AI moderation. Without human oversight, AI can <a href=\"https:\/\/www.theverge.com\/2022\/2\/25\/22949293\/tumblr-nycchr-settlement-adult-content-ban-algorithmic-bias-lgbtq\" rel=\"nofollow noopener\" target=\"_blank\">end up penalizing<\/a> the very communities most vulnerable to the hateful content the systems are designed to combat.<\/p>\n<p>Gilbert, who is also the research director of Cornell\u2019s Citizens and Technology Lab, says that \u201cmarginalized and vulnerable populations are among those who experience the highest rates of moderation, and that typically this is a result of \u2018false-positives,\u2019\u201d often driven by instances of <a href=\"https:\/\/www.dangerousspeech.org\/counterspeech\" rel=\"nofollow noopener\" target=\"_blank\">counter-speech<\/a>, <a href=\"https:\/\/www.amacad.org\/publication\/daedalus\/hear-our-languages-hear-our-voices-storywork-theory-praxis-indigenous-language-reclamation\" rel=\"nofollow noopener\" target=\"_blank\">language reclamation<\/a>, and \u201cresponses to hateful content.\u201d<\/p>\n<p>\u201cFalse positives are an equity issue. They mean that groups that are already marginalized are further silenced and censored,\u201d she added.<\/p>\n<p>AI moderators can also make communities less effective at moderating themselves. On Reddit, for example, some subreddit moderators would prefer to ban users who use hateful or violent rhetoric. But if Reddit\u2019s AI removes such content before a human moderator sees it, those moderators lose the ability to assess whether a ban is warranted.<\/p>\n<p>In terms of giving human mods more control, Reddit this week <a href=\"https:\/\/redditinc.com\/news\/modernizing-reddits-infrastructure-and-moderation-tools\" rel=\"nofollow noopener\" target=\"_blank\">announced<\/a> expanding testing for Rules Hub, a suite of tools that lets human mods \u201cchoose which rules should be automatically enforced, decide what happens when a rule is triggered (send to queue, filter, or remove), preview the experience before enabling it, and review logs and insights.\u201d Reddit expects Rules Hub to eventually replace the <a href=\"https:\/\/support.reddithelp.com\/hc\/en-us\/articles\/15484574206484-Automoderator#h_01G7N78R3914C5QBT6X2E90S3Y\" rel=\"nofollow noopener\" target=\"_blank\">Automod <\/a>tool, which relies primarily on exact keywords.<\/p>\n<p>AI is a tool, not the solution<\/p>\n<p>Mods I\u2019ve spoken with have <a href=\"https:\/\/arstechnica.com\/gadgets\/2025\/02\/reddit-mods-are-fighting-to-keep-ai-slop-off-subreddits-they-could-use-help\/\" rel=\"nofollow noopener\" target=\"_blank\">repeatedly blamed<\/a> the generative AI boom for a spike in content that breaks community-specific or broader platform rules. That\u2019s a serious problem for social media sites that rely on user contributions.<\/p>\n<p>Companies will continue to try new methods of moderating more reliably and effectively, but reducing human input is a step backward. Low-effort AI-generated content is changing the challenges moderation teams face, but that makes stronger approaches more necessary, where machine-scale detection can be combined with human judgment and expertise.<\/p>\n<p>Just as social media has no value without people, content moderation can\u2019t succeed without human judgment at the forefront.<\/p>\n<p>Advance Publications, which owns Ars Technica parent Cond\u00e9 Nast, is the largest shareholder in Reddit.<\/p>\n","protected":false},"excerpt":{"rendered":"AI\u2019s biases Typical social media AI-based moderating systems use machine learning classifiers to analyze posts and identify and&hellip;\n","protected":false},"author":2,"featured_media":576572,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[345,343,344,85,46,125],"class_list":["post-576571","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-il","tag-israel","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/posts\/576571","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/comments?post=576571"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/posts\/576571\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/media\/576572"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/media?parent=576571"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/categories?post=576571"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/tags?post=576571"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}