Tag
The author comments that the X platform (now part of XAI/SpaceX) is flooded with scam posts, arguing that because X's user data is a key basis for SpaceX's valuation, the platform is unwilling to remove scam traffic, and questions the management's ethics.
Ars Technica reports on the limitations of AI-based content moderation, highlighting biases against marginalized groups and the need for human oversight, while noting Reddit's expansion of Rules Hub to give human moderators more control.
Reddit is introducing Rules Hub, an AI-powered moderation suite that uses LLMs to automatically enforce community rules, expanding to all new communities ahead of a wider 2026 launch. The company also announced plans to require third-party apps to use its developer platform and continue efforts against unauthorized scraping.
Discord acknowledged that a bug in its AI moderation system wrongfully banned over 8,000 users for harmless images like chessboards and game textures, highlighting challenges in automated content moderation.
Meta was secretly using Google's Gemini for customer service, ad tools, and content moderation because it outperformed their own Llama models, until Google cut off access due to excessive capacity usage.
A new report from the Center for Countering Digital Hate (CCDH) reveals that racist comments targeting politicians tripled after Meta relaxed its content moderation rules, with violent threats and hate speech quadrupling and bullying doubling.
South Korea will mandate that online forums use AI tools to scan every image for censorship, marking a significant expansion of automated content moderation.
Introduction of an AI-powered moderation tool designed to enhance safety in chat experiences by detecting malware and inappropriate content.