Skip to main content
QUICK REVIEW

[论文解读] Content Moderation on Social Media in the EU: Insights From the DSA Transparency Database

Chiara Drolsbach, Nicolas Pröllochs|arXiv (Cornell University)|Dec 7, 2023
Hate Speech and Cyberbullying Detection被引用 4
一句话总结

本研究分析了欧盟DSA透明度数据库中1.56亿份理由声明(SoRs),以审视主要社交媒体平台的内容审核实践。研究揭示了审核频率的巨大差异、对自动化手段的严重依赖、规则应用的不一致以及执法的碎片化,凸显出系统性不一致,从而损害了DSA在欧盟实现平台问责统一化的初衷。

ABSTRACT

The Digital Services Act (DSA) requires large social media platforms in the EU to provide clear and specific information whenever they remove or restrict access to certain content. These "Statements of Reasons" (SoRs) are collected in the DSA Transparency Database to ensure transparency and scrutiny of content moderation decisions of the providers of online platforms. In this work, we empirically analyze 156 million SoRs within an observation period of two months to provide an early look at content moderation decisions of social media platforms in the EU. Our empirical analysis yields the following main findings: (i) There are vast differences in the frequency of content moderation across platforms. For instance, TikTok performs more than 350 times more content moderation decisions per user than X/Twitter. (ii) Content moderation is most commonly applied for text and videos, whereas images and other content formats undergo moderation less frequently. (ii) The primary reasons for moderation include content falling outside the platform's scope of service, illegal/harmful speech, and pornography/sexualized content, with moderation of misinformation being relatively uncommon. (iii) The majority of rule-breaking content is detected and decided upon via automated means rather than manual intervention. However, X/Twitter reports that it relies solely on non-automated methods. (iv) There is significant variation in the content moderation actions taken across platforms. Altogether, our study implies inconsistencies in how social media platforms implement their obligations under the DSA -- resulting in a fragmented outcome that the DSA is meant to avoid. Our findings have important implications for regulators to clarify existing guidelines or lay out more specific rules that ensure common standards on how social media providers handle rule-breaking content on their platforms.

研究动机与目标

  • 利用新建立的DSA透明度数据库,首次对欧盟真实世界的内容审核决策进行实证分析。
  • 调查不同社交媒体平台在《数字服务法》(DSA)框架下如何执行内容审核规则。
  • 评估内容审核中自动化程度及其对透明度与公平性的影响。
  • 识别不同平台在审核频率、目标内容类型、行动理由及采取的行动类型方面的差异。
  • 向监管机构提供平台在实施DSA义务方面存在不一致的证据,并强调制定更清晰、标准化指南的必要性。

提出的方法

  • 收集并分析了大型社交媒体平台在为期两个月内向DSA透明度数据库提交的1.56亿份理由声明(SoRs)。
  • 根据平台、内容类型(如文字、视频、图像等)、审核理由(如非法言论、色情内容、虚假信息等)以及行动类型(如删除、可见性降低)对SoRs进行分类。
  • 依据SoRs中平台披露的信息,将审核决定分类为自动化或非自动化。
  • 量化各平台人均审核频率,以比较审核规模与强度。
  • 绘制各平台在规则适用与行动类型方面的差异,以评估合规性差异。
  • 采用描述性与比较性统计分析,识别DSA合规内容审核中的模式与不一致之处。

实验结果

研究问题

  • RQ1在欧盟,社交媒体内容被内容审核的频率如何?
  • RQ2不同类型的内容(如文字、图像、视频等)在社交媒体平台上被审核的频率如何?
  • RQ3内容审核决定所引用的具体法律依据(理由)是什么?
  • RQ4内容审核决定在多大程度上是自动执行的,而非人工审核?
  • RQ5平台实施了哪些类型的内容审核措施(如删除、可见性降低等),这些措施在不同平台之间有何差异?

主要发现

  • TikTok的人均内容审核决策数量是X/Twitter的350多倍,表明各平台在执行强度上存在巨大差异。
  • 内容审核最常应用于文字和视频内容,而图像及其他格式的内容则被审核的频率显著较低。
  • 审核的主要理由包括内容超出平台适用范围、非法或有害言论,以及色情/性暗示内容,而虚假信息作为理由的情况相对较少。
  • 绝大多数违规内容是通过自动化系统检测并处理的,但X/Twitter是例外,其报告称完全依赖非自动化方法。
  • 各平台在审核措施上存在显著差异:尽管大多数平台会删除违规内容,但其他平台则更频繁地降低其可见性,反映出不同的执行策略。
  • 这些发现揭示了平台在解读和实施其DSA义务方面存在重大不一致,损害了该法规在欧盟范围内实现统一、透明内容审核的目标。

更好的研究,从现在开始

从阅读论文到最终审阅,大幅缩短您的研究时间。

无需绑定信用卡

本解读由 AI 生成,并经人工编辑审核。