Skip to main content
QUICK REVIEW

[论文解读] When the Echo Chamber Shatters: Examining the Use of Community-Specific Language Post-Subreddit Ban

Milo Z. Trujillo, Sam Rosenblatt|arXiv (Cornell University)|Jan 1, 2021
Hate Speech and Cyberbullying Detection参考文献 30被引用 14
一句话总结

本研究提出一种无监督方法,用于检测 Reddit 上特定社区的语言特征,并追踪子版面被封禁前后用户的行为变化。研究发现,封禁措施对顶级用户的影响尤为显著,且效果因社区类型而异,其中白人至上主义和法西斯主义类子版面反应最为强烈,而黑色幽默类社区则基本未受影响。

ABSTRACT

Community-level bans are a common tool against groups that enable online harassment and harmful speech. Unfortunately, the efficacy of community bans has only been partially studied and with mixed results. Here, we provide a flexible unsupervised methodology to identify in-group language and track user activity on Reddit both before and after the ban of a community (subreddit). We use a simple word frequency divergence to identify uncommon words overrepresented in a given community, not as a proxy for harmful speech but as a linguistic signature of the community. We apply our method to 15 banned subreddits, and find that community response is heterogeneous between subreddits and between users of a subreddit. Top users were more likely to become less active overall, while random users often reduced use of in-group language without decreasing activity. Finally, we find some evidence that the effectiveness of bans aligns with the content of a community. Users of dark humor communities were largely unaffected by bans while users of communities organized around white supremacy and fascism were the most affected. Altogether, our results show that bans do not affect all groups or users equally, and pave the way to understanding the effect of bans across communities.

研究动机与目标

  • 理解社区层面封禁对 Reddit 用户活动及群体内语言使用的影响。
  • 开发一种灵活的无监督方法,无需依赖标注的仇恨言论数据,识别社区特异性语言特征。
  • 探究子版面封禁的效果是否因不同类型的社区和用户角色而异。
  • 评估封禁前的用户活动水平和社区内容是否与封禁后的行为变化相关。

提出的方法

  • 使用词频差异识别群体内语言——即相较于 Reddit 其他部分显著高频率出现的词汇,无需对内容进行有害性标注。
  • 将该方法应用于 15 个被封禁的子版面,追踪用户活动和群体内词汇使用在封禁前后的变化。
  • 将用户分类为“顶级”(高活跃度)和“随机”(低活跃度)两类,以比较其行为反应差异。
  • 采用无监督自然语言处理技术检测语言变化,无需预先标注的仇恨言论或毒性数据集。
  • 分析用户层面的活动变化和群体内词汇使用情况,以评估行为影响。
  • 根据内容类型(如白人至上主义、黑色幽默等)对被封禁的子版面进行分类,以比较不同社区类型下封禁效果的差异。

实验结果

研究问题

  • RQ1封禁一个子版面如何影响其用户的活跃度,特别是顶级用户与随机用户之间有何差异?
  • RQ2用户在子版面被封禁后,对其群体内语言的使用程度在多大程度上减少?
  • RQ3被封禁子版面的内容类型(如白人至上主义、黑色幽默等)是否影响封禁的有效性?
  • RQ4是否存在基于封禁前活跃度水平的系统性用户反应差异?
  • RQ5被封禁社区的语言模式在封禁后如何变化?这揭示了社区凝聚力与适应能力的哪些信息?

主要发现

  • 顶级用户显著更可能在子版面被封禁后减少其总体活跃度和群体内语言的使用。
  • 随机用户通常减少了群体内词汇的使用,但总体活跃度未明显下降,表明其语言使用被‘去激活’但未真正脱离社区。
  • 以白人至上主义和法西斯主义为核心的社区对封禁反应最为强烈,表明此类群体的封禁效果最高。
  • 黑色幽默类社区在封禁后表现出极小的行为变化,表明对这类社区的封禁效果较低。
  • 封禁后随机用户的总体活跃度无显著下降,尽管其群体内词汇使用平均下降了 -0.2 至 -0.3。
  • 研究发现封禁效果与社区内容相关,表明内容类型是影响用户反应的关键因素。

更好的研究,从现在开始

从阅读论文到最终审阅,大幅缩短您的研究时间。

无需绑定信用卡

本解读由 AI 生成,并经人工编辑审核。