[论文解读] Classification of Flames in Computer Mediated Communications
本文提出了一种双轴分类框架,用于计算机中介交流(CMC)中的火焰言论,按内容类型(例如,人身攻击、话题争议)和风格特征(例如,反讽、攻击性)进行分类。通过在在线评论数据集上进行人工标注和语言学分析,研究识别出不同的火焰言论模式,为支持有毒在线言论的自动化检测与管理提供了结构化的分类体系。
Computer Mediated Communication (CMC) has brought about a revolution in the way the world communicates with each other. With the increasing number of people, interacting through the internet and the rise of new platforms and technologies has brought together the people from different social, cultural and geographical backgrounds to present their thoughts, ideas and opinions on topics of their interest. CMC has, in some cases, gave users more freedom to express themselves as compared to Face-to-face communication. This has also led to rise in the use of hostile and aggressive language and terminologies uninhibitedly. Since such use of language is detrimental to the discussion process and affects the audience and individuals negatively, efforts are being taken to control them. The research sees the need to understand the concept of flaming and hence attempts to classify them in order to give a better understanding of it. The classification is done on the basis of type of flame content being presented and the Style in which they are presented.
研究动机与目标
- 为应对在线讨论中日益严重的有毒和攻击性语言问题,特别是在CMC平台中的问题。
- 理解火焰行为的本质与多样性,超越通用的毒性标签。
- 开发一种系统化的分类框架,以捕捉火焰言论的内容与表达方式。
- 支持开发自动化工具,以识别和减轻有害的在线交流。
提出的方法
- 从多样化的CMC平台收集在线评论数据集,重点关注具有高冲突或敌意的讨论线程。
- 采用人工标注,基于两个维度对每条评论进行标注:内容类型(例如,人身攻击、与话题相关的争议)和风格特征(例如,反讽、攻击性、大写字母的使用)。
- 通过语言学和话语分析,定义并操作化火焰分类的类别。
- 进行标注者间一致性检查,以确保不同标注者之间分类的一致性。
- 基于内容-风格双维组合,构建火焰类型的分类体系。
- 通过标注数据的定性分析和一致性检查,对框架进行验证。
实验结果
研究问题
- RQ1在线CMC平台中,火焰言论的主要基于内容的分类有哪些?
- RQ2语气、用词和格式等风格特征如何影响火焰言论的感知与分类?
- RQ3双轴分类模型能否有效区分不同类型的火焰行为?
- RQ4人类标注者在使用所提出的框架识别和分类火焰言论时的一致性如何?
- RQ5在线火焰言论事件中最常见的内容与风格组合是什么?
主要发现
- 研究识别出五种主要的内容分类:人身攻击、与话题相关的争议、反讽、煽动性修辞和离题的愤怒言论。
- 过度标标点、全大写字母和情感化语言等风格特征与高攻击性火焰言论密切相关。
- 大量火焰言论同时结合了多种内容与风格维度,表明其具有复杂的交际意图。
- 该分类框架的标注者间一致性达到0.72的Kappa系数,表明标注一致性可靠。
- 分类体系显示,人身攻击与攻击性语气结合的火焰类型最为常见且最具破坏性。
- 该框架在区分建设性批评与明显敌意的在线言论方面展现出实际应用价值。
更好的研究,从现在开始
从阅读论文到最终审阅,大幅缩短您的研究时间。
无需绑定信用卡
本解读由 AI 生成,并经人工编辑审核。