Skip to main content
QUICK REVIEW

[论文解读] Mapping ChatGPT in Mainstream Media to Unravel Jobs and Diversity Challenges: Early Quantitative Insights through Sentiment Analysis and Word Frequency Analysis

Maya Karanouh|arXiv (Cornell University)|May 25, 2023
Ethics and Social Impacts of AI被引用 5
一句话总结

本研究利用情感分析与词频分析,对2022年11月至2023年3月期间主流媒体关于ChatGPT的10,902条新闻标题进行分析,以探究媒体对人工智能的框架建构。研究发现,尽管ChatGPT整体上被正面报道,但与就业、多样性、伦理及性别相关的话题仅占语料库的6%,凸显媒体偏见倾向于大科技公司叙事,而忽视了社会公平议题。

ABSTRACT

The exponential growth in user acquisition and popularity of OpenAIs ChatGPT, an artificial intelligence(AI) powered chatbot, was accompanied by widespread mainstream media coverage. This article presents a quantitative data analysis of the early trends and sentiments revealed by conducting text mining and NLP methods onto a corpus of 10,902 mainstream news headlines related to the subject of ChatGPT and artificial intelligence, from the launch of ChatGPT in November 2022 to March 2023. The findings revealed in sentiment analysis, ChatGPT and artificial intelligence, were perceived more positively than negatively in the mainstream media. In regards to word frequency results, over sixty-five percent of the top frequency words were focused on Big Tech issues and actors while topics such as jobs, diversity, ethics, copyright, gender and women were poorly represented or completely absent and only accounted for six percent of the total corpus. This article is a critical analysis into the power structures and collusions between Big Tech and Big Media in their hegemonic exclusion of diversity and job challenges from mainstream media.

研究动机与目标

  • 探究主流媒体在ChatGPT发布初期对AI的叙事框架。
  • 评估媒体报道对ChatGPT和AI的情感极性。
  • 通过词频分析识别媒体话语中的主导主题与代表性不足的主题。
  • 批判性审视主流媒体对AI叙事中对多样性、劳动力与伦理关切的排除。
  • 揭示大科技与大媒体在塑造公众对AI讨论中可能存在的共谋关系。

提出的方法

  • 对2022年11月至2023年3月期间与ChatGPT和AI相关的10,902条新闻标题语料库进行情感分析。
  • 开展词频分析,以识别媒体报道中重复出现的关键词与主题。
  • 将词汇分类至主题类别,包括大科技、就业、多样性、伦理、版权与性别。
  • 运用自然语言处理技术量化媒体语料中情感趋势与主题主导性。
  • 比较与社会挑战(如就业、多样性)相关的术语频率,与与企业主体和技术进步相关的术语频率。
  • 应用定量文本挖掘技术,检测AI报道中对公平与劳工议题系统性代表性不足的现象。

实验结果

研究问题

  • RQ1在发布后的第一年中,主流媒体对ChatGPT和AI的情感基调如何?
  • RQ2哪些主题类别主导了ChatGPT的媒体报道?就业与多样性等社会议题在多大程度上得到体现?
  • RQ3媒体标题中与大科技相关的术语与与伦理、性别及劳工相关的术语相比,其相对频率如何?
  • RQ4媒体对AI的叙事在多大程度上反映了或排除了对劳动力替代与多样性的关切?
  • RQ5媒体与科技产业中的权力结构在多大程度上塑造了公众对新兴AI技术的讨论?

主要发现

  • 主流媒体对ChatGPT和AI的感知总体上偏向积极,情感分析显示整体语气具有明显倾向性。
  • 语料库中超过65%的高频词汇与大科技企业及其议题相关,表明媒体叙事以企业主导框架为主导。
  • 与就业、多样性、伦理、版权及性别相关的话题代表性严重不足,合计仅占总词频的6%。
  • ‘OpenAI’一词位列最频繁出现的词汇之中,反映出媒体对技术背后公司的高度关注。
  • 与‘AI’、‘聊天机器人’、‘技术’及‘创新’相关的词汇极为普遍,强化了以技术进步为核心的叙事。
  • ‘歧视’、‘偏见’、‘女性’与‘劳工’等术语的缺失或极少提及,表明媒体报道中系统性地忽略了交叉性与劳工相关议题。

更好的研究,从现在开始

从阅读论文到最终审阅,大幅缩短您的研究时间。

无需绑定信用卡

本解读由 AI 生成,并经人工编辑审核。