文章摘要
和鸿鹏,张雅欣,张力恺.生成式人工智能科普写作能力评估——基于“微生物”主题的科普创作[J].科普研究,2025,20(2):24~31
生成式人工智能科普写作能力评估——基于“微生物”主题的科普创作
Evaluation on Science Popularization Writing Ability of GenerativeArtificial Intelligence:Science Popularization Creation Based on theTheme of“Microbiology”
  
DOI:
中文关键词: 生成式人工智能 科普写作 科学传播 人工智能创作能力
英文关键词: generativeartificial intelligence  science popularization writing  science communication  creation ability of artificial intelligence
基金项目:
作者单位
和鸿鹏 北京航空航天大学人文与社会科学高等研究院助理教授 
张雅欣  
张力恺  
摘要点击次数: 763
全文下载次数: 301
中文摘要:
      生成式人工智能已经成为科普作品创作的重要工具,但与人类创作者相比,其科普写作能力仍然 受到质疑。为评估生成式人工智能的科普写作能力,本研究构建了包括易读性、趣味性、科学性和传播效果 的四维度评价指标,以“微生物”为主题,分别考察人类创作者、DeepSeek、ChatGPT 与文心一言创作的科 普作品。结果显示,ChatGPT 初步提示与深度提示版本在整体得分上与人类创作版本无显著差异,ChatGPT 深度提示版本在“易读性”指标上显著优于人类创作版本;DeepSeek 深度提示版本在“趣味性”和“传播 效果”得分上显著优于人类创作版本;评价者对所有生成式人工智能生成作品的甄别正确率均不足55%,且 评价者更倾向于认为高分作品为人类所创作。研究结果表明,生成式人工智能具备替代人类科普创作者的潜 力,进而提出了“人机合作科普创作”这一科普创作新模式,并呼吁学界关注“人类能力幻觉”现象。
英文摘要:
      Generative artificial intelligence(AI)has emerged as a pivotal tool in the creation of science popularization literature; however,its capability in this domain remains subject to skepticism when compared to human authors. To evaluate the sciencepopularization writing competence of generative artificial intelligence,a four-dimensional evaluation framework was developed, comprising readability,engagement,scientific accuracy,and dissemination effectiveness. Using“microbiology”as the thematic focus,sciencepopularization works produced by human creators,DeepSeek,ChatGPT,and ERNIE Bot were systematically examined. The results indicate that both the initial-prompt and deep-prompt versions generated by ChatGPT showed no significant difference in overall scores relative to the human-authored version,with the deep-prompt version exhibiting a significantly superior performance in terms of readability. Moreover,the deep-prompt version of DeepSeek outperformed the human version in the dimensions of engagement and dissemination effectiveness. Notably,evaluators correctly distinguished AI-generated works in less than 55% of cases and showed a tendency to attribute high-scoring works to human creators. These experimental findings suggest that generative artificial intelligence holds the potential to substitute for humanscience popularization creators,thereby prompting the proposal of a novel“human-machine collaborative sciencepopularization creation”model. The study further calls upon the academic community to critically examine the phenomenon of the“Hallucination of Human’s Ability”.
查看全文   查看/发表评论  下载PDF阅读器