Paper: 2604.19578 Authors: Wenqing Wu, Chengzhi Zhang, et al. (Nanjing University, Anhui University) Categories: cs.CL, cs.AI

Background

With the rapid advancement of LLMs, the academic community has faced unprecedented disruptions, particularly in academic communication. The primary function of peer review is improving manuscript quality.

Key Questions

  1. Are LLMs altering the core evaluative functions of peer review?
  2. How do LLMs affect linguistic form, evaluative focus, and recommendation signals?
  3. What is the impact on informativeness of paper decisions?

Methodology

Analyzed changes in peer review reports for academic articles following LLM emergence:

  • Linguistic features: Length, word/sentence complexity
  • Evaluation aspects: Automatically annotated review sentences
  • LLM-assisted detection: MLE method to identify LLM-generated reviews

Key Findings

Following LLM emergence, peer review texts have become:

  • Longer and more fluent
  • Increased emphasis on summaries and surface-level clarity
  • More standardized linguistic patterns

Particularly for reviewers with lower confidence scores.

Decline in Deep Evaluation

Attention to deeper evaluative dimensions has declined:

  • Originality
  • Replicability
  • Nuanced critical reasoning

Takeaways

  • LLM assistance affects not just form but substantive evaluation focus
  • Need to preserve human expertise in assessing novelty and impact
  • Quality over quantity: longer reviews aren’t necessarily better reviews

论文: 2604.19578 作者: 吴文清、张成志等(南京大学、安徽大学) 分类: cs.CL, cs.AI

背景

随着LLM的快速发展,学术界面临前所未有的变革,特别是在学术交流领域。同行评审的主要功能是提高手稿质量。

关键问题

  1. LLM是否在改变同行评审的核心评估功能?
  2. LLM如何影响语言形式、评估重点和推荐信号?
  3. 对论文决策信息量的影响是什么?

方法论

分析了LLM出现后学术论文评审报告的变化:

  • 语言特征:长度、词/句复杂度
  • 评估方面:自动标注的评审句子
  • LLM辅助检测:MLE方法识别LLM生成的评审

关键发现

LLM出现后,同行评审文本:

  • 更长更流畅
  • 更强调摘要和表面清晰度
  • 语言模式更标准化

特别是在信心分数较低的评审者中。

深度评估的衰退

对更深层评估维度的关注减少:

  • 原创性
  • 可复现性
  • 细致批判性推理

要点总结

  • LLM辅助不仅影响形式还影响实质性评估重点
  • 需要保留人类在评估新颖性和影响方面的专业知识
  • 质量重于数量:更长的评审不一定是更好的评审