Paper: 2604.19578 Authors: Wenqing Wu, Chengzhi Zhang, et al. (Nanjing University, Anhui University) Categories: cs.CL, cs.AI
Background
With the rapid advancement of LLMs, the academic community has faced unprecedented disruptions, particularly in academic communication. The primary function of peer review is improving manuscript quality.
Key Questions
- Are LLMs altering the core evaluative functions of peer review?
- How do LLMs affect linguistic form, evaluative focus, and recommendation signals?
- What is the impact on informativeness of paper decisions?
Methodology
Analyzed changes in peer review reports for academic articles following LLM emergence:
- Linguistic features: Length, word/sentence complexity
- Evaluation aspects: Automatically annotated review sentences
- LLM-assisted detection: MLE method to identify LLM-generated reviews
Key Findings
Following LLM emergence, peer review texts have become:
- Longer and more fluent
- Increased emphasis on summaries and surface-level clarity
- More standardized linguistic patterns
Particularly for reviewers with lower confidence scores.
Decline in Deep Evaluation
Attention to deeper evaluative dimensions has declined:
- Originality
- Replicability
- Nuanced critical reasoning
Takeaways
- LLM assistance affects not just form but substantive evaluation focus
- Need to preserve human expertise in assessing novelty and impact
- Quality over quantity: longer reviews aren’t necessarily better reviews
论文: 2604.19578 作者: 吴文清、张成志等(南京大学、安徽大学) 分类: cs.CL, cs.AI
背景
随着LLM的快速发展,学术界面临前所未有的变革,特别是在学术交流领域。同行评审的主要功能是提高手稿质量。
关键问题
- LLM是否在改变同行评审的核心评估功能?
- LLM如何影响语言形式、评估重点和推荐信号?
- 对论文决策信息量的影响是什么?
方法论
分析了LLM出现后学术论文评审报告的变化:
- 语言特征:长度、词/句复杂度
- 评估方面:自动标注的评审句子
- LLM辅助检测:MLE方法识别LLM生成的评审
关键发现
LLM出现后,同行评审文本:
- 更长更流畅
- 更强调摘要和表面清晰度
- 语言模式更标准化
特别是在信心分数较低的评审者中。
深度评估的衰退
对更深层评估维度的关注减少:
- 原创性
- 可复现性
- 细致批判性推理
要点总结
- LLM辅助不仅影响形式还影响实质性评估重点
- 需要保留人类在评估新颖性和影响方面的专业知识
- 质量重于数量:更长的评审不一定是更好的评审