| [1] |
Dance A. Stop the peer-review treadmill. I want to get off[J]. Nature, 2023, 614(7948): 581-583.
|
| [2] |
Dimitrova D. Evolution and challenges for peer review[J]. Journalism & Mass Communication Quarterly, 2024, 101(1): 5-12.
|
| [3] |
Hanson M A, Barreiro P G, Crosetto P, et al. The strain on scientific publishing[J]. Quantitative Science Studies, 2024, 5(4): 823-843.
|
| [4] |
Hosseini M, Horbach S P J M. Fighting reviewer fatigue or amplifying bias? Considerations and recommendations for use of ChatGPT and other large language models in scholarly peer review[J]. Research Integrity and Peer Review, 2023, 8(1): 4.
|
| [5] |
Guo Daya, Yang Dejian, Zhang Haowei, et al. DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning[J]. Nature, 2025, 645(8081): 633-638.
|
| [6] |
van Dis E A M, Bollen J, Zuidema W, et al. ChatGPT: Five priorities for research[J]. Nature, 2023, 614(7947): 224-226.
|
| [7] |
Donker T. The dangers of using large language models for peer review[J]. The Lancet Infectious Diseases, 2023, 23(7): 781.
|
| [8] |
Russo G, Horta Ribeiro M, Davidson T R, et al. The AI review lottery: Widespread AI-assisted peer reviews boost paper scores and acceptance rates[J]. Proceedings of the ACM on Human-Computer Interaction, 2025, 9(7): 1-28.
|
| [9] |
Liang Weixin, Zhang Yuhui, Cao Hancheng, et al. Can large language models provide useful feedback on research papers? A large-scale empirical analysis[J/OL]. NEJM AI, 2024, 1(8): AIoa2400196. DOI:10.1056/AIoa2400196 .
|
| [10] |
Ye Rui, Pang Xianghe, Chai Jingyi, et al. Are we there yet? Revealing the risks of utilizing large language models in scholarly peer review[PP/OL]. arXiv (2024-12-02)[2026-03-20]. .
|
| [11] |
Lo Vecchio N. Personal experience with AI-generated peer reviews: A case study[J]. Research Integrity and Peer Review, 2025, 10(1): 4.
|
| [12] |
中国人民大学人文社会科学学术成果评价研究中心. 人文社会科学论文质量评估指标体系实施方案(试行)[R/OL]. 北京: 中国人民大学书报资料中心, 2014[2026-05-21]. .
|
| [13] |
Zhuang Zhenzhen, Chen Jiandong, Xu Hongfeng, et al. Large language models for automated scholarly paper review: A survey[J]. Information Fusion, 2025, 124: 103332.
|
| [14] |
张智雄, 于改红, 刘熠, 等. ChatGPT对文献情报工作的影响[J]. 数据分析与知识发现, 2023, 7(3): 36-42.
|
|
Zhang Zhixiong, Yu Gaihong, Liu Yi, et al. The influence of ChatGPT on library & information services[J]. Data Analysis and Knowledge Discovery, 2023, 7(3): 36-42.
|
| [15] |
Thelwall M. Research quality evaluation by AI in the era of large language models: Advantages, disadvantages, and systemic effects-An opinion paper[J]. Scientometrics, 2025, 130(10): 5309-5321.
|
| [16] |
Boell S K, Cecez-Kecmanovic D. A hermeneutic approach for conducting literature reviews and literature searches[J]. Communications of the Association for Information Systems, 2014, 34: 257-286.
|
| [17] |
Paré G, Kitsiou S. Methods for literature reviews[M]//Lau F, Kuziemsky C, editors. Handbook of eHealth Evaluation: An Evidence-Based Approach. Victoria (BC): University of Victoria, 2017: 157-180. (2017-11-13)[2026-05-22]. .
|
| [18] |
Takkinen P. Post-sustainability: A hermeneutic literature review[J]. The Anthropocene Review, 2025, 12(3): 436-464.
|
| [19] |
Liu R, Shah N B. ReviewerGPT? an exploratory study on using large language models for paper reviewing[PP/OL]. arXiv (2023-06-01)[2026-03-20]. .
|
| [20] |
Thakkar N, Yuksekgonul M, Silberg J, et al. A large-scale randomized study of large language model feedback in peer review[J]. Nature Machine Intelligence, 2026, 8(3): 326-336.
|
| [21] |
Goldberg A, Ullah I, Khuong T G H, et al. Usefulness of LLMs as an author checklist assistant for scientific papers: NeurIPS'24 experiment[PP/OL]. V2. arXiv (2024-11-08)[2026-03-20]. .
|
| [22] |
焦丽珍, 刘选, 刘俊杰, 等. 走向人机协同:大语言模型赋能学术论文评审的效能与边界[J]. 重庆高教研究, 2025, 13(5): 119-127.
|
|
Jiao Lizhen, Liu Xuan, Liu Junjie, et al. Towards human-machine collaboration: Efficacy and boundaries of large language models in empowering academic paper review[J]. Chongqing Higher Education Research, 2025, 13(5): 119-127.
|
| [23] |
Gao Zhaolin, Brantley K, Joachims T. Reviewer2: Optimizing review generation through prompt generation[PP/OL]. V2. arXiv (2024-12-02)[2026-03-20]. .
|
| [24] |
Tan Cheng, Dongxin Lyu, Li Siyuan, et al. Peer review as a multi-turn and long-context dialogue with role-based interactions[PP/OL]. arXiv (2024-06-09)[2026-03-20]. .
|
| [25] |
Taechoyotin P, Wang G, Zeng T, et al. MAMORX: Multi-agent multi-modal scientific review generation with external knowledge[C/OL]//NeurIPS 2024 Workshop on Foundation Models for Science (FM4Science). Vancouver: NeurIPS, 2024[2026-03-20]. .
|
| [26] |
Zhou Ruiyang, Chen Lu, Yu Kai. Is LLM a reliable reviewer? A comprehensive evaluation of LLM on automatic paper reviewing tasks[C]//Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024). European Language Resources Association (ELRA) and ICCL, 2024: 9340-9351.
|
| [27] |
Liang Weixin, Izzo Z, Zhang Yaohui, et al. Monitoring AI-modified content at scale: A case study on the impact of ChatGPT on AI conference peer reviews[C]//Proceedings of the 41st International Conference on Machine Learning. Vienna: PMLR, 2024: 29575-29620.
|
| [28] |
Lin Zhicheng. Hidden prompts in manuscripts exploit AI-assisted peer review[PP/OL]. arXiv (2025-07-08)[2026-03-20]. .
|
| [29] |
张重毅, 牛欣悦, 孙君艳, 等. ChatGPT探析: AI大型语言模型下学术出版的机遇与挑战[J]. 中国科技期刊研究, 2023, 34(4): 446-453.
|
|
Zhang Chongyi, Niu Xinyue, Sun Junyan, et al. ChatGPT: Opportunities and challenges of large language models for academic publishing[J]. Chinese Journal of Scientific and Technical Periodicals, 2023, 34(4): 446-453.
|
| [30] |
Scherbakov D, Hubig N, Jansari V, et al. The emergence of large language models as tools in literature reviews: A large language model-assisted systematic review[J]. Journal of the American Medical Informatics Association, 2025, 32(6): 1071-1086.
|
| [31] |
陆伟, 刘寅鹏, 石湘, 等. 大模型驱动的学术文本挖掘——推理端指令策略构建及能力评测[J]. 情报学报, 2024, 43(8): 946-959.
|
|
Lu Wei, Liu Yinpeng, Shi Xiang, et al. Large language model-driven academic text mining: Construction and evaluation of inference-end prompting strategy[J]. Journal of the China Society for Scientific and Technical Information, 2024, 43(8): 946-959.
|
| [32] |
王亮. 检索增强生成(RAG)驱动的知识服务:原理、范式及评估[J]. 科技与出版, 2025(4): 37-46.
|
|
Wang Liang. Retrieval-augmented generation (RAG)-driven knowledge service: Principles, paradigms, and evaluation[J]. Science-Technology & Publication, 2025(4): 37-46.
|
| [33] |
王婷, 王娜, 崔运鹏, 等. 基于人工智能大模型技术的果蔬农技知识智能问答系统[J]. 智慧农业(中英文), 2023, 5(4): 105-116.
|
|
Wang Ting, Wang Na, Cui Yunpeng, et al. Agricultural technology knowledge intelligent question-answering system based on large language model[J]. Smart Agriculture, 2023, 5(4): 105-116.
|
| [34] |
Lála J, O'Donoghue O, Shtedritski A, et al. PaperQA: Retrieval-augmented generative agent for scientific research[PP/OL]. V2. arXiv (2023-12-14)[2026-03-20]. .
|
| [35] |
Asai A, He J, Shao Rulin, et al. Synthesizing scientific literature with retrieval-augmented language models[J]. Nature, 2026, 650(8103): 857-863.
|
| [36] |
史忠艳, 雷洁, 孙坦, 等. DeepSeek赋能领域知识图谱低成本构建研究[J]. 农业图书情报学报, 2025, 37(3): 4-17.
|
|
Shi Zhongyan, Lei Jie, Sun Tan, et al. Research on DeepSeek-empowered low-cost construction of domain-specific knowledge graphs[J]. Journal of Library and Information Science in Agriculture, 2025, 37(3): 4-17.
|
| [37] |
ELICIT. Elicit: AI for scientific research[EB/OL]. [2026-02-22]. .
|
| [38] |
ELICIT. How we evaluated Elicit Systematic Review[EB/OL]. (2025-03-18)[2026-02-22]. .
|
| [39] |
ELICIT. Systematic Literature Reviews[EB/OL]. [2026-02-22]. .
|
| [40] |
Nicholson J M, Mordaunt M, Lopez P, et al. Scite: A smart citation index that displays the context of citations and classifies their intent using deep learning[J]. Quantitative Science Studies, 2021, 2(3): 882-898.
|
| [41] |
Apata O E, Kwok O M, Lee Y H. The use of generative artificial intelligence (AI) in academic research: A review of the consensus app[J]. Cureus, 2025, 17(7): e87297.
|
| [42] |
Thelwall M, Yaghi A. In which fields can ChatGPT detect journal article quality? An evaluation of REF2021 results[J]. Trends in Information Management, 2025, 13(1): 1-29.
|
| [43] |
叶继元, 郭卫兵. 生成式人工智能参与学术评价的反思[J]. 中国社会科学评价, 2024(1): 37-48, 158.
|
|
Ye Jiyuan, Guo Weibing. Reflections on the participation of generative AI in academic evaluation[J]. China Social Science Review, 2024(1): 37-48, 158.
|
| [44] |
程秀峰, 李嘉琦, 杨金庆, 等. 大语言模型对学术论文评价的可利用性探讨[J]. 图书情报工作, 2024, 68(18): 41-49.
|
|
Cheng Xiufeng, Li Jiaqi, Yang Jinqing, et al. Exploring the application of large language models in academic paper evaluation[J]. Library and Information Service, 2024, 68(18): 41-49.
|
| [45] |
Thelwall M. Evaluating research quality with Large Language Models: An analysis of ChatGPT's effectiveness with different settings and inputs[J]. Journal of Data and Information Science, 2025, 10(1): 7-25.
|
| [46] |
Du Jiangshu, Wang Yibo, Zhao Wenting, et al. LLMs assist NLP researchers: Critique paper (meta-) reviewing[C]//Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing. ACL, 2024: 5081-5099.
|