A Review of AIGC Text Detection in Academic Papers
DOI:
https://doi.org/10.70088/cazwht74Keywords:
artificial intelligence-generated content, academic papers, AI-generated text detection, large language modelsAbstract
With the widespread adoption of large language models in academic writing, the detection of artificial intelligence-generated content has become an important research topic. This paper reviews the detection of AI-generated text in academic papers. It defines the relevant concepts, surveys detection methods and commercial detection tools, and introduces experimental datasets and evaluation metrics. The review shows that current detection approaches face several challenges, including a shortage of suitable datasets, ambiguous labels for human–AI collaborative text, insufficient coverage of generative models, limited generalization, and inadequate interpretability. Future research should develop datasets covering multiple languages, disciplines, models, and generation methods; strengthen evaluation on unseen models and adversarial samples; and improve explainable detection and human-review mechanisms, thereby supporting the governance of academic integrity.References
China Academy of Information and Communications Technology and JD Discovery Research Institute, White Paper on Artificial Intelligence Generated Content (AIGC), Beijing, China, 2022.
Y. Wu, X. Qiu, Y. Wang, et al., “Deconstruction of the causes and governance of information pollution triggered by large language model hallucinations,” Journal of Intelligence, 2025. [Online]. Available: https://link.cnki.net/urlid/61.1167.G3.20251208.0922.012
J. Qiu, T. Zhang, and Z. Xu, “Collaborative governance strategies for AIGC-generated false information based on a tripartite evolutionary game,” Library and Information Service, vol. 70, no. 2, pp. 71–85, 2026, doi: 10.13266/j.issn.0252-3116.2026.02.006.
J. Chu and X. Fan, “Multidimensional causes and governance strategies of information pollution in the AIGC environment,” Information Studies: Theory & Application, vol. 49, no. 2, pp. 38–46, 2026, doi: 10.16353/j.cnki.1000-7490.2026.02.005.
Cyberspace Administration of China, *Measures for the Management of Generative Artificial Intelligence Services (Draft for Comment)*, Apr. 11, 2023. [Online]. Available: https://www.cac.gov.cn/2023-04/11/c_1682854275475410.htm
J. Wu, W. Gan, Z. Chen, et al., “AI-generated content (AIGC): A survey,” arXiv preprint arXiv:2304.06632, 2023. [Online]. Available: https://arxiv.org/abs/2304.06632
L. G. Foo, H. Rahmani, and J. Liu, “AI-generated content (AIGC) for various data modalities: A survey,” ACM Computing Surveys, vol. 57, no. 9, Art. no. 243, 2025, doi: 10.1145/3728633.
J. Wu, S. Yang, R. Zhan, et al., “A survey on LLM-generated text detection: Necessity, methods, and future directions,” Computational Linguistics, vol. 51, no. 1, pp. 275–338, 2025, doi: 10.1162/coli_a_00549.
K. C. Fraser, H. Dawkins, and S. Kiritchenko, “Detecting AI-generated text: Factors influencing detectability with current methods,” Journal of Artificial Intelligence Research, vol. 82, pp. 2233–2278, 2025, doi: 10.1613/jair.1.16665.
T. Hong, “Functional reflection and normative application of AIGC detection,” Studies in Science of Science, 2026. [Online]. Available: https://doi.org/10.16192/j.cnki.1003-2053.20260224.002
I. Solaiman, M. Brundage, J. Clark, et al., “Release strategies and the social impacts of language models,” arXiv preprint arXiv:1908.09203, 2019. [Online]. Available: https://arxiv.org/abs/1908.09203
B. Guo, X. Zhang, Z. Wang, et al., “How close is ChatGPT to human experts? Comparison corpus, evaluation, and detection,” arXiv preprint arXiv:2301.07597, 2023. [Online]. Available: https://arxiv.org/abs/2301.07597
Y. Tian, H. Chen, and X. Wang, “Multiscale positive-unlabeled detection of AI-generated texts,” in Proc. 12th Int. Conf. Learn. Represent., 2024.
Y. Wang, X. Guo, Z. Liu, et al., “Detection and comparative analysis of Chinese paper abstracts generated by AI and written by scholars,” Journal of Intelligence, vol. 42, no. 9, pp. 127–134, 2023.
Q. Zhang, X. Wang, and Y. Gao, “A comparative study of literature abstracts generated by ChatGPT and written by scholars: A case study of the information resource management field,” Library and Information Service, vol. 68, no. 8, pp. 35–47, 2024, doi: 10.13266/j.issn.0252-3116.2024.08.004.
E. Mitchell, Y. Lee, A. Khazatsky, et al., “DetectGPT: Zero-shot machine-generated text detection using probability curvature,” in Proc. 40th Int. Conf. Mach. Learn., Honolulu, HI, USA: PMLR, 2023, pp. 24950–24962.
K. Wu, L. Pang, H. Shen, et al., “LLMDet: A third party large language models generated text detection tool,” in Findings of the Association for Computational Linguistics: EMNLP 2023, Singapore: Assoc. Comput. Linguistics, 2023, pp. 2113–2133, doi: 10.18653/v1/2023.findings-emnlp.139.
X. Shen and L. Wang, “Detection of artificial-intelligence-generated academic journal texts,” Science-Technology & Publication, no. 8, pp. 56–62, 2023, doi: 10.16510/j.cnki.kjycb.2023.08.002.
S. Abdelnabi and M. Fritz, “Adversarial watermarking transformer: Towards tracing text provenance with data hiding,” in 2021 IEEE Symp. Secur. Privacy, Los Alamitos, CA, USA: IEEE Comput. Soc., 2021, pp. 121–140, doi: 10.1109/SP40001.2021.00083.
U. Topkara, M. Topkara, and M. J. Atallah, “The hiding virtues of ambiguity: Quantifiably resilient watermarking of natural language text through synonym substitutions,” in Proc. 8th Workshop Multimedia Secur., New York, NY, USA: ACM, 2006, pp. 164–174, doi: 10.1145/1161366.1161397.
X. Yang, J. Zhang, K. Chen, et al., “Tracing text provenance via context-aware lexical substitution,” in Proc. AAAI Conf. Artif. Intell., vol. 36, no. 10, 2022, pp. 11613–11621, doi: 10.1609/aaai.v36i10.21415.
X. Yang, K. Chen, W. Zhang, et al., “Watermarking text generated by black-box language models,” arXiv preprint arXiv:2305.08883, 2023. [Online]. Available: https://arxiv.org/abs/2305.08883
Z. Hu, L. Chen, X. Wu, et al., “Unbiased watermark for large language models,” arXiv preprint arXiv:2310.10669, 2023. [Online]. Available: https://doi.org/10.48550/arXiv.2310.10669
P. Yu, J. Chen, X. Feng, et al., “CHEAT: A large-scale dataset for detecting ChatGPT-written abstracts,” IEEE Trans. Big Data, vol. 11, no. 3, pp. 898–906, 2025, doi: 10.1109/TBDATA.2025.3536929.
E. Mosca, M. H. I. Abdalla, P. Basso, et al., “Distinguishing fact from fiction: A benchmark dataset for identifying machine-generated scientific papers in the LLM era,” in Proc. 3rd Workshop Trustworthy Nat. Lang. Process., Toronto, ON, Canada: Assoc. Comput. Linguistics, 2023, pp. 190–207, doi: 10.18653/v1/2023.trustnlp-1.17.
P. C. Theocharopoulos, P. Anagnostou, P. Tsakanikas, et al., “Detection of fake generated scientific abstracts,” arXiv preprint arXiv:2304.06148, 2023. [Online]. Available: https://arxiv.org/abs/2304.06148
B. Alhijawi, R. Jarrar, A. AbuAlRub, et al., “Deep learning detection method for large language models-generated scientific content,” Neural Comput. Appl., vol. 37, pp. 91–104, 2025, doi: 10.1007/s00521-024-10538-y.
Y. Guo, Z. Dou, H. H. Nguyen, et al., “Measuring human involvement in AI-generated text: A case study on academic writing,” arXiv preprint arXiv:2506.03501, 2025. [Online]. Available: https://arxiv.org/abs/2506.03501
M. Sokolova, N. Japkowicz, and S. Szpakowicz, “Beyond accuracy, F-score and ROC: A family of discriminant measures for performance evaluation,” in AI 2006: Advances in Artificial Intelligence, Berlin, Germany: Springer, 2006, pp. 1015–1021, doi: 10.1007/11941439_114.
B. Tufts, X. Zhao, and L. Li, “A practical examination of AI-generated text detectors for large language models,” in Findings of the Association for Computational Linguistics: NAACL 2025, Albuquerque, NM, USA: Assoc. Comput. Linguistics, 2025, pp. 4839–4856, doi: 10.18653/v1/2025.findings-naacl.271.
Downloads
Published
Issue
Section
License
Copyright (c) 2026 Jie Wang (Author)

This work is licensed under a Creative Commons Attribution 4.0 International License.









