A Study on Translation Quality Assessment of English Versions of the Shang Han Lun Based on Corpus Linguistics and Large Language Models
DOI:
https://doi.org/10.70088/29dek341Keywords:
Large language models, Corpus linguistics, TCM classics translation, Shang Han Lun, Translation quality assessmentAbstract
Driven by the dual national strategies of TCM internationalization and cultural digitization, high-quality English translation of TCM classics has become a core component of TCM cross-cultural communication. However, the quality of existing English translations of TCM classics varies greatly, and there is a lack of quantifiable and reusable evaluation standards. Traditional evaluation models often rely excessively on experts' subjective judgments, which hinders the scientific selection of translations and the improvement of communication effectiveness. Taking the core text of Shang Han Lun as the research object, this study selects five translations (Luo Xiwen, Wiseman, Cheng Zhaozhi and Chen Jiaxu, Greta Young Jie De, Huang Hai) to construct a Chinese-English parallel corpus, and uses the GPT-4o large language model to generate reference translations. By integrating automated evaluation metrics with in-depth case analysis, this study constructs a composite evaluation system covering three dimensions: accuracy, fluency and domain adaptability. Empirical results show that there are significant differences among the translations in terminology consistency and the handling of culture-loaded words. Among them, Greta Young's translation has the highest comprehensive score, showing the best fit with the reference translation in terminology standardization and semantic transmission, and demonstrating strong potential for standardized dissemination. At the methodological level, this study couples the quantitative advantages of corpus linguistics with the semantic understanding capabilities of large language models, providing a reusable intelligent research paradigm for the translation quality assessment of TCM classics.References
State Council, Strategic Plan for the Development of Traditional Chinese Medicine (2016–2030), 2016.
General Office of the State Council, *Notice of the General Office of the State Council on Issuing the Implementation Plan for Major Projects of Revitalization and Development of Traditional Chinese Medicine*, State Council General Office Doc. No. 3 [2023], 2023.
OpenAI, GPT-4o Technical Report, 2024. [Online]. Available: https://openai.com/research/gpt-4o. [Accessed: Aug. 7, 2026].
Z. Luo, “A case study on English translations of the Huangdi Neijing under the translation quality assessment model,” M.S. thesis, Jiangxi University of Chinese Medicine, 2019. doi: 10.27180/d.cnki.gjxzc.2019.000087.
J. Ning, “Comparison between machine translation and human translation—Taking the terminology of Huangdi Neijing as an example,” Masterpieces Review, no. 36, pp. 146–147, 2021.
F. Dong and W. Xu, “Current status and trends of translation research on Shang Han Za Bing Lun based on CiteSpace,” English Square, no. 10, pp. 23–28, 2026. doi: 10.16723/j.cnki.yygc.2026.10.032.
R. Han, “Research on machine translation models for ancient Chinese and their quality assessment methods,” M.S. thesis, North China University of Science and Technology, Tangshan, 2025.
C. Zhang, F. Chen, L. Hu, et al., “Research on English translation issues of Shang Han Lun,” Western Journal of Traditional Chinese Medicine, vol. 37, no. 6, pp. 103–107, 2024.
T. Kocmi and C. Federmann, “Large language models are state-of-the-art evaluators of translation quality,” in Proc. 17th Conf. Eur. Chapter Assoc. Comput. Linguistics, Dubrovnik, Croatia: ACL, 2023, pp. 193–203.
Beijing International Studies University, BISU-AiTQA: Large Language Model Translation Quality Evaluation Report (v1.0), 2025.
M. Baker, “Corpus linguistics and translation studies: Implications and applications,” in Text and Technology: In Honour of John Sinclair, M. Baker, G. Francis, and E. Tognini-Bonelli, Eds. Amsterdam, The Netherlands: John Benjamins, 1993, pp. 233–250.
Y. Ren, Z. Chen, Y. Lin, et al., “Research on machine translation quality assessment of TCM terminology in the era of artificial intelligence—Taking ChatGPT-4 and Google Translate as examples,” Guiding Journal of Traditional Chinese Medicine and Pharmacy, vol. 31, no. 12, pp. 284–293, 2025.
L. Wen and Y. Liu, “Quality evaluation of GPT-4o translations based on NLP metrics—Taking The Analects of Confucius as an example,” Journal of Yichun University, vol. 47, no. 7, pp. 92–98, 2025.
K. Papineni, S. Roukos, T. Ward, et al., “BLEU: A method for automatic evaluation of machine translation,” in Proc. 40th Annu. Meeting Assoc. Comput. Linguistics, Philadelphia, PA, USA: ACL, 2002, pp. 311–318.
M. Snover, B. Dorr, R. Schwartz, et al., “A study of translation edit rate with targeted human annotation,” Machine Translation, vol. 20, no. 2, pp. 85–112, 2006.
A. Lavie and A. Agarwal, “METEOR: An automatic metric for MT evaluation with high levels of correlation with human judgments,” in Proc. 2nd Workshop Stat. Machine Translation, Prague, Czech Republic: ACL, 2007, pp. 228–231.
J. Jiang, “A corpus-based comparative study on translation styles of English versions of Shang Han Lun,” Journal of Basic Chinese Medicine, no. 9, pp. 1531–1534, 2023.
B. Gao and Y. Yang, “A comparative study of English translations of Shang Han Lun from the perspective of Skopos theory,” Journal of Yangling Vocational and Technical College, no. 6, pp. 110–118, 2025.
J. Huang and X. Li, “A corpus-based study on translator styles of three English versions of Shang Han Lun,” Medicine & Philosophy, no. 13, pp. 73–77+81, 2024.
Downloads
Published
Issue
Section
License
Copyright (c) 2026 Xintong Zhang (Author)

This work is licensed under a Creative Commons Attribution 4.0 International License.









