magazinelogo

Journal of Humanities, Arts and Social Science

ISSN Online: 2576-0548 ISSN Print: 2576-0556 CODEN: JHASAY
Frequency: monthly Email: jhass@hillpublisher.com
Total View: 7432052 Downloads: 2198999 Citations: 441 (From Dimensions)
ArticleOpen Access http://dx.doi.org/10.26855/jhass.2026.08.004

Multi-modal English Subtitling for Regional Cultural Tourism Short Videos: Evidence from Chizhou’s Overseas Promotional Clips

Tingting Hou

School of Foreign Languages, Chizhou University, Chizhou 247000, Anhui, China.

*Corresponding author: Tingting Hou

2026 Chizhou Social Science Innovation and Development Research Project “A Study on English Subtitle Translation of Chizhou Cultural-Tourism Short Videos from the Multi-modal Perspective” (Project No.: 2026YB79).
Published: August 13, 2026

Abstract

The global prevalence of vertical short-video platforms has reconstructed the overseas communication logic of Chinese local cities. As a distinctive tourist city in the southern bank of the Yangtze River, Chizhou boasts irreplaceable cultural and ecological resources dominated by Jiuhua Mountain, one of China’s four great Buddhist sacred mountains, picturesque Jiangnan (the cultural region south of the lower Yangtze River in East China, celebrated for classical landscapes, traditional literature and folk heritage) pastoral scenery and diverse local intangible cultural heritage. A great number of tourism-themed short videos released on Douyin and TikTok integrate visual frames, background audio and on-screen English subtitles as three interdependent semiotic channels to deliver local cultural connotations to international audiences. Nevertheless, most subtitle production for Chizhou’s promotional short videos merely relies on crude machine translation without synchronous matching of picture and audio information. Text-only translation logic easily triggers three prominent communication barriers: obscure cultural symbols caused by literal translation, redundant descriptive content conflicting with visual information, and overlong sentences that cannot be read within limited shot duration. Taking social-semiotic Multi-modal Discourse Analysis (MDA) as the core theoretical framework, this paper carries out qualitative multi-modal case analysis based on a self-built corpus of 24 original short-video clips collected from domestic and overseas short-video platforms. After sorting and comparing representative translation samples of Buddhist culture, natural landscape, and folk custom themes, the study summarizes four typical multi-modal defects in current English subtitles. Corresponding targeted optimization strategies are proposed from three dimensions: visual information compression, image-assisted cultural compensation, and audio-oriented stylistic adaptation. This research verifies that qualified tourism subtitles should prioritize cross-modal coherent meaning transmission rather than rigid literal loyalty to source texts. It enriches empirical research materials of audiovisual multi-modal translation for prefecture-level tourist cities in China and provides operable reference standards for Chizhou and similar small-scale cities to polish overseas promotional subtitles and boost cross-cultural tourism communication efficiency.

Keyword

Multi-modal discourse analysis; cultural tourism short videos; English subtitling; cross-cultural translation; Chizhou

References

Chen, Y. X. (2026). Multi-modal coherence in Chinese cultural audiovisual subtitles. Modern Linguistics Review, 14(1), 73-79.

Díaz-Cintas, J., & Remael, A. (2021). Subtitling: Concepts and practices (2nd expanded ed.). Routledge.

Kress, G. (2010). Multi-modality: A social semiotic approach to contemporary communication (1st ed.). Routledge.

Kress, G., & Van Leeuwen, T. (2006). Reading images: The grammar of visual design (2nd fully revised ed.). Routledge.

Lu, S. (2025). Latest trends of audiovisual translation in East Asia’s short-video ecosystem. Translation Studies Forum, 17(2), 89-106.

Pang, S. N., Jin, W., & Zhang, L. J. (2025). Visual cultural compensation in animated film subtitling. Journal of Multi-modal Communication, 13(10), 1089-1095.

Selwa, R. (2026). A cognitive-multi-modal framework for audiovisual translation training. African Journal of Second Language Teaching, 18(1), 45-62.

Szarkowska, A., Gerber-Morón, O., & Piatkowska, M. (2025). The impact of subtitle speed on non-native viewers’ cognitive load. The Translator, 31(4), 411-430.

Valdeón, R. A. (2024). The translation of multi-modal texts: Challenges and theoretical approaches. Perspectives: Studies in Translation Theory and Practice, 32(1), 1-13.

Zhang, D. L. (2009). A comprehensive theoretical framework for multi-modal discourse analysis. Foreign Languages in China, 6(1), 24-30.

Copyright

© 2026 by the author(s).
This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution-NonCommercial-NoDerivatives (CC BY-NC-ND) license, which permits non-commercial use, distribution, and reproduction in any medium, provided the original work is properly cited and is not modified or adapted.
https://creativecommons.org/licenses/by-nc-nd/4.0/

How to cite this paper

Multi-modal English Subtitling for Regional Cultural Tourism Short Videos: Evidence from Chizhou’s Overseas Promotional Clips

How to cite this paper: Tingting Hou. (2026) Multi-modal English Subtitling for Regional Cultural Tourism Short Videos: Evidence from Chizhou’s Overseas Promotional Clips. Journal of Humanities, Arts and Social Science10(8), 870-875.

DOI: http://dx.doi.org/10.26855/jhass.2026.08.004