Jul 2026· Global Knowledge Memory and Communication· pp. 1-20· 0 citations· 50 references
TL;DR
The study’s primary contribution lies in its cross-linguistic comparative design, which integrates topic modeling, keyword clustering and social network analysis across Chinese- and English-language scholarly corpora.
Abstract
This study aims to examine how digital humanities methodologies, particularly natural language processing and network analysis, have influenced archival scholarship across Chinese- and English-language contexts from 2004 to 2023.
The study uses latent Dirichlet allocation for topic modeling, term frequency-inverse document frequency for keyword extraction and social network analysis to examine thematic patterns in archival journal articles. Chinese knowledge information processing tagger is used for Chinese-language data, while Gensim with Chinese and Gensim natural language toolkit is applied to English-language data, enabling cross-linguistic comparison of thematic patterns.
The results reveal distinct thematic divergence. Chinese-language journals emphasize state-led initiatives, such as national identity construction and digital infrastructure, whereas English-language journals focus on community archives, Indigenous rights and archival justice. Despite these differences, both corpora converge on themes such as digital archive management, policy dissemination and archival education, reflecting a shared emphasis on archival infrastructure and governance. Topic modeling identifies six coherent themes in Chinese articles and seven in English. Keyword analysis shows that Chinese literature prioritizes institutional roles, whereas English texts emphasize social justice and community engagement.
The study’s primary contribution lies in its cross-linguistic comparative design, which integrates topic modeling, keyword clustering and social network analysis across Chinese- and English-language scholarly corpora. This approach highlights divergent orientations in archival knowledge production across the two communities, a dimension that has received limited systematic attention in prior single-language or single-method studies.
This study incorporates text mining into critical discourse analysis to examine how government science agencies in China and the United States position themselves in space science communication on social media. Two specialized corpora have been built by collecting posts from government science agencies on Weibo from China and X (formerly Twitter) from the U.S. With the help of the text mining tool KH Coder, this study gives a corpus-assisted discourse analysis of the particular ways of positioning at different levels of discourse: (1) topics/themes, (2) addressing terms, and (3) those words that co-occur with self-addressing terms in a sentence. The findings reveal significant differences in their preferential ways of positioning. Chinese government science agencies present themselves as state-affiliated yet approachable institutions, blending achievements and operational efficiency with patriotism and collective pride. Their use of diverse addressing terms and co-occurrence patterns portrays them as experienced, supportive guides, balancing national identity with interpersonal closeness. In contrast, U.S. government science agencies emphasize professionalism, focusing on research, space exploration, mission execution, and audience engagement, with minimal reference to state affiliation. Their addressing terms are formal and standardized, with pride centred on mission success and discovery, highlighting expertise and scientific leadership. Their preferential ways of positioning are further explained in their respective contexts in order to present a proper understanding of these differences.
Fangfang Chen, C. Ngai, Ming Liu· PLoS ONE· 0 citations
A bibliometric analysis of English-language, Scopus-indexed scholarship on cultural bias and ethical concerns in artificial intelligence (AI)-driven communication, covering 1,919 documents published between 2015 and 2025, suggests exponential growth in scholarly output, particularly from 2023 onward.
V. Muriira, Venoth Nallisamy, J. Gikonyo et al.· 0 citations
This study integrates bibliometric visualization with content coding to clarify the development, knowledge structure, and research gaps of digital discourse studies, offering directions for future interdisciplinary inquiry.
Yifan Liu, Omar Ali Al-Smadi, Siti Soraya Lin Abdullah Kamal et al.· Journal of Nusantara Studies...· 0 citations
This study explores the role of media linguistics in shaping the linguistic norms of contemporary mass media in Kazakhstan, an area that remains insufficiently examined in relation to bilingual media practices, genre variation, and digital communication platforms. It focuses on the influence of global communication trends, digital technologies, and national language policy on linguistic change. The research examines the introduction of English-language borrowings, their adaptation in Kazakh and Russian, and the modification of traditional linguistic structures. These shifts are driven by the widespread use of English-language platforms such as social media (Facebook, Instagram, Telegram), streaming services, and international content formats. The findings show that English loanwords are most common in advertising and news, where there is a need for rapid adaptation to global trends. In contrast, analytical and official publications adhere to traditional linguistic norms, highlighting a balance between formal and informal communication. The study also emphasizes the importance of Kazakhstan’s national language policy in preserving linguistic identity, with measures regulating foreign words in the media to develop sustainable language standards. Digital technologies also shape informal media discourse through memes, hashtags, and hybrid language forms. These changes are particularly evident among younger audiences, who adapt more quickly to language innovations, while older generations tend to be more critical of these shifts. In conclusion, media linguistics proves to be an effective tool for analyzing the intersection of globalization, national identity, and digital transformation, offering valuable insights into the adaptation of linguistic norms amid rapid technological development.
A. Mukhanbetzhanova,, S. Tapanova· The International Journal of...· 0 citations
This paper examines linguistic features and language use trends in Tanzanian Facebook discourse. As in many African countries, social media has become an integral part of everyday communication in Tanzania, particularly among young people, making it important to understand its linguistic characteristics. The study employed an online survey and analysed 50 Facebook posts drawn from different users’ accounts. The findings reveal that Tanzanian Facebook discourse is characterised by graphological features, including images and emojis, and orthographic features such as creative punctuation, vowel deletion, and number-letter substitutions, reflecting brevity and informality. The study also identified translanguaging practices, including slang, code-switching, language shifts, and vowel lengthening, which express linguistic creativity, identity, and social belonging. In addition, translation practices facilitate communication across linguistic boundaries and promote cross-cultural interaction in multilingual online environments. These findings demonstrate the dynamic relationship between language, digital communication, and cultural identity in contemporary Tanzania. The study concludes that social media has reshaped language use by encouraging innovative linguistic practices while maintaining communicative effectiveness. It recommends digital literacy programmes to enhance awareness of how these linguistic features influence readability and engagement and calls for further research on language use across different social media platforms and their long-term impact on Swahili and English in Tanzania.
Atanas M. Ndunguru, G. Mapunda, M. I. Choyo· East African Journal of Comm...· 0 citations
This article examines the influence of English on Vietnamese in the context of globalisation, with particular attention to lexical borrowing, code-switching, digital discourse, and language attitudes. Grounded in contact linguistics, sociolinguistic theory, and language ideology studies, the paper uses a qualitative design that combines naturalistic observation of public and digital language practices with a systematic review of recent scholarship on English-Vietnamese contact. The analysis shows that English influence is most visible in the lexical domain, especially in technology, education, business, and popular culture. English-Vietnamese code-switching is also widespread in youth-oriented and digitally mediated communication, where it serves pragmatic functions such as emphasis, brevity, and topic framing, as well as social functions such as identity performance and in-group solidarity. Syntactic and phonological influences appear more limited and register-specific, although Englishinfluenced discourse markers, passive constructions, and loanword pronunciation patterns are increasingly noticeable among proficient bilingual users. The study further indicates that attitudes toward English influence remain ambivalent: younger speakers often treat bilingual hybridity as a sign of modernity and global identity, while purist concerns persist in public and institutional discourse. The paper concludes with implications for Vietnamese language policy and English education.
Vu Mai Phuong· International Journal of Edu...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.