Skip to content

Spanish–Guaraní Code-Switching in Social Media: A Dependency Parsing Study

Aug 2026 · International Journal of Bilingualism · 0 citations · 20 references

Abstract

This study examines how Spanish–Guaraní bilinguals structure code-switching (CS) on Twitter/X and investigates what these patterns reveal about competing models of bilingual grammar. Specifically, it explores language dominance, switch direction, syntactic position, switch span, and register effects in Paraguayan bilingual discourse. The study adopts a corpus-based and dependency-analytic approach using Universal Dependencies (UD) annotation to examine CS patterns in a Spanish–Guaraní social media corpus. The analysis is based on an extended UD-annotated Spanish–Guaraní Twitter/X corpus. Switch points were examined according to direction, dependency relations (DEPREL), part-of-speech categories, span length, and register (formal vs. informal tweets). Quantitative analyses were complemented by qualitative examination of representative examples. Results show that Guaraní, including the variant Jopará, functions as the dominant matrix language overall, especially in formal tweets. Switches into Spanish occur disproportionately at core grammatical positions such as subjects, objects, and obliques, whereas switches into Guaraní frequently occur at clause-edge and discourse-related positions. Switch spans are generally short but vary systematically according to syntactic environment, with clause-structuring elements licensing longer spans than argument-level switches. The study introduces switch span as a measure of structural scope in CS and combines dependency-based syntactic analysis with register-sensitive investigation in a low-resource bilingual setting. The findings support Matrix-Language Frame predictions in formal registers while demonstrating the importance of hierarchical and scope-based accounts of bilingual grammar. More broadly, the study shows how dependency annotation can provide fine-grained insights into CS in Indigenous and mixed-language repertoires.

View source

Similar papers

Open access Aug 2026

Linguistic frame switching in German–English bilinguals: implications for neuroscience

Introduction: Multilingual speakers frequently report that operating in a different language involves more than a change in grammatical system. The phenomenon at issue is linguistic frame switching (LFS): the involuntary reorganisation of cognitive, emotional, and identity-relevant schemas when a multilingual speaker moves between languages. This study investigated whether German–English bilinguals exhibit systematic language-dependent differences in self-perception and narrative content. Materials and methods: This report presents findings from an original mixed-methods study of 27 German–English bi- and multilinguals, combining self-report questionnaire data with computational text analysis (linguistic inquiry and word count, LIWC) of matched narrative pairs elicited by identical picture prompts in both languages at a six-week interval. Results: Approximately 78% of the participants reported feeling distinctly different when using their respective languages (21 of 27; χ2(1) = 8.33, p = 0.004). LIWC analysis revealed systematic cross-linguistic divergences: German narratives were consistently richer in cognitive processing, occupation-related content, and achievement themes; English narratives were higher in social process language. At the individual level, the contrasts were sometimes dramatic—the same person, the same picture, but two narratives that read as if written by different authors. Conclusions: It is argued that these patterns cannot be explained by code-switching or translation equivalence effects alone. LFS, the present report suggests, has largely untapped implications for how neuroscience conceptualises the brain–mind–environment relationship. These findings are exploratory and subject to important limitations: the sample was small (N = 27) and self-selected, language proficiency was self-rated rather than objectively measured, and no monolingual control group was included.

Judith Zangerle · 0 citations
Open access Aug 2026

Exploring Mixed Copies: Evidence From English-Estonian Bilingual Speech

Recently, contact linguistics has become increasingly interested in multiword units. At the same time, the code-copying framework (CCF) includes the notion of mixed copies (MCs) that are in-between global copies (‘borrowing’) and selective copies (‘structural change’) and illustrate the transition between the lexicon and grammar. The research question is: What types of MCs occur in English-Estonian bilingual speech? The data were transcribed, and MCs identified, annotated, and classified according to their structure. English items were searched for in Estonian dictionaries to establish their Estonian equivalents or conventionalization of such items. The frequencies of MCs and their Estonian equivalents were also searched on Google to determine whether the MCs occur outside the corpus. Three datasets were analysed: written texts from 44 blogs (385,124 tokens), spoken data from 10 vlogs (117,555 tokens), and 8 podcasts (77,277 tokens). Quantitative analyses of the various MC types were conducted, followed by a qualitative analysis of representative examples. Compound nouns constitute the majority of MCs, followed by idioms, phrasal compounds, and a small number of compound verbs. No frame-changing MCs (i.e., MCs resulting in grammatical change) were attested. Since compound nouns and analytic verbs occur in both languages, structural similarity may be a facilitating factor in copying. The notion of MCs is not widely used. Research typically focuses on particular types of items (e.g., compound nouns or verbs); here, however, the question is reversed: which types of items yield MCs? It was established that the proportion of MCs in the data is comparable to that of selective copies. Within MCs, the globally copied element renders the remaining part more specific, highlighting the importance of meaning in contact-induced language change. MCs are also present on the Estonian internet and, in some cases, outnumber their Estonian equivalents, if such equivalents exist.

A. Verschik, H. Kask · 0 citations
Open access Aug 2026

The Impact of English on Other Languages and Linguistic Justice

This article explores the implications of English's role as a global lingua franca for other languages, with a particular focus on the transformations these languages undergo due to English influence. Such changes are most evident in the lexical domain, where they are typically examined under the framework of lexical borrowing or Anglicisms. However, they also extend to grammatical, stylistic, and pragmatic levels, affecting formulation structures, argumentation strategies, and behavioural patterns, often without speakers’ conscious awareness. This article aims to highlight that both the growing prevalence of English across diverse domains and the resulting shifts in speakers’ native languages are interconnected manifestations of the same global linguistic development. The observed linguistic changes are the result of an asymmetrical relationship between languages, which raises questions of linguistic justice. The study draws on corpus analyses of the German language, supplemented by data from cross‐linguistic research initiatives, such as the Global Anglicism Database Network project (GLAD, https://www.nhh.no/en/research‐centres/global‐anglicism‐database‐network/ ).

Sabine Fiedler · 0 citations
Review Open access Aug 2026

Bridging constructions in Sino-Tibetan languages

Abstract This article presents the first systematic typological survey of bridging constructions in the Sino-Tibetan language family. Bridging constructions are discourse-linking devices in which material from a preceding clause is repeated or summarized at the onset of a following sentence, thereby contributing to cohesion, discourse continuity, and information backgrounding. Despite growing typological interest in such constructions, previous cross-linguistic research has relied heavily on data from Papunesia and South America, leaving Eurasia – and Sino-Tibetan in particular – underrepresented. The study is based on a source-driven sample of 50 Sino-Tibetan languages and combines two complementary datasets: a primary corpus of unedited oral narratives from five languages and a broader comparative survey based on secondary grammatical descriptions. This mixed methodology allows for both fine-grained analysis of attested bridging constructions and broader assessment of their distribution across the family. The analysis examines their occurrence, formal realization, and relationship to morphological type, clause-linking strategies, participant-tracking resources, zero anaphora, and areal patterning. The results show that bridging constructions are widespread across Sino-Tibetan, but that their syntactic realization varies considerably. The data provide limited support for strong typological correlations. Instead, formal variation reflects the interaction of language-specific clause-linkage resources, discourse-packaging strategies, and broader areal tendencies along the Indosphere-Sinosphere continuum.

Katia kat ᶨ a Chirkova tʃirkova, Hongdi xuŋti Ding tiŋ, Paul paul van Els van ɛls · 0 citations

Gender Bias Evaluation in English-Portuguese Automated Translation Outputs

This work contributes a preliminary cross-system comparison for an underexplored language pair and lays the groundwork for larger-scale evaluations of gender-inclusive AT, indicating that all four models struggle to faithfully convey gender information from source to target.

Xiao-Lan Xu, Sara Mendes, F. Batista et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.