Skip to content

Balancing accuracy and efficiency: Evaluating encoder-and decoder-based models for word sense disambiguation and regular polysemy detection

· 0 citations · 44 references

TL;DR

On a large-scale all-words WSD task, the encoder model not only outperformed the decoder model but also generated substantially lower carbon emissions – an eight-fold reduction.

View source

Similar papers

Preprint Aug 2026

Generative vs. Encoder Large Language Models for ASR Evaluation: A Comparative Study

The results show that encoder-based metrics remain highly competitive, while generative LLMs perform strongly in hypothesis comparison and improve the interpretability of ASR evaluation.

Thibault Bañeras-Roux, Shashi Kumar, Driss Khalil et al. · 0 citations

Sahara Tokenizers at MWE-2026 PARSEME 2.0 Subtask 1: Combining Contextual Embeddings with Structural Decoding for Multi-Word Expression Detection

Alation studies reveal a strong synergy between POS features and CRF decoding, with the combined approach yielding the best single-model performance, and ensembling models trained with different objectives improves both overall F1 score and discontinuous MWE scores, demonstrating the importance of training diversity fo...

Yunus Karatepe, Mert Sülük, Zeynep Tu˘gçe Kırımlı et al. · 0 citations
Aug 2026

A semi-automated LLM-based framework for word sense disambiguation in Serbian

LLM-assisted sense assignment with a Serbian WordNet-based custom inventory, iterative inventory expansion, and expert validation is combined with a constrained JSON-formatted output to support the practical construction and refinement of sense-annotated resources in a low-resource setting.

Saša Petalinkar, R. Stanković, Milica Ikonić Nešić et al. · 0 citations
Jul 2026

A data-centric analysis for efficient semantic knowledge acquisition in word embeddings

A data-centric analysis of semantic knowledge acquisition in word embeddings, focusing on word analogy and semantic similarity shows that, for relational semantics, training-data quality outweighs quantity, and that simple proxy models remain a practical, interpretable tool for efficient data selection.

Aishwarya Jadhav, Mark Anderson, J. Camacho-Collados et al. · 0 citations
#natural language process... Preprint Sep 2026

Discourse Dependency: A Continuous Criterion for Translation Difficulty

Recent calls for harder machine translation benchmarks have not clarified what difficulty should mean. We argue that one meaningful and currently unmeasured axis is referential reach, the distance a segment must look back into its document to resolve the entities and pronouns it contains. We formalize this as discourse...

Ahrii Kim, Chanjun Park, Seong-heum Kim · 0 citations
2026

The Impact of Tokenization Algorithms on Hungarian Language Model Performance

Results show that BPE produces the most compact and morphologically aligned subword representations, while the modified Unigram LM achieved the best overall downstream performance across tasks, underscore that tokenizer choice and vocabulary design are critical determinants of language model efficiency and performance...

Mátyás Osváth, Matej Molnár, Roland Gunics et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.