Skip to content

Category

small language model

2,728 papers

#small language model Review Oct 2026

Grounding language models with deterministic verifiers for fraud detection

VR-FraudNet, a five-stage framework combining a time-conditioned spectral graph encoder, LightGBM triage and threshold routing, a schema-constrained rationale language model, a fixed deterministic verifier, and an isotonic probability mixer with split-conformal calibration, does not establish universal adversarial safe...

Md Sultanul Arefin Sourav, Evha Rozario, Md Ashiqul Islam et al. · 0 citations
#small language model Review Oct 2026

Author metadata affects Large Language Model scores in scientific peer review

A controlled counterfactual audit of 80 recent English-language arXiv manuscripts shows that metadata perturbations can displace manuscripts and alter top-K shortlist membership, identifying a concrete reliability risk for LLM-assisted scientific evaluation pipelines.

Marco Rospocher · 0 citations
#small language model Book Oct 2026

Who Should Own the Loop? Harness Decomposition for Small-Model Repository Repair

The results show that harness-managed control flow can substantially improve the effectiveness of the smallest models and suggest that small models can be useful for repository repair when responsibilities are divided across explicit stages that can be independently assigned to the component best suited to each.

Francesco Dente, Dario Satriani, Donatello Santoro et al. · 0 citations

Breaking Memory Wall for Fast Edge LLM Inference Using Contextual Sparsity

Deploying Large Language Models (LLMs) on memory-constrained edge servers to serve requests from mobile devices is challenging due to their substantial resource demands. The Key-Value (KV) cache and Feed-Forward Network (FFN) parameters consume the majority of available memory. However, existing methods typically rely...

Zhong-Xiang Wei, Yi-Peng Zhou, Jin Zhao et al. · 0 citations
#small language model Preprint Oct 2026

Purifying Backdoored Large Vision-Language Models by Removing Hijacked Directions

OrthoPurify is proposed, a more efficient method to purify backdoored model weights via one-step orthogonal projection, which reduces the attack success rate to near zero while preserving the original performance across diverse benchmarks, without retraining the backdoored model or introducing inference-time overhead.

Bo-Jun Yang, Hao-Chen Zhou, Zhi-Fang Zhang et al. · 0 citations
#small language model Book Oct 2026

Multi-modal Boundary Testing of Vision–Language Models

This work constructs a signal by force-decoding a fixed set of candidate answers and develops a framework that manipulates the image and question text separately, making the boundary searchable for any behaviour formulated as a choice between candidate answers.

R. Kaiser · 0 citations
#small language model Preprint Oct 2026

Beyond Anonymous Captions: Grounding Character Identity in Video Captioning and Question Answering

Linking people's appearance and actions to character identities is essential for understanding video narratives. We present a framework for identity-aware video captioning and person-centric question answering that combines automatic character identification, explicit spatial grounding, and task-specific adaptation. St...

A. F. Razzouki, Killian Steunou, Khalil Guetari et al. · 0 citations
#small language model Preprint Oct 2026

BanglaBox: A Phonetically-Balanced Corpus and Data-Efficient Foundation-Model Adaptation for Bangla Text-to-Speech with Zero-Shot Voice Cloning

This work contributes a phonetically- and gender-balanced two-tier Bangladeshi Bangla corpus balanced via a tiered Jensen-Shannon divergence objective over conjunct clusters (juktakkhor), together with three fine-tuning changes: a merge-consistent tokenizer extension, Bangla text normalization, and a prompt-masked dual...

Emtiaz Uddin Ahmed, Araf Mahmud, S. Hossain et al. · 0 citations
#small language model Preprint Oct 2026

Why VLMs Miss Small Objects, and When Zooming In Is Safe

A theory built on two quantities of the image interface: S, the number of visual tokens across an object's side, and L, the content a call must cover, which concludes that recognition improves gradually with S, and any search strategy, zoom agents included, obeys a recall-cost frontier.

Jun-Zhe Shi, Yuan Gan, Shi-Da Jiang · 0 citations

Goldsmith: Gold-Loss-Guided Definition Optimization with an Agentic Annotation Harness

Many annotation projects begin before experts have a stable guideline or enough labels to train a task-specific model. We present Goldsmith, an agentic pipeline that turns a small gold set---expert-annotated calibration examples representing the intended task boundaries---into a reusable structured annotation definitio...

Yi-Han Li, Han-Yi Zhang, Xiao-Xi Jiang et al. · 0 citations
#machine learning Preprint Oct 2026

OrBIT: Structure-Guided Embedding Compression

Embedding tables are among the largest components of modern language models. Most compression methods fix a coding geometry such as coordinate blocks, low-rank subspaces, or unrestricted codebooks, and optimize within it. We instead ask whether the coding geometry can itself be discovered. We introduce \emph{OrBIT}, a...

Yunied Puig, Amit Kumar Jaiswal · 0 citations

From tech blogs

See all →

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.