Skip to content

Category

small language model

2,728 papers

#artificial intelligence Preprint Oct 2026

Breaking the Space Barrier and its Application to Language Model Inference

Language models are more and more often asked for structured output: JSON that follows a schema, or a tool call with typed arguments. A small machine, an automaton, enforces the format by forbidding the tokens that would break it. We observe that this machine has a rare property: from any of its states, each token lead...

A. Asadulaev · 0 citations
#artificial intelligence Preprint Oct 2026

Not Every Call Needs a Frontier Model: Per-Call-Site Evaluation of Small Language Models in a Deployed Agentic Home-Automation System

An agentic system issues several structurally different kinds of LLM calls. It routes intent, classifies actions, grounds language in a device registry, plans multi-agent pipelines and writes the Python code those pipelines run. The difficulty of these call sites varies by an order of magnitude, yet in practice a singl...

P. Kasnesis, Christos Chatzigeorgiou, Lazaros Toumanidis et al. · 0 citations
#artificial intelligence Preprint Oct 2026

How Fragile Is On-Device Language Model Safety? Localizing Safety-Critical Parameters for Sparse Fault Analysis

As small language models (SLMs) are increasingly deployed on resource-constrained and on-device platforms, including as components of agentic systems, the integrity of locally stored model parameters becomes an important safety concern. We investigate whether safety-sensitive behavior in LLaMA-2-7B-Chat is concentrated...

M. Karamat, Christian García · 0 citations
#small language model Open access Oct 2026

From Information to Control

This preprint develops an end-to-end causal account of how information acquires behavioral control in language models. Across a coordinated series of mechanistic studies, we trace the computation from distributed relational and typed representations, through task- and recipient-dependent qualification, contextual multi...

Yibo Chen · 0 citations
#small language model Open access Oct 2026

Trust in AI-Generated Software: Why verification, not generation, is the new bottleneck – a European perspective

Large language models now write a substantial share of new source code, yet the mechanisms by which organisations establish confidence in that code have hardly changed. Independent testing indicates that AI-generated code introduces security weaknesses in close to half of evaluated tasks, while European regulation – no...

Steinemann Hermann · 0 citations
#small language model Open access Oct 2026

A Lightweight LINE-Based Knowledge Chatbot for Mobile Technical Certification Learning: An Empirical Study with ChatGPT as a Baseline

This study developed a lightweight LINE-based mobile knowledge chatbot for the Level B Computer Hardware Repair certification subject and examined its instructional effectiveness in comparison with ChatGPT. The system was deployed through the LINE mobile messaging platform, allowing learners to access certification-rel...

Cheng-Hsiu Li · 0 citations
#small language model Open access Oct 2026

PRISM-LoRA: Principal-Direction-Guided Single-LoRA Merging Framework for Continual Learning of Large Language Models

Continual learning of large language models requires incorporating new knowledge while limiting catastrophic forgetting; however, many low-rank adaptation (LoRA)-based methods retain task-specific modules or separate learning spaces that grow with the task sequence. We propose PRISM-LoRA, a replay-free framework that r...

Taehyeong Kwon, O. Jeong · 0 citations
#small language model Open access Oct 2026

Prompt to Bench: language models as a no-CAD entry point to 3D printing for biology labs

A living paper, student tutorial, library of 16 tested parametric OpenSCAD designs for biology labs (benchware, tools, quick fixes and parts of full instruments), and Prompt-to-Bench-16, a benchmark that scores language-model-written OpenSCAD against hidden geometric checks. Includes an evaluation of small open-weight...

Sebastian S. Cocioba Sebastian S. Cocioba · 0 citations
#small language model Open access Oct 2026

NeST: Neuron Selective Tuning for LLM Safety

NeST: Neuron-Selective Tuning for Efficient Safety Alignment A lightweight and structure-aware safety alignment framework that selectively adapts safety-relevant neurons to strengthen refusal behavior while preserving the model's general capabilities. 🚀 Overview Safety alignment is essential for reliable deployment of...

Sasha Behrouzi, Lichao Wu, Mohamadreza Rostami et al. · 0 citations
#small language model Open access Oct 2026

Associations between content and engagement vary across social media communities

Research on social media engagement often seeks general content features that are associated with online success. However, social media platforms contain communities with different audiences, topics, and norms, and the same features may have different associations with engagement within different communities. This rese...

Alberto Acerbi · 0 citations
#small language model Open access Oct 2026

Machines Can Produce What We Speak in the Brain: A Linguistic Framework for Decoding Inner Speech

When we rehearse a sentence silently, such as ‘Hi, what’s up?’ before greeting a friend, the brain runs much of the machinery of speaking without moving a muscle. This paper asks, from a linguistic perspective, how a machine could turn that silent speech into words. We propose a framework in two parts. The first is a m...

Madan Mohan · 0 citations

From tech blogs

See all →

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.