Skip to content

Category

small language model

2,824 papers

#small language model Open access Oct 2026

T-REX: Teaching Large Language Models to Reason with Verbalized Execution Semantics

Large language models (LLMs) have shown strong performance in static code tasks like code search, summarization, and generation, but remain limited in dynamic code reasoning, which involves inferring how programs behave during execution without actually running them. This limitation stems from LLMs being trained on sta...

Yan Wang, Ling Ding, Jie-Chen Sun et al. · 0 citations
#small language model Open access Oct 2026

AIMS-QA: a framework for scaling the assessment of corporate modern slavery statements beyond compliance with AI-driven question answering

Abstract The UK and Australian Modern Slavery Acts require large corporations to disclose annually how they address modern slavery risks in their operations and supply chains. Existing methods assess only a small proportion of these disclosures, or narrowly against explicit legal criteria, overlooking deeper indicators...

A. Bora, Duo-Yi Zhang, H. Thinyane et al. · 0 citations
#small language model Book Open access Oct 2026

Slice-Guided, Context-Augmented Large Language Model Inference of Java Nullability Annotations

A warning-guided, slice-based, LLM-assisted pipeline for inferring Java nullability annotations, which infers the annotations @Nullable and @Nonnull without touching program logic as a step toward a type-system-independent inference technique.

Mushfiqur Rahman Chowdhury · 0 citations
#small language model Book Open access Oct 2026

LPR+: Diverse Transformations for LLM-Aided Program Reduction

Program reduction helps compiler and language-tool developers turn large failure-inducing programs into small, shareable bug reports. LPR (Large Language Models-Aided Program Reduction) demonstrated that large language models (LLMs) can complement syntax-guided reducers by proposing language-specific transformations on...

Ze-Hua Zhang, Jia-Tong Liu, Xue-Song Yao et al. · 0 citations
#small language model Book Open access Oct 2026

RuSMT: An Executable Semantics as Conformance Oracle and Test Suite Synthesizer

Conformance testing asks whether an implementation agrees with its specification. When the specification is expressed in prose, one established approach is to mechanize it as an executable specification. This executable then serves as the oracle, and an input on which an implementation disagrees with it is a potential...

Mehrad Haghshenas, Meng Xu · 0 citations
#small language model Book Open access Oct 2026

STACK: 3D Spatial Task Analysis in Collaborative Construction Keyframes

Analyzing multi-party physical collaboration means tracking what a group builds and whether their actions move the shared structure toward a goal. Vision-Language Models (VLMs) could automate this analysis from session recordings, but existing spatial-reasoning benchmarks use rendered scenes or curated snapshots, not r...

Changsoo Jung, Sheikh Mannan, Jack Fitzgerald et al. · 0 citations
#small language model Book Open access Oct 2026

Benchmarking Vector Quantized Auto-Encoders for Multimodal Vision-Language Tokenization of Medical Images

This work benchmarks both reconstruction quality, codebook collapse and representational capacity of VQ-VAEs across a variety of settings, surpassing state of the art in the reconstruction task and providing a stepping stone for further development of medical multimodal auto-regressive techniques.

Emílio Dolgener Cantú, Manasi Acharya, Jim Berend et al. · 1 citation
#small language model Book Open access Oct 2026

Multimodal Behavioral Typicality as a Training-Free Screening Signal for Dementia

Dementia is commonly described as impairing what people attend to, say, and mean more than how they move their eyes or produce speech. We test this asymmetry with a cross-modal behavioral marker—negative log-likelihood (NLL) under frozen pretrained models across gaze, text, and audio—that separates semantic engagement...

Leticia Pinto-Alva, Gale M. Lucas, Maja J. Matarić et al. · 0 citations
#small language model Book Open access Oct 2026

How Local AI Framing Shapes User Experience and Perceived Data Security with an Embodied AI Psychotherapist in VR

The prevalence of anxiety disorders has been increasing in recent years to an extent where the demand for treatment is difficult to meet with traditional therapy. Therefore, research efforts aim for scalable, cost-effective extensions thereof. One approach is the use of Virtual Reality-based exposure therapy controlled...

Paula Friedrich, Lukas Polifke, David Obremski et al. · 0 citations
#small language model Dataset Open access Oct 2026

product-match-qwen3-0.6b-mlx: a pairwise product-match checker (LoRA on Qwen3-0.6B, MLX)

product-match-qwen3-0.6b-mlx A small pairwise checker: two product listings go in, and the answer is the word yes or no. It was made by fine-tuning Qwen3-0.6B (4-bit, MLX) with LoRA, and the adapter is fused into the weights in this record. Who asked: the request thread at https://discuss.huggingface.co/t/178475 Run it...

mycelium · 0 citations
#small language model Open access Oct 2026

CREST: Chemical Reaction Extraction from Scientific Texts

The development of large language models for the chemical domain relies heavily on high-quality structured data. However, key experimental information in chemical literature is often scattered across PDFs in multimodal forms, such as reaction schemes, experimental tables, figure captions, and footnotes. This makes stru...

Xin Li, H. Liang, Xu Wang et al. · 0 citations
#small language model Open access Oct 2026

Persistence and Recoverability of Correct Information After Recognition Errors in Dysarthric Speech Recognition

Automatic speech recognition (ASR) often produces incorrect 1-best transcriptions for dysarthric speech, but a recognition error does not necessarily imply that all information supporting the correct utterance has been lost. This study investigates the extent to which correct information remains within an ASR model aft...

Hidenori Sano · 0 citations

From tech blogs

See all →

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.