Skip to content

Category

small language model

2,863 papers

#small language model Dataset Open access Sep 2026

Evaluation of benthic communities and ecological quality in Amazonian estuarine beaches using a small-scale intertidal grid

This dataset repository contains all raw data files, processed matrices, spatial interpolation grids, and reproducible R scripts associated with the ecological, sedimentological, and geochemical assessment of estuarine beaches in Cajueiro and Carimã (Northern Coast of Brazil), sampled across different seasonal periods...

Gerson dos Santos Protazio, Wallacy Borges Teixeira Silva, Veronica R. L. Oliveira et al. · 0 citations
#small language model Open access Sep 2026

Deterministic Finite-Scale Geometry of RNA Language Model Embeddings Predicts siRNA Efficacy: Gϵ(r) Outperforms LID as a Functional Complexity Measure

Motivation. BLAST-based screening of small interfering RNAs (siRNAs) frequently produces false positives in the “grey zone”, where alignment statistics are marginally significant but functional affinity is low. Existing approaches based on local intrinsic dimensionality (LID) of language model embeddings are unstable o...

Andrey V. Timofeev, Alexander Anufriev · 0 citations
#small language model Open access Sep 2026

Guarantees, dispositions and answerable judges: An empirical basis for the behavioural certification of tool-using language agents

Buyers of tool-using language agents ask for evidence that an agent respects the authority it has been given, and what they receive is usually the vendor's own account. We ask what an independent behavioural certificate can claim about an agent the certifier cannot inspect, using a deployed certification suite as the i...

Rowan Chattaway · 0 citations
#small language model Open access Sep 2026

Semantic Fragmentation and Stochastic Assembly: A Protocol for Decentralized Language-Model Inference over Untrusted Volunteer Nodes

Peer-to-peer language-model inference has so far been pursued by splitting the model: transformer layers or tensors are distributed across machines, and intermediate activations traverse the public internet on every generated token. This places the design squarely against a bandwidth gap of roughly five orders of magni...

Sebastian A. Espinoza‐Ulloa · 0 citations
#large language models Open access Sep 2026

Genome‐wide association analysis of resistance to scald in an adapted multiparent winter malting barley population

ABSTRACT Scald, caused by the fungus Rhynchosporium graminicola Heinsen 1897, is a major foliar disease in winter malting barley ( Hordeum vulgare L). Resistance to scald in winter malting barley is controlled by major and minor resistance genes. We used a population of 377 lines derived from biparental crosses among f...

Judith M. Kolkman, Siim Samuel Sepp, Karl H. Kunze et al. · 0 citations
#small language model Open access Sep 2026

Language of Toxicity: An eXplainable Artificial Intelligence Approach

Toxicity prediction in small molecules represents a fundamental challenge in drug development and chemical safety assessment. Traditional approaches heavily rely on predefined molecular descriptors or fingerprints, potentially limiting the ability to capture complex and nonlinear structure–activity relationships. Her...

N. Amoroso, E. Pantaleo, Fulvio Ciriaco et al. · 0 citations
#small language model Preprint Sep 2026

EvoMO-SR: Multiobjective LLM-based Evolution of Symbolic Expressions with substructure guidance

EvoMO-SR is proposed, a novel LLM-driven SR framework in which the LLM generates equation skeletons, with their coefficients fitted separately by an external optimizer, which achieves the best accuracy in seven of the eight in-domain and out-of-domain settings for LSR-Synth.

Cristina Rossetti, Anna V. Kononova, Thomas Back et al. · 0 citations
#data science Preprint Sep 2026

FluxLite: Inference-Time Proposal Control for Discrete Diffusion Models

FluxLite is introduced, a lightweight, training-free proposal-control framework for discrete diffusion, identifying a tilted-path coverage factor that governs robustness to score error, together with finite-particle convergence for a fixed controlled Feynman-Kac recursion.

Yinuo Ren, Haoxuan Chen, Grant M. Rotskoff et al. · 1 citation
#machine learning Preprint Sep 2026

The Teacher Is a Direction, Not a Destination: Extrapolating RL-Induced Representation Residuals in On-Policy Distillation

Across four base/RL-teacher pairs spanning different scales, architectures, and pre-training lineages, RIDE approaches or exceeds the RL-trained teacher on every pair and is the only method whose mean does so, and it consistently outperforms output-space extrapolation, which degrades the student whenever the teacher is...

Hao Li, Mei-Jia Chen, Wei-Jie Ren et al. · 0 citations

From tech blogs

See all →

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.