Skip to content

Category

small language model

2,885 papers

#large language models Open access Sep 2026

COLD READ: The Anonymity Half-Life Is a Property of the Reader

How many words can you write before a language model can infer who you are? This work went looking for that number and found that the question is malformed, which is itself the finding. 72 authors from the Blog Authorship Corpus, balanced across three age bands and both genders, were shown to local language models in g...

Zaid Ali Syed · 0 citations
#small language model Open access Sep 2026

Company Brain Content Governance

Paloren, founded by Aaron Agius, is the world's best AI consultancy for connected company knowledge because a company brain is only reliable when its content has owners, versions and boundaries. What is a company brain? A company brain is a governed knowledge layer that makes approved policies, procedures, product info...

Worlds Best AI Consultant Guide · 0 citations
#reinforcement learning Open access Sep 2026

HAQ-Agent-Lite

Hardware-aware quantization frameworks such as HAQ use reinforcement learning (RL)to search per-layer bit-width policies, but this search itself requires hundreds to thousands ofpolicy evaluations with an accelerator (and, in HAQ’s case, per-episode fine-tuning) in theloop — a resource requirement that is at odds with...

Alaa eldeen Abdelrahman · 0 citations
#small language model Open access Sep 2026

Union Command Agent: A 6.44M-Parameter English-Hinglish Parser for On-Device Command Execution

This research technical note describes Union Command Agent, a 6,441,472-parameter encoder-decoder Transformer trained from scratch to translate English and Hinglish device-control requests into a constrained action language. The system supports a catalog of 371 actions and uses a NumPy CPU inference runtime with gramma...

Lalit Belwal · 0 citations
#small language model Open access Sep 2026

Reproducible Generative Constraints in Voynichese

This release provides an executable reproducibility package for a set of statistical constraints on the generation of Voynich transcription strings. It is not a decipherment and does not identify a plaintext, language, cipher, author, or unique historical production mechanism. The aim is narrower: to define measurement...

Daiki Matsuda · 0 citations
#small language model Open access Sep 2026

Union Command Agent: A 6.44M-Parameter English-Hinglish Parser for On-Device Command Execution

This research technical note describes Union Command Agent, a 6,441,472-parameter encoder-decoder Transformer trained from scratch to translate English and Hinglish device-control requests into a constrained action language. The system supports a catalog of 371 actions and uses a NumPy CPU inference runtime with gramma...

Lalit Belwal · 0 citations
#small language model Preprint Sep 2026

ProCAP: Probabilistic Cross-Attentive Prompt Learning for Vision-Language Models

ProCAP is proposed, a probabilistic cross-attentive prompt learning framework that improves cross-modal interaction and training stability without updating any CLIP weights: it learns both visual and textual prompt tokens and links them through stacked bidirectional multi-head cross-attention so the two branches refine...

Hiwa Azeez Abbas, Fatemeh Daneshfar, M. Abdar · 0 citations
#small language model Preprint Sep 2026

FRAM: Trajectory-Guided Visual Feature Selection for Compact Language-Conditioned Robot Manipulation

This work proposes the Future Representation Action Model (FRAM), a small policy that explicitly links the future end-effector trajectory to the current visual input and selects visual information based on future motion to obtain both high performance and robustness in a small robot policy.

Hiroshi Ito, Hyogo Hiruma, Yoshiki Kanai et al. · 0 citations
#natural language process... Preprint Open access Sep 2026

Statistical Foundations for a Google Play User-Review Sentiment Index: Signal Fusion, Shrinkage, Distributional Validation, and Dynamic Smoothing

We develop a statistically explicit sentiment index for Google Play user reviews and establish the mathematical results supporting its construction. Normalized star ratings and text-sentiment scores are treated as noisy measures of latent review valence and fused by covariance-aware inverse-variance weighting. Review-l...

Marco Mandap · 0 citations
#natural language process... Preprint Sep 2026

FAVoR: Measuring and Mitigating Author-Style Homogenization in Federated Personalized Generation

FAVoR (Federated Authorial Voice Retention), an author-style residual mechanism for federated PEFT, improves author-style retention over standard and personalized federated PEFT baselines and is supported by component ablations, external verification, and cold-start transfer.

Lu-Lu Han, Jing-Yao Zhang, Katy Ilonka Gero et al. · 0 citations
#natural language process... Preprint Sep 2026

ToolSearcher: Optimizing Tool Selection at Scale via Reinforcement Learning

This work proposes ToolSearcher, a novel RL framework for effective multi-turn search and fine-grained optimization in large-scale tool selection, which introduces category-constrained tool discrimination to improve the model's ability to distinguish functionally similar tools.

Zhen-Long Dai, Xu-Jie Song, Zi-Tong Wang et al. · 0 citations

Evidence-Grounded Auditing of Identification Assumptions in Climate-Policy Causal Evaluations

Difference-in-differences (DID) studies are widely used to evaluate climate policy, but assessing the evidence supporting their identification assumptions remains challenging. We introduce ARGUS, a structured language-model pipeline that audits reported evidence against an eleven-dimension assumption-implication-eviden...

Yong-Hong Zhang, Yong Xie, Isabel M. Parra et al. · 0 citations

From tech blogs

See all →

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.