Skip to content

Category

small language model

2,807 papers

#machine learning Preprint Oct 2026

Robust Parameter-Efficient LLM Adaptation on Analog Hardware

Analog in-memory computing is a promising platform for on-device execution of large language models because it performs matrix--vector multiplications (MVMs) in memory and in parallel, reducing data movement. However, limited digital-to-analog converter precision, input noise, and finite conductance states can degrade...

Jin-Dan Li, Zhao-Xian Wu, Tian-Yi Chen · 0 citations
#machine learning Preprint Oct 2026

How Should Teachers Be Prepared? RL on Student-Induced States for On-Policy Distillation

On-policy distillation (OPD) improves the reasoning capabilities of small language models through token-level teacher supervision on student-generated trajectories. Yet can teachers that excel at solving problems independently also guide student reasoning effectively? Prior work shows that when student prefixes follow...

Xiao-Yu Ma, Hao-Yue Liu, Zhi-Chao Wang et al. · 0 citations
#artificial intelligence Preprint Oct 2026

Choosing an energy-efficient software architecture for building system diagnostic support

Around 30\% of global energy expenditure can be attributed to the building sector, where a large portion of energy-consumption could be avoided by repairing existing faults. Fault detection and diagnosis (FDD) software addresses this issue; however, its creation and operation also have an environmental impact. The magn...

Roxane Koitz-Hristov, Franz Wotawa · 0 citations
#artificial intelligence Review Oct 2026

Scaling Down the Scaling Laws: Parameter Efficiency and Compute-Optimal Training in Resource-Constrained Large Language Models

Large language models (LLMs) have achieved substantial performance gains through increases in model size, training data, and computational resources. However, traditional scaling approaches produce diminishing returns, rising financial and environmental costs, and barriers to participation for researchers operating out...

J. Dwyer · 0 citations
#artificial intelligence Preprint Oct 2026

Backdooring Sparse Autoencoders

Sparse autoencoders (SAEs) are increasingly used not only to interpret language models but also to intervene on their internal representations. We show that this creates a supply-chain attack surface: a maliciously modified SAE can induce attacker-chosen behavior when inserted into the forward pass of an otherwise unch...

E. Ahlers, Daniel Passon, Tobias Kiecker et al. · 0 citations
#artificial intelligence Preprint Oct 2026

Nexus: An Execution Fabric for AI Agents Across Cloud, Edge, and Devices

Language-model agents are evolving into long-running services that interact with models, tools, computers, mobile devices, and distributed environments. Existing agent frameworks simplify reasoning and tool invocation. However, cloud-centric designs face three limitations: centralized execution increases failure impact...

C. Chang, Jia-Lin Zhou · 0 citations
#artificial intelligence Preprint Oct 2026

G-CARB: Graph-Localized Conformal Agent Risk Budget for Compositional Harm

Small language model (SLM) agents need safety controls that track consequences across tool calls with little monitoring overhead. A private read, for example, becomes a leak when a later action sends that data outside the system. We introduce CARB (Conformal Agent Risk Budget), which calibrates when to stop an agent us...

Zi-Jun Yu, Yu-Tong Gu, Vahid Partovi Nia et al. · 0 citations
#artificial intelligence Preprint Oct 2026

IRSTD-Agent: Agentic Infrared Small Target Detection via Zoom-Guided Interaction Learning

Infrared small-target detection plays an important role in maritime monitoring and aerial surveillance. Although multimodal large language models (MLLMs) offer promising capabilities for visual understanding, existing MLLM-based approaches struggle to precisely localize infrared small targets. In this paper, we propose...

Jia-Wen Xi, Yu Zhang, Tian-Yi Zhao et al. · 0 citations
#artificial intelligence Preprint Oct 2026

Grammar-Guided Code Watermarking with Green Temperature

Large language model watermarking embeds detectable statistical signals during decoding, but the resulting changes to token probabilities can degrade generation quality. This trade-off is particularly important for code, where small changes in token selection can break syntax or alter program behavior. Existing code wa...

Hyundong Jin, Hyeseon An, Soohan Lim et al. · 0 citations
#artificial intelligence Preprint Oct 2026

Clean: Second-order LLM Training at Linear Memory Cost via Nystr\"om Sketching

Training large language models (LLMs) entails a fundamental trade-off: memory-efficient optimizers such as Adam discard cross-parameter curvature, whereas full-curvature methods such as SOAP can accelerate convergence at prohibitive memory costs. We introduce Clean, a memory-efficient and full-curvature optimizer desig...

Beheshteh T. Rakhshan, S. Rajabi, Maziar Sargordi Shikai Fang et al. · 0 citations
#artificial intelligence Preprint Oct 2026

HERA: Harness-Environment Co-Evolution for Reliable Agentic Abstention

Large language model (LLM) agents are increasingly capable of acting in complex tool-use environments, yet they often fail to recognize when tasks are infeasible and no valid solution exists. Recent work has formalized this reliability gap as the problem of agentic abstention, and existing approaches typically optimize...

Hang Luo, Bing-Bing Wen, Guang Yang et al. · 0 citations
#artificial intelligence Preprint Oct 2026

ImproveAnyTask: An Autonomous Post-Training Harness for Iterative Model Self-Improvement

Adapting general-purpose large language models to specific tasks requires substantial human effort in designing data and training strategies. Sustaining improvement is especially challenging because model updates change the error distribution, requiring strategies to be continually refined. We introduce ImproveAnyTask,...

Xing-Bo Yao, Xiao-Man Wang, Zheng-Wu Lei et al. · 0 citations

From tech blogs

See all →

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.