Skip to content

Category

small language model

2,884 papers

#artificial intelligence Review Sep 2026

Simulating Respondents, Not Single Questions: Coherent Survey Generation with Large Language Models

Large language models are increasingly used to simulate response distributions in social surveys. Prior work has achieved accurate population-level simulation for individual questions. Real questionnaires, however, ask each respondent a sequence of related questions. A simulated respondent should show coherent preferen...

Ji Huang, Meng-Fei Li, Shuai Shao · 0 citations
#artificial intelligence Preprint Sep 2026

RSI-Router: Evolving Subtask-Level LLM Routing and Skills for Cost-Efficient Agents

This paper introduces RSI-router, a routing framework that constructs subtask-level model assignments and model-specific skills through recursive self-improvement over accumulated experience and establishes a stronger performance--cost Pareto frontier than 9 routing methods.

Hao Li, Hang-Fan Zhang, Zhi-Yao Cui et al. · 0 citations
#artificial intelligence Review Sep 2026

ControlScope: Workflow Revision and Reliability in LLM Agents

ControlScope compares continuing generated code, editing the next tool call's data arguments, and replacing the unfinished workflow from the same public execution state to evaluate one-time and repeated reviews across filesystem tasks, ALFWorld, and AppWorld.

Jing-Jie Ning, Xue-Qi Li, Yi-Bo Kong et al. · 0 citations
#artificial intelligence Preprint Sep 2026

READ-Bench: Benchmarking Historical Instance Retrieval for Time-Series Diagnosis

Time-series diagnostic systems rarely rely on retrieving relevant historical cases, and when they do, retrieval is evaluated only indirectly through downstream prediction. We introduce READ-Bench, a benchmark for historical-case retrieval across 12 diagnostic datasets, centered on multivariate time series, that defines...

Gerardo Pastrana, Hao-Jun Li, Dhruv Mehta et al. · 0 citations
#artificial intelligence Review Sep 2026

A Benchmark for LLM's Understanding of Middle School and High School Science Topics

Large language models (LLMs) are increasingly integrated into educational settings, yet educators lack robust, standards-aligned tools to evaluate their effectiveness in K-12 science contexts. Existing benchmarks predominantly assess general language or advanced scientific reasoning, leaving a critical gap in understan...

Noah L. Schroeder, Yessy Eka Ambarwati, Yu-Ji Zhang et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Improving Medical Calculation of LLMs with Embedded Coding

Large Language Models (LLMs) perform well on medical examinations and question-answering benchmarks, but remain unreliable on medical calculation tasks that require exact numerical outputs. These calculations support high-stakes decisions such as medication dosing, organ-function assessment, and prognostic scoring, for...

Tianshi Ming, Ying-Ying Zhang, Xian Wu · 0 citations

Translating Post-Quantum Cryptography Roadmaps into an Actionable SME Migration Framework

The results show that PQC migration is not simply an algorithm-replacement exercise; it is an organisational transformation process requiring governance, cryptographic visibility, vendor coordination, phased implementation, and continuous monitoring.

Babatunde Oladoja, S. Tanev · 0 citations
#small language model Review Open access Sep 2026

Role of artificial intelligence in the diagnosis of amyotrophic lateral sclerosis via biomarkers, neuroimaging, and wearable health technologies: a narrative review

Amyotrophic lateral sclerosis (ALS) is a progressive neurodegenerative disorder characterised by upper and lower motor neuron loss. Diagnosis is frequently delayed because of phenotypic heterogeneity, overlap with ALS mimics, and the absence of a single definitive biomarker. Artificial intelligence (AI) offers compleme...

Dhinesh Selvaraju, Kishore Durairaj, Krishna Ravi · 0 citations

From tech blogs

See all →

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.