Skip to content

Category

small language model

2,746 papers

#small language model Open access Oct 2026

Machines Can Produce What We Speak in the Brain: A Linguistic Framework for Decoding Inner Speech

When we rehearse a sentence silently, such as ‘Hi, what’s up?’ before greeting a friend, the brain runs much of the machinery of speaking without moving a muscle. This paper asks, from a linguistic perspective, how a machine could turn that silent speech into words. We propose a framework in two parts. The first is a m...

Madan Mohan · 0 citations
#small language model Dataset Open access Oct 2026

A Dataset for Multidimensional Evaluation of French Synthetic Texts Generated by SLMs

This dataset comprises synthetic news descriptions in French, generated from real newspaper headlines using a controlled Small Language Models (SLM) pipeline. The generation process follows three distinct configurations: Retrieval-Augmented Generation (RAG) and generation without external context (NO RAG) and QLoRA-bas...

Ayman Abdellaoui, Jose Manuel García-Campos · 0 citations
#small language model Open access Oct 2026

Reporting adherence of meridian-sinew Tuina for non-specific neck pain to a 2026 expert consensus

This registration is retrospective. It records a systematic appraisal that was completed between 24 September and 1 November 2026. The protocol, the search strategy and the eligibility criteria were finalised in writing before data extraction began; the appraisal itself, however, was finished before this registration w...

RongKai Li · 0 citations
#small language model Open access Oct 2026

Фактори продуктивності YOLO11 у реальному часі на Raspberry Pi 5 за обмеженого бюджету ресурсів: протокол даних і ключові результати

An object detector for small smart cameras and mobile robots must fit within the resources of an onboard computer that it shares with other tasks. On a Raspberry Pi 5 with a budget of two inference threads, we study how the frame rate, latency, accuracy and heating of the YOLO11n detector are affected by the following...

Alla N. Lavrenyuk, Олександр Вікторович Коломієць, Sergii Lavreniuk · 0 citations
#small language model Open access Oct 2026

Content-Matched Is Not Lexically Matched: Text and Untrained-Network Baselines for Evaluation-Awareness Probing

Linear probes are increasingly read as evidence that language models represent whether they are being evaluated. Content-matched designs, which hold a task fixed and switch one cue at a time, are meant to make that evidence strong. The assumption is that a probe separating the two versions has detected the model recogn...

Karan Singh · 0 citations
#small language model Open access Oct 2026

From Information to Control

This preprint develops an end-to-end causal account of how information acquires behavioral control in language models. Across a coordinated series of mechanistic studies, we trace the computation from distributed relational and typed representations, through task- and recipient-dependent qualification, contextual multi...

Yibo Chen · 0 citations
#small language model Open access Oct 2026

Dosezy

v2.5.9: Custom Multi-Week Recurrence, Android 15 Edge-to-Edge, and Cross-Platform Schema Parity Summary of Changes Release v2.5.9 introduces custom multi-week recurrence scheduling (e.g. taking medications every 2 or 3 weeks on specific days), enforces strict portrait orientation across all activities, normalizes devic...

Saaduddin Mohammad, Md Rahif Uddin Khan, Khwaja Mohammed · 0 citations
#small language model Open access Oct 2026

PREreview of "Measured Joules, Learned Routes: Learning to Route for Energy-Efficient LLM Serving"

This Zenodo record is a permanently preserved version of a PREreview. You can view the complete PREreview at https://prereview.org/reviews/23177620. ## Summary Routing papers usually optimize proxy costs — parameter counts, API prices, model counts. This one measures the thing itself: per-query GPU energy. The authors...

Karmendra Pandey · 0 citations
#small language model Open access Oct 2026

SpecEdge: Adaptive Speculative Decoding and Quantization Dynamics for Edge Small Language Models

Auto-regressive sequence generation in modern transformer language models is severely memory-bandwidth bound, yielding low arithmetic intensity on resource-constrained edge devices. While speculative decoding mitigates this bottleneck by utilizing a compact draft model to propose candidates verified concurrently by a t...

Benedict Baah · 0 citations
#small language model Dataset Open access Oct 2026

Dataset for "Correcting or Confirming? A Pilot Study of LLM Responses to Simulated Student Pressure in Higher Education"

This dataset accompanies the study “Correcting or Confirming? A Pilot Study of LLM Responses to Simulated Student Pressure in Higher Education”. The workbook documents 60 conversations and 120 model responses collected from ChatGPT, Gemini, and Claude. The protocol comprises ten English-language academic scenarios, two...

Georgios Roussos, Georgios Alexandridis · 0 citations
#small language model Open access Oct 2026

Evaluating Chinese large language models for HIV/AIDS health information: a multidimensional comparative study of response quality

Background Chinese large language models (LLMs) are increasingly used for HIV/AIDS health information, yet their multidimensional performance remains unevaluated. This study systematically assessed five mainstream Chinese LLMs across accuracy, caring, completeness, understandability, and actionability. Methods A cross-...

Jinhao Gu, Jianwei Liu, Yunrong Liu et al. · 0 citations

From tech blogs

See all →

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.