Skip to content
Preprint

Deep Learning Models Also Recall Features

Aug 2026 · 2 citations · 19 references
Computer Science

Abstract

Recent work in mechanistic interpretability has studied how large language models recall facts stored in their weights. This paper argues that factual recall points to something broader: a general kind of operation in deep learning models, which I call feature recall. The core observation is that a linear projection can be read as retrieving stored information scaled by input activations. I define feature recall, show it applies across architectures, and contrast it with the established paradigm of feature combination. I also consider how cases of feature recall might be mechanistically identified. The account gives philosophers a new conceptual tool for understanding deep learning, and points to empirical directions for mechanistic interpretability research.

View source

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.