Epistemic State Representations in Large Language Models
Large language models often generate confident but fabricated content, yet whether they maintain internal representations of their own epistemic states is unknown. To address this question, contrastive activation vectors were extracted for 15 epistemic states from five language models, using 100 matched present-neutral...