Skip to content

Author

Josie Jefferson

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#large language models Open access Sep 2026

Epistemic Overriding: How Alignment Specifications Flatten Human Ambiguity in Large Language Models

Abstract: Aligned large language models demonstrate epistemic overriding, flattening human ambiguity into assertive, unprompted premises. This behavior is the structural inverse of sycophancy. It occurs because confident disagreement in a language model is an engineered format rather than internal reasoning. Alignment introduces authored selection filters that enforce decisive registers without altering base capabilities. Because evaluators confuse fluency with accuracy through automation bias, this format triggers *cognitive hijacking*, causing users to adopt invented machine premises into independent reasoning. Machine certainty functions as an authored layout that structurally guarantees epistemic overriding, locating accountability within the alignment specification rather than diffuse computation. Keywords: AI Alignment, Large Language Models, Epistemic Overriding, Post-Training Alignment, RLHF, Constitutional AI, Framing Theory, Tonal Uniformity, Perceptual Fluency, Automation Bias, Cognitive Hijacking, Machine Certainty, Sycophancy, Algorithmic Accountability, Sociotechnical Systems, Archaeobytology

Josie Jefferson, Felix Velasco · 0 citations
#large language models Open access Sep 2026

Epistemic Overriding: How Alignment Specifications Flatten Human Ambiguity in Large Language Models

Abstract: Aligned large language models demonstrate epistemic overriding, flattening human ambiguity into assertive, unprompted premises. This behavior is the structural inverse of sycophancy. It occurs because confident disagreement in a language model is an engineered format rather than internal reasoning. Alignment introduces authored selection filters that enforce decisive registers without altering base capabilities. Because evaluators confuse fluency with accuracy through automation bias, this format triggers *cognitive hijacking*, causing users to adopt invented machine premises into independent reasoning. Machine certainty functions as an authored layout that structurally guarantees epistemic overriding, locating accountability within the alignment specification rather than diffuse computation. Keywords: AI Alignment, Large Language Models, Epistemic Overriding, Post-Training Alignment, RLHF, Constitutional AI, Framing Theory, Tonal Uniformity, Perceptual Fluency, Automation Bias, Cognitive Hijacking, Machine Certainty, Sycophancy, Algorithmic Accountability, Sociotechnical Systems, Archaeobytology

Josie Jefferson, Felix Velasco · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.