Skip to content
Preprint

Your AI, On a Dial: Controlling Investment Bias in LLMs with a Single Neuron

Aug 2026 · 0 citations · 33 references
Computer Science Economics

TL;DR

The results show that an LLM's aggregate investment stance can be calibrated toward a specified target at inference time.

Abstract

Large language models (LLMs) are increasingly used in investment decision-making, yet prior work shows that they exhibit systematic, model-specific investment preferences. We study whether a model's overall investment stance can be calibrated to a specified direction and strength. We introduce an investment-bias dial, an inference-time intervention on a single neuron that continuously adjusts a model-level decision prior---its overall tendency toward buying or selling---without targeting specific firms or investment attributes. Using matched positive and negative evidence, we evaluate five open-weight LLMs and find that the dial produces monotonic changes in investment stance without modifying prompts or model parameters. At the response level, the dial shifts both investment decisions and the evidential emphasis of generated rationales under identical inputs. In an agentic retrieval setting, the dial also changes what information the model searches for, which evidence it selects, and which evidence is reflected in its final analysis. In a long-context evaluation, the dial maintains stable stance control as context length increases, whereas a matched system-prompt instruction progressively attenuates. We further show that changes in the dial propagate to security rankings and downstream portfolio composition in an exploratory backtest. Overall, our results show that an LLM's aggregate investment stance can be calibrated toward a specified target at inference time.

View source

Similar papers

Review Jul 2026

Analyzing and Correcting Benevolence Bias in Large Language Models

Benevolence bias is identified and measure, a small but consistent tendency for aligned LLMs to lean toward the kinder, safer, more socially approved answer on value-laden survey questions, and is easy to diagnose and straightforward to fix.

Yuanzi Li, Jun-Hao Wang, Minghui Liu et al. · 0 citations
#natural language process... Preprint Sep 2026

The Analyst in the Prompt: Role, Retrieval, and Memory Biases in LLM Financial Analysis

Large Language Models (LLMs) increasingly use user context such as memory, profiles, and role prompts to personalize their responses. This personalization can affect evidence-based judgment: the same evidence may lead to different conclusions under different user contexts. Finance provides a high-stakes setting to study this problem because decisions often depend on interpreting long and complex documents. We test this using 3,575 SEC filings across twelve LLMs. We compare persona-conditioned retrieval, neutral retrieval, and memory-framed context to separate the effect of evidence selection from the effect of interpretation. We find that most user-context spillover comes from how models interpret the same evidence under different roles, rather than from retrieving different evidence. We then test two simple mitigation strategies: expressing the same investor mindset as a user profile instead of an assistant role, and separating evidence-based and personalized outputs. Both reduce spillover, but neither removes it completely, and their effectiveness varies substantially across models.

Ahmed Asaad, Amr Mohamed, Yang Zhang et al. · 0 citations
Jul 2026

How generative AI behaves in the newsvendor problem: a behavioral experimental study

A replicable methodology is introduced, findings across two architecturally distinct LLMs from different developers are extended, and it is demonstrated that deliberate prompt design meaningfully reduces AI decision bias.

Jing-Jie Su, Yan Lang, Kay-Yut Chen · 0 citations
Jul 2026

More Data, Worse Decisions? Preference Reversals in Neural Networks under Gram Incompatibility

This work shows that pooled refitting recomputes the inverse-Gram geometry used to weight source evidence, which can reverse shared preferences, and derive exact and approximate preservation conditions, and develops a three-stage audit that traces strict pairwise reversals through decision changes to task-defined utility loss.

Yanli Yan, Yuanzheng Li, Yong Zhao et al. · 0 citations
Review Aug 2026

Mind the Gaps: Mixture-of-Minds for Human Simulation

Anacreon is introduced, an audience simulation model that targets the individual level within a narrow, well-specified domain and reaches a state-of-the-art ordinal alignment of 0.775, the individual-level accuracy measure on which the field has converged, with a small residual bias.

P. Dahiya · 0 citations
Jul 2026

The Computational Basis of Confidence in Large Language Models

A computational account of confidence in multimodal language models is provided, when answer logits behave as readouts of a latent decision variable is delineated, and statistical decision confidence is established as a unifying framework for studying confidence across biological and artificial intelligence.

D. Kumaran, Viorica Patraucean, M. Ovsjanikov et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.