Financial forecasting from earnings conference calls requires models to reason over complex corporate disclosures, market expectations, and subtle communication signals. However, existing financial benchmarks are often limited to unimodal inputs or single-task settings, making it difficult to evaluate whether multimoda...
Dong Shu, Yan-Guang Liu, Huo-Pu Zhang et al.· 0 citations
Reward models score responses from large language models (LLMs) and guide LLM training toward human preferences. However, reward models can favor superficial attributes such as length or confidence, leading LLMs to produce higher-scoring but not more correct responses. Existing mitigation methods either retrain the rew...
Shuang Liu, Yongliang Miao, Yan-Guang Liu et al.· 0 citations
Universal Activation Verbalizer (UAV), a framework that uses a shared decoder to explain activations from heterogeneous donor models, and provides the activation-grounded factual and semantic information needed for faithful explanations.
Hai-Yan Zhao, Zirui He, Guan-Chun Wang et al.· arXiv.org· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.