Skip to content

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

Self-Confirming Superposition Traps in Reinforcement Learning

It is shown that this loop can sustain a lower-return policy even when representation fitting is globally optimal on data selected by the agent, which then uses the resulting returns to guide its next choices.

Dai Shi, Andi Han, Feng Chen et al. · 0 citations
#machine learning Preprint Sep 2026

SupportCal: Label-Free Calibration of Post-Trained LLMs via Reference Support and Corroboration

SupportCal is introduced, a label-free post-hoc method that retains agreement examples at unit weight and assigns disagreement examples continuous weights based on the own-base PLM's relative support and corroboration from pretrained references selected from a size-compatible candidate pool.

Linhan Luo, Le-Quan Lin, Dai Shi et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.