Skip to content

Author

Baolong Bi

We have 5 of 50 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Oct 2026

SkillForge: Co-Evolving Skills and Agents via Dynamic Skill Lifecycles

Memory-augmented reinforcement learning strengthens LLM agents'ability to solve complex long-horizon tasks. Skills are one such form of memory, pairing instructions with an applicability condition over task types. However, retaining every skill indiscriminately as the policy improves lets obsolete or harmful entries ac...

Yuyao Ge, Yi-Wei Wang, Yu-Chen He et al. · 0 citations
Review Jul 2026

AISPA: User-Centric System Prompt Auditing for Large Language Model Applications

A user-centric framework for systematically auditing system prompts in AI systems, AISPA is introduced, a user-centric framework for systematically auditing system prompts in AI systems that examines specific parts of a system prompt and evaluates them along eight dimensions that matter to users.

Xiangning Lin, Shenzhe Zhu, Shu Yang et al. · 0 citations
Jul 2026

Audio-Zero: Label-Free Self-Evolution for Fine-Grained Audio Reasoning

This work introduces Audio-Zero, the first label-free self-evolution framework in the field of LALMs that improves fine-grained auditory perception and reasoning and reveals that increasingly fine-grained auditory descriptions emerge naturally from game pressure.

Siqian Tong, Xuan Li, Chao-Zhuo Li et al. · 1 citation
Aug 2026

PopCD: Test-Time Behavior Enhancement via Polarity-Prompt Contrastive Decoding.

This work introduces Polarity-Prompt Contrastive Decoding (PopCD), a test-time behavior control method that generalizes contrastive decoding to broader enhancement settings and is applicable to both LLMs and Vision-Language Models without additional training.

Bao-Long Bi, Yuyao Ge, Shenghua Liu et al. · 0 citations
#machine learning Preprint Jul 2026

LP-SFT: Local-Preserving Supervised Fine-Tuning via Multimodal Entropy Structure

LP-SFT, a Local-Preserving Supervised Fine-Tuning objective designed to explicitly protect this inherent entropy structure, improves overall performance over vanilla SFT and recent SFT-enhancement baselines, suggesting that local preservation helps mitigate capability degradation without collapsing sampling-accessible...

Yueyang Wang, Baolong Bi, Shuo Lu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.