Pathology vision-language models (VLMs) have shown strong visual perception ability, but their robustness in the language domain remains poorly characterized. Existing pathology VLM benchmarks largely rely on canonical closed-set prompts or perturb only generic templates, treating language as a fixed evaluation compone...
Fang-Qi Cheng, Kuo Gong, Shan Liu et al.· 0 citations
SlideBank is introduced, a training-free framework that represents each WSI as a persistent, concept-indexed, and spatially grounded evidence bank and achieves over 99% rephrasing consistency and substantially reduces amortized inference cost through persistent evidence reuse.
Bei-Di Zhao, Gexin Huang, Ciro Zhang et al.· 1 citation
SkillHEX is introduced, a closed-loop framework coupling hypothesis-driven self-verification with evidence-guided tree search that translates falsifiable failure hypotheses into executable tests, producing diagnostic evidence as dense reward without additional environment attempts.
Yuru Feng, Yaoqi Chen, Bei-Di Zhao et al.· 0 citations
This work proposes MESA (a Multi-structure Evidence Selection framework for long-horizon Agent), which builds five complementary structure views of each trajectory and learns from end-to-end answer-level feedback to select and fuse a query-specific subset for a frozen answer model.
Bei-Di Zhao, Yaoqi Chen, Yuru Feng et al.· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.