A unified comparison of attribute-based ZSL, episodic meta-learning, metric and prototype estimators, graph and optimal-transport inference, generative any-shot models, and adaptation of vision–language models is analyzed.
Jie Li, Yu-Bo Sun, Xun Du et al.· Mathematics· 0 citations
ConceptFormer models query-relevant evidence as continuous, query-conditioned latent concepts that explicitly bridge localized visual evidence and semantic relevance, without requiring either textual intermediate representations or direct reliance on raw visual annotations.
Chunyi Peng, Zhi-Peng Xu, Yu-Kun Yan et al.· 0 citations
Deep research requires models to retrieve, connect, and synthesize evidence from large-scale heterogeneous sources to answer complex queries and produce analytical reports. Existing benchmarks mainly evaluate final outcomes, such as answer correctness, report quality, or citation alignment, while providing limited visi...
Yubo Sun, Chunyi Peng, Yukun Yan et al.· arXiv.org· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.