Multimodal Large Language Models (MLLMs) have demonstrated remarkable reasoning capabilities across vision and language tasks. However, their massive computational and memory demands hinder real-world deployment. While recent efforts reduce costs by employing lightweight language backbones, existing paradigms remain co...
Peng-Cheng Zheng, Chao-Ning Zhang, Jia-Xing Yan et al.· 0 citations
A spatially adaptive neural operator (SANO), which replaces this spatially shared parameterization with a spatially continuous field of location-dependent operator parameters, which consistently outperforms competitive neural-operator, hypernetwork-based, and physics-informed baselines.
Jia-Quan Zhang, Chao-Ning Zhang, Shu-Xu Chen et al.· 0 citations
A budget-aware framework that systematically orchestrates when and what to teach and integrates a solvability-aware teacher gate to dictate the teacher model and a score-guided turn selection mechanism to decide what informative turns to retain is proposed.
Jian-Wei Zhang, Si-Han Cao, Peng-Cheng Zheng et al.· 0 citations
A geometry-aware incremental neural operator (GeoIncNO) for stable long-horizon PDE prediction and a mean--fluctuation decoupled reconstruction mechanism, where stable mean structures and dynamic fluctuations are fused separately, and phase correction is applied only to the zero-mean fluctuation component.
Jia-Quan Zhang, Shuxu Chen, Haifan Meng et al.· 0 citations
This paper proposes salient-residual decoupled multi-view learning for clustering, SRDMVC, introducing a novel decomposition-fusion iterative optimization, which separates the feature space into a salient space and a residual subspace effectively and fuses them using a novel attention mechanism.
Gao-Kai Wang, Yazhou Ren, Feng-Yu Zhang et al.· Proceedings of the Thirty-Fi...· 0 citations
Experiments show that HERO consistently improves long-horizon accuracy, stable rollout length, and out-of-distribution robustness at no inference-time cost, indicating that history-enriched relative supervision is effective for stabilizing long-horizon autoregressive prediction.