Skip to content

Author

Quan Z. Sheng

We have 3 of 42 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Book Open access Aug 2026

Causal Abstraction Learning for Multi-Modal Grounded Planning

Causal Abstraction Learning for Multi-Modal Grounded Planning (CALM) is proposed, a framework that enhances planning agents with the ability to discover and exploit causal regularities across tasks.

Xin-Shu Li, Shiyi Yang, Ziqi Xu et al. · 0 citations
Aug 2026

Self-supervised Causal Effects Estimation

Self-supervised Causal Effects Estimation is proposed, a novel framework that integrates causal priors with self-supervised learning to construct balanced and predictive representations for causal effects estimation that consistently outperforms state-of-the-art methods.

Xin-Shu Li, Shiyi Yang, Venus Haghighi et al. · 0 citations
Book Open access Aug 2026

Causal Abstraction Learning for Multi-Modal Grounded Planning

Recent advances in multimodal embodied agents have enabled long-horizon planning in visually rich environments via natural language. Yet, their generalization remains brittle when task instructions deviate from familiar examples, exposing a reliance on surface imitation rather than structural understanding. We propose Causal Abstraction Learning for Multi-Modal Grounded Planning (CALM), a framework that enhances planning agents with the ability to discover and exploit causal regularities across tasks. CALM incrementally develops a causal library by abstracting precondition–effect structure from successful executions, yielding compact representations that emphasize stable dependencies beyond incidental context. When execution diverges from expectation, these abstractions are refined through contrastive causal reasoning, enabling targeted adjustments that resolve underlying mechanism mismatch. The resulting structure serves as a transferable prior for planning in novel settings, integrating perceptual cues with mechanism-informed knowledge. Without retraining or task-specific heuristics, CALM generalizes robustly and efficiently to linguistic and perceptual variation. Experiments on ALFRED and VirtualHome demonstrate consistent gains, highlighting causal abstraction as a scalable inductive bias for grounded planning.

Xinshu Li, Shiyi Yang, Ziqi Xu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.