Skip to content

Author

Ang Li

We have 3 of 38 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Oct 2026

PAIR: Bridging Perception and Action in Vision-Language-Action Models

Vision-language-action (VLA) models map visual observations and language instructions to continuous robot actions. This task requires a transition from representations that describe the scene and instruction to representations that support action generation. Many continuous-action VLAs leave this transition implicit an...

Kai Feng, Guoheng Sun, Ang Li · 0 citations
#machine learning Preprint Oct 2026

XGenAct: Geometry-Enhanced World Action Models through Cross-Task Generation

World action models (WAMs) have advanced robot control by predicting how observations and actions evolve over time. Despite this progress, RGB and action based future prediction does not explicitly address the spatial understanding needed for robot manipulation. Existing efforts often add a limited set of spatial predi...

Ting-Ting Du, Zi-Yao Wang, Guoheng Sun et al. · 0 citations
Jun 2026

Drop-Then-Recovery: How Redundant Are Vision-Language-Action Models?

The results suggest that current VLA benchmarks may exert limited pressure on deep language grounding and compositional instruction understanding, and that future VLA architectures should allocate capacity more deliberately across language, vision, and action components.

Guoheng Sun, Kai Feng, Shwai He et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.