Skip to content

Author

Shwetak N. Patel

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#computer vision Preprint Sep 2026

Can Vision-Language Models Analyze Human-Centered Video? Mapping Model Capabilities and Human-AI Collaborative Workflows

This work systematically analyzes all 1,702 CHI 2026 full papers and identifies 125 that annotate videos, and derives a five-dimensional taxonomy spanning analytic purpose, viewpoint, phenomenon, reasoning requirement, and annotation authority to map the capabilities and limitations of a general-purpose VLM.

Xi-Yuan Shen, Jiuyang Lyu, Seokhyun Hwang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.