Skip to content

Author

Robby T. Tan

We have 2 of 34 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

UC-VLM: Consistency-Driven Learning for AI-Generated Image Detection with Vision-Language Large Models

UC-VLM is a unified multi-stage binary-supervised framework that consistently reuses the same authenticity labels for visual adaptation and label-conditioned text generation, while leveraging automatically optimized instructions to reduce prompt sensitivity without requiring human-written rationales or hand-crafted prompts.

Lei Tan, Shuwei Li, Mohan S. Kankanhalli et al. · 0 citations
Preprint Jul 2026

Token-Based Affordance Grounding with Large Vision-Language Models

TokAG, a zero-shot affordance grounding framework that exploits the token-level semantic-spatial signals in LVLMs to localize action-relevant regions without external supervision, and introduces a spatial-aware token-selection mechanism to systematically evaluate each output token.

Seung Il Lee, Qinqian Lei, Daguang Xu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.