Adversarial patches to Vision-Language-Action (VLA) policies can cause both immediate action corruption and persistent state effects that remain after the patch is removed. Existing evaluations largely focus on continuous attacks and do not separate these two effects. We introduce a state-restoration protocol that remo...
Enjia Wu, Fu-Sen Guo, Yu-Xin Cao et al.· 0 citations
Video watermarking underpins copyright protection and provenance for generated media, yet almost every video is compressed by a codec before it is stored or shared. A codec discards precisely the perceptually redundant components that most watermarks rely on, so the payload is often lost even when the marked video look...
Yu-Xin Cao, Hao Yang, Zi-Qi Ding et al.· 0 citations
KeyBound is presented, a learned audio watermark that restores the two ingredients classical watermarking supplied and learned schemes set aside, a secret key and a host-aware carrier, so the key governs payload access while the host-conditioned carrier resists direct transplantation.
Bang-Shuo Zhu, Yu-Xin Cao, Wei-Fei Jin et al.· 0 citations
Video Large Language Models (VideoLLMs) are increasingly deployed in safety-critical applications such as content moderation and video analytics. To process long videos efficiently, VideoLLMs rely on frame sampling, token compression, and modality fusion, which together form an observation pipeline that reduces the raw...
Bang-Shuo Zhu, Wei Song, Yu-Xin Cao et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.