Representation autoencoders (RAEs) reuse features from a pretrained visual encoder as reconstruction and diffusion latents, integrating strong visual representations into image generation. However, RAEs still need to decide which encoder layers form the shared latent space for the generator and pixel decoder. This choi...
Hong-Yang Du, Yun-Fei Xie, Jun-Jie Ye et al.· 0 citations
The feasibility of using generative AI tools such as Grok, ChatGPT and Gemini to create images of imagined monozygotic twins as a means to increase representation of twins in face recognition training sets is discussed.
Michael Zang, Haiyu Wu, Mrinal Sharma et al.· 0 citations
The Veracity-Influence-Sobriety score (VIScore), a metric that quantifies the reachability and capacity of a predictor given the encoded feature, and the hallucination of the searching-based planner, is proposed, showcasing the importance of these three aspects in planning success.
Hai-Yu Wu, Randall Balestriero, Morgan E. Levine· 3 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.