Skip to content

Author

Matthieu Cord

We have 3 of 22 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

JoLT: Joint Latent Trajectories for Context-Guided High-Resolution Tiled Generation

Although text-to-image generative models produce impressive results, they struggle to generate densely detailed, high-resolution (HR) images. Current literature addresses this issue with a low-to-high-resolution approach. First, a low-resolution (LR) image is generated. Then, an upsampled version is generated using the LR image as an additional cue. In this paper, we present Joint Latent Trajectories (JoLT). To generate an image, JoLT uses two streams that jointly denoise LR and HR latent images at each sampling step. The LR latent controls the overall layout, while the HR latent controls the details. We interconnect both branches to jointly integrate their information. We extensively validate our method, demonstrating its advantages over competing baselines. The resulting images are not only richly detailed but also visually pleasing, opening new avenues for artistic creation.

Mathis Koroglu, Guillaume Jeanneret, Hugo Caselles-Dupré et al. · 0 citations
Preprint Aug 2026

Spatially-Grounded Text-to-Video Generation via Inference-Time Gradient-Free Optimization

This work presents Gradient-free Analytical Trajectory Optimization Video Generation (GATO-Vid), a novel training-free and gradient-free approach for precise spatial guidance that significantly outperforms existing baselines in localization accuracy while introducing minimal computational overhead.

Guillaume Jeanneret, Mathis Koroglu, Hugo Caselles-Dupré et al. · 0 citations
Jul 2026

DiMaS: Distribution Matching for Steering Vision-Language-Action Models

DiMaS is proposed, a Distribution-Matching Steering strategy tailored to flow-matching VLAs, which transports between representation distributions rather than shifting along a fixed direction, and it effectively controls behavior across two state-of-the-art VLAs.

Pegah Khayatan, Sara Meziane, Jayneel Parekh et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.