Which layers of a feed-forward geometry model are needed to preserve both camera poses and dense 3D structure? We study layer redundancy in VGGT, $\pi^3$, and VGGT-$\Omega$: 3,018 pruned configurations, scored on seven camera-pose and dense-geometry metrics across four indoor and outdoor datasets. Four findings follow:...
Feng-Yi Zhang, Holger Caesar, Xiang-Yu Sun et al.· 0 citations
While 2D Vision Foundation Models offer a pathway to automate 3D semantic pseudo-labelling, translating these priors into robust 3D representations typically requires complex heuristics or multi-model ensembles. We introduce SplatLabel, an automated pipeline that leverages a 4D Gaussian representation to extract LiDAR...
Nitya Nanvani, Andras Palffy, H. Caesar· 0 citations
FedCKA is proposed, a Centered Kernel Alignment (CKA)-based strategy that dynamically handles the personalization-globalization trade-off, and outperforms established federated baselines, including FedBN, FedRep, and FedSelect, improving average NDS by 7 percentage points over the strongest baseline.
Jolle Verhoog, Ali Burak Ünal, Holger Caesar· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.