G3Ego: Gaze-Guided Graphs for Egocentric Action Understanding
G3Ego, a graph-based framework for egocentric action understanding that uses gaze as a structural cue to identify action-relevant entities in the scene, achieves competitive performance compared with video-based approaches and consistently improves Macro-F1 under class-imbalanced evaluation, while avoiding reliance on computationally expensive video pretraining.
Marko Haralović, Akash Ramakrishnan, E. T. Martínez
· 0 citations