Object Detection and Scene Perception for Connected and Autonomous Vehicles Using LM-JEPA
Highlights What are the main findings? A latent-space LM-JEPA framework enables resource-efficient multi-modal object detection and scene perception for connected and autonomous vehicles, achieving higher perception accuracy with lower inference latency compared to conventional LLM and VLM-based methods. Context-aware...