Skip to content

Category

robotics

1,156 papers

#artificial intelligence Preprint Sep 2026

BEE: Intervention-Adaptive Real-World Reinforcement Learning with Vision-Language-Action Models

Vision-language-action (VLA) models handle long-horizon manipulation, yet success hinges on a few precision-critical phases where millimeter-scale errors undo all prior progress. Online reinforcement learning (RL) can optimize exactly these actions, but free exploration is far too costly on real robots, which makes hum...

Wei-Hui Zhao, Xiao Yan, Zu-Nian Wan et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Automotive mmWave Spinning Radar Place Recognition with Spatially Gated Feature-Correlation Representation

Automotive spinning FMCW radar provides dense, $360^\circ$ sensing and remains reliable under poor illumination and adverse weather, making it well-suited to autonomous navigation. Place recognition uses these observations to identify previously visited locations for re-localization and long-term navigation. However, h...

Saimunur Rahman, Sagun Shrestha, Abdelwahed Khamis et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Geometry-Conditioned Visual Place Recognition in Natural Environments

Visual Place Recognition (VPR) in natural environments remains challenging due to repetitive vegetation, sparse distinctive landmarks, and substantial appearance and viewpoint variation across traversals. While visual observations of the same place can change considerably, their underlying spatial structure is often mo...

Walter Nedov, Saimunur Rahman, Kavindie Katuwandeniya et al. · 0 citations
#artificial intelligence Preprint Sep 2026

FLINT: Fast Lightweight Inference for Traversability

Navigation in off-road conditions is challenging due to the lack of structure. There is no fixed vocabulary for what is traversable. The traversability depends on both the environment and the embodiment's dynamics. Neither of these two variables can be hand-labeled at scale. Thus, traversability has to be learned by th...

W. Bonilla, M. Boisvert, David-Alexandre Poissant et al. · 0 citations
#artificial intelligence Preprint Sep 2026

SlackDrive: Reclaiming Runtime Slack for Adaptive Driving Inference

SlackDrive is proposed, a pre-inference compute allocator that reuses realized latency to select the compute budget of each control step before model execution, complementing existing profiling and resource scheduling while preserving the driving backbone and its compute actuator.

Xiao-Huan Pei, Heng-Guang Zhou, Yuan-Hao Ban et al. · 0 citations
#robotics Preprint Sep 2026

Talk2Escape: Conversational Grounding for Vision-and-Language Navigation

Talk2Escape is introduced, a proactive and model-agnostic dialogue intervention framework that reframes navigation as a closed-loop interactive process and proves that proactive dialogue drastically improves navigation robustness in physical environments.

Ze-Rui Li, Si-Hao Lin, Yan-Yan Shao et al. · 0 citations
#robotics Preprint Aug 2026

Remote Surfaces at Your Fingertips: Electrovibration-Based Tactile Feedback for Robot Teleoperation via Touchscreen Interfaces

Electrovibration-based tactile feedback is demonstrated to be a viable and effective modality for robot teleoperation, improving operator responsiveness and sense of presence in contact-rich manipulation tasks, with direct applicability to safety-critical domains such as nuclear maintenance.

Alperen Kenan, Juan Jose Garcia Cardenas, Adriana Tapus et al. · 0 citations

DreamAvoid: Critical-Phase Test-Time Dreaming to Avoid Failures in VLA Policies

This work proposes DreamAvoid, a critical-phase test-time dreaming framework that enables VLA models to anticipate and avoid failures, and introduces an autonomous boundary learning paradigm to refine the system's understanding of the subtle boundary between success and failure.

Xianzhe Fan, Yuxiang Lu, Shen-Yuan Gao et al. · 1 citation · ⚡1

Never Too Late for Force: Accelerating VLA Post-Training with Reactive Force Injection

Across towel folding, book insertion, and Hanoi ring placement, LIFT learns faster and reaches higher performance than vision-only post-training, while ablations show that reactive force memory and online corrective data are both important for robust contact-rich manipulation.

Yi Wang, Wen-Di Chen, Zi-Mo Wen et al. · 4 citations
#machine learning Preprint Open access Sep 2026

A 3D-Printable Dataset for Fair Testing and Comparisons of Tactile Sensors

Existing texture datasets for tactile sensing primarily consist of sensor readings from a specific sensor interacting with available surfaces/objects rather than describing the textures themselves, limiting fair comparison between tactile sensors and hindering reproducible research. In this work, we introduce a 3D-prin...

Dexter R. Shepherd, Nicolas Herzig, Phil Husbands et al. · 0 citations
#artificial intelligence Preprint Open access Sep 2026

A Very Big Video Reasoning Suite

Rapid progress in video models has largely focused on visual quality, leaving their reasoning capabilities underexplored. Video reasoning grounds intelligence in spatiotemporally consistent visual environments that go beyond what text can naturally capture, enabling intuitive reasoning over spatiotemporal structure suc...

Maijunxian Wang, Ruisi Wang, Juyi Lin et al. · 0 citations

From tech blogs

See all →
Microsoft Research Blog Sep 23, 2026

Offloaded inference for real-world physical AI robotics

Robots are getting smarter, but how can their hardware match that growth? New Microsoft Research findings show that moving AI inference beyond the robot can improve task success, boost efficiency, and support more advanced physical AI workloads. The post Offloaded inference for real-world physical AI robotics appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.