Skip to content

Category

robotics

1,156 papers

DroneShield-AI: A Multi-Modal Sensor Fusion Framework for Real-Time Autonomous Drone Threat Detection, Behavioral Intent Classification, and Swarm Intelligence in Contested Airspace

This v2 revision reports measured results on the completed implementation of DroneShield-AI, a unified open framework integrating six processing layers: RF signal classification, acoustic motor-signature detection, YOLOv8-based visual detection, evidence-weighted sensor fusion, a Behavioral Intent Classification Engine...

Marius Bayizere · 0 citations
#machine learning Preprint Open access Sep 2026

ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation

Long-horizon robotic manipulation requires dense feedback that reflects how a task advances through its procedural stages, not merely whether the final outcome is successful. Existing reward models often rely on trajectory-level success labels or time-based interpolation, which can conflate elapsed time with true task...

Youhe Feng, Hansen Shi, Haoyang Li et al. · 0 citations

Universal Pose Pretraining for Generalizable Vision-Language-Action Policies

Pose-VLA is proposed, a decoupled paradigm that separates VLA training into a pre-training phase for extracting universal 3D spatial priors in a unified camera-centric space, and a post-training phase for efficient embodiment alignment within robot-specific action space.

Haitao Lin, Hanyang Yu, Jingshun Huang et al. · 11 citations
#machine learning Review Aug 2025

AI-driven Dispensing of Coral Reseeding Devices for Broad-scale Restoration of the Great Barrier Reef

Coral reefs are on the brink of collapse, with climate change, ocean acidification, and pollution leading to a projected 70-90% loss of coral species within the next decade. Reef restoration is crucial, but its success hinges on introducing automation to upscale efforts. In this work, we present a highly configurable A...

Scarlett Raine, Emilio Olivastri, Benjamin Moshirian et al. · 3 citations
#machine learning Preprint Open access Sep 2026

A Survey on Reinforcement Learning Applications in SLAM

Simultaneous localization and mapping (SLAM) allows a mobile robot or autonomous vehicle to build a map of an unknown environment while estimating its own pose within that map. Reinforcement learning (RL), in which an agent learns a decision policy from interaction and reward, has been applied to decide how such system...

Mohammad Dehghani Tezerjani, Mohammad Khoshnazar, Mohammadhamed Tangestanizadeh et al. · 0 citations
#machine learning Preprint Open access Sep 2026

Fast LeWorldModel

Joint-Embedding Predictive Architectures (JEPAs), including recent LeWorldModel (LeWM), have become a promising foundation for reconstruction-free visual world models. For visual planning, however, LeWM evaluates candidate action sequences by repeatedly applying a local one-step latent transition model. This autoregres...

Yuntian Gao, Xiangyu Xu · 0 citations
#machine learning Preprint Open access Sep 2026

ACSAC: Adaptive Chunk Size Actor-Critic with Causal Transformer Q-Network

Long-horizon, sparse-reward tasks pose a fundamental challenge for reinforcement learning, since single-step TD learning suffers from bootstrapping error accumulation across successive Bellman updates. Actor-critic methods with action chunking address this by operating over temporally extended actions, which reduce the...

Qian Chen, Junqiao Zhao, Hongtu Zhou et al. · 0 citations
#machine learning Preprint Sep 2026

Statistical Learning of Contractive Dynamical Representations for Composite Adaptive Control

A statistically principled hard expectation-maximization procedure is introduced, with a Kalman smoother in the hard E-step, to identify dynamical representations of disturbance whose latent evolution is uniformly contractive, which yields a composite adaptive tracking controller with predictive capability and provable...

Min Kim, José Leonardo Brenes, Fred Y. Hadaegh et al. · 0 citations
#machine learning Preprint Open access Sep 2026

Natural State-Prediction Accuracy can Hide Weak Controlled Responsiveness in VLA Readouts

Accurately decoding object states from the internal representations of vision-language-action (VLA) models does not establish that the predictions respond faithfully to changes in the target physical state. In natural observations, object state, robot configuration, occlusion, and task progress vary together, allowing...

Hyungjoon Kim, Wonbin Son, Mi Young Lee et al. · 0 citations
#machine learning Preprint Sep 2026

On the Numerical Reliability of Differentiable Physics-Based Optimization for Robotic Material Manipulation

Differentiable physics is increasingly used in robotic material manipulation for system identification, trajectory or skill optimization, demonstration generation, and robot or end-effector design. These applications depend on gradients propagated through long, contact-rich simulation rollouts. We study the numerical r...

Xin-Tong Yang, Ming-Lun Wei, Yu-Kun Lai et al. · 0 citations

From tech blogs

See all →
Microsoft Research Blog Sep 23, 2026

Offloaded inference for real-world physical AI robotics

Robots are getting smarter, but how can their hardware match that growth? New Microsoft Research findings show that moving AI inference beyond the robot can improve task success, boost efficiency, and support more advanced physical AI workloads. The post Offloaded inference for real-world physical AI robotics appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.