Skip to content
Conference

Vision-Based Predictive Control for Dual-Arm Nonprehensile Transportation

Aug 2026 · 2026 IEEE International Conference on Mechatronics and Automation (ICMA) · pp. 1788-1793 · 0 citations · 20 references

Abstract

Dual-arm robots often encounter difficulties when handling easily deformable or structurally complex objects using traditional grasping-based manipulation. In addition, grasping and releasing operations introduce significant time overhead. To address these limitations, this paper proposes a vision-based predictive control framework for dual-arm nonprehensile transportation. The proposed method employs a hybrid end effector design that integrates an elastic tether with a tray, enabling flexible and stable transportation without direct grasping. A predictive control strategy is adopted to optimize dual-arm motion trajectories on the move under kinematic and safety constraints. To further enhance coordination accuracy, a direct visual servoing scheme is incorporated to dynamically regulate the arm velocities, minimizing relative motion between the end effectors and the object. This effectively suppresses oscillations induced by the elastic tether. Both simulation and experimental results demonstrate that the proposed approach ensures convergence to desired states and achieves continuous, stable, and safe object transportation, even in the presence of disturbances.

View source

Similar papers

Aug 2026

Grasping force modulation for controlled slip in object pivoting: an RL-based approach for efficient manipulation

Quantitative experiments showed that the proposed method generally outperformed the baselines in mass and CoM variations, particularly in terms of success rate, while maintaining robust performance across the evaluated conditions.

Jinseok Kim, Iksu Choi, Hunjo Lee et al. · 0 citations
Jul 2026

Optimization of sim-to-real transfer in the humanoid robot NICO

This work added YOLO-based object and hand detection, stereo vision-based localization using the robot's built-in low-resolution fisheye cameras, and task-specific corrections for grasp execution to form a novel calibration-based grasping pipeline that does not require RGB-D cameras, motion capture, or external tracking systems.

J. Gavura, Igor Farkas · 0 citations
Preprint Aug 2026

Spatiotemporal Agility: Time-Constrained Reinforcement Learning for Vision-Guided Dynamic Quadrupedal Interception

An integrated framework that combines a vision module for landing point and time prediction with a direct position and time conditioned RL locomotion policy, instead of intermediate velocity commands is proposed, which mitigates perception latency during dynamic interception.

Yi-Dong Zhu, Zibo Dai, Tong-Ning Zhang et al. · 0 citations
Preprint Aug 2026

Contact-Guided Exploration for Non-Prehensile Locomanipulation with Multi-Critic RL

Non-prehensile manipulation offers versatile skills for moving and rearranging heavy or bulky objects, particularly when combined with a mobile manipulation platform. However, both model-based and model-free approaches struggle with the complex hybrid dynamics and the sparsity of the contact in these tasks. To address these challenges, we propose a contact-guided exploration strategy implemented within a Multi-Critic Reinforcement Learning (RL) framework. A dedicated exploration critic is trained with a dense contact-seeking reward that guides the end-effector toward meaningful contact points; its influence is progressively decayed to recover a task-optimal policy. We obtain candidate interaction points from a general-purpose grasping algorithm, enabling the exploration mechanism to generalise across various object geometries. We evaluate the approach on multiple tasks, including box pushing, chair transportation, and a dishwasher opening task. Finally, we validate the chair transportation policy through extensive experiments on a quadrupedal mobile manipulator, demonstrating deployable non-prehensile manipulation in the real world.

Simone Tolomei, Mayank Mittal, F. Angelini et al. · 0 citations
Conference Jul 2026

Stability-Aware Closed-Loop Recovery for Robotic Grasping in Complex Simulation Environments

Language-guided robotic grasping has made significant progress in semantic understanding, but existing methods often rely on open-loop execution strategies and struggle to handle physical disturbances such as object collisions, target displacement, and transportation slippage. To address this problem, this paper proposes a stability-aware dynamic recovery mechanism, named SADR. Based on a multi-threaded decoupled architecture, SADR decouples semantic planning, target tracking, and execution control, and constructs a two-stage stability criterion through pre-closure displacement checking and post-closure force/current feedback verification. When target instability, missed grasping, or slippage is detected, the system performs local trajectory correction and re-grasping based on real-time tracking results, without restarting global semantic planning. Experiments in PyBullet show that SADR significantly improves the grasping success rate under high-density disturbance scenarios while reducing the average task completion time. This study provides an effective closed-loop recovery solution for improving the reliability of robotic grasping tasks in complex simulation environments.

Chun-Cheng Zhang, Lei Sun · 0 citations
Preprint Aug 2026

Predictive Relative-Velocity Steering for Safe Robotic Manipulator Teleoperation in Dynamic Environments

Recent advances in teleoperation have enabled robotic manipulators to perform dexterous, human-arm-like motions. However, human operators may fail to avoid suddenly appearing obstacles promptly and effectively, particularly under network latency or limited attention, thereby creating safety risks. To address this issue, we propose a lightweight and modular framework for proactive collision avoidance, operating directly at the end-effector velocity-command level. After preprocessing the point cloud, the framework first predicts potential collisions based on time-to-collision (TTC) with integrated overshoot protection, and subsequently rotates the relative-velocity vector using Rodrigues'rotation formula. The deflection changes only the direction of the relative velocity while preserving its magnitude, thereby mitigating the deadlock problem commonly encountered by conventional artificial potential field (APF) methods. The prediction module compensates for point-cloud processing latency introduced by complex teleoperation pipelines, while the lightweight design enables the high-frequency control required for teleoperation. Simulations across diverse scenarios show that the proposed method achieves a higher end-effector collision avoidance rate than the baseline methods. Experiments on a physical robotic system further validate its collision-avoidance effectiveness.

Changhao Hu, Zeyi Liu, Song Hu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.