Back to #diffusion models
#diffusion models Open access

CHAI – Compliant Human-centered Adaptive Interaction through Diffusion-based Language Trajectory Transformer

Sep 2026 · IEEE Robotics and Automation Letters · Vol 11, pp. 10266-10273 · 0 citations · 38 references

TL;DR

CHAI (Compliant Human-centered Adaptive Interaction), a novel language-driven framework for real-time modulation of a robot’s kinematics and mechanical compliance in real-world environments, introduces on-the-fly language-driven impedance (compliance) modulation along both translational and rotational directions.

Abstract

We present CHAI (Compliant Human-centered Adaptive Interaction), a novel language-driven framework for real-time modulation of a robot’s kinematics and mechanical compliance in real-world environments. By interpreting natural language instructions and visual context through pre-trained vision-language encoders, CHAI combines a transformer-based geometry encoder with a conditional diffusion model to iteratively refine a nominal kinematic trajectory and its associated compliance profile. CHAI introduces on-the-fly language-driven impedance (compliance) modulation along both translational and rotational directions—including motion-aligned and radial axes—executed through a passivity-aware and therefore stable Cartesian impedance controller. This capability is key for supporting compliant interaction, improving adaptability and reducing hazardous contact in physical interaction tasks. Comprehensive experiments demonstrate significant gains in scalability, adaptability, trajectory accuracy, and interactive behaviour over prior methods. Ablation studies further validate the contributions of stiffness control and multi-modal conditioning.

Read PDF

Similar papers

Preprint Aug 2026

Accelerating Human-Aware Robot Trajectory Generation via Diffusion and Consistency Distillation

This research proposes a constrained motion planning framework for robot manipulators in human-robot interaction (HRI). For a non-redundant manipulator with a fully specified end-effector pose, additional requirements such as collision avoidance and self-collision avoidance are difficult to handle as simple null-space secondary tasks. This limitation makes it challenging to generate feasible joint-space trajectories in HRI environments where safety and kinematic constraints must be considered simultaneously. To address this limitation, collision- and self-collision-aware trajectories are generated using Rapidly-exploring Random Tree (RRT) and RRT* algorithms, and the resulting dataset is used to train a diffusion model that generates constraint-satisfying trajectories through guided sampling. To reduce the inference time required for iterative diffusion sampling, consistency distillation is applied, and a joint-weighted jerk regularization term is incorporated into the loss function to promote smoother trajectories by penalizing abrupt changes in joint acceleration. Simulation results show that the consistency model generates 150 trajectory candidates in less than 100 ms, maintains a high episode success rate, and substantially reduces joint and end-effector jerk when jerk regularization is applied.

Byeong-Il Ham, Hyunbin Kim, Kyung-Soo Kim · 0 citations
Jun 2026

Kinematic optimization and adaptive compliant synchronization control of a neck-shoulder massage robot

Traditional neck-shoulder rehabilitation devices exhibit limited interaction fidelity due to rigid mechanical structures and imprecise multi-axis synchronization when interfacing with nonlinear human soft tissue. This study aims to develop a dual-motor neck-shoulder massage robot featuring kinematic optimization and a hierarchical adaptive control architecture to achieve compliant force tracking and high-precision motion coordination. A crank-rocker mechanism, synthesized via kinematic optimization, generates biomimetic trajectories, complemented by a scissor-lift module for active depth adjustment. Within the control framework, an extended Kalman filter fuses multi-modal sensor data for real-time contact state estimation. An incremental Fuzzy PID controller accommodates the nonlinear stiffness of muscle tissue to ensure active compliance. Concurrently, a Robust Adaptive Cross-Coupling Synchronization (RACCS) algorithm regulates dual-motor coordination under variable loads. Validation with 45 subjects demonstrates a 30% tracking error reduction compared to open-loop baselines on the same hardware. Compliant force control precision is maintained within a 0.2 N margin, yielding over 90% target area coverage across the cervical and shoulder regions. Stability analysis confirms the robustness of the RACCS algorithm against heterogeneous load disturbances. This study contributes a kinematically optimized crank rocker and scissor lift mechanism together with a hierarchical adaptive control architecture for distributed dual motor systems. The integrated design manages nonlinear soft tissue impedance and offers a scalable platform for precise cervical fatigue relief in healthy subjects.

Xiujuan Sun, Weijie Zhang, Handi Bian et al. · 0 citations
2026

Neural Adaptive Admittance Control With Guaranteed Performance for Physical Human–Robot Interaction

Physical human-robot interaction (pHRI) offers considerable potential for improving task efficiency and alleviating operator workload. Nevertheless, the intrinsic variability of human motion intention (HMI) and robot model uncertainties pose substantial challenges to achieving accurate coordinated control. To address these issues, this paper proposes a guaranteed-performance neural adaptive admittance control framework. First, the damping coefficient is dynamically tuned using real-time interaction force feedback, while a neural network (NN) is employed to estimate HMI-induced uncertainties in the coupled human-robot system. These two components are then integrated into the admittance model to construct a high-level interaction strategy that generates compliant reference trajectories for smooth and stable collaboration. Subsequently, low-level motion control with error transformation is developed to enforce prescribed output constraints, thereby ensuring unified regulation of transient and steady-state performance. Moreover, another NN is introduced to approximate the lumped robot dynamics for improved tracking accuracy. Finally, the effectiveness and superiority of the proposed method are validated through trajectory tracking, circle drawing, and obstacle avoidance tasks. Note to Practitioners—This paper focuses on developing an active interaction control approach that enables high-performance tracking for robots subject to model uncertainties while providing high-quality assistance to operators with unknown motion intention. The proposed framework is well-suited to industrial applications such as human-robot cooperative assembly and co-transportation. By incorporating output-constraint-based neural adaptive admittance control, safe, reliable, and compliant physical interaction can be achieved. Consequently, the controller supports further extension to medical rehabilitation and exoskeleton systems, demonstrating broad promise across a wide range of interaction-intensive scenarios.

Chengguo Liu, Hefu Ye, Kai Zhao · 0 citations
Preprint Jul 2026

User-Driven Learning from Demonstration: A Trajectory and Impedance Learning Method

This paper presents a method for user-driven robot Learning from Demonstration (LfD) that reduces user effort while ensuring compliant and precise reproduction. The method eliminates repeated teaching for the same task and enables real-time learning from a single demonstration. Demonstrated motions are reproduced with high precision, while impedance variations are learned in real time to provide both compliance and robustness against perturbations. This mitigates potential safety issues in Human-Robot Interaction (HRI) that arise from conventional time-indexed trajectories lacking compliance. The proposed approach integrates a three-dimensional (3D) Fast Diffeomorphic Matching (FDM) algorithm with a Dynamical System (DS)-based motion generator to achieve real-time single-shot demonstration learning and reproduction. An Extended Kalman Filter (EKF) framework compensates for reproduction errors and recovers from external interactions. Furthermore, an impedance parameterization function is incorporated to learn impedance variations from demonstrations and maintain surface contact for specific applications. The proposed approach is validated through comprehensive experiments on a 7 Degree-of-Freedom (DOF) KUKA LWR IV+ robot.

Ziren Yang, M. Kermani · 0 citations

AI-Enabled Force and Torque Control for Human Robot Interaction

Safe and intuitive human robot interaction (HRI) requires precise regulation of contact forces and torques while adapting to dynamic and uncertain human behavior. Traditional impedance and admittance control strategies rely on fixed parameters and accurate system modeling, which often limit their performance in unstructured or collaborative environments. This paper presents an AI-enabled force and torque control framework that integrates machine learning techniques with conventional control methods to enhance adaptability, compliance, and safety in physical human robot interaction. The proposed approach employs deep neural networks and reinforcement learning to learn human intent and interaction dynamics directly from multi-modal sensor data, including force torque sensors, joint encoders, and inertial measurements. By continuously adjusting control gains in real time, the system achieves stable interaction while minimizing excessive contact forces and undesired torques. Experimental evaluations conducted on a collaborative robotic platform demonstrate significant improvements over classical control schemes, including reduced interaction force peaks, smoother torque profiles, and improved task execution efficiency during cooperative manipulation tasks. The results indicate that AI-driven force and torque control can substantially improve robustness, adaptability, and user comfort in human robot collaboration, making it a promising solution for applications in rehabilitation robotics, assistive devices, and industrial cobots.

Vishal Khanna · 0 citations
Preprint Jul 2026

Semantic Audio-driven Understanding for Dynamic Humanoid Whole Body Control

Recent advances in humanoid robotics and reinforcement learning have enabled the acquisition of highly expressive whole-body motion policies. However, most robotic performances remain based on pre-scripted sequences or externally triggered behaviors, limiting autonomy and responsiveness to dynamic environments. In this work, we introduce a novel multi-modal orchestration framework for semantic audio-driven humanoid control, enabling robots to autonomously select and execute appropriate motion skills in real time. The system processes continuous audio streams and routes them into music or speech branches. Music input is handled via audio fingerprinting and semantic embeddings to retrieve track identity and temporal alignment, allowing dynamic mapping between musical segments and motion policies. Speech input is grounded into a discrete library of imitation-learned skills, enabling direct human-robot interaction. Both modalities share a unified interface that schedules skill execution over a reinforcement learning control pipeline. We validate the approach in simulation and on a Unitree G1 humanoid, showing robust sim-to-real transfer and consistent audio-conditioned policy selection. Supplementary materials are available at the following site: https://lab-rococo-sapienza.github.io/semantic-WBC/

J. Marcelo, M. Brienza, E. Bugli et al. · 0 citations

Related blog posts