#reinforcement learning
Jul 2026
From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning
LA-MAML (Language Adapted MAML), which modifies the inner loop by adapting the global policy parameters in a single step through a learned embedding of the task instruction, replacing the inner loop trajectory collection and gradient-based updates.
Garvit Singla, U. M. Natarajan, Raghuram Bharadwaj Diddigi
· arXiv.org · 0 citations