Intelligent optimization of personalized learning path based on transformer and reinforcement learning.
Findings suggest that closed-loop coordination among state representation, policy optimization, and educational constraints contributes to improvements in personalized learning path recommendation quality.