Skip to content

Author

Runjun Mao

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Aug 2026

Quality-Diversity Reinforcement Learning using Behavior Regulated Policy Gradient

BRPG-MAP-Elites is introduced, a new QD-RL algorithm adopting a strictly Markovian actor-critic architecture within the MAP-Elites framework that demonstrates a 43% improvement in average QD-scores over DCRL-MAP-Elites and achieves higher robustness in the generated policies.

Runjun Mao, Antoine Cully · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.