Skip to content

Author

Roger Girgis

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Constrained Group Relative Policy Optimization

This work introduces Constrained GRPO, a Lagrangian-based extension of GRPO for constrained policy optimization, and addresses the coupling induced by reward scalarization by scalarizing standardized advantages rather than rewards.

Roger Girgis, Rodrigue de Schaetzen, Luke Rowe et al. · 2 citations · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.