Metanormative Theory for RL-Based Moral Agents
This paper aims to draw out ideas from recent work in metanormative theory that can be useful for designing artificial moral and value-aligned agents and examine the RL architecture through the lens of these ideas.
Aleks Knoks, Marija Slavkovik
· 0 citations