Skip to content

Author

Armaan A. Abraham

We have 1 of 8 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Reducing Scalar Rewards to Binary Success: General Off-Policy Learning with Success Functions

It is shown that any discounted-reward problem can be recast as a modified, reward-free problem with a single success state and an equivalent optimal policy, and any value-learning problem can be trained by classification with cross-entropy, in a way that is in principle exact and suffers no loss of precision.

Armaan A. Abraham · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.