A strong, fixed, rule-based expert is built for Gin Rummy and used only as a yardstick, never for training, and the result is a lightweight, game-agnostic recipe that trains competitive agents without training on the expert, for any game a small model can handle, reported with robust statistics and released as a reusable package.
Nima Kelidari, M. Haghi, Mahdi Salmani· arXiv.org· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.