Off-Policy Evaluation for Semantic ID Recommenders: Does the Model's Own Code Hierarchy Help?
This work asks a simple question: can the model's own SID tree serve as the action abstraction for that OPE, and explains how resolution depth is the operative knob and a conditional bias bound links the coarsening bias to the quantizer's worst-case reconstruction residual and the target-logging divergence.