Skip to content

Author

Alexander Koller

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Oct 2026

On Language Drift during RLVR Post-Training

Recent advances in LLM reasoning models---driven primarily by the paradigm of post-training via reinforcement learning with verifiable reward (RLVR)---have enabled them to accomplish impressively complex tasks. However, in parallel with their rising capabilities, LLMs have increasingly displayed signs of language drift...

Michael Sullivan, Alexander Koller · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.