Skip to content

Author

Iryna Gurevych

We have 14 of 134 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Feb 2026

From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves

The results show that improving IF in LRMs can significantly enhance privacy, suggesting a promising direction for future privacy-aware LRMs, and introduces an SFT dataset that teaches models to follow general instructions throughout their reasoning process.

Haritz Puerto, Haonan Li, Xudong Han et al. · 0 citations
Preprint Aug 2026

Parameter Exploration for RLVR via Variational Learning

Evidence that parameter-space exploration can improve reinforcement learning for LLMs is presented, and a family of methods called Perturbed Parameter Policy Optimization (3PO) is introduced which use different sampling strategies and different rollout grouping for reward estimation.

Vatsal Venkatkrishna, Nico Daheim, Iryna Gurevych · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.