Skip to content

Author

Mehmet Kerem Türkcan

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Oct 2026

Dynamic Minimax Regret Optimization for Robust LLM Post-Training

Modern LLM training increasingly relies on heterogeneous data sources spanning different domains, tasks, preference distributions, and difficulty levels. We study dynamic minimax regret for group-distributionally robust LLM post-training under instantaneous mini-batch-only bandit feedback. The framework views the train...

Cheng-Bo Zang, Hao-Yu Dong, Mehmet Kerem Türkcan et al. · 0 citations
#machine learning Preprint Sep 2026

Reinforcement Learning of Communication in a Mesh of Small Language Models

TalkMesh, a decentralized mesh of small language model agents that learns when and what to communicate is presented, a decentralized mesh of small language model agents that reaches the accuracy of majority voting over 32 samples with each of three models.

Mehmet Kerem Türkcan · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.