Skip to content

Author

Minseong Sim

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#large language models Open access Oct 2026

Interoperable Modular Language Models I: From Decomposition to Interchangeability

> Large language models are usually trained, distributed, and consumed as monolithic systems even though their internal computations are structurally modular. This paper asks a stronger question than model splitting, routing, merging, or adapter composition: can internal language-model subsystems become independently p...

Minseong Sim · 0 citations
#reinforcement learning Open access Sep 2026

Diagnosing Estimator Homomorphism Compatibility in Reinforcement Learning with Verifiable Rewards

Reinforcement learning with verifiable rewards (RLVR) evaluates decoded outcomes, whereas policy-gradient estimators operate on concrete trajectories and batch-dependent training state. These two levels need not respect the same equivalence relation. We formalize estimator homomorphism compatibility: when several concr...

Minseong Sim · 0 citations
#reinforcement learning Open access Sep 2026

Diagnosing Estimator Homomorphism Compatibility in Reinforcement Learning with Verifiable Rewards

Reinforcement learning with verifiable rewards (RLVR) evaluates decoded outcomes, whereas policy-gradient estimators operate on concrete trajectories and batch-dependent training state. These two levels need not respect the same equivalence relation. We formalize estimator homomorphism compatibility: when several concr...

Minseong Sim · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.