Skip to content

Author

Bodhisatta Maiti

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#large language models Open access Sep 2026

Probing Instruction Execution Stability of Large Language Models under Semantic Paraphrase

Large Language Models (LLMs) are commonly controlled through natural-language instructions, yet equivalentprompts can yield inconsistent behavior. While prior work has largely evaluated prompt effectiveness throughtask accuracy, less attention has been paid to the stability of instruction execution under paraphrase. In thiswork, we analyze how instruction-preserving paraphrases—prompts that retain identical task semantics andconstraints—affect the reliability with which LLMs execute structured instructions. We conduct a controlled studyusing a classification task with a strict JSON output contract, evaluating multiple paraphrases across repeated runsto separate paraphrase-induced effects from stochastic variation. Our analysis distinguishes between semanticagreement (label consistency) and execution agreement (format and constraint adherence). Experiments on eightopen-source instruction-tuned models show that semantic decisions generally remain stable under paraphrasing,while structured output reliability varies substantially depending on both the paraphrase and the model. Thesefindings indicate that execution robustness is a distinct, model-dependent property not captured by accuracyalone and highlight the importance of evaluating prompt behavior under paraphrase in structured-output settings.

Bodhisatta Maiti, Debshree Chowdhury · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.