MindForge is introduced, an automated pipeline that converts open-source command-line programs into source-free environments that expose only a compiled reference executable and its documentation that consistently improves over the base model across all seven unseen software engineering benchmarks, spanning long-horizon repository generation and translation.
Yihao Chen, Shi Chang, Khaled Chawa et al.· arXiv.org· 0 citations
From an industrial code-generation improvement effort, a maintainer's perspective on why this work is hard in practice is offered, distilling three recurring challenges, zero-sum mixture design, yield as the binding metric, and end-to-end integration under uncertainty, and arguing that progress depends less on one-off recipes than on an engineering discipline for programming dataware.
Gopi Krishnan Rajbahadur, A. M. Ebrahimi, Boyuan Chen et al.· 0 citations
This work proposes LLMSafeGuard, a lightweight real-time framework that integrates an external validator into decoding, rejecting unsafe outputs while allowing valid ones, and introduces a similarity-based validation approach, simplifying safety constraint validation and eliminating the need for external control model training.
Ximing Dong, Shaowei Wang, Dayi Lin et al.· SIGSOFT FSE Companion· 0 citations
Teaching deletion during post-training reduces deletion avoidance and improves broader code-editing performance, suggesting the behavior is undertrained rather than beyond reach.
A. M. Ebrahimi, M. M. Hasan, Aaditya Bhatia et al.· arXiv.org· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.