Large language models (LLMs) have shown strong performance in static code tasks like code search, summarization, and generation, but remain limited in dynamic code reasoning, which involves inferring how programs behave during execution without actually running them. This limitation stems from LLMs being trained on sta...
Yan Wang, Ling Ding, Jie-Chen Sun et al.· Proceedings of the ACM on Pr...· 0 citations
Abstract The UK and Australian Modern Slavery Acts require large corporations to disclose annually how they address modern slavery risks in their operations and supply chains. Existing methods assess only a small proportion of these disclosures, or narrowly against explicit legal criteria, overlooking deeper indicators...
A. Bora, Duo-Yi Zhang, H. Thinyane et al.· Data & Policy· 0 citations
A warning-guided, slice-based, LLM-assisted pipeline for inferring Java nullability annotations, which infers the annotations @Nullable and @Nonnull without touching program logic as a step toward a type-system-independent inference technique.
Mushfiqur Rahman Chowdhury· Companion Proceedings of the...· 0 citations
Program reduction helps compiler and language-tool developers turn large failure-inducing programs into small, shareable bug reports. LPR (Large Language Models-Aided Program Reduction) demonstrated that large language models (LLMs) can complement syntax-guided reducers by proposing language-specific transformations on...
Ze-Hua Zhang, Jia-Tong Liu, Xue-Song Yao et al.· Companion Proceedings of the...· 0 citations
Reach audiences
Advertise in front of researchers, engineers, and readers.
Conformance testing asks whether an implementation agrees with its specification. When the specification is expressed in prose, one established approach is to mechanize it as an executable specification. This executable then serves as the oracle, and an input on which an implementation disagrees with it is a potential...
Mehrad Haghshenas, Meng Xu· Proceedings of the 1st Inter...· 0 citations
Analyzing multi-party physical collaboration means tracking what a group builds and whether their actions move the shared structure toward a goal. Vision-Language Models (VLMs) could automate this analysis from session recordings, but existing spatial-reasoning benchmarks use rendered scenes or curated snapshots, not r...
Changsoo Jung, Sheikh Mannan, Jack Fitzgerald et al.· Proceedings of the 28th Inte...· 0 citations
This work benchmarks both reconstruction quality, codebook collapse and representational capacity of VQ-VAEs across a variety of settings, surpassing state of the art in the reconstruction task and providing a stepping stone for further development of medical multimodal auto-regressive techniques.
Emílio Dolgener Cantú, Manasi Acharya, Jim Berend et al.· Companion Publication of the...· 1 citation
Dementia is commonly described as impairing what people attend to, say, and mean more than how they move their eyes or produce speech. We test this asymmetry with a cross-modal behavioral marker—negative log-likelihood (NLL) under frozen pretrained models across gaze, text, and audio—that separates semantic engagement...
Leticia Pinto-Alva, Gale M. Lucas, Maja J. Matarić et al.· Proceedings of the 28th Inte...· 0 citations
The prevalence of anxiety disorders has been increasing in recent years to an extent where the demand for treatment is difficult to meet with traditional therapy. Therefore, research efforts aim for scalable, cost-effective extensions thereof. One approach is the use of Virtual Reality-based exposure therapy controlled...
Paula Friedrich, Lukas Polifke, David Obremski et al.· Companion Publication of the...· 0 citations
product-match-qwen3-0.6b-mlx A small pairwise checker: two product listings go in, and the answer is the word yes or no. It was made by fine-tuning Qwen3-0.6B (4-bit, MLX) with LoRA, and the adapter is fused into the weights in this record. Who asked: the request thread at https://discuss.huggingface.co/t/178475 Run it...
mycelium· Zenodo (CERN European Organi...· 0 citations
The development of large language models for the chemical domain relies heavily on high-quality structured data. However, key experimental information in chemical literature is often scattered across PDFs in multimodal forms, such as reaction schemes, experimental tables, figure captions, and footnotes. This makes stru...
Xin Li, H. Liang, Xu Wang et al.· Journal on Image and Video P...· 0 citations
Automatic speech recognition (ASR) often produces incorrect 1-best transcriptions for dysarthric speech, but a recognition error does not necessarily imply that all information supporting the correct utterance has been lost. This study investigates the extent to which correct information remains within an ASR model aft...
Hidenori Sano· Zenodo (CERN European Organi...· 0 citations