Code, prompts, raw model outputs and analysis for the paper Detecting erroneous records in information extraction by small language models. Three local models (Qwen2.5-3B, Gemma 3 4B, Qwen2.5-7B in Ollama) turn 1000 MASSIVE en-US requests into JSON records. Each record gets an error score from rule-based validation, di...
Oleksandr Kholodniak· Zenodo (CERN European Organi...· 0 citations
Code, prompts, raw model outputs and analysis for the paper Detecting erroneous records in information extraction by small language models. Three local models (Qwen2.5-3B, Gemma 3 4B, Qwen2.5-7B in Ollama) turn 1000 MASSIVE en-US requests into JSON records. Each record gets an error score from rule-based validation, di...
Oleksandr Kholodniak· Zenodo (CERN European Organi...· 0 citations
The Intelligent Microbial Application Platform (IMicAP), an AI-driven, open-access, AI-powered microbial platform designed to bridge the gap between fragmented microbial data and actionable biological insights, is developed.
Chao-Yu Zhu, Xing Wang, Li-Hui Feng et al.· Frontiers in Microbiology· 0 citations
The United Nations sustainable development goals emphasize mental health as essential for inclusive and sustainable development, making stress, an important factor affecting well-being and societal participation, a critical area of study. Leveraging social media data, we propose the emotion-aware stress cause analysis...
Soumitra Ghosh, G. Singh, Khanjan Pathak et al.· IEEE Transactions on Computa...· 0 citations
Reach audiences
Advertise in front of researchers, engineers, and readers.
This paper argues for SQEs as a complementary evaluation strategy for AI systems operating in dynamic, contested, and interdisciplinary settings and introduces small qualitative evaluations (SQEs) as a human-in-the-loop framework for assessing LLM performance in such less-bounded domains.
Michael Simeone, Jacki Hyatt, E. Bienenstock et al.· Neural computing & applicati...· 0 citations
This work presents Omni-Embed-Mini, a 0.9B-parameter model that maps text, speech, audio, images, video, and visually-rich documents into a single shared cosine space without updating any text-side parameter, and is competitive with the closed gemini-embedding-2, edging ahead of it on the overall-modality average.
Mohammed Irfan Kurpath, Jaseel Muhammad Kaithakkodan, Sahal Shaji Mullappilly et al.· 0 citations
This work develops OSCAR (Optimization modeling by Simulator, Coder, and Reviewer), which uses an offline Simulator certified against labeled decision examples to compare candidates and continues searching beyond feasibility, and model the search for the next certified improvement as sequential decisions under unobserv...
Jin-Zhi Bu, Hai-Xin Tang, Hua-Nan Zhang· 0 citations
Results show that attributes comparable to those identified by humans can be discovered automatically, enabling the creation of high-quality structured datasets economically and at scale.
Relational Agenda Programming (RAP) is presented, a uniform execution model in which reasoning, control flow, and I/O all flow through the same relational machinery with no privileged escape hatches.
Andrew Chen· Proceedings of the 2026 ACM...· 0 citations
Large language models (LLMs) have shown strong performance in static code tasks like code search, summarization, and generation, but remain limited in dynamic code reasoning, which involves inferring how programs behave during execution without actually running them. This limitation stems from LLMs being trained on sta...
Yan Wang, Ling Ding, Jie-Chen Sun et al.· Proceedings of the ACM on Pr...· 0 citations
Abstract The UK and Australian Modern Slavery Acts require large corporations to disclose annually how they address modern slavery risks in their operations and supply chains. Existing methods assess only a small proportion of these disclosures, or narrowly against explicit legal criteria, overlooking deeper indicators...
A. Bora, Duo-Yi Zhang, H. Thinyane et al.· Data & Policy· 0 citations