Skip to content
Review

Scientific Knowledge Discovery in the Age of Large Language Models

Jul 2026 · arXiv.org · Vol abs/2607.26670 · 0 citations · 46 references
Computer Science

TL;DR

This chapter surveys 34 peer-reviewed papers applying generative LLMs to literature retrieval and the screening of candidate studies against eligibility criteria, identified via a Boolean search over the OpenAIRE Graph.

Abstract

The rapid growth of scholarly literature has made identifying relevant publications increasingly difficult, and conventional search systems still depend heavily on manually formulated queries and effortful manual inspection. Generative large language models (LLMs) offer a more flexible alternative, supporting literature retrieval and the screening of candidate studies against eligibility criteria. This chapter surveys 34 peer-reviewed papers applying generative LLMs to these two tasks, identified via a Boolean search over the OpenAIRE Graph (1,589 records screened to 34 inclusions). Reviewed studies are characterised by LLMs employed, model access and adaptation, prompting and architectural techniques, ground-truth sources, and evaluation metrics.

View source

Similar papers

Review Jul 2026

Pruning large language models: a systematic literature review

This systematic literature review (SLR) provides a comprehensive overview of pruning techniques applied to LLMs, based on 60 peer-reviewed studies and preprints published between 2022 and 2025, sourced from major digital libraries.

F. Bazikar, Atefeh Hemmati, Akram Reza et al. · 0 citations
Review Open access Sep 2026

Operationalising LLM-assisted screening of literature to support systematic reviews

Large language models (LLMs) can ease the work of screening titles and abstracts for systematic reviews, but obtaining reliable results requires researchers to make practical choices about which LLMs to use, how to combine their scores into a ranking, and how far down that ranking to read. We aimed to identify a genera...

S. Spillias, Laura Avila-Turriago, C. Brown et al. · 0 citations
Conference Open access Sep 2026

Graph4LLM: A Systematic Survey of Graph-Enhanced Large Language Models

This survey examines how these methods integrate graphs into various stages of the LLM pipeline, including the input, model, and output phases, and outlines the challenges and future research directions for developing more efficient and interpretable solutions.

Xin-Yan Zhu, Cheng Yang, Qiu-Yue Wang et al. · 0 citations
Preprint Aug 2026

COCI: Conference Organisers and Content Identifier

Despite the critical role of grey literature in scholarly communication, artefacts such as Calls for Papers (CfPs) remain largely isolated from modern Scholarly Knowledge Graphs. The unstructured and highly heterogeneous nature of these documents has traditionally hindered their large-scale processing. In this demo pap...

Angelo Salatino, Francesco Osborne, Alexis Vizcaino et al. · 0 citations
Review Aug 2026

Information modeling of scientific articles for semantic publishing: theory, models and applications

The digital proliferation of scientific articles since the 1980s has made strategic reading an essential skill for researchers. To support automatic filtering, linking and analysis of scientific literature, fine-grained scientific content, extensive semantic links and machine-readable formats are required. This rev...

Mengjuan Weng, Xiaoguang Wang, Ning-Yuan Song et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.