Guideline-grounded retrieval versus unconfigured generation in obstructive sleep apnea: a pilot comparison of notebookLM and ChatGPT for AASM v3.0 PSG scoring-rule questions
Large language models (LLMs) are increasingly used in sleep medicine, but their reliability for polysomnography (PSG) scoring-rule tasks that require up-to-date guideline knowledge remains uncertain. This pilot study compared response quality when identical AASM v3.0 PSG scoring-rule questions were answered by Notebook...