Retrieval-augmented generation (RAG) improves knowledge-intensive large language model (LLM) applications by conditioning generation on retrieved documents, but longer contexts increase latency, key-value (KV) cache memory, and token cost. Post-retrieval compression can reduce this cost, yet existing compressors often...
T. Nguyen, Qi-Ran Hu, Ban-Ruo Liu et al.· 0 citations
This work introduces PRISM (Peer Review Intelligence via Structured Multi-dimensional assessment), a benchmarking framework that evaluates review quality across four dimensions: Depth of Analysis, Novelty Assessment, Law Identification&Major Issues Prioritization, and Multi-dimensional Constructiveness.
Ngoc Phan Phuoc Loc, Toan Huynh La Viet, Thanh Tran Khanh et al.· arXiv.org· 6 citations· ⚡1
This paper investigates the cause of reasoning shrinkage under SFT-based post-training and identifies data-centric factors as a key driver of shrinkage in reasoning models and highlights diversity-aware designs as an effective lever for controlling it.
N. Nguyen, P. Shojaee, Phuc Minh Nguyen et al.· arXiv.org· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.