Open access
Sep 2026
Prompt Escalation for Lightweight Large Language Models: An Empirical Evaluation of Cost–Performance Trade-Offs
Results support ZS as a low-overhead reference within the four evaluated benchmarks and the stated model, quantization, prompt, and decoding settings, with model–task exceptions.
Seyoung Kim, Bonggyun Ko
· Applied Sciences · 0 citations