Trajectory-Derived Confidence for Reliable, Resource-Aware Clinical Text-to-SQL Agents
LLM agents for clinical text-to-SQL applications reason autonomously over multiple steps but cannot assess whether their own reasoning or outputs can be trusted. In high leverage applications such as healthcare, this presents a critical risk where system mistakes can be costly. These reliability failures are also resou...