Ask, Condition or Abstain: Reinforcement Learning for Missing-Premise Reasoning
ACA-RL supports a new mission for NLP evaluation: measuring whether models can recognize when a task is underdetermined and handle uncertainty, not only whether they can answer fully specified questions.