Tron-1B: Fast, Calibrated Typed Decisions with a Set-Attention Option Head
Applications built around large language models make many small decisions: which queue a ticket belongs to, whether a message is a prompt-injection attempt, how urgent a request is, whether an agent's step succeeded. Routing each of these through a generative LLM is slow, costly and hard to calibrate. We present Tron-1...