now you can compose both. a small model handles routing and guardrails, and the llm steps in only when needed.
openinference's new decision span captures the question, choice, confidence, and probabilities, so system one calls show up right next to the reasoning they gate.