Squeezing more out of small models with evidence-grounded reasoning

How to squeeze more performance out of small models by making grounding a constraint, not a prompt trick: evidence for claims, checkable outputs, and measuring correctness separately from fluency.



