7adcfff0fa
enable_thinking=False was a silent no-op for minimax and openai sources. The shape of the switch now stays at provider level while a model-level capability table says whether a given model can honour it at all, and reasoning_tokens is collected so the cost of thinking can be told apart from the cost of answering. Verified against the live gateway: a seventeen-row matrix over 137 real calls, kept out of the CI gate behind the slow marker.