Do you have config recommandations ? Apparently you need to set up the following to have proper thinking in models.json : "compat": { "supportsDeveloperRole": false, "supportsReasoningEffort": true "thinkingFormat": "qwen-chat-template", "supportsStrictMode": false, "maxTokensField": "max_tokens" },