Ollama API supports "think": false to disable extended thinking mode on qwen3.5. Reduces response time from 95s to ~3s per comparison. Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com>
Ollama API supports "think": false to disable extended thinking mode on qwen3.5. Reduces response time from 95s to ~3s per comparison. Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com>