Issue evidence · Serve · ollama/ollama
GitHub issue: https://github.com/ollama/ollama/issues/18916
Reproduced, workaround proven
On a roughly 19.5k-token prompt, qwen3.5:4b finishes with done_reason "stop", HTTP 200, and nothing in message.content, only reasoning in message.thinking.
Served qwen3.5:4b with Ollama on an RTX 4090 through Badgr (64 s to ready, $0.0074 spend) and replayed the issue's request 3 times per variant: the original request, the same request with think=false, and sanitized and reduced copies.
Badgr's outcome check treats a 200 with an empty message.content as a failed reply, so a pipeline using it retries or fails loudly instead of passing an empty string on. The workaround is think=false.
Checked 2026-10-11 in Badgr’s local development environment.
← All issue evidence