== tools/call query_model ============================================ arguments {"max_tokens":32,"prompt":"In one sentence, what is a TPU?"} latency 499 ms isError false result: 📡 **Reasoning only — no answer yet.** `finish_reason: length` after 32 tokens, all of them thinking. This is Gemma 4 reasoning, not a broken server. Re-run with a larger `max_tokens` (currently 32). ... No answer: the model was still reasoning when it stopped. Raise --max-tokens. exit=2