Wire
Cline gives local models five minutes to load
Cline 4.1.3 raised its default Ollama response-start timeout from 30 seconds to five minutes, a 10x larger window for local models to cold-load before an agent task fails. The coordinated SDK release also retries responses containing no text, reasoning, or tool call, while unreachable servers still fail immediately and explicit timeout settings still win. For teams benchmarking the whole coding-agent harness rather than model weights alone, cold-start tolerance belongs in the reliability score: local inference cannot save API spend if the orchestrator mistakes loading for failure.