ec2caea840
llama.cpp and Ollama model discovery probed /models and /props with a 250ms timeout tuned for a loopback server. That cap also applied to a host reached over the network, so a remote or LAN LLAMA_CPP_BASE_URL (or OLLAMA_BASE_URL/OLLAMA_HOST) with normal round-trip latency timed out, discovery returned no models, and the picker fell back to stale 127.0.0.1:8080 entries. Select the probe timeout by host: strictly-loopback base URLs keep the fast fail so a busy or foreign service on the default port never stalls startup; every non-loopback host gets a generous discovery budget. Fixes #7087