mock: disable LLM keep-alive by default, opt in with ?warm=1
The 2-minute warmLLM() probe pinned Qwen3.8-27B-dual on the LM Studio server (192.168.3.7:1234), preventing other models from loading reliably. Default is now off; add ?warm=1 to a tab's URL to restore the warm first-turn behavior for that session.
This commit is contained in:
parent
6457501520
commit
c05e35cc08
|
|
@ -122,7 +122,11 @@ async function warmLLM() {
|
||||||
} catch {}
|
} catch {}
|
||||||
warming = false;
|
warming = false;
|
||||||
}
|
}
|
||||||
setInterval(() => { if (llmOn && !scriptBusy) warmLLM(); }, 120000); // keep the model warm while the tab is open
|
// Keep-alive is DISABLED by default: the probe pins the big model (Qwen3.8-27B-dual)
|
||||||
|
// on the LM Studio server and prevents loading other models. Opt in per session
|
||||||
|
// with ?warm=1 in the URL (accepts the cold first turn cost, see above).
|
||||||
|
const warmEnabled = new URLSearchParams(location.search).has('warm');
|
||||||
|
setInterval(() => { if (warmEnabled && llmOn && !scriptBusy) warmLLM(); }, 120000);
|
||||||
// ---------------- web search client (SearXNG via the /search proxy) ------
|
// ---------------- web search client (SearXNG via the /search proxy) ------
|
||||||
// Live facts (opening hours, fees, seasonal notes) with provenance URLs —
|
// Live facts (opening hours, fees, seasonal notes) with provenance URLs —
|
||||||
// the third leg of the 3-source blend. Enrichment APPENDS to suggestion
|
// the third leg of the 3-source blend. Enrichment APPENDS to suggestion
|
||||||
|
|
|
||||||
Loading…
Reference in New Issue
Block a user