[Bug] /v1/internal/model/load silently ignores `args` (loader flags like ctx-size, cache-type) — endpoint returns OK but model loads with UI
![[Bug] /v1/internal/model/load silently ignores `args` (loader flags like ctx-size, cache-type) — endpoint returns OK but model loads with UI](https://www.chat-gpts.plus/wp-content/uploads/2026/06/7577-f376545a-768x403.jpg)
用户在使用 TextGen WebUI 的 OpenAI API 扩展时,通过 POST /v1/internal/model/load 接口传递 args 字典(例如设置 ctx-size 为 32768),希望覆盖 UI 默认的 loader 参数。接口返回 HTTP 200 和 "OK" ,但
![[SYCL][DOCKER] Increased Gemma 4 26b performance with updated docker dependencies](https://www.chat-gpts.plus/wp-content/uploads/2026/06/24045-65cf0f68-768x403.jpg)




![[Question]: Failed to build knowledge graph](https://www.chat-gpts.plus/wp-content/uploads/2026/06/6886-32ff5dfd-768x403.jpg)


![[Bug]: `--reasoning-parser gemma4` silently disables structured output (xgrammar) when `enable_thinking=false`](https://www.chat-gpts.plus/wp-content/uploads/2026/06/39130-0cfbe746-768x403.jpg)