Misc. bug: Webui causing high CPU load in Firefox

该问题出现在使用 llama.cpp 的 llama-server 运行推理速度约 100 T/s 的 reasoning 模型时,前端 Firefox 打开 WebUI 并生成较长的 reasoning 输出,导致客户端 Firefox 占用 100% 单核 CPU。优先排查方向是 WebUI 前

该问题出现在使用 llama.cpp 的 llama-server 运行推理速度约 100 T/s 的 reasoning 模型时,前端 Firefox 打开 WebUI 并生成较长的 reasoning 输出,导致客户端 Firefox 占用 100% 单核 CPU。优先排查方向是 WebUI 前
![Dify python code execution error: No usable temporary directory found in ['/tmp', '/var/tmp', '/usr/tmp', '/'] error: exit status 255](https://www.chat-gpts.plus/wp-content/uploads/2026/09/18678-8dcba7c2-768x403.jpg)
该报错通常出现在 Dify 自托管(Docker)环境中运行 Python 代码节点时,Dify Sandbox 找不到可用的临时目录。优先排查 Sandbox 容器内临时目录的权限与配置。
![[Bug]: --otlp-traces-endpoint initializes tracer but never sends spans (instrument_otel/manual_instrument_otel never invoked)](https://www.chat-gpts.plus/wp-content/uploads/2026/09/56696-8c43ad90-768x403.jpg)
在 vLLM 启动时传入 --otlp-traces-endpoint (可选叠加 --collect-detailed-traces all )后,服务能正常推理、Prometheus 指标正常,但 OTLP 端点始终收不到任何 span,日志里也没有报错。通常发生在 vLLM 已 fork 出

这个报错通常发生在使用 langchain-openai 调用 gpt-6-astra 并绑定 function tools 时,SDK 默认仍然走 /v1/chat/completions ,而该端点不支持 gpt-6-astra 带 function tools 的请求;优先排查 LangCha
![[bug]: On windows: After organizing outputs images in sub-folders, clearing intermediates becomes impossible](https://www.chat-gpts.plus/wp-content/uploads/2026/09/9504-fbb443e3-768x403.jpg)
这个报错通常出现在 Windows 上用 InvokeAI 6.14.0rc1、并且执行过“Image Storage Maintenance”把输出图片整理进日期子文件夹之后;此时“Clear Intermediates”会因数据库外键约束拒绝删除图片而失败,优先排查数据库而非 Windows 权
![[Feature]: Configurable deployment and credential database reload intervals](https://www.chat-gpts.plus/wp-content/uploads/2026/09/40972-eeac5495-768x403.jpg)
这是在 LiteLLM Proxy 中希望调整 add_deployment_job (模型与凭据重载)以及定时重载、search tools、MCP 重载任务的调度间隔时遇到的配置问题。上游已支持该配置项,优先确认版本是否达到 v1.95.0,并使用 general_settings.proxy_

当 LiteLLM Proxy 通过团队级别名( team_public_model_name 或 model_aliases )对外暴露模型时, GET /v1/models/{id} 或 /v1/models 可能丢失 mode 、 max_input_tokens 、 max_output_t
![[Bug]: annotation list / hit-history has_more uses raw limit while query caps at 100 (missed sibling of #41780/#41776)](https://www.chat-gpts.plus/wp-content/uploads/2026/09/41875-511ce6cc-768x403.jpg)
当 Dify 自托管实例调用标注列表或命中历史分页接口时,若传入的 limit 恰好等于总数、或大于 100,分页返回值 has_more 会算错,导致前端出现“幽灵加载更多”或提前停止分页。优先排查这三个接口的 has_more 计算是否使用了未经 100 上限裁剪的原始 limit 。

在 Dify v1.17.0 自托管环境中,Agent V2(dify-agent 运行时)在一次运行以“未处理完的工具调用”结束时(停止、取消或超时),同一会话继续发送新消息会报 Agent V2 cannot continue conversation after a run ends with

当 Open WebUI 中的多模态 Agent 通过 Open Terminal 的 read_file 查看图片时,上下文计量器会把持久化消息里的 base64 图片数据当作文本按“字符数 ÷ 4”来估算 token,结果单条 tool 消息被算成约 1M “token”,远超引擎实际上下文,从