[bug]: “Unknown Qwen3 variant” Error when generating offline using Anima.
![[bug]: "Unknown Qwen3 variant" Error when generating offline using Anima.](https://www.chat-gpts.plus/wp-content/uploads/2026/09/9332-fb6feefe-768x403.jpg)
该报错发生在 InvokeAI 断网状态下使用 Anima Base V1 或微调模型进行文生图时,因为 Qwen3 文本编码器的配置文件(config)未持久化缓存,启动时需要联网从 Hugging Face 下载配置。优先排查 InvokeAI 模型缓存目录中是否存在已下载的 Qwen3 con
![[bug]: "Unknown Qwen3 variant" Error when generating offline using Anima.](https://www.chat-gpts.plus/wp-content/uploads/2026/09/9332-fb6feefe-768x403.jpg)
该报错发生在 InvokeAI 断网状态下使用 Anima Base V1 或微调模型进行文生图时,因为 Qwen3 文本编码器的配置文件(config)未持久化缓存,启动时需要联网从 Hugging Face 下载配置。优先排查 InvokeAI 模型缓存目录中是否存在已下载的 Qwen3 con
![[Feature]: publish a signature for the remote model cost map so installs can verify what they fetched](https://www.chat-gpts.plus/wp-content/uploads/2026/09/40051-5fdbcb08-768x403.jpg)
该 Issue 是一个功能请求(Feature Request),而非缺陷报错。请求方希望 LiteLLM 为远程拉取的 model_prices_and_context_window.json 发布签名,以便安装方验证数据真实性。最终该请求被发起者主动关闭,因为深入调查后发现该请求针对的层级有误—

该报错发生在 LlamaIndex 的 Ollama 集成中,当调用 `chat`/`stream_chat`/`achat`/`astream_chat` 方法并传入 `think=False` 时,该参数被内部逻辑错误地忽略,导致仍然返回思考模型的推理 token。优先排查并修复 `llama-
![[Bug] Service API / console segment update clears attachments when attachment_ids is omitted](https://www.chat-gpts.plus/wp-content/uploads/2026/09/41774-dd68358e-768x403.jpg)
当通过 Service API 或控制台更新知识库分段时,如果请求体只修改了 content 而未携带 attachment_ids 字段,Dify 会误将附件与多模态向量全部清空。优先排查更新接口的入参中是否显式传入了 attachment_ids,或升级到修复版本。

当 Open WebUI 的 Function(函数/过滤器)在加载执行时抛错,系统会自动把该 Function 的 is_active 置为 False (切换为停用),这是官方设计行为而非 Bug。优先排查:修复 Function 代码中的依赖或语法错误后,前往“管理面板 → Functions

这个报错发生在 SwarmUI 在 CachyOS(AMD ROCm 环境)上通过 Linux 脚本安装 ComfyUI 后端依赖时。优先排查 comfy-install-linux.sh 中先安装 AMD PyTorch、随后执行 requirements.txt 安装导致 PyTorch 被覆盖

该报错是 LiteLLM 自适应路由器(adaptive router)的持久化状态中出现了 alpha 或 beta 参数小于等于 0 的“毒化(poisoned)单元格”,导致采样时抛出 `ValueError: gammavariate: alpha and beta must be > 0.

该报错是 llama.cpp 在 Windows 下多 GPU(2× RTX 3090)CUDA 张量并行(tensor split)或层并行(layer split)推理时出现的稳定性问题,表现为 CUDA illegal memory access 或进程无提示崩溃。优先排查多 GPU 异步 C
![[Bug]: Strict tool calling attaches no structural tag when the reasoning and tool parsers share a parser engine](https://www.chat-gpts.plus/wp-content/uploads/2026/09/53745-ee48e740-768x403.jpg)
当 vLLM 的推理解析器(reasoning parser)与工具调用解析器(tool parser)指向同一个解析器引擎时,Strict Tool Calling 模式下可能不会附加结构约束标签(structural tag),导致模型输出不受 JSON Schema 约束。优先排查 --rea

此报错发生在 Transformers 使用 MetalConfig 量化(bits=4/8)的模型上,当以 batch size > 1 进行生成时,只有第一行输出正确,后续所有行解码结果坍缩为 token 0(输出 !!!!... )。优先排查 affine_qmm_t Metal kernel