apply_chat_template returns all-zero assistant_masks for multimodal inputs

当你在 Transformers 里对多模态对话(含 image/video/audio)调用 apply_chat_template 并开启 return_assistant_tokens_mask=True 时,如果图像占位 token 被展开,assistant mask 会全为 0,优先排查

当你在 Transformers 里对多模态对话(含 image/video/audio)调用 apply_chat_template 并开启 return_assistant_tokens_mask=True 时,如果图像占位 token 被展开,assistant mask 会全为 0,优先排查

当你在 Ollama 上运行 Gemma 4 系列模型做 OCR、文档解析或读取小字号文本时,如果发现识别结果缺字、串字、准确率明显下降,通常不是模型本身质量问题,而是 Ollama 把图像 token 预算 max_soft_tokens 硬编码为 280 ,导致高分辨率视觉任务细节丢失。优先排查

这个报错通常出现在用 llama-server 跑 MiMo-V2.6-Distill-Qwen-9B(GGUF)并调用 OpenAI 兼容 tools API 时,模型自带的 chat template 被误判为 Qwen3-Coder 模板,导致工具调用永远走不到正常结束。优先排查 llama.
![[Bug][ROCm/gfx942]: DeepSeek-V4-Flash silent retrieval corruption for prompts ≥ ~4-5k tokens (AITER sparse indexer)](https://www.chat-gpts.plus/wp-content/uploads/2026/09/52109-0c477c70-768x403.jpg)
这类静默检索损坏通常出现在 ROCm/gfx942 上运行 DeepSeek-V4-Flash/Pro 并启用 AITER 稀疏索引器(DSA indexer)的长上下文场景中;当压缩后的候选 token 超过 index_topk 、开始真正发生 top-k 选择后,服务器不会报错,但 needl

该问题通常出现在使用 AutoProcessor.apply_chat_template 处理多模态对话(图文混合)且开启 return_assistant_tokens_mask=True 时,返回的 assistant_masks 全为 0,需要优先排查多模态占位符展开导致的字符索引错位。核心英

当你在 Ollama 中通过 top_logprobs 请求超过 20 个候选 token 时,请求会在进入后端前被 Ollama 自己的请求校验拦截,报错信息为 top_logprobs is capped at 20, but nothing downstream requires that 。

在 Langfuse 自托管版开启 Settings → Feature Previews → Improved Message Rendering(normalizedIoPreview)后,Formatted 视图会丢弃 OpenAI 风格 ChatML assistant 消息里 messag

当你在 Langfuse 仪表盘上查看分数分析时,如果使用的是带小数的分数(0–1 置信度、LLM-judge 评分、cost、latency 等),数值分数直方图会把值分到错误的桶里,透视表的 subtotal/grand total 行也会给出偏大或偏小的平均值。这两个错误都不会报错,只会静默返

当你在 LobeChat 中配置了多个 Telegram 机器人并分别绑定到不同 Agent 时,消息工具执行路径调用 sendMessage 会忽略传入的 botId ,转而使用“最近保存的那个已启用 Telegram 机器人”发送消息,导致消息通过错误的机器人发出。优先排查消息工具运行时的凭据解
![[Bug]: rag/res/term.freq is loaded by term_weight but not shipped — every lowercase Latin token gets the same IDF](https://www.chat-gpts.plus/wp-content/uploads/2026/09/18414-d3efad56-768x403.jpg)
当你在 RAGFlow 中启用权重计算(term_weight)并处理英文、葡萄牙语、西班牙语、法语等拉丁字母语言时,启动日志出现 Load term.freq FAIL! ,且检索排序出现“停用词和内容词权重几乎相同”的异常,优先排查 rag/res/term.freq 是否存在、你的 RAGFl