[Bug]: GLM tool-call streaming final chunks repeat metadata and combine arguments with finish_reason
![[Bug]: GLM tool-call streaming final chunks repeat metadata and combine arguments with finish_reason](https://www.chat-gpts.plus/wp-content/uploads/2026/07/44098-762f8462-768x403.jpg)
用户在 vLLM main 分支(commit 6bdabbad5 / 023808c23 )上,通过 OpenAIServingChat 服务层,使用 GLM 工具解析器(如 glm45 / glm47 )进行工具调用的流式输出时触发。问题不依赖具体 GLM 模型权重加载,可通过 unit tes

![[Question]: build docker images error on Mac M4](https://www.chat-gpts.plus/wp-content/uploads/2026/07/10073-6f796378-768x403.jpg)
![[Bug]: Streaming output segmentation (Qwen3-ASR)](https://www.chat-gpts.plus/wp-content/uploads/2026/07/47421-9196bf15-768x403.jpg)
![[Bug]: presentation parsing bug](https://www.chat-gpts.plus/wp-content/uploads/2026/07/13060-7fd73526-768x403.jpg)
![[Bug]: Tool schema marks **kwargs as a required (untyped) parameter, forcing the LLM to fill it](https://www.chat-gpts.plus/wp-content/uploads/2026/07/22134-dfdd0514-768x403.jpg)


![[Bug]: Vllm + Gemma 4 + claude code: tool calling problems](https://www.chat-gpts.plus/wp-content/uploads/2026/07/39043-5bb1c48d-768x403.jpg)
