Eval bug: Running llama-server only possible with single AMD GPU, running multiple always causes Segmentation fault regardless of model size

这个报错通常出现在 Ubuntu + ROCm 环境下、机器装有 2 张及以上 AMD GPU 时,llama-server 加载任意模型都会直接崩溃。优先排查多 GPU 环境与当前 llama.cpp 构建是否匹配(重新构建后是否仍能识别全部设备、是否有 I2C/EEPROM 读取失败)。








