跳到正文
原文
Reddit · r/LocalLLaMA· /u/Impossible_Art9151·· 2 小时前AI 评分8

求助:在 2 x DGX Spark 集群上用 llama-server 启动 Unsloth Qwen3.8-Flash-Next MTP GGUF 时遇到 stall

struggling with llama.cpp 2 x dgx spark mtp files start command (unsloth)

AI 导读

用户在 2 张 DGX Spark 组成的集群上部署 `deepseek-flash` 成功后,尝试用 llama.cpp 的 `./llama-server` 启动 Unsloth 发布的 `Qwen3.8-Flash-Next-GGUF:Q8_0` 多 token 预测(MTP)版本时遭遇 DGX stall。

来源:Reddit · r/LocalLLaMA · reddit.com