跳到正文
原文
Reddit · r/LocalLLaMA· /u/FantasticNature7590·· 5 天前AI 评分55

本地 Qwen3.8 三版本对比:3 周测试揭示推理优化与能力差异

I spent 3 weeks testing local Qwen3.8 on the new low-latency SGLang/vLLM recipes: DFlash2 2.8x. Builds: RadixArk + Inferact 27B NVFP4, 27B BF16, orcarouter 27B Uncensored, Flash-Next NVFP4

AI 导读

作者花费3周在 RTX PRO 6000 上测试了 Qwen3.8 的三个版本:RadixArk 27B NVFP4、orcarouter 27B Uncensored NVFP4 和 RadixArk Flash-Next NVFP4。

来源:Reddit · r/LocalLLaMA · reddit.com