跳到正文
原文
Reddit · r/LocalLLaMA· /u/MD_Reptile·· 5 天前AI 评分21

用户用 3 张 RTX 3060 12GB 搭建 "Flash Next" 推理卡:速度达 38-40 t/s

Flash next rig born from mining parts.

AI 导读

Reddit 用户在 r/LocalLLaMA 分享使用 3 张 RTX 3060 12GB 显卡运行 Flash Next 框架的测试结果,在 IQ3 模型上达到 38-40 t/s 推理速度,同条件 llama.cpp 仅能实现 13.2 t/s。

来源:Reddit · r/LocalLLaMA · reddit.com