Reddit · r/LocalLLaMA· /u/GodComplecs·· 6 天前AI 评分28
开源大语言模型指令模式性能引争议:新模型下降更严重,3.6 多个 coding benchmarks 领先 3.8
Should we plead opensource labs to still produce great non thinking (instruct) models?
AI 导读
近期测试表明,新一代开源大语言模型在非思考/指令模式下的性能下降幅度大于旧模型,例如 3.6 版本在多个 coding benchmarks 的指令模式中超越了 3.8 版本。许多用户仍需无需长推理的指令模型,呼吁开源实验室继续优化这一方向。
来源:Reddit · r/LocalLLaMA · reddit.com