Reddit · r/LocalLLaMA· /u/ricyoung·· 2 天前AI 评分23
基于 Qwen3.5-9B 微调的"Bev"决策模型:故意选错答案的对照测试用例
I trained a model to be wrong 98% of the time and 96% sure about it. It took three tries.
AI 导读
作者基于 Qwen3.5-9B 微调出决策模型 Bev,借助 Bespoke Nimble adapter 微调 51 分钟后,使她在 324 条 held-out 决策上答对率仅 1.9%、平均置信度 96%,置信度 ≥90% 时答对率仅 1.4%。
来源:Reddit · r/LocalLLaMA · reddit.com