跳到正文
原文
Reddit · r/LocalLLaMA· /u/firstcenturyman·· 5 小时前AI 评分57

Hirundo 通过机器反学习移除 Qwen3.6-35B-A3B 中的政治对齐并开源权重

We unlearned CCP alignment from Qwen3.6-35B-A3B: censored/propaganda answers 89.8% → 2.8%, general benchmarks within ~1 point (open weights)

AI 导读

Hirundo 团队发布 Qwen3.6-35B-A3B-Westernized 和 Qwen3.5-4B-Westernized 两个开源权重模型,用机器反学习(behavioral-unlearning + LoRA)替换原模型中的政治对齐。

来源:Reddit · r/LocalLLaMA · reddit.com