Researchers at Hirundo applied machine unlearning to Qwen3.6-35B-A3B, reducing the model’s CCP‑aligned responses from 89.8% to 2.8% while preserving general performance within ~1 point on standard benchmarks. The unlearning specifically targets the model’s internal weights, eliminating refusals such as the “I don't know what you are referring to” answer to queries about June 4, 1989. The resulting Westernized model is released as open weights on HuggingFace.
Read original
reddit/r/LocalLLaMA