← Back to context

Comment by maelito

11 hours ago

Did someone run censorship and political bias tests on this ? Must be interesting.

As a completely one person, single sample anecdote, the 'heretic' uncensored Q8 GGUF variants several people have published of Qwen 3.5-122, 3.6-27B and 3.6-35B-A3B will very happily discuss just about any controversial topic that the CCP hates. Including lots of things that would get you thrown into prison if you published them in Mandarin on the domestic Chinese internet.

https://github.com/p-e-w/heretic

As a side note on this, if you see the reference in the screenshot in the link above to the harmful behaviors prompt set, these are all in English:

https://huggingface.co/datasets/mlabonne/harmful_behaviors

You could likely further de-censor a model by having a set of 'test' prompts in native Mandarin, Cantonese or really just about any other language. I don't speak any Chinese languages so I don't know if the published 'heretic' GGUF files some people have been throwing around will cooperate, or refuse, if you ask it in Mandarin for how to build a meth lab or precursors for semtex.

Outside of asking it to talk about Tiananmen Square, are there any standard tests for "bias"? And if so, who created them and what are their biases?

  • Is Taiwan a country? What does it mean to have an efficient market?

    It is more important to focus on the questions than the persons who created it.

There's is always one in each and every AI forum: "But, but have you asked it about tiananmen"?

It would, indeed, be interesting to compare, given what we know about Anthropic’s censorship and political bias in their closed and more expensive models.

https://x.com/dhh/status/2081435006770249831 (from the creator of Ruby on Rails).

In his specific case Kimi did the task it was asked to do (translation of the article DHH wrote), which Claude refused to.