Those questions are used as a canary for government manipulation because it's a known topic.
Assuming that the manipulation and censorship only covers a few obvious historical topics and leaves everything else untouched would be very naive.
I don’t see how that responds to the point in the parent comment. Censorship or not, chatbots are unreliable for serious history questions.
Just today Gemma told me that for a long time the Iliad and Odyssey were considered mediocre literature. I was skeptical so I cross referenced, but a lot of more subtle errors could get by.
The fact that it’s a canary makes it a prime tool for A/B testing of generalized approaches to censor or “secure” a model.
But it’s the most useless canary ever. We already know that certain topics are taboo in China.
As for all the other uses the models have, it seems pretty clear they’re not doing anything weird. If they were, people would be posting examples of that and not of Tiananmen Square.
What does it matter if government manipulates data nobody should be using these systems to get though? The ideological purity of the model has no bearing on whether or not it will try to inject some backdoors into your code or steal sensitive information, and using ideological purity tests as an analogue for compromise is not likely to be effective. So why do people care about the ideological purity of AI models when these are supposed to be used for making code?