Any benchmark of political censorship would, invariably, just measure (assuming the benchmark itself was constructed perfectly, though realistically it would only be an approximation) against the benchmark creators preferred bias.
Selecting topics that are frequently commonly censored in Chinese media is a reasonable area to focus on, because this is a model produced by a Chinese company. People are interested in whether the typical patterns of Chinese censorship are being applied to open source LLMs.