This seems like mildly interesting distillation work wrapped up in a nonsense attempt to drag censorship into the discussion.
There's no way your <200 examples for SFT would ever change how the model thinks of Holodomor unless you'd very intentionally crafted examples to do so.
It feels like you're expecting rubes to draw conclusions that are irrelevant to the actual work you did.
The examples you're talking about are not involved in the training process, so their number is irrelevant. As stated in the post, the goal of this work is to determine whether a teacher's unrelated behaviors are inherited by the student distilled on a different task. Changing how the model thinks about the Holodomor is completely irrelevant.