huh, sounds like they’re talking about over fitting and datasets etc, it seems like this is almost more like a fine tune/distill than just a pure quantization