logoalt Hacker News

janalsncmyesterday at 7:33 PM0 repliesview on HN

Right, if a model says it is Qwen there is no way to distinguish a ModernBert fine tuned with Qwen completion data from a Qwen model fine tuned with completion data.

It’s also entirely possible that they used completions from a pool of open weight models.