logoalt Hacker News

rco8786today at 5:16 PM1 replyview on HN

If they’re reaching the same results across a variety of the most popular public models, it doesn’t seem like that big a deal to know if it was Opus 4 or Opus 4.5


Replies

hn_throwaway_99today at 5:52 PM

Reproducibility is (supposed to be) a cornerstone of science. Model versions are absolutely critical to understand what was actually tested and how to reproduce it.

show 1 reply