Interesting. Chart is on this page:
https://artificialanalysis.ai/evaluations/omniscience
"AA-Omniscience Hallucination Rate (lower is better) measures how often the model answers incorrectly when it should have refused or admitted to not knowing the answer." - Argon is currently the best model by this metric.