The Code/X ArchiveView on X
aipulsedaily

@aipulseda1ly

Gemini 4 Argon has an insanely low hallucination rate on Artificial Analysis. 15%.

Grok 4.7 is at 29%. GPT-6 Astra 45%. Opus 5.5 59%. Fable 5.1 69%.

The only models below it barely answer anything. None of them get more than 15% right.

It gets fewer answers right than Opus 5.5 on max, 50% against 66%. But when it doesnt know, it says so instead of making something up.

Cant wait to get my hands on it.
Image from the post
662603.9K499