upvote
Imo omniscience correlates better to how useful the model is in practice than the intelligence index. But you have to use both together of course.
reply
Too late to edit, now including AA-omni Score [0] and 'by Domain' Software Engineering [1]. Also added comparison to the eyeballs-median of the 'top 10' models and then you see that indeed Mistral Large 4 Preview scores miserable in the AA-Omni indices.

                          Inte         Cost     AA-   Omni
                          llig          per    Omni  Softw
  Open Weight model       ence  Speed  Task   score    Eng

  Mimo-V.26-Pro            46     47   $0.13      8     33
  GLM-5.3 (max)            45     73   $2.01     14     37
  DeepSeek 4.1 Flash Max   39    227   $0.27     -5     34
  Mistral Large 4 Preview  38    116   $1.13     -5      5

  Closed/proprietary      ~50   110-  $1.50-    ~43    ~85
     median top 10               242   $7.50
[0] https://artificialanalysis.ai/evaluations/omniscience [1] https://artificialanalysis.ai/evaluations/omniscience?detail...
reply