← Back to context

Comment by smartbit

9 hours ago

https://artificialanalysis.ai/models/mistral-large-4 for the main stats

                          Inte         Cost
                          llig          per
  Open Weight model       ence  Speed  Task

  Mimo-V.26-Pro            46     47   $0.13
  GLM-5.3 (max)            45     73   $2.01
  DeepSeek 4.1 Flash Max   39    227   $0.27
  Mistral Large 4 Preview  38    116   $1.13

Imo omniscience correlates better to how useful the model is in practice than the intelligence index. But you have to use both together of course.

  • Too late to edit, now including AA-omni Score [0] and 'by Domain' Software Engineering [1]. Also added comparison to the eyeballs-median of the 'top 10' models and then you see that indeed Mistral Large 4 Preview scores miserable in the AA-Omni indices.

                              Inte         Cost     AA-   Omni
                              llig          per    Omni  Softw
      Open Weight model       ence  Speed  Task   score    Eng
    
      Mimo-V.26-Pro            46     47   $0.13      8     33
      GLM-5.3 (max)            45     73   $2.01     14     37
      DeepSeek 4.1 Flash Max   39    227   $0.27     -5     34
      Mistral Large 4 Preview  38    116   $1.13     -5      5
    
      Closed/proprietary      ~50   110-  $1.50-    ~43    ~85
         median top 10               242   $7.50
    

    [0] https://artificialanalysis.ai/evaluations/omniscience [1] https://artificialanalysis.ai/evaluations/omniscience?detail...