Performance
Equal-weight mean of each model’s normalized benchmark scores within the selected use case. Offers for the same model share one rank.
Model decision tool
Choose a task, then compare normalized benchmark evidence with dated provider pricing. Missing evidence stays missing.
Step 1
Loading comparison data...
Step 2
Step 3
Cost scenario: 1M input + 1M output tokens in USD. Table headers sort with click, Enter, or Space.
Equal-weight mean of each model’s normalized benchmark scores within the selected use case. Offers for the same model share one rank.
Best value = 2 × performance × cost efficiency ÷ (performance + cost efficiency). Cost efficiency = 100 × the lowest positive scenario cost ÷ the model’s scenario cost.
Use cases need at least two comparable models. The table retains every priced model plus the 100 highest-scoring unpriced models; prices, dates, and sources are never inferred.