readable AI benchmarks (simplified)
Quality Score · correct answers and intelligence breakdowns · higher is better
loading live catalog · GET /api/recommend
Blended price · USD per 1M tokens · log scale · lower is better
most attractive Pareto frontierPoints use exact values. labels move · data points do not. only exact duplicates grouped.
#Model
loading live ranking · GET /api/recommend
Value Score (v1) always uses blendedPrice, not selected graph cost. Quality Score (v1) gives answerQuality a pre-normalization weight of 0.50; available weights always normalize. AA indices are combined without source-overlap correction. Both are this site’s decision scores, not Artificial Analysis benchmarks.
ranked pick: GET /api/recommend · graph is not the ranking · API