· State of the Art

Daily AI Decision Intelligence

Multi-source rankings blended into one score. See which model is truly State of the Art — today.

Data as of 2026-08-14Updated Aug 14, 05:21 PM

Overall Intelligence

Multi-source blend of general intelligence benchmarks into a single AISOTA score.

#ModelArtificial AnalysisLiveBenchAISOTA Score
1SOTAClaude Fable 5 (with fallback)2 sources90.097.793.8
22gpt-5.5-xhigh93.093.0
33Claude Opus 5 (max)2 sources95.090.792.8
4GPT-5.6 Sol (max)2 sources85.095.390.2
5smaug-agentic86.086.0
6Grok 4.6 (high)2 sources80.079.179.5
7Kimi K3 (max)2 sources75.081.478.2
8Qwen3.8 Max2 sources70.083.776.8
9gpt-5.4-xhigh74.474.4
10Gemini 3.7 Flash (high)2 sources55.088.471.7

Coding Ability

Multi-source blend of coding benchmarks into a single AISOTA score.

#ModelArtificial Analysis AgenticLiveBench CodingAISOTA Score
1SOTAsmaug-agentic97.797.7
22Claude Opus 5 (max)2 sources95.093.094.0
33Qwen3.8 Max2 sources85.088.486.7
4claude-sonnet-5-xhigh-effort86.086.0
5Claude Fable 5 (with fallback)2 sources75.095.385.2
6GPT-5.6 Sol (max)2 sources80.083.781.8
7Grok 4.6 (high)2 sources90.072.181.0
8Kimi K3 (max)2 sources70.090.780.3
9muse-spark-1.1-xhigh79.179.1
10gpt-5.5-xhigh74.474.4

Value for Money

Intelligence per dollar, based on median input + output price.

#ModelAA IntelligenceOpenRouter PriceAISOTA Score
1SOTAGemini 3.7 Flash (high)2 sources55.065.060.0
22GPT-5.6 Luna (max)2 sources40.074.857.4
33DeepSeek V4 Pro 0813 (max)2 sources50.059.254.6
4Claude Opus 5 (max)2 sources95.012.253.6
5MiniMax-M32 sources30.072.851.4
6Grok 4.6 (high)2 sources80.020.450.2
7Claude Fable 5 (with fallback)2 sources90.06.148.0
8GPT-5.6 Sol (max)2 sources85.08.846.9
9Muse Spark 1.2 (xhigh)2 sources65.026.245.6
10Qwen3.8 Max2 sources70.020.145.0

Model Usage

API request volume per model over the last 7 days, sourced from OpenRouter public rankings.

#ModelOpenRouter UsageAISOTA Score
1SOTAdeepseek-v4-flash843.17M99.8
22gemini-2-5-flash-lite254.84M99.5
33gpt-5-6-luna237.25M99.3
4gemini-2-5-flash152.47M99.0
5hy3132.91M98.8
6gpt-4o-mini119.03M98.6
7gemini-3-1-flash-lite117.7M98.3
8deepseek-v4-pro107.73M98.1
9gemini-3-flash104.59M97.9
10gemma-4-31b-it104.07M97.6

How AISOTA scoring works

Each model is scored against the sources on that board. Within each source, models are ranked and converted to a 0–100 percentile. A model's AISOTA score is the weighted average of its percentile ranks across sources. SOTA marks today's #1.

Data is collected automatically from public APIs each day. Missing cells mean the source does not list that model.