· State of the Art

Daily AI Decision Intelligence

Multi-source rankings blended into one score. See which model is truly State of the Art — today.

Data as of 2026-09-05Updated Sep 5, 05:06 PM

Overall Intelligence

Multi-source blend of general intelligence benchmarks into a single AISOTA score.

#ModelArtificial AnalysisLiveBenchAISOTA Score
1SOTAClaude Fable 5.1 (max with fallback)2 sources95.098.196.5
22GPT-6 Astra (max)2 sources90.094.492.2
33gpt-5.5-xhigh88.988.9
4Claude Fable 5 (with fallback)2 sources80.096.388.2
5Claude Opus 5 (max)2 sources85.087.086.0
6gemini-3.7-flash-high85.285.2
7Muse Spark 1.3 (max)2 sources75.092.683.8
8smaug-agentic83.383.3
9qwen3.8-max81.581.5
10GPT-5.6 Sol (max)2 sources70.090.780.3

Coding Ability

Multi-source blend of coding benchmarks into a single AISOTA score.

#ModelArtificial Analysis AgenticLiveBench CodingAISOTA Score
1SOTAclaude-fable-5-1-max-effort98.198.1
22smaug-agentic96.396.3
33claude-fable-5-max-effort94.494.4
4claude-opus-5-max-effort92.692.6
5muse-spark-1.3-xhigh90.790.7
6kimi-k388.988.9
7glm-5.387.087.0
8qwen3.8-max85.285.2
9claude-sonnet-5-xhigh-effort83.383.3
10gpt-5.6-sol-max81.581.5

Value for Money

Intelligence per dollar, based on median input + output price.

#ModelAA IntelligenceOpenRouter PriceAISOTA Score
1SOTAGLM-5.3-Flash2 sources35.075.555.3
22GPT-5.6 Luna (max)2 sources30.074.252.1
33Claude Fable 5.1 (max with fallback)2 sources95.05.050.0
4Muse Spark 1.3 (max)2 sources75.024.849.9
5Gemini 3.8 Flash (high)2 sources50.047.048.5
6Claude Opus 5 (max)2 sources85.010.948.0
7GPT-6 Astra (max)2 sources90.05.647.8
8GPT-5.6 Sol (max)2 sources70.022.246.1
9Claude Fable 5 (with fallback)2 sources80.04.642.3
10Grok 4.6 (high)2 sources65.017.941.5

Model Usage

API request volume per model over the last 7 days, sourced from OpenRouter public rankings.

#ModelOpenRouter UsageAISOTA Score
1SOTAdeepseek-v4-flash963.46M99.8
22gpt-5-6-luna368.71M99.5
33glm-5-3-flash358.56M99.3
4gemini-2-5-flash-lite271.27M99.1
5gemini-3-1-flash-lite133.12M98.9
6text-embedding-3-small130.44M98.6
7gemini-2-5-flash130.28M98.4
8gpt-oss-20b126.42M98.2
9hy4115.69M98.0
10gpt-4o-mini105.24M97.7

AI Tools Leaderboard

Ranked lists of the best AI tools across the most popular categories, with editorial notes on each tool.

AI Compute & API Pricing

Live AI compute prices: model API inference costs, cloud GPU benchmark pricing, and best value models per quality point. Updated daily from public sources.

API Inference Pricing

2796 entries

  1. 1SOTAdeepinfra/meta-llama/Llama-3.2-3B-Instruct$0.02/M
  2. 2lambda_ai/llama3.2-3b-instruct$0.02/M
  3. 3nscale/Qwen/Qwen2.5-Coder-3B-Instruct$0.02/M
  4. 4nscale/Qwen/Qwen2.5-Coder-7B-Instruct$0.02/M
  5. 5nebius/Qwen/Qwen2.5-Coder-7B$0.02/M

Chinese AI Models Pricing

873 entries

  1. 1SOTAnscale/Qwen/Qwen2.5-Coder-3B-Instruct$0.02/M
  2. 2nscale/Qwen/Qwen2.5-Coder-7B-Instruct$0.02/M
  3. 3nebius/Qwen/Qwen2.5-Coder-7B$0.02/M
  4. 4nscale/deepseek-ai/DeepSeek-R1-Distill-Llama-8B$0.03/M
  5. 5novita/qwen/qwen3-4b-fp8$0.03/M

Cloud GPU Benchmark Pricing

18 entries

  1. 1SOTAStandard_NV4as_v4$0.054/h
  2. 2Standard_NC36ds_xl_RTXPRO6000BSE_v6$0.333/h
  3. 3Standard_NV24s_v3$0.584/h
  4. 4Standard_NV16as_v4$1.901/h
  5. 5Standard_NC320lds_xl_RTXPRO6000BSE_v6$2.643/h

How AISOTA scoring works

Each model is scored against the sources on that board. Within each source, models are ranked and converted to a 0–100 percentile. A model's AISOTA score is the weighted average of its percentile ranks across sources. SOTA marks today's #1.

Data is collected automatically from public APIs each day. Missing cells mean the source does not list that model.