· State of the Art

Daily AI Decision Intelligence

Multi-source rankings blended into one score. See which model is truly State of the Art — today.

Data as of 2026-08-27Updated Aug 27, 11:07 AM

Overall Intelligence

Multi-source blend of general intelligence benchmarks into a single AISOTA score.

#ModelArtificial AnalysisLiveBenchAISOTA Score
1SOTAClaude Fable 5 (with fallback)2 sources90.098.094.0
22gpt-5.5-xhigh93.993.9
33Claude Opus 5 (max)2 sources95.091.893.4
4GPT-5.6 Sol (max)2 sources85.095.990.5
5smaug-agentic87.887.8
6qwen3.8-max85.785.7
7Grok 4.6 (high)2 sources80.081.680.8
8Kimi K3 (max)2 sources75.083.779.3
9gpt-5.4-xhigh77.677.6
10gemini-3.1-pro-preview-high71.471.4

Coding Ability

Multi-source blend of coding benchmarks into a single AISOTA score.

#ModelArtificial Analysis AgenticLiveBench CodingAISOTA Score
1SOTAsmaug-agentic98.098.0
22Claude Opus 5 (max)2 sources95.093.994.5
33GLM-5.3 (max)2 sources90.089.889.9
4qwen3.8-max87.887.8
5claude-sonnet-5-xhigh-effort85.785.7
6Claude Fable 5 (with fallback)2 sources65.095.980.5
7GPT-5.6 Sol (max)2 sources75.083.779.3
8deepseek-v4-flash-vision-exp77.677.6
9Kimi K3 (max)2 sources60.091.875.9
10GLM-5.3-Flash2 sources80.071.475.7

Value for Money

Intelligence per dollar, based on median input + output price.

#ModelAA IntelligenceOpenRouter PriceAISOTA Score
1SOTAGLM-5.3-Flash2 sources60.086.073.0
22Gemini 3.7 Flash (high)2 sources45.065.755.4
33GPT-5.6 Luna (max)2 sources35.074.354.6
4GPT-5.6 Sol (max)2 sources85.022.753.9
5Claude Opus 5 (max)2 sources95.011.053.0
6Grok 4.6 (high)2 sources80.019.049.5
7Claude Fable 5 (with fallback)2 sources90.05.747.9
8GLM-5.3 (max)2 sources70.023.746.9
9Qwen3.8 2.4T A95B2 sources65.019.342.1
10Kimi K3 (max)2 sources75.07.741.4

Model Usage

API request volume per model over the last 7 days, sourced from OpenRouter public rankings.

#ModelOpenRouter UsageAISOTA Score
1SOTAdeepseek-v4-flash970.62M99.8
22ox-alpha343.49M99.5
33gpt-5-6-luna316.57M99.3
4gemini-2-5-flash-lite271.88M99.1
5gpt-oss-120b135.59M98.9
6gpt-4o-mini129.58M98.6
7mimo-v2-5122.75M98.4
8gemini-3-1-flash-lite116.93M98.2
9gemini-2-5-flash114.07M97.9
10text-embedding-3-small108.43M97.7

AI Tools Leaderboard

Ranked lists of the best AI tools across the most popular categories, with editorial notes on each tool.

AI Compute & API Pricing

Live AI compute prices: model API inference costs, cloud GPU benchmark pricing, and best value models per quality point. Updated daily from public sources.

API Inference Pricing

2530 entries

  1. 1SOTAdeepinfra/meta-llama/Llama-3.2-3B-Instruct$0.02/M
  2. 2lambda_ai/llama3.2-3b-instruct$0.02/M
  3. 3nscale/Qwen/Qwen2.5-Coder-3B-Instruct$0.02/M
  4. 4nscale/Qwen/Qwen2.5-Coder-7B-Instruct$0.02/M
  5. 5nebius/Qwen/Qwen2.5-Coder-7B$0.02/M

Chinese AI Models Pricing

712 entries

  1. 1SOTAnscale/Qwen/Qwen2.5-Coder-3B-Instruct$0.02/M
  2. 2nscale/Qwen/Qwen2.5-Coder-7B-Instruct$0.02/M
  3. 3nebius/Qwen/Qwen2.5-Coder-7B$0.02/M
  4. 4nscale/deepseek-ai/DeepSeek-R1-Distill-Llama-8B$0.03/M
  5. 5novita/qwen/qwen3-4b-fp8$0.03/M

Cloud GPU Benchmark Pricing

18 entries

  1. 1SOTAStandard_NV4as_v4$0.054/h
  2. 2Standard_NC36ds_xl_RTXPRO6000BSE_v6$0.333/h
  3. 3Standard_NV24s_v3$0.584/h
  4. 4Standard_NV16as_v4$1.901/h
  5. 5Standard_NC320lds_xl_RTXPRO6000BSE_v6$2.643/h

How AISOTA scoring works

Each model is scored against the sources on that board. Within each source, models are ranked and converted to a 0–100 percentile. A model's AISOTA score is the weighted average of its percentile ranks across sources. SOTA marks today's #1.

Data is collected automatically from public APIs each day. Missing cells mean the source does not list that model.