· State of the Art

Daily AI Decision Intelligence

Multi-source rankings blended into one score. See which model is truly State of the Art — today.

Data as of 2026-10-01Updated Oct 2, 05:00 AM

Overall Intelligence

Multi-source blend of general intelligence benchmarks into a single AISOTA score.

#ModelArtificial AnalysisLiveBenchAISOTA Score
1SOTAclaude-fable-5-1-max-effort—98.498.4
22claude-fable-5-max-effort—96.896.8
33gpt-6-astra-max—95.295.2
4muse-spark-1.3-xhigh—93.793.7
5claude-opus-5-5-xhigh-effort—92.192.1
6gpt-6.1-sol-xhigh—90.590.5
7gpt-5.6-sol-max—88.988.9
8deepseek-v4.1-flash-max—87.387.3
9gpt-5.5-xhigh—85.785.7
10claude-opus-5-max-effort—84.184.1

Coding Ability

Multi-source blend of coding benchmarks into a single AISOTA score.

#ModelArtificial Analysis AgenticLiveBench CodingAISOTA Score
1SOTAdeepseek-v4.1-flash-max—98.498.4
22claude-opus-5-5-xhigh-effort—96.896.8
33claude-fable-5-1-max-effort—95.295.2
4smaug-agentic—93.793.7
5claude-fable-5-max-effort—92.192.1
6claude-opus-5-max-effort—90.590.5
7muse-spark-1.3-xhigh—88.988.9
8kimi-k3—87.387.3
9glm-5.3—85.785.7
10qwen3.8-max—84.184.1

Model Usage

API request volume per model over the last 7 days, sourced from OpenRouter public rankings.

#ModelOpenRouter UsageAISOTA Score
1SOTAjev-1-13944.05M99.8
22deepseek-v4-flash788.41M99.6
33space-bunny-alpha505.25M99.4
4deepseek-v4-1-flash395.56M99.2
5glm-5-3-flash355.41M99.0
6gemini-2-5-flash-lite297.09M98.8
7gpt-5-6-luna221.34M98.6
8gpt-6-luna218.64M98.4
9gemma-4-31b-it164.94M98.2
10text-embedding-3-small159.65M98.0

AI Tools Leaderboard

Ranked lists of the best AI tools across the most popular categories, with editorial notes on each tool.

AI Compute & API Pricing

Live AI compute prices: model API inference costs, cloud GPU benchmark pricing, and best value models per quality point. Updated daily from public sources.

API Inference Pricing

3448 entries

  1. 1SOTAdeepinfra/meta-llama/Llama-3.2-3B-Instruct$0.02/M
  2. 2lambda_ai/llama3.2-3b-instruct$0.02/M
  3. 3nscale/Qwen/Qwen2.5-Coder-3B-Instruct$0.02/M
  4. 4nscale/Qwen/Qwen2.5-Coder-7B-Instruct$0.02/M
  5. 5nebius/Qwen/Qwen2.5-Coder-7B$0.02/M

Chinese AI Models Pricing

1069 entries

  1. 1SOTAnscale/Qwen/Qwen2.5-Coder-3B-Instruct$0.02/M
  2. 2nscale/Qwen/Qwen2.5-Coder-7B-Instruct$0.02/M
  3. 3nebius/Qwen/Qwen2.5-Coder-7B$0.02/M
  4. 4together_ai/Qwen/Qwen2-1.5B-Instruct$0.02/M
  5. 5nscale/deepseek-ai/DeepSeek-R1-Distill-Llama-8B$0.03/M

Cloud GPU Benchmark Pricing

19 entries

  1. 1SOTAStandard_NV4as_v4$0.054/h
  2. 2Standard_NC36ds_xl_RTXPRO6000BSE_v6$0.333/h
  3. 3Standard_NV24s_v3$0.584/h
  4. 4Standard_NC72lds_xl_RTXPRO6000BSE_v6$0.610/h
  5. 5Standard_NV16as_v4$1.901/h

How AISOTA scoring works

Each model is scored against the sources on that board. Within each source, models are ranked and converted to a 0–100 percentile. A model's AISOTA score is the weighted average of its percentile ranks across sources. SOTA marks today's #1.

Data is collected automatically from public APIs each day. Missing cells mean the source does not list that model.