Providers

Compare models across different providers, sorted by Cost Index for a clear view of their relative cost. The index is calculated using the formula below.

Cost Index

2 × +=$0.50= ×1

How do you use AI?

Real AI workloads can vary widely in their input-to-output ratio. Chat is relatively balanced, while RAG, agents, and long-context tasks often process far more input tokens than they generate. Because input and output are priced differently, the same model can rank very differently depending on the workload you choose.

AI Developers

Amazon Bedrock Aggregator

50 models · 50 flat comparable offers

Provider fee: Prices are the US East (N. Virginia) on-demand rates from the AWS Price List. Other AWS regions are priced differently, and provisioned throughput is bought by the hour rather than by the token. Official source

Website · Pricing source

Cerebras Aggregator

3 models · 3 flat comparable offers

Provider fee: Public hosted language models and exact structured per-token prices from the official public catalog. Official source

Website · Pricing source

Cloudflare Aggregator

38 models · 38 flat comparable offers

Provider fee: Workers AI is billed in Neurons against a paid plan, with a daily free allocation. The per-token prices here are the dollar equivalents Cloudflare states on the same page. Official source

Website · Pricing source

DeepInfra Aggregator

70 models · 70 flat comparable offers

Website · Pricing source

Fireworks AI Aggregator

16 models · 16 flat comparable offers

Provider fee: Only Fireworks Serverless text and vision-language models are included; dedicated and on-demand deployments are excluded. Official source

Website · Pricing source

Groq Aggregator

6 models · 3 flat comparable offers

Provider fee: Production and Preview general text-output listings are included; Contact Sales remains unpriced. Official source

Website · Pricing source

Hugging Face Inference Providers Aggregator

87 models · 159 flat comparable offers

Provider fee: Hugging Face routes requests to the selected inference provider and states that it adds no markup to the provider rate. Official source

Website · Pricing source

Novita AI Aggregator

54 models · 0 flat comparable offers

Provider fee: Public Serverless LLM listings include exact representable tier bands; unsupported complex prices remain unpriced rather than flattened. Official source

Website · Pricing source

OpenRouter Aggregator

228 models · 225 flat comparable offers

Provider fee: OpenRouter charges a fee on credit purchases. It is not part of the per-token prices shown here; see the official FAQ. Official source

Website · Pricing source

Together AI Aggregator

21 models · 21 flat comparable offers

Provider fee: Only the public Serverless Chat catalog is included; dedicated endpoints and other inference products are excluded. Official source

Website · Pricing source

Anthropic AI Developer

15 models · 15 flat comparable offers

Website · Pricing source

Cohere AI Developer

14 models · 2 flat comparable offers

Provider fee: Only exact public production API token prices are included; Trial access, private deployment, and custom sales pricing are excluded. Official source

Website · Pricing source

DeepSeek AI Developer

2 models · 2 flat comparable offers

Website · Pricing source

Google AI Developer

18 models · 18 flat comparable offers

Provider fee: Prices are the Gemini Developer API paid-tier rates from Google AI Studio. Vertex AI sells the same models under a separate price list. Official source

Website · Pricing source

Meta AI Developer

1 models · 1 flat comparable offers

  • $1.25 / $4.25 · Arena rank #4×25 / ×11

Website · Pricing source

MiniMax AI Developer

2 models · 1 flat comparable offers

  • $0.30 / $1.20 · Arena rank #119×6.0 / ×3.0
  • $0.30 / $1.20 · Arena rank #76×6.0 / ×3.0

Website · Pricing source

Mistral AI AI Developer

8 models · 8 flat comparable offers

Website · Pricing source

Moonshot / Kimi AI Developer

3 models · 3 flat comparable offers

  • $0.95 / $4.00 · Arena rank #48×19 / ×10
  • $0.95 / $4.00×19 / ×10
  • $3.00 / $15.00 · Arena rank #10×60 / ×38

Website · Pricing source

OpenAI AI Developer

39 models · 39 flat comparable offers

  • $0.05 / $0.40 · Arena rank #215×1.0 / ×1.0
  • $0.10 / $0.40×2.0 / ×1.0
  • $0.15 / $0.60×3.0 / ×1.5
  • $0.40 / $0.40×8.0 / ×1.0
  • $0.20 / $1.20 · Arena rank #63×4.0 / ×3.0
  • $0.20 / $1.25 · Arena rank #139×4.0 / ×3.1
  • $0.40 / $1.60×8.0 / ×4.0
  • $0.50 / $1.50 · Arena rank #312×10 / ×3.8
  • $0.50 / $1.50×10 / ×3.8
  • $0.25 / $2.00 · Arena rank #156×5.0 / ×5.0

Website · Pricing source

Alibaba / Qwen AI Developer

58 models · 16 flat comparable offers

Provider fee: Only Qwen-authored models in the Alibaba Cloud Model Studio Singapore deployment are included. Official source

Website · Pricing source

xAI AI Developer

7 models · 7 flat comparable offers

Website · Pricing source

Xiaomi MiMo AI Developer

2 models · 2 flat comparable offers

  • $0.14 / $0.28 · Arena rank #90×2.8 / ×0.7
  • $0.435 / $0.87 · Arena rank #40×8.7 / ×2.2

Provider fee: Only the official Overseas Pricing token rates are included; domestic RMB pricing and hosted web-search fees are excluded. Official source

Website · Pricing source

Z.AI AI Developer

15 models · 15 flat comparable offers

Website · Pricing source

Arena quality & value

Current model quality against the same Cost Index basis used above.

Quality vs cost

Higher Arena rating is better. Arena rank #1 is the top rank.

11001200130014001500×0.5×1×5×10×50×100Cost Index (log scale)Arena quality rating ↑ betterAnthropic: Claude Fable 5 Arena #1 · Rating 1507 Cost Index ×140 Cheapest at Anthropic Arena Top 10 #11Anthropic: Claude Opus 4.6 Arena #5 · Rating 1497 Cost Index ×70 Cheapest at Anthropic Arena Top 10 #33Anthropic: Claude Opus 4.7 Arena #6 · Rating 1494 Cost Index ×70 Cheapest at Anthropic Arena Top 10 #44Anthropic: Claude Opus 4.8 Arena #29 · Rating 1473 Cost Index ×70 Cheapest at AnthropicAnthropic: Claude Sonnet 4.6 Arena #32 · Rating 1472 Cost Index ×42 Cheapest at AnthropicAnthropic: Claude Sonnet 5 Arena #47 · Rating 1461 Cost Index ×28 Cheapest at Anthropic Arena evaluated claude-sonnet-5-high (High), not the base model. This reviewed representative evaluation is used for Anthropic: Claude Sonnet 5's rank, filters and quality comparisons.c4ai-aya-expanse-32b Arena #289 · Rating 1267 Cost Index ×5.0 Cheapest at CohereClaude Opus 5 Arena #12 · Rating 1487 Cost Index ×70 Cheapest at Anthropic Arena Top 10 #7 Arena has 2 reviewed evaluations for Claude Opus 5. AI for Less conservatively uses the lowest-performing result: claude-opus-5-max (Max). claude-opus-5-high (High): Arena #7, rating 1493. claude-opus-5-max (Max): Arena #12, rating 1487 — used.7command-r-08-2024 Arena #299 · Rating 1250 Cost Index ×1.8 Cheapest at OpenRoutercommand-r-plus-08-2024 Arena #283 · Rating 1276 Cost Index ×30 Cheapest at CohereDeepSeek: DeepSeek V3.1 Terminus Arena #122 · Rating 1415 Cost Index ×3.0 Cheapest at DeepInfraDeepSeek: DeepSeek V3.2 Arena #103 · Rating 1425 Cost Index ×1.8 Cheapest at DeepInfraDeepSeek: DeepSeek V3.2 Exp Arena #110 · Rating 1422 Cost Index ×1.9 Cheapest at Hugging Face Inference ProvidersDeepSeek: DeepSeek V4 Flash 0423 Arena #85 · Rating 1436 Cost Index ×0.7 Cheapest at OpenRouterDeepSeek: DeepSeek V4 Pro Arena #53 · Rating 1458 Cost Index ×6.6 Cheapest at DeepSeekDeepSeek: R1 Arena #145 · Rating 1398 Cost Index ×7.8 Cheapest at Hugging Face Inference ProvidersDeepSeek: R1 0528 Arena #111 · Rating 1422 Cost Index ×6.3 Cheapest at DeepInfraGemini 3.5 Flash-Lite Arena #55 · Rating 1457 Cost Index ×6.2 Cheapest at GoogleGemini 3.6 Flash Arena #20 · Rating 1481 Cost Index ×11 Cheapest at Google Arena Top 10 #10 Arena evaluated gemini-3.6-flash-high (High), not the base model. This reviewed representative evaluation is used for Gemini 3.6 Flash's rank, filters and quality comparisons.10gemini-2.5-flash Arena #131 · Rating 1410 Cost Index ×6.2 Cheapest at DeepInfragemini-2.5-pro Arena #72 · Rating 1445 Cost Index ×25 Cheapest at DeepInfragemini-3.5-flash Arena #25 · Rating 1475 Cost Index ×24 Cheapest at DeepInfra Arena has 2 reviewed evaluations for gemini-3.5-flash. AI for Less conservatively uses the lowest-performing result: gemini-3.5-flash-medium (Medium). gemini-3.5-flash-high (High): Arena #21, rating 1479. gemini-3.5-flash-medium (Medium): Arena #25, rating 1475 — used.Gemma 3 12b It Arena #213 · Rating 1342 Cost Index ×0.5 Cheapest at DeepInfraGemma 4 31B Arena #62 · Rating 1451 Cost Index ×6.9 Cheapest at Cerebrasgemma-3-27b-it Arena #181 · Rating 1366 Cost Index ×0.6 Cheapest at DeepInfragemma-3-4b-it Arena #264 · Rating 1303 Cost Index ×0.3 Cheapest at Amazon BedrockGoogle: Gemini 3.1 Flash Lite Preview Arena #92 · Rating 1432 Cost Index ×4.0 Cheapest at OpenRouterGoogle: Gemma 2 27B Arena #271 · Rating 1289 Cost Index ×3.9 Cheapest at OpenRouterGoogle: Gemma 3n 4B Arena #243 · Rating 1318 Cost Index ×0.5 Cheapest at OpenRouterGpt 3.5 Turbo 0125 Arena #312 · Rating 1225 Cost Index ×5.0 Cheapest at OpenAIGpt 3.5 Turbo 1106 Arena #323 · Rating 1204 Cost Index ×8.0 Cheapest at OpenAIGpt 4 0613 Arena #284 · Rating 1276 Cost Index ×240 Cheapest at OpenAIGpt 4 Turbo 2024 04 09 Arena #232 · Rating 1324 Cost Index ×100 Cheapest at OpenAIgrok-4.3 Arena #78 · Rating 1442 Cost Index ×10 Cheapest at OpenRoutergrok-4.5 Arena #36 · Rating 1470 Cost Index ×20 Cheapest at OpenRouterKimi K2.5 Arena #93 · Rating 1431 Cost Index ×6.3 Cheapest at DeepInfra Arena has 2 reviewed evaluations for Kimi K2.5. AI for Less conservatively uses the lowest-performing result: kimi-k2.5-instant (Instant). kimi-k2.5-thinking (Thinking): Arena #64, rating 1450. kimi-k2.5-instant (Instant): Arena #93, rating 1431 — used.Kimi K2.6 Arena #48 · Rating 1460 Cost Index ×10 Cheapest at DeepInfraKimi-K3 Arena #10 · Rating 1490 Cost Index ×40 Cheapest at DeepInfra Arena Top 10 #6 Arena evaluated kimi-k3-max (Max), not the base model. This reviewed representative evaluation is used for Kimi-K3's rank, filters and quality comparisons.6Llama 3 8b Instruct Arena #313 · Rating 1223 Cost Index ×2.8 Cheapest at CloudflareLlama 3.1 8b Instruct Arena #320 · Rating 1211 Cost Index ×0.2 Cheapest at Hugging Face Inference ProvidersLlama 3.2 1b Instruct Arena #375 · Rating 1111 Cost Index ×0.5 Cheapest at CloudflareLlama 3.2 3b Instruct Arena #346 · Rating 1167 Cost Index ×0.9 Cheapest at CloudflareLlama 3.3 70b Instruct Arena #242 · Rating 1318 Cost Index ×1.3 Cheapest at Hugging Face Inference ProvidersLlama 4 Scout 17b 16e Instruct Arena #236 · Rating 1322 Cost Index ×0.9 Cheapest at Hugging Face Inference ProvidersMeta: Muse Spark 1.1 Arena #8 · Rating 1491 Cost Index ×14 Cheapest at OpenRouter Arena Top 10 #55mimo-v2.5 Arena #90 · Rating 1434 Cost Index ×1.1 Cheapest at OpenRoutermimo-v2.5-pro Arena #40 · Rating 1467 Cost Index ×3.5 Cheapest at OpenRouterMinimax M2 Arena #209 · Rating 1346 Cost Index ×3.1 Cheapest at OpenRouterMinimax M3 Arena #76 · Rating 1442 Cost Index ×3.3 Cheapest at Hugging Face Inference ProvidersMiniMax-M2.7 Arena #119 · Rating 1416 Cost Index ×3.0 Cheapest at Hugging Face Inference ProvidersMistral Large 2407 Arena #249 · Rating 1314 Cost Index ×20 Cheapest at OpenRouterMistral Large 3 Arena #124 · Rating 1414 Cost Index ×5.0 Cheapest at Amazon BedrockMistral: Mistral Medium 3.5 Arena #98 · Rating 1427 Cost Index ×21 Cheapest at Mistral AIMistral: Mistral Small 3 Arena #285 · Rating 1274 Cost Index ×0.4 Cheapest at DeepInframuse-spark-1.2 Arena #4 · Rating 1498 Cost Index ×14 Cheapest at Meta Arena Top 10 #2 Arena evaluated muse-spark-1.2 (xHigh) (xHigh), not the base model. This reviewed representative evaluation is used for muse-spark-1.2's rank, filters and quality comparisons.2Nvidia Nemotron 3 Ultra 550b A55b Nvfp4 Arena #101 · Rating 1426 Cost Index ×7.2 Cheapest at Hugging Face Inference ProvidersOpenAI: GPT-4o (2024-05-13) Arena #207 · Rating 1346 Cost Index ×50 Cheapest at OpenAIOpenAI: GPT-4o (2024-08-06) Arena #221 · Rating 1335 Cost Index ×30 Cheapest at OpenRouterOpenAI: GPT-4o-mini (2024-07-18) Arena #245 · Rating 1318 Cost Index ×1.8 Cheapest at OpenRouterOpenAI: GPT-5 Arena #89 · Rating 1434 Cost Index ×25 Cheapest at OpenAI Arena evaluated gpt-5-high (High), not the base model. This reviewed representative evaluation is used for OpenAI: GPT-5's rank, filters and quality comparisons.OpenAI: GPT-5 Mini Arena #156 · Rating 1390 Cost Index ×5.0 Cheapest at OpenAI Arena evaluated gpt-5-mini-high (High), not the base model. This reviewed representative evaluation is used for OpenAI: GPT-5 Mini's rank, filters and quality comparisons.OpenAI: GPT-5 Nano Arena #215 · Rating 1337 Cost Index ×1.0 Cheapest at OpenAI Arena evaluated gpt-5-nano-high (High), not the base model. This reviewed representative evaluation is used for OpenAI: GPT-5 Nano's rank, filters and quality comparisons.OpenAI: GPT-5.1 Arena #81 · Rating 1439 Cost Index ×25 Cheapest at OpenAIOpenAI: GPT-5.2 Arena #87 · Rating 1435 Cost Index ×35 Cheapest at OpenAIOpenAI: GPT-5.4 Arena #41 · Rating 1466 Cost Index ×40 Cheapest at OpenAIOpenAI: GPT-5.4 Mini Arena #69 · Rating 1448 Cost Index ×12 Cheapest at OpenAI Arena evaluated gpt-5.4-mini-high (High), not the base model. This reviewed representative evaluation is used for OpenAI: GPT-5.4 Mini's rank, filters and quality comparisons.OpenAI: GPT-5.4 Nano Arena #139 · Rating 1402 Cost Index ×3.3 Cheapest at OpenAI Arena evaluated gpt-5.4-nano-high (High), not the base model. This reviewed representative evaluation is used for OpenAI: GPT-5.4 Nano's rank, filters and quality comparisons.OpenAI: GPT-5.5 Arena #24 · Rating 1476 Cost Index ×80 Cheapest at OpenAIOpenAI: GPT-5.6 Luna Arena #63 · Rating 1451 Cost Index ×3.2 Cheapest at OpenAI Arena evaluated gpt-5.6-luna-xhigh (xHigh), not the base model. This reviewed representative evaluation is used for OpenAI: GPT-5.6 Luna's rank, filters and quality comparisons.OpenAI: GPT-5.6 Sol Arena #18 · Rating 1482 Cost Index ×40 Cheapest at OpenRouter Arena Top 10 #9 Arena evaluated gpt-5.6-sol-xhigh (xHigh), not the base model. This reviewed representative evaluation is used for OpenAI: GPT-5.6 Sol's rank, filters and quality comparisons.9OpenAI: GPT-5.6 Terra Arena #44 · Rating 1465 Cost Index ×32 Cheapest at OpenAI Arena evaluated gpt-5.6-terra-xhigh (xHigh), not the base model. This reviewed representative evaluation is used for OpenAI: GPT-5.6 Terra's rank, filters and quality comparisons.OpenAI: gpt-oss-120b Arena #195 · Rating 1352 Cost Index ×0.5 Cheapest at OpenRouterOpenAI: gpt-oss-20b Arena #246 · Rating 1318 Cost Index ×0.4 Cheapest at OpenRouterOpenAI: o3 Mini Arena #201 · Rating 1348 Cost Index ×13 Cheapest at OpenAIOpenAI: o3 Mini High Arena #183 · Rating 1363 Cost Index ×13 Cheapest at OpenRouterphi-4 Arena #297 · Rating 1256 Cost Index ×0.6 Cheapest at DeepInfraQwen: Qwen3 VL 235B A22B Thinking Arena #149 · Rating 1395 Cost Index ×9.6 Cheapest at OpenRouterqwen3-235b-a22b Arena #172 · Rating 1375 Cost Index ×2.0 Cheapest at Hugging Face Inference ProvidersQwen3-235B-A22B-Thinking-2507 Arena #144 · Rating 1399 Cost Index ×5.5 Cheapest at Alibaba / QwenQwen3-30B-A3B Arena #230 · Rating 1327 Cost Index ×1.5 Cheapest at DeepInfraqwen3-30b-a3b-instruct-2507 Arena #165 · Rating 1383 Cost Index ×0.6 Cheapest at OpenRouterQwen3-32B Arena #204 · Rating 1347 Cost Index ×0.8 Cheapest at Hugging Face Inference Providersqwen3-coder-480b-a35b-instruct Arena #160 · Rating 1388 Cost Index ×4.6 Cheapest at Hugging Face Inference ProvidersQwen3-Next-80B-A3B-Instruct Arena #141 · Rating 1401 Cost Index ×2.6 Cheapest at DeepInfraqwen3-next-80b-a3b-thinking Arena #178 · Rating 1369 Cost Index ×3.0 Cheapest at Alibaba / QwenQwen3-VL-235B-A22B-Instruct Arena #123 · Rating 1414 Cost Index ×2.6 Cheapest at DeepInfraQwen3.5-122B-A10B Arena #118 · Rating 1417 Cost Index ×5.2 Cheapest at OpenRouterQwen3.5-27B Arena #134 · Rating 1408 Cost Index ×3.9 Cheapest at OpenRouterQwen3.5-35B-A3B Arena #148 · Rating 1395 Cost Index ×2.6 Cheapest at DeepInfraQwen3.5-397B-A17B Arena #77 · Rating 1442 Cost Index ×6.2 Cheapest at OpenRouterqwen3.6-max-preview Arena #49 · Rating 1460 Cost Index ×16 Cheapest at OpenRouterqwen3.6-plus Arena #74 · Rating 1444 Cost Index ×5.2 Cheapest at OpenRouterqwen3.7-plus Arena #52 · Rating 1458 Cost Index ×3.8 Cheapest at OpenRouterQwen3.8-Max Arena #17 · Rating 1482 Cost Index ×17 Cheapest at DeepInfra Arena Top 10 #88Qwq 32b Arena #220 · Rating 1336 Cost Index ×4.6 Cheapest at CloudflareSpaceXAI: Grok 4.6 Arena #46 · Rating 1462 Cost Index ×20 Cheapest at OpenRouter Arena evaluated grok-4.6-high (High), not the base model. This reviewed representative evaluation is used for SpaceXAI: Grok 4.6's rank, filters and quality comparisons.Z.ai: GLM 4.5 Arena #130 · Rating 1411 Cost Index ×6.8 Cheapest at OpenRouterZ.ai: GLM 4.5 Air Arena #176 · Rating 1373 Cost Index ×2.2 Cheapest at Hugging Face Inference ProvidersZ.ai: GLM 4.5V Arena #194 · Rating 1353 Cost Index ×6.0 Cheapest at Hugging Face Inference ProvidersZ.ai: GLM 4.6 Arena #106 · Rating 1424 Cost Index ×6.0 Cheapest at Hugging Face Inference ProvidersZ.ai: GLM 4.6V Arena #170 · Rating 1377 Cost Index ×3.0 Cheapest at OpenRouterZ.ai: GLM 5.1 Arena #39 · Rating 1468 Cost Index ×9.9 Cheapest at OpenRouterZ.ai: GLM 5.2 Arena #35 · Rating 1470 Cost Index ×9.9 Cheapest at OpenRouter Arena evaluated glm-5.2-max (Max), not the base model. This reviewed representative evaluation is used for Z.ai: GLM 5.2's rank, filters and quality comparisons.Z.ai: GLM 5V Turbo Arena #91 · Rating 1433 Cost Index ×13 Cheapest at OpenRouter

All Arena-ranked models Arena Top 10

Arena Top 10

The highest-ranked Arena models tracked by AI for Less, shown with their current Cost Index.

  1. Anthropic: Claude Fable 5
  2. muse-spark-1.2
  3. Anthropic: Claude Opus 4.6
  4. Anthropic: Claude Opus 4.7
  5. Meta: Muse Spark 1.1
  6. Kimi-K3
  7. Claude Opus 5
  8. Qwen3.8-Max
  9. OpenAI: GPT-5.6 Sol
  10. Gemini 3.6 Flash

Quality data: Arena leaderboard dataset · CC BY 4.0 · text_style_control · published 2026-08-19