AI Model Intelligence: Rankings, Benchmarks & Pricing

RuntimeWire's model intelligence desk: our own benchmark standings, head-to-head verdicts, and newsroom coverage — plus live pricing, context windows, and usage for every major AI model across text, image, video, audio, and speech.

Browse models by output type: text, image, video, audio, and speech. Each listing includes context window, supported input and output modalities, and current pricing — plus RuntimeWire's own signals: Composite Index rank, head-to-head record, and newsroom coverage.

RuntimeWire Composite Index — current standings

  1. Z.ai: GLM 5.3 Flash — 81.9 (2/3 suites)
  2. Claude Opus 5 (Fast) — 80.5 (2/3 suites)
  3. Qwen: Qwen3.8 Flash — 79.6 (2/3 suites)
  4. Google: Gemini 3.7 Flash — 78.1 (2/3 suites)
  5. step-3.7-flash — 75.3 (2/3 suites)
  6. DeepSeek-V4-Flash-0731 — 71.7 (1/3 suites)
  7. SpaceXAI: Grok 4.6 — 66.5 (1/3 suites)
  8. Anthropic: Claude Opus 4.8 — 62.0 (1/3 suites)

Benchmark methodology and full runs

Latest head-to-head verdicts

All head-to-head matchups · Elo leaderboard

Most covered in RuntimeWire stories (last 30 days)

Pricing and usage data is sourced from OpenRouter; benchmark standings, head-to-head verdicts, and story coverage are RuntimeWire's own reporting. Prices change frequently — confirm on the provider's site before relying on them.