Skip to main content
ZICQ

Overall LLM Rankings

ZICQ LLM rankings update daily across intelligence, coding, price, and context for selection and comparison.

Weighted composite of OpenRouter benchmarks Intelligence / Coding / Agentic (60% / 25% / 15%), covering all 338+ models.

Last updated: 15 min ago (2026-10-03 15:30)
Models tracked 466
Free & tunable 22
Vision 295
Tool calling 398
Reasoning 333
Max context Auto Router (Beta) 2M tokens
Avg price (input) — USD / 1M tokens
Cheapest model IBM: Granite 4.0 Micro $0.017 / 1M
TOP by Overall score Anthropic: Claude Opus 5.5 57.6
Clear
  1. #1
    Anthropic: Claude Fable 5.1 Anthropic Doc & Vision Benchmarked

    Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

    Tools Vision Docs Reasoning Structured
    Intelligence 53.4
    Coding 81.6
    Agentic 57.9
    Context 1M Output ≤ 128K
    $10.00/1M $50.00/1M out cache $0.25/1M
    TOP Flagship
  2. #2
    Anthropic: Claude Fable 5.1 (batch) Anthropic Doc & Vision Benchmarked

    Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

    Tools Vision Docs Reasoning Structured
    Intelligence 53.4
    Coding 81.6
    Agentic 57.9
    Context 1M Output ≤ 128K
    $5.00/1M $25.00/1M out cache $0.12/1M
    TOP 3 Overall
  3. #3
    OpenAI: GPT-6 Astra OpenAI Doc & Vision Benchmarked

    GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

    Tools Vision Docs Reasoning Structured
    Intelligence 52.7
    Coding 76.9
    Agentic 51.0
    Context 1.1M Output ≤ 128K
    $10.00/1M $50.00/1M out cache $1.00/1M
    TOP 3 Overall
  4. #4
    OpenAI: GPT-6 Astra (batch) OpenAI Doc & Vision Benchmarked

    GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

    Tools Vision Docs Reasoning Structured
    Intelligence 52.7
    Coding 76.9
    Agentic 51.0
    Context 1.1M Output ≤ 128K
    $5.00/1M $25.00/1M out cache $0.50/1M
    1M Context Pick
  5. #5
    Anthropic: Claude Opus 5 Anthropic Doc & Vision Benchmarked

    Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

    Tools Vision Docs Reasoning Structured
    Intelligence 50.8
    Coding 78.0
    Agentic 56.5
    Context 1M Output ≤ 128K
    $5.00/1M $25.00/1M out cache $0.50/1M
    1M Context Pick
  6. #6
    Anthropic: Claude Opus 5 (batch) Anthropic Doc & Vision Benchmarked

    Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

    Tools Vision Docs Reasoning Structured
    Intelligence 50.8
    Coding 78.0
    Agentic 56.5
    Context 1M Output ≤ 128K
    $2.50/1M $12.50/1M out cache $0.25/1M
    1M Context Pick
  7. #7
    Anthropic: Claude Fable 5 Anthropic Doc & Vision Benchmarked

    Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

    Tools Vision Docs Reasoning Structured
    Intelligence 49.6
    Coding 76.5
    Agentic 50.7
    Context 1M Output ≤ 128K
    $10.00/1M $50.00/1M out cache $1.00/1M
    1M Context Pick
  8. #8
    Anthropic: Claude Fable 5 (batch) Anthropic Doc & Vision Benchmarked

    Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

    Tools Vision Docs Reasoning Structured
    Intelligence 49.6
    Coding 76.5
    Agentic 50.7
    Context 1M Output ≤ 128K
    $5.00/1M $25.00/1M out cache $0.50/1M
    1M Context Pick
  9. #9
    OpenAI: GPT-5.6 Sol OpenAI Doc & Vision Benchmarked

    GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

    Tools Vision Docs Reasoning Structured
    Intelligence 47.0
    Coding 77.4
    Agentic 50.2
    Context 1.1M Output ≤ 128K
    $2.00/1M $10.00/1M out cache $0.20/1M
    1M Context Pick
  10. #10
    OpenAI: GPT-5.6 Sol (batch) OpenAI Doc & Vision Benchmarked

    GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

    Tools Vision Docs Reasoning Structured
    Intelligence 47.0
    Coding 77.4
    Agentic 50.2
    Context 1.1M Output ≤ 128K
    $1.00/1M $5.00/1M out cache $0.10/1M
    1M Context Pick
  11. #11
    Qwen: Qwen3.8 Max (0902) Alibaba 通义 视觉多模态 Benchmarked

    Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

    Tools Vision Video Reasoning Structured
    Intelligence 45.4
    Coding 76.2
    Agentic 56.0
    Context 1M Output ≤ 131K
    $2.00/1M $6.00/1M out cache $0.25/1M
    1M Context Pick
  12. #12
    SpaceXAI: Grok 4.6 xAI Doc & Vision Benchmarked

    Grok 4.6 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM. It is succeeded by [Grok 4.7](/x-ai/grok-4.7).

    Tools Vision Docs Reasoning Structured
    Intelligence 44.3
    Coding 76.8
    Agentic 53.0
    Context 500K Output ≤ 450K
    $2.00/1M $6.00/1M out cache $0.50/1M
    Recommended
  13. #13
    Z.ai: GLM 5.3 Zhipu 智谱 General LLM Benchmarked

    GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

    Tools Reasoning Structured
    Intelligence 44.8
    Coding 74.8
    Agentic 53.1
    Context 1.0M Output ≤ 131K
    $1.40/1M $4.40/1M out cache $0.14/1M
    1M Context Pick
  14. #14
    Z.ai: GLM 5.3 (batch) Zhipu 智谱 General LLM Benchmarked

    GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

    Tools Reasoning Structured
    Intelligence 44.8
    Coding 74.8
    Agentic 53.1
    Context 1.0M Output ≤ 131K
    $0.45/1M $2.00/1M out cache $0.10/1M
    1M Context Pick
  15. #15
    MoonshotAI: Kimi K3 Moonshotai 视觉多模态 Benchmarked

    Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

    Tools Vision Video Reasoning Structured
    Intelligence 43.6
    Coding 76.2
    Agentic 50.0
    Context 1.0M Output ≤ 943K
    $2.70/1M $13.50/1M out cache $0.27/1M
    1M Context Pick
  16. #16
    MoonshotAI: Kimi K3 (batch) Moonshotai 视觉多模态 Benchmarked

    Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

    Tools Vision Video Reasoning Structured
    Intelligence 43.6
    Coding 76.2
    Agentic 50.0
    Context 1.0M Output ≤ 16K
    $2.28/1M $11.40/1M out cache $0.23/1M
    1M Context Pick
  17. #17
    OpenAI: GPT-5.6 Terra OpenAI Doc & Vision Benchmarked

    GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

    Tools Vision Docs Reasoning Structured
    Intelligence 42.1
    Coding 76.7
    Agentic 43.2
    Context 1.1M Output ≤ 128K
    $2.00/1M $12.00/1M out cache $0.20/1M
    1M Context Pick
  18. #18
    OpenAI: GPT-5.6 Terra (batch) OpenAI Doc & Vision Benchmarked

    GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

    Tools Vision Docs Reasoning Structured
    Intelligence 42.1
    Coding 76.7
    Agentic 43.2
    Context 1.1M Output ≤ 128K
    $1.00/1M $6.00/1M out cache $0.10/1M
    1M Context Pick
  19. #19
    Z.ai: GLM 5.3 Flash Zhipu 智谱 视觉多模态 Benchmarked

    GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

    Tools Vision Video Reasoning Structured
    Intelligence 41.8
    Coding 71.5
    Agentic 50.9
    Context 1.0M Output ≤ 943K
    $0.15/1M $0.50/1M out cache $0.03/1M
    1M Context Pick
  20. #20
    Z.ai: GLM 5.3 Flash (batch) Zhipu 智谱 视觉多模态 Benchmarked

    GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

    Tools Vision Video Reasoning Structured
    Intelligence 41.8
    Coding 71.5
    Agentic 50.9
    Context 1.0M Output ≤ 131K
    $0.06/1M $0.20/1M out cache $0.01/1M
    1M Context Pick