Overall LLM Rankings
ZICQ LLM rankings update daily across intelligence, coding, price, and context for selection and comparison.
Weighted composite of OpenRouter benchmarks Intelligence / Coding / Agentic (60% / 25% / 15%), covering all 338+ models.
Last updated: 15 min ago (2026-10-03 15:30)-
#1
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Tools Vision Docs Reasoning StructuredIntelligence 53.4Coding 81.6Agentic 57.9Context 1M Output ≤ 128K$10.00/1M $50.00/1M out cache $0.25/1MTOP Flagship -
#2
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Tools Vision Docs Reasoning StructuredIntelligence 53.4Coding 81.6Agentic 57.9Context 1M Output ≤ 128K$5.00/1M $25.00/1M out cache $0.12/1MTOP 3 Overall -
#3
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
Tools Vision Docs Reasoning StructuredIntelligence 52.7Coding 76.9Agentic 51.0Context 1.1M Output ≤ 128K$10.00/1M $50.00/1M out cache $1.00/1MTOP 3 Overall -
#4
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
Tools Vision Docs Reasoning StructuredIntelligence 52.7Coding 76.9Agentic 51.0Context 1.1M Output ≤ 128K$5.00/1M $25.00/1M out cache $0.50/1M1M Context Pick -
#5
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Tools Vision Docs Reasoning StructuredIntelligence 50.8Coding 78.0Agentic 56.5Context 1M Output ≤ 128K$5.00/1M $25.00/1M out cache $0.50/1M1M Context Pick -
#6
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Tools Vision Docs Reasoning StructuredIntelligence 50.8Coding 78.0Agentic 56.5Context 1M Output ≤ 128K$2.50/1M $12.50/1M out cache $0.25/1M1M Context Pick -
#7
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Tools Vision Docs Reasoning StructuredIntelligence 49.6Coding 76.5Agentic 50.7Context 1M Output ≤ 128K$10.00/1M $50.00/1M out cache $1.00/1M1M Context Pick -
#8
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Tools Vision Docs Reasoning StructuredIntelligence 49.6Coding 76.5Agentic 50.7Context 1M Output ≤ 128K$5.00/1M $25.00/1M out cache $0.50/1M1M Context Pick -
#9
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
Tools Vision Docs Reasoning StructuredIntelligence 47.0Coding 77.4Agentic 50.2Context 1.1M Output ≤ 128K$2.00/1M $10.00/1M out cache $0.20/1M1M Context Pick -
#10
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
Tools Vision Docs Reasoning StructuredIntelligence 47.0Coding 77.4Agentic 50.2Context 1.1M Output ≤ 128K$1.00/1M $5.00/1M out cache $0.10/1M1M Context Pick -
#11
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
Tools Vision Video Reasoning StructuredIntelligence 45.4Coding 76.2Agentic 56.0Context 1M Output ≤ 131K$2.00/1M $6.00/1M out cache $0.25/1M1M Context Pick -
#12
Grok 4.6 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM. It is succeeded by [Grok 4.7](/x-ai/grok-4.7).
Tools Vision Docs Reasoning StructuredIntelligence 44.3Coding 76.8Agentic 53.0Context 500K Output ≤ 450K$2.00/1M $6.00/1M out cache $0.50/1MRecommended -
#13
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
Tools Reasoning StructuredIntelligence 44.8Coding 74.8Agentic 53.1Context 1.0M Output ≤ 131K$1.40/1M $4.40/1M out cache $0.14/1M1M Context Pick -
#14
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
Tools Reasoning StructuredIntelligence 44.8Coding 74.8Agentic 53.1Context 1.0M Output ≤ 131K$0.45/1M $2.00/1M out cache $0.10/1M1M Context Pick -
#15
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
Tools Vision Video Reasoning StructuredIntelligence 43.6Coding 76.2Agentic 50.0Context 1.0M Output ≤ 943K$2.70/1M $13.50/1M out cache $0.27/1M1M Context Pick -
#16
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
Tools Vision Video Reasoning StructuredIntelligence 43.6Coding 76.2Agentic 50.0Context 1.0M Output ≤ 16K$2.28/1M $11.40/1M out cache $0.23/1M1M Context Pick -
#17
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
Tools Vision Docs Reasoning StructuredIntelligence 42.1Coding 76.7Agentic 43.2Context 1.1M Output ≤ 128K$2.00/1M $12.00/1M out cache $0.20/1M1M Context Pick -
#18
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
Tools Vision Docs Reasoning StructuredIntelligence 42.1Coding 76.7Agentic 43.2Context 1.1M Output ≤ 128K$1.00/1M $6.00/1M out cache $0.10/1M1M Context Pick -
#19
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Tools Vision Video Reasoning StructuredIntelligence 41.8Coding 71.5Agentic 50.9Context 1.0M Output ≤ 943K$0.15/1M $0.50/1M out cache $0.03/1M1M Context Pick -
#20
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Tools Vision Video Reasoning StructuredIntelligence 41.8Coding 71.5Agentic 50.9Context 1.0M Output ≤ 131K$0.06/1M $0.20/1M out cache $0.01/1M1M Context Pick