Overall LLM Rankings
ZICQ LLM rankings update daily across intelligence, coding, price, and context for selection and comparison.
Weighted composite of OpenRouter benchmarks Intelligence / Coding / Agentic (60% / 25% / 15%), covering all 338+ models.
Last updated: 29 min ago (2026-10-04 19:36)-
#81
Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering tasks, verifying its own work, and...
Tools Vision Docs Reasoning StructuredIntelligence 46.4Coding 0.0Agentic 0.0Context 500K Output ≤ 450K$2.00/1M $6.00/1M out cache $0.50/1MRecommended -
#82
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Tools Vision Docs Reasoning StructuredIntelligence 20.7Coding 52.1Agentic 15.8Context 1M Output ≤ 64K$3.00/1M $15.00/1M out cache $0.30/1M1M Context Pick -
#83
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Tools Vision Docs Reasoning StructuredIntelligence 20.7Coding 52.1Agentic 15.8Context 1M Output ≤ 64K$1.50/1M $7.50/1M out cache $0.15/1M1M Context Pick -
#84
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Tools Vision Docs Reasoning StructuredIntelligence 24.9Coding 42.2Agentic 15.5Context 1M Output ≤ 900K$1.25/1M $2.50/1M out cache $0.20/1M1M Context Pick -
#85
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Tools Vision Docs Reasoning StructuredIntelligence 24.9Coding 42.2Agentic 15.5Context 1M Output ≤ 900K$1.00/1M $2.00/1M out cache $0.16/1M1M Context Pick -
#86
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
Tools Vision Audio Video Docs Reasoning StructuredIntelligence 22.2Coding 49.3Agentic 14.3Context 1.0M Output ≤ 65K$0.30/1M $2.50/1M out cache $0.03/1M1M Context Pick -
#87
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
Tools Vision Audio Video Docs Reasoning StructuredIntelligence 22.2Coding 49.3Agentic 14.3Context 1.0M Output ≤ 65K$0.15/1M $1.25/1M out cache $0.01/1M1M Context Pick -
#88
MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi. Built at a scale of over 1T parameters, it is designed to push the ceiling of capability for the most demanding...
Tools Vision Audio Video Reasoning StructuredIntelligence 46.3Coding 0.0Agentic 0.0Context 1.1M Output ≤ 131K$0.43/1M $0.87/1M out cache $0.00/1M1M Context Pick -
#89
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Tools Reasoning StructuredIntelligence 20.1Coding 50.6Agentic 19.3Context 262K Output ≤ 32K$0.02/1M $0.06/1M out cache $0.00/1MHigh Value Pick -
#90
LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic...
Tools ReasoningIntelligence 19.1Coding 45.3Agentic 14.0Context 1.0M Output ≤ 262K$0.30/1M $1.20/1M out cache $0.01/1M1M Context Pick -
#91
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
Tools Vision Video Reasoning StructuredIntelligence 18.4Coding 48.2Agentic 8.3Context 262K Output ≤ 235K$0.55/1M $3.50/1M out cache $0.22/1MRecommended -
#92
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
Tools Vision Docs Reasoning StructuredIntelligence 0.0Coding 73.6Agentic 38.6Context 1M Output ≤ 128K$5.00/1M $25.00/1M out cache $0.50/1M1M Context Pick -
#93
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
Tools Vision Docs Reasoning StructuredIntelligence 0.0Coding 73.6Agentic 38.6Context 1M Output ≤ 128K$2.50/1M $12.50/1M out cache $0.25/1M1M Context Pick -
#94
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
Tools Vision Reasoning StructuredIntelligence 39.5Coding 0.0Agentic 0.0Context 1.0M Output ≤ 943K$0.00/1M $2.40/1M out cache $0.00/1M1M Context Pick -
#95
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
Tools Vision Reasoning StructuredIntelligence 39.5Coding 0.0Agentic 0.0Context 1.0M Output ≤ 131K$0.11/1M $0.34/1M out cache $0.00/1M1M Context Pick -
#96
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
Tools Vision Video Reasoning StructuredIntelligence 18.2Coding 41.9Agentic 13.1Context 262K Output ≤ 235K$0.15/1M $1.00/1M out cache $0.05/1MHigh Value Pick -
#97
GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...
Tools Vision Docs Reasoning StructuredIntelligence 38.1Coding 0.0Agentic 0.0Context 1.1M Output ≤ 128K$0.10/1M $0.50/1M out cache $0.01/1M1M Context Pick -
#98
GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...
Tools Vision Docs Reasoning StructuredIntelligence 38.1Coding 0.0Agentic 0.0Context 1.1M Output ≤ 128K$0.05/1M $0.25/1M out cache $0.01/1M1M Context Pick -
#99
MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for...
Tools Vision Audio Video Reasoning StructuredIntelligence 37.9Coding 0.0Agentic 0.0Context 1.1M Output ≤ 131K$0.14/1M $0.28/1M out cache $0.00/1M1M Context Pick -
#100
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Tools Vision Docs Reasoning StructuredIntelligence 16.9Coding 43.9Agentic 8.0Context 200K Output ≤ 64K$1.00/1M $5.00/1M out cache $0.10/1MRecommended