Overall LLM Rankings
ZICQ LLM rankings update daily across intelligence, coding, price, and context for selection and comparison.
Weighted composite of OpenRouter benchmarks Intelligence / Coding / Agentic (60% / 25% / 15%), covering all 338+ models.
Last updated: 28 min ago (2026-10-03 14:30)-
#11
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
Tools Vision Video Reasoning StructuredIntelligence 45.4Coding 76.2Agentic 56.0Context 1M Output ≤ 131K$2.00/1M $6.00/1M out cache $0.25/1M1M Context Pick -
#25
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
Tools Reasoning StructuredIntelligence 39.9Coding 71.9Agentic 50.1Context 1.0M Output ≤ 131K$2.00/1M $6.00/1M out cache $0.25/1M1M Context Pick -
#37
Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...
Tools Vision Video Reasoning StructuredIntelligence 33.7Coding 68.1Agentic 45.8Context 1M Output ≤ 131K$0.42/1M $3.00/1M out cache $0.08/1M1M Context Pick -
#38
Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...
Tools Vision Video Reasoning StructuredIntelligence 33.7Coding 68.1Agentic 45.8Context 262K Output ≤ 235KFreeFree & Tunable -
#48
Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...
Tools Reasoning StructuredIntelligence 29.5Coding 66.0Agentic 22.5Context 1M Output ≤ 131K$1.48/1M $4.42/1M out cache $0.29/1M1M Context Pick -
#65
Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...
Tools Vision Reasoning StructuredIntelligence 25.2Coding 55.9Agentic 17.5Context 1M Output ≤ 131K$0.32/1M $1.28/1M out cache $0.06/1M1M Context Pick -
#76
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
Tools Vision Video Reasoning StructuredIntelligence 21.4Coding 53.7Agentic 18.5Context 262K Output ≤ 81K$0.32/1M $3.20/1M outRecommended -
#91
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
Tools Vision Video Reasoning StructuredIntelligence 18.4Coding 48.2Agentic 8.3Context 262K Output ≤ 235K$0.55/1M $3.50/1M out cache $0.22/1MRecommended -
#96
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
Tools Vision Video Reasoning StructuredIntelligence 18.2Coding 41.9Agentic 13.1Context 262K Output ≤ 235K$0.15/1M $1.00/1M out cache $0.05/1MHigh Value Pick -
#102
The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...
Tools Vision Video Reasoning StructuredIntelligence 15.6Coding 45.7Agentic 7.6Context 262K Output ≤ 65K$0.26/1M $2.08/1M outRecommended -
#124
Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...
Tools StructuredIntelligence 9.2Coding 36.2Agentic 0.9Context 262K Output ≤ 235K$0.12/1M $0.80/1M out cache $0.07/1MHigh Value Pick -
#126
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
Tools Vision Video Reasoning StructuredIntelligence 11.2Coding 28.7Agentic 1.2Context 262K Output ≤ 32K$0.10/1M $0.15/1M outHigh Value Pick -
#127
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
Tools Vision Video Reasoning StructuredIntelligence 0.0Coding 54.5Agentic 0.0Context 1M Output ≤ 65K$0.33/1M $1.95/1M out1M Context Pick -
#130
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
Tools Reasoning StructuredIntelligence 12.7Coding 22.1Agentic 1.3Context 131K Output ≤ 117K$0.23/1M $2.30/1M outRecommended -
#158
The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...
Tools Vision Video Reasoning StructuredIntelligence 0.0Coding 37.0Agentic 0.0Context 262K Output ≤ 235K$0.15/1M $1.00/1M out cache $0.05/1MHigh Value Pick -
#160
Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal reasoning traces are separated...
Tools Reasoning StructuredIntelligence 9.8Coding 12.1Agentic 0.9Context 81K Output ≤ 32K$0.20/1M $2.40/1M outHigh Value Pick -
#175
Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...
Tools Reasoning StructuredIntelligence 0.0Coding 17.4Agentic 0.0Context 262K Output ≤ 32K$0.15/1M $1.20/1M outHigh Value Pick -
#178
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
Tools Reasoning StructuredIntelligence 0.0Coding 15.3Agentic 0.9Context 131K Output ≤ 16K$0.08/1M $0.28/1M outHigh Value Pick -
#180
Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
Tools Reasoning StructuredIntelligence 0.0Coding 13.8Agentic 0.9Context 131K Output ≤ 16K$0.12/1M $0.24/1M outHigh Value Pick -
#190
Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...
Tools Reasoning StructuredIntelligence 0.0Coding 9.0Agentic 0.8Context 131K Output ≤ 8K$0.12/1M $0.45/1M outHigh Value Pick