综合大模型榜
智客 ZICQ 大模型榜单每日更新,覆盖智能、编程、价格与上下文,便于选型与比价。
依据 OpenRouter benchmarks 智能指数 / 编码 / 智能体加权综合排序(60% / 25% / 15%),覆盖全部 338+ 个模型。
最近更新:12 分钟前(2026-10-04 18:36)-
#81
Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering tasks, verifying its own work, and...
工具 视觉 文档 推理 结构化智能 46.4编码 0.0智能体 0.0上下文 500K 输出 ≤ 450K$2.00/1M $6.00/1M 输出 cache $0.50/1M推荐实时模型 -
#82
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
工具 视觉 文档 推理 结构化智能 20.7编码 52.1智能体 15.8上下文 1M 输出 ≤ 64K$3.00/1M $15.00/1M 输出 cache $0.30/1M百万上下文极客首选 -
#83
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
工具 视觉 文档 推理 结构化智能 20.7编码 52.1智能体 15.8上下文 1M 输出 ≤ 64K$1.50/1M $7.50/1M 输出 cache $0.15/1M百万上下文极客首选 -
#84
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
工具 视觉 文档 推理 结构化智能 24.9编码 42.2智能体 15.5上下文 1M 输出 ≤ 900K$1.25/1M $2.50/1M 输出 cache $0.20/1M百万上下文极客首选 -
#85
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
工具 视觉 文档 推理 结构化智能 24.9编码 42.2智能体 15.5上下文 1M 输出 ≤ 900K$1.00/1M $2.00/1M 输出 cache $0.16/1M百万上下文极客首选 -
#86
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
工具 视觉 音频 视频 文档 推理 结构化智能 22.2编码 49.3智能体 14.3上下文 1.0M 输出 ≤ 65K$0.30/1M $2.50/1M 输出 cache $0.03/1M百万上下文极客首选 -
#87
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
工具 视觉 音频 视频 文档 推理 结构化智能 22.2编码 49.3智能体 14.3上下文 1.0M 输出 ≤ 65K$0.15/1M $1.25/1M 输出 cache $0.01/1M百万上下文极客首选 -
#88
MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi. Built at a scale of over 1T parameters, it is designed to push the ceiling of capability for the most demanding...
工具 视觉 音频 视频 推理 结构化智能 46.3编码 0.0智能体 0.0上下文 1.1M 输出 ≤ 131K$0.43/1M $0.87/1M 输出 cache $0.00/1M百万上下文极客首选 -
#89
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
工具 推理 结构化智能 20.1编码 50.6智能体 19.3上下文 262K 输出 ≤ 32K$0.02/1M $0.06/1M 输出 cache $0.00/1M高性价比优选 -
#90
LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic...
工具 推理智能 19.1编码 45.3智能体 14.0上下文 1.0M 输出 ≤ 262K$0.30/1M $1.20/1M 输出 cache $0.01/1M百万上下文极客首选 -
#91
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
工具 视觉 视频 推理 结构化智能 18.4编码 48.2智能体 8.3上下文 262K 输出 ≤ 235K$0.55/1M $3.50/1M 输出 cache $0.22/1M推荐实时模型 -
#92
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
工具 视觉 文档 推理 结构化智能 0.0编码 73.6智能体 38.6上下文 1M 输出 ≤ 128K$5.00/1M $25.00/1M 输出 cache $0.50/1M百万上下文极客首选 -
#93
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
工具 视觉 文档 推理 结构化智能 0.0编码 73.6智能体 38.6上下文 1M 输出 ≤ 128K$2.50/1M $12.50/1M 输出 cache $0.25/1M百万上下文极客首选 -
#94
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
工具 视觉 推理 结构化智能 39.5编码 0.0智能体 0.0上下文 1.0M 输出 ≤ 943K$0.00/1M $2.40/1M 输出 cache $0.00/1M百万上下文极客首选 -
#95
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
工具 视觉 推理 结构化智能 39.5编码 0.0智能体 0.0上下文 1.0M 输出 ≤ 131K$0.11/1M $0.34/1M 输出 cache $0.00/1M百万上下文极客首选 -
#96
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
工具 视觉 视频 推理 结构化智能 18.2编码 41.9智能体 13.1上下文 262K 输出 ≤ 235K$0.15/1M $1.00/1M 输出 cache $0.05/1M高性价比优选 -
#97
GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...
工具 视觉 文档 推理 结构化智能 38.1编码 0.0智能体 0.0上下文 1.1M 输出 ≤ 128K$0.10/1M $0.50/1M 输出 cache $0.01/1M百万上下文极客首选 -
#98
GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...
工具 视觉 文档 推理 结构化智能 38.1编码 0.0智能体 0.0上下文 1.1M 输出 ≤ 128K$0.05/1M $0.25/1M 输出 cache $0.01/1M百万上下文极客首选 -
#99
MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for...
工具 视觉 音频 视频 推理 结构化智能 37.9编码 0.0智能体 0.0上下文 1.1M 输出 ≤ 131K$0.14/1M $0.28/1M 输出 cache $0.00/1M百万上下文极客首选 -
#100
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
工具 视觉 文档 推理 结构化智能 16.9编码 43.9智能体 8.0上下文 200K 输出 ≤ 64K$1.00/1M $5.00/1M 输出 cache $0.10/1M推荐实时模型