跳到主内容
智客 ZICQ

综合大模型榜

智客 ZICQ 大模型榜单每日更新,覆盖智能、编程、价格与上下文,便于选型与比价。

依据 OpenRouter benchmarks 智能指数 / 编码 / 智能体加权综合排序(60% / 25% / 15%),覆盖全部 338+ 个模型。

最近更新:6 分钟前(2026-10-04 18:36)
当前榜单模型数 466
免费可调 22
支持视觉 295
支持工具调用 398
支持推理 333
最长上下文 Auto Router (Beta) 2M tokens
平均价格 (输入) — 美元 / 百万 token
最低价模型 DeepSeek: DeepSeek V4.1 Flash $0.003 / 1M
当前维度 TOP · 综合智能分 Anthropic: Claude Opus 5.5 57.6
清除
  1. #21
    Anthropic: Claude Opus 4.8 Anthropic 文档视觉模型 有基准

    Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

    工具 视觉 文档 推理 结构化
    智能 41.8
    编码 74.3
    智能体 41.9
    上下文 1M 输出 ≤ 128K
    $5.00/1M $25.00/1M 输出 cache $0.50/1M
    百万上下文极客首选
  2. #22
    Anthropic: Claude Opus 4.8 (batch) Anthropic 文档视觉模型 有基准

    Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

    工具 视觉 文档 推理 结构化
    智能 41.8
    编码 74.3
    智能体 41.9
    上下文 1M 输出 ≤ 128K
    $2.50/1M $12.50/1M 输出 cache $0.25/1M
    百万上下文极客首选
  3. #23
    Google: Gemini 3.8 Flash Google DeepMind 音视频多模态 有基准

    Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

    工具 视觉 音频 视频 文档 推理 结构化
    智能 40.9
    编码 76.3
    智能体 40.2
    上下文 1.0M 输出 ≤ 65K
    $0.75/1M $3.75/1M 输出 cache $0.07/1M
    百万上下文极客首选
  4. #24
    Google: Gemini 3.8 Flash (batch) Google DeepMind 音视频多模态 有基准

    Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

    工具 视觉 音频 视频 文档 推理 结构化
    智能 40.9
    编码 76.3
    智能体 40.2
    上下文 1.0M 输出 ≤ 65K
    $0.38/1M $1.88/1M 输出 cache $0.04/1M
    百万上下文极客首选
  5. #25
    Qwen: Qwen3.8 2.4T A95B Alibaba 通义 通用大语言模型 有基准

    Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

    工具 推理 结构化
    智能 39.9
    编码 71.9
    智能体 50.1
    上下文 1.0M 输出 ≤ 131K
    $2.00/1M $6.00/1M 输出 cache $0.25/1M
    百万上下文极客首选
  6. #26
    Meta: Muse Spark 1.2 Meta 文档视觉模型 有基准

    Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, and PDF documents, returns text, and offers a 1M-token context window....

    工具 视觉 视频 文档 推理 结构化
    智能 39.6
    编码 72.2
    智能体 43.2
    上下文 1.0M 输出 ≤ 943K
    $1.25/1M $4.25/1M 输出 cache $0.15/1M
    百万上下文极客首选
  7. #27
    Google: Gemini 3.7 Flash Google DeepMind 音视频多模态 有基准

    Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

    工具 视觉 音频 视频 文档 推理 结构化
    智能 39.1
    编码 76.1
    智能体 35.2
    上下文 1.0M 输出 ≤ 65K
    $0.75/1M $3.75/1M 输出 cache $0.07/1M
    百万上下文极客首选
  8. #28
    Google: Gemini 3.7 Flash (batch) Google DeepMind 音视频多模态 有基准

    Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

    工具 视觉 音频 视频 文档 推理 结构化
    智能 39.1
    编码 76.1
    智能体 35.2
    上下文 1.0M 输出 ≤ 65K
    $0.38/1M $1.88/1M 输出 cache $0.04/1M
    百万上下文极客首选
  9. #29
    SpaceXAI: Grok 4.5 xAI 文档视觉模型 有基准

    Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

    工具 视觉 文档 推理 结构化
    智能 38.8
    编码 72.4
    智能体 41.2
    上下文 500K 输出 ≤ 450K
    $2.00/1M $6.00/1M 输出 cache $0.30/1M
    推荐实时模型
  10. #30
    Anthropic: Claude Sonnet 5 Anthropic 文档视觉模型 有基准

    Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

    工具 视觉 文档 推理 结构化
    智能 38.2
    编码 71.5
    智能体 43.6
    上下文 1M 输出 ≤ 128K
    $2.00/1M $10.00/1M 输出 cache $0.20/1M
    百万上下文极客首选
  11. #31
    Anthropic: Claude Sonnet 5 (batch) Anthropic 文档视觉模型 有基准

    Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

    工具 视觉 文档 推理 结构化
    智能 38.2
    编码 71.5
    智能体 43.6
    上下文 1M 输出 ≤ 128K
    $1.00/1M $5.00/1M 输出 cache $0.10/1M
    百万上下文极客首选
  12. #32
    OpenAI: GPT-5.5 OpenAI 文档视觉模型 有基准

    GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

    工具 视觉 文档 推理 结构化
    智能 38.4
    编码 74.9
    智能体 36.4
    上下文 1.1M 输出 ≤ 128K
    $5.00/1M $30.00/1M 输出 cache $0.50/1M
    百万上下文极客首选
  13. #33
    OpenAI: GPT-5.5 (batch) OpenAI 文档视觉模型 有基准

    GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

    工具 视觉 文档 推理 结构化
    智能 38.4
    编码 74.9
    智能体 36.4
    上下文 1.1M 输出 ≤ 128K
    $2.50/1M $15.00/1M 输出 cache $0.25/1M
    百万上下文极客首选
  14. #34
    OpenAI: GPT-5.6 Luna OpenAI 文档视觉模型 有基准

    GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

    工具 视觉 文档 推理 结构化
    智能 37.3
    编码 71.4
    智能体 42.1
    上下文 1.1M 输出 ≤ 128K
    $0.20/1M $1.20/1M 输出 cache $0.02/1M
    百万上下文极客首选
  15. #35
    OpenAI: GPT-5.6 Luna (batch) OpenAI 文档视觉模型 有基准

    GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

    工具 视觉 文档 推理 结构化
    智能 37.3
    编码 71.4
    智能体 42.1
    上下文 1.1M 输出 ≤ 128K
    $0.10/1M $0.60/1M 输出 cache $0.01/1M
    百万上下文极客首选
  16. #36
    DeepSeek: DeepSeek V4 Pro 0813 DeepSeek 通用大语言模型 有基准

    DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

    工具 推理 结构化
    智能 36.0
    编码 68.8
    智能体 41.3
    上下文 1.0M 输出 ≤ 943K
    $0.85/1M $5.00/1M 输出 cache $0.70/1M
    百万上下文极客首选
  17. #37
    Qwen: Qwen3.8 27B Alibaba 通义 视觉多模态 有基准

    Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

    工具 视觉 视频 推理 结构化
    智能 33.7
    编码 68.1
    智能体 45.8
    上下文 1M 输出 ≤ 131K
    $0.42/1M $2.55/1M 输出 cache $0.08/1M
    百万上下文极客首选
  18. #38
    Qwen: Qwen3.8 27B (free) Alibaba 通义 视觉多模态 免费 有基准

    Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

    工具 视觉 视频 推理 结构化
    智能 33.7
    编码 68.1
    智能体 45.8
    上下文 262K 输出 ≤ 235K
    免费
    免费可调 · 限时白嫖
  19. #39
    DeepSeek: DeepSeek V4 Flash 0731 DeepSeek 通用大语言模型 有基准

    DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

    工具 推理 结构化
    智能 34.3
    编码 69.1
    智能体 41.0
    上下文 1.0M 输出 ≤ 943K
    $0.02/1M $1.28/1M 输出 cache $0.02/1M
    百万上下文极客首选
  20. #40
    Z.ai: GLM 5.2 Zhipu 智谱 通用大语言模型 有基准

    GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

    工具 推理 结构化
    智能 33.7
    编码 68.8
    智能体 38.4
    上下文 1.0M 输出 ≤ 943K
    $0.38/1M $3.49/1M 输出 cache $0.26/1M
    百万上下文极客首选