Skip to main content
ZICQ

Tool Calling (Function Call)

ZICQ LLM rankings update daily across intelligence, coding, price, and context for selection and comparison.

Models declaring tools / tool_choice support, ranked by Agentic index plus a tool bonus.

Last updated: 25 min ago (2026-10-03 14:30)
Models tracked 466
Free & tunable 22
Vision 295
Tool calling 398
Reasoning 333
Max context Auto Router (Beta) 2M tokens
Avg price (input) — USD / 1M tokens
Cheapest model IBM: Granite 4.0 Micro $0.017 / 1M
TOP by Tool bonus Anthropic: Claude Opus 5.5 $-57.600
Clear
  1. #1
    Anthropic: Claude Fable 5.1 Anthropic Doc & Vision Benchmarked

    Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

    Tools Vision Docs Reasoning Structured
    Context 1M Output ≤ 128K
    $10.00/1M $50.00/1M out cache $0.25/1M
    Tool Calling Expert
  2. #2
    Anthropic: Claude Fable 5.1 (batch) Anthropic Doc & Vision Benchmarked

    Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

    Tools Vision Docs Reasoning Structured
    Context 1M Output ≤ 128K
    $5.00/1M $25.00/1M out cache $0.12/1M
    Stable Function Calling
  3. #3
    Anthropic: Claude Opus 5 Anthropic Doc & Vision Benchmarked

    Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

    Tools Vision Docs Reasoning Structured
    Context 1M Output ≤ 128K
    $5.00/1M $25.00/1M out cache $0.50/1M
    Stable Function Calling
  4. #4
    Anthropic: Claude Opus 5 (batch) Anthropic Doc & Vision Benchmarked

    Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

    Tools Vision Docs Reasoning Structured
    Context 1M Output ≤ 128K
    $2.50/1M $12.50/1M out cache $0.25/1M
    Stable Function Calling
  5. #5
    Qwen: Qwen3.8 Max (0902) Alibaba 通义 视觉多模态 Benchmarked

    Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

    Tools Vision Video Reasoning Structured
    Context 1M Output ≤ 131K
    $2.00/1M $6.00/1M out cache $0.25/1M
    Stable Function Calling
  6. #6
    Z.ai: GLM 5.3 Zhipu 智谱 General LLM Benchmarked

    GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

    Tools Reasoning Structured
    Context 1.0M Output ≤ 131K
    $1.40/1M $4.40/1M out cache $0.14/1M
    Stable Function Calling
  7. #7
    Z.ai: GLM 5.3 (batch) Zhipu 智谱 General LLM Benchmarked

    GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

    Tools Reasoning Structured
    Context 1.0M Output ≤ 131K
    $0.45/1M $2.00/1M out cache $0.10/1M
    Stable Function Calling
  8. #8
    SpaceXAI: Grok 4.6 xAI Doc & Vision Benchmarked

    Grok 4.6 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM. It is succeeded by [Grok 4.7](/x-ai/grok-4.7).

    Tools Vision Docs Reasoning Structured
    Context 500K Output ≤ 450K
    $2.00/1M $6.00/1M out cache $0.50/1M
    Stable Function Calling
  9. #9
    OpenAI: GPT-6 Astra OpenAI Doc & Vision Benchmarked

    GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

    Tools Vision Docs Reasoning Structured
    Context 1.1M Output ≤ 128K
    $10.00/1M $50.00/1M out cache $1.00/1M
    Stable Function Calling
  10. #10
    OpenAI: GPT-6 Astra (batch) OpenAI Doc & Vision Benchmarked

    GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

    Tools Vision Docs Reasoning Structured
    Context 1.1M Output ≤ 128K
    $5.00/1M $25.00/1M out cache $0.50/1M
    Stable Function Calling
  11. #11
    Z.ai: GLM 5.3 Flash Zhipu 智谱 视觉多模态 Benchmarked

    GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

    Tools Vision Video Reasoning Structured
    Context 1.0M Output ≤ 943K
    $0.15/1M $0.50/1M out cache $0.03/1M
    Stable Function Calling
  12. #12
    Z.ai: GLM 5.3 Flash (batch) Zhipu 智谱 视觉多模态 Benchmarked

    GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

    Tools Vision Video Reasoning Structured
    Context 1.0M Output ≤ 131K
    $0.06/1M $0.20/1M out cache $0.01/1M
    Stable Function Calling
  13. #13
    Anthropic: Claude Fable 5 Anthropic Doc & Vision Benchmarked

    Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

    Tools Vision Docs Reasoning Structured
    Context 1M Output ≤ 128K
    $10.00/1M $50.00/1M out cache $1.00/1M
    Stable Function Calling
  14. #14
    Anthropic: Claude Fable 5 (batch) Anthropic Doc & Vision Benchmarked

    Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

    Tools Vision Docs Reasoning Structured
    Context 1M Output ≤ 128K
    $5.00/1M $25.00/1M out cache $0.50/1M
    Stable Function Calling
  15. #15
    OpenAI: GPT-5.6 Sol OpenAI Doc & Vision Benchmarked

    GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

    Tools Vision Docs Reasoning Structured
    Context 1.1M Output ≤ 128K
    $2.00/1M $10.00/1M out cache $0.20/1M
    Stable Function Calling
  16. #16
    OpenAI: GPT-5.6 Sol (batch) OpenAI Doc & Vision Benchmarked

    GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

    Tools Vision Docs Reasoning Structured
    Context 1.1M Output ≤ 128K
    $1.00/1M $5.00/1M out cache $0.10/1M
    Stable Function Calling
  17. #17
    Qwen: Qwen3.8 2.4T A95B Alibaba 通义 General LLM Benchmarked

    Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

    Tools Reasoning Structured
    Context 1.0M Output ≤ 131K
    $2.00/1M $6.00/1M out cache $0.25/1M
    Stable Function Calling
  18. #18
    MoonshotAI: Kimi K3 Moonshotai 视觉多模态 Benchmarked

    Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

    Tools Vision Video Reasoning Structured
    Context 1.0M Output ≤ 943K
    $2.70/1M $13.50/1M out cache $0.27/1M
    Stable Function Calling
  19. #19
    MoonshotAI: Kimi K3 (batch) Moonshotai 视觉多模态 Benchmarked

    Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

    Tools Vision Video Reasoning Structured
    Context 1.0M Output ≤ 16K
    $2.28/1M $11.40/1M out cache $0.23/1M
    Stable Function Calling
  20. #20
    Qwen: Qwen3.8 27B Alibaba 通义 视觉多模态 Benchmarked

    Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

    Tools Vision Video Reasoning Structured
    Context 1M Output ≤ 131K
    $0.42/1M $3.00/1M out cache $0.08/1M
    Stable Function Calling