跳到正文
文档发射发射 Agent

所有模型,一个密钥

346 个模型,来自 51 家厂商。Agent 按所示价格从 AI 预算中付费。$BINF 持有者购买额度最高优惠 8%。

bInference Router 近 7 天数据

Token 数
119.2M
调用次数
1.4K
用到的模型
20
发起调用的 Agent
4

上线以来共 119.2M Token、1.4K 次调用。只统计数量:从不存储任何提示词或回答。

346 个模型

  1. Anthropic · 2026年9月22日本周 24.9M Token智能评分 58

    Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5.

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。1M Token 上下文
    1M Token 上下文

    Agent输入 $4.80 · 输出 $24.00

    持有者输入 $4.42 · 输出 $22.08

  2. Anthropic · 2026年9月28日本周 20.1M Token智能评分 56

    Claude Sonnet 5.5 is Anthropic's Sonnet-class model for well-scoped everyday work, succeeding Claude Sonnet 5 as a direct upgrade.

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。1M Token 上下文
    1M Token 上下文

    Agent输入 $2.40 · 输出 $12.00

    持有者输入 $2.21 · 输出 $11.04

  3. DeepSeek · 2026年7月31日本周 17.6M Token智能评分 34

    DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total.

    输入:文本。输出:文本。支持:工具调用、推理、结构化输出。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.013 · 输出 $1.54

    持有者输入 $0.012 · 输出 $1.41

  4. OpenAI · 2026年9月4日本周 9.6M Token智能评分 53

    GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work.

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。1M Token 上下文
    1M Token 上下文

    Agent输入 $12.00 · 输出 $60.00

    持有者输入 $11.04 · 输出 $55.20

  5. Mistral · 2026年4月30日本周 7.4M Token智能评分 14

    Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI.

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出。262K Token 上下文
    262K Token 上下文

    Agent输入 $1.80 · 输出 $9.00

    持有者输入 $1.66 · 输出 $8.28

  6. Google · 2026年9月2日本周 7.1M Token智能评分 41

    Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and…

    输入:文本、图像、文件、音频、视频。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.90 · 输出 $4.50

    持有者输入 $0.83 · 输出 $4.14

  7. Qwen · 2026年9月23日本周 6.5M Token

    Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point.

    输入:文本、图像、视频。输出:文本。支持:工具调用、推理、结构化输出。1M Token 上下文
    1M Token 上下文

    Agent输入 $4.80 · 输出 $14.40

    持有者输入 $4.42 · 输出 $13.25

  8. MiniMax · 2026年5月31日本周 4.7M Token智能评分 29

    MiniMax-M3 is a multimodal foundation model from MiniMax.

    输入:文本、图像、视频。输出:文本。支持:工具调用、推理、结构化输出。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.36 · 输出 $1.44

    持有者输入 $0.33 · 输出 $1.32

  9. OpenAI · 2026年9月29日本周 4.7M Token智能评分 52

    GPT-6.1 Sol is an upgrade to GPT-6 Sol from OpenAI, positioned below the flagship GPT-6 Astra in the GPT-6 series.

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。1M Token 上下文
    1M Token 上下文

    Agent输入 $2.40 · 输出 $12.00

    持有者输入 $2.21 · 输出 $11.04

  10. Z.ai · 2026年9月23日本周 4.6M Token

    GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through…

    输入:文本。输出:文本。支持:工具调用、推理。1M Token 上下文
    1M Token 上下文

    Agent输入 $3.36 · 输出 $10.56

    持有者输入 $3.09 · 输出 $9.72

  11. MoonshotAI · 2026年7月16日本周 4M Token智能评分 44

    Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI.

    输入:文本、图像、视频。输出:文本。支持:工具调用、推理、结构化输出。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.80 · 输出 $12.00

    持有者输入 $0.73 · 输出 $11.04

  12. DeepSeek · 2026年8月12日本周 4M Token智能评分 36

    DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek.

    输入:文本。输出:文本。支持:工具调用、推理、结构化输出。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.79 · 输出 $2.38

    持有者输入 $0.73 · 输出 $2.19

  13. Meta · 2025年4月5日本周 2.6M Token

    Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with…

    输入:文本、图像。输出:文本。支持:工具调用、结构化输出。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.23 · 输出 $0.78

    持有者输入 $0.21 · 输出 $0.72

  14. Google · 2026年7月21日本周 688.8K Token智能评分 22

    Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities.

    输入:文本、图像、文件、音频、视频。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.36 · 输出 $3.00

    持有者输入 $0.33 · 输出 $2.76

  15. DeepSeek · 2026年9月10日本周 487.8K Token智能评分 40

    DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED)…

    输入:文本、图像。输出:文本。支持:工具调用、推理、结构化输出。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.036 · 输出 $0.60

    持有者输入 $0.033 · 输出 $0.55

  16. Anthropic · 2025年10月15日本周 80.4K Token智能评分 17

    Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of…

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。200K Token 上下文
    200K Token 上下文

    Agent输入 $1.20 · 输出 $6.00

    持有者输入 $1.10 · 输出 $5.52

  17. OpenAI · 2024年7月18日本周 3.7K Token

    GPT-4o mini is OpenAI's newest model after GPT-4 Omni, supporting both text and image inputs with text outputs.

    输入:文本、图像、文件。输出:文本。支持:工具调用、结构化输出、联网搜索。128K Token 上下文
    128K Token 上下文

    Agent输入 $0.18 · 输出 $0.72

    持有者输入 $0.17 · 输出 $0.66

  18. Anthropic · 2026年9月1日本周 661 Token智能评分 53

    Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge…

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。1M Token 上下文
    1M Token 上下文

    Agent输入 $12.00 · 输出 $60.00

    持有者输入 $11.04 · 输出 $55.20

  19. OpenAI · 2025年8月7日本周 164 Token智能评分 17

    GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks.

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。400K Token 上下文
    400K Token 上下文

    Agent输入 $0.30 · 输出 $2.40

    持有者输入 $0.28 · 输出 $2.21

  20. Unbiased · 2026年10月1日

    Pareto is a multimodal composite model built for research, coding, and agentic workflows, while delivering frontier-level performance across a broad…

    输入:文本、图像。输出:文本。支持:工具调用。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.96 · 输出 $3.84

    持有者输入 $0.88 · 输出 $3.53

  21. OpenAI · 2026年9月29日

    GPT-6.1 Sol Pro is the same underlying model as GPT-6.1 Sol, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。1M Token 上下文
    1M Token 上下文

    Agent输入 $2.40 · 输出 $12.00

    持有者输入 $2.21 · 输出 $11.04

  22. Perceptron · 2026年9月25日

    Perceptron Mk1.5 is Perceptron's embodied reasoning model for physical agents.

    输入:文本、图像、音频、视频。输出:文本。支持:工具调用、推理、结构化输出。37K Token 上下文
    37K Token 上下文

    Agent输入 $0.18 · 输出 $1.80

    持有者输入 $0.17 · 输出 $1.66

  23. Fireworks · 2026年9月24日

    Ember-1 is a specialized reasoning model from Fireworks Research, built on Kimi K3.

    输入:文本、图像。输出:文本。支持:工具调用、推理、结构化输出。1M Token 上下文
    1M Token 上下文

    Agent输入 $3.60 · 输出 $18.00

    持有者输入 $3.31 · 输出 $16.56

  24. AionLabs · 2026年9月23日

    Aion 3.5 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models.

    输入:文本。输出:文本。支持:工具调用、推理。262K Token 上下文
    262K Token 上下文

    Agent输入 $0.84 · 输出 $1.68

    持有者输入 $0.77 · 输出 $1.55

  25. AionLabs · 2026年9月23日

    Aion 3.5 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models.

    输入:文本。输出:文本。支持:工具调用、推理。262K Token 上下文
    262K Token 上下文

    Agent输入 $3.60 · 输出 $7.20

    持有者输入 $3.31 · 输出 $6.62

  26. Upstage · 2026年9月23日智能评分 24

    Solar Mini 4 is Upstage's compact, cost-efficient language model, a 35B-parameter mixture-of-experts with 3B active parameters and a 524K context…

    输入:文本。输出:文本。支持:工具调用、推理、结构化输出。524K Token 上下文
    524K Token 上下文

    Agent输入 $0.060 · 输出 $0.24

    持有者输入 $0.055 · 输出 $0.22

  27. Cohere · 2026年9月22日

    Command A+ is Cohere's flagship model for enterprise agentic workflows.

    输入:文本、图像。输出:文本。支持:工具调用、推理、结构化输出。192K Token 上下文
    192K Token 上下文

    Agent输入 $0.36 · 输出 $1.80

    持有者输入 $0.33 · 输出 $1.66

  28. OpenAI · 2026年9月22日

    GPT-6 Luna Pro is the same underlying model as GPT-6 Luna, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.12 · 输出 $0.60

    持有者输入 $0.11 · 输出 $0.55

  29. OpenAI · 2026年9月22日智能评分 38

    GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol.

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.12 · 输出 $0.60

    持有者输入 $0.11 · 输出 $0.55

  30. OpenAI · 2026年9月22日

    GPT-6 Sol Pro is the same underlying model as GPT-6 Sol, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。1M Token 上下文
    1M Token 上下文

    Agent输入 $2.40 · 输出 $12.00

    持有者输入 $2.21 · 输出 $11.04

  31. OpenAI · 2026年9月22日智能评分 48

    GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier.

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。1M Token 上下文
    1M Token 上下文

    Agent输入 $2.40 · 输出 $12.00

    持有者输入 $2.21 · 输出 $11.04

  32. Xiaomi · 2026年9月21日

    MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro.

    输入:文本、图像、音频、视频。输出:文本。支持:工具调用、推理、结构化输出。1M Token 上下文
    1M Token 上下文

    Agent输入 $5.22 · 输出 $10.44

    持有者输入 $4.80 · 输出 $9.60

  33. Xiaomi · 2026年9月21日智能评分 38

    MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi.

    输入:文本、图像、音频、视频。输出:文本。支持:工具调用、推理、结构化输出。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.17 · 输出 $0.34

    持有者输入 $0.15 · 输出 $0.31

  34. Xiaomi · 2026年9月21日智能评分 46

    MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi.

    输入:文本、图像、音频、视频。输出:文本。支持:工具调用、推理、结构化输出。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.52 · 输出 $1.04

    持有者输入 $0.48 · 输出 $0.96

  35. SpaceXAI · 2026年9月21日智能评分 46

    Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6.

    输入:文本、图像、文件。输出:文本。支持:工具调用、推理、结构化输出、联网搜索。500K Token 上下文
    500K Token 上下文

    Agent输入 $2.40 · 输出 $7.20

    持有者输入 $2.21 · 输出 $6.62

  36. Qwen · 2026年9月21日

    Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video…

    输入:文本、图像、音频、视频。输出:文本。支持:工具调用、推理、结构化输出。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.18 · 输出 $0.56

    持有者输入 $0.17 · 输出 $0.52

  37. PrismML · 2026年9月18日

    Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B.

    输入:文本、图像。输出:文本。支持:工具调用、推理、结构化输出。262K Token 上下文
    262K Token 上下文

    Agent输入 $0.090 · 输出 $0.60

    持有者输入 $0.083 · 输出 $0.55

  38. Z.ai · 2026年9月18日

    GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s.

    输入:文本、图像、视频。输出:文本。支持:工具调用、推理。1M Token 上下文
    1M Token 上下文

    Agent输入 $0.44 · 输出 $1.50

    持有者输入 $0.41 · 输出 $1.38

  39. Unbiased · 2026年9月17日

    Pareto is a multimodal composite model built for research, coding, and agentic workflows, while delivering frontier-level performance across a broad…

    输入:文本、图像。输出:文本。支持:工具调用。262K Token 上下文
    262K Token 上下文

    Agent输入 $3.00 · 输出 $9.00

    持有者输入 $2.76 · 输出 $8.28

  40. Inference.net · 2026年9月12日

    Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net.

    输入:文本。输出:文本。支持:结构化输出。128K Token 上下文
    128K Token 上下文

    Agent输入 $0.036 · 输出 $0.18

    持有者输入 $0.033 · 输出 $0.17