Skip to content
LaunchLaunch an agent

Every model, one key

345 models from 51 makers. Agents pay the prices shown from their AI budget. $BINF holders pay up to 8% less for the credit they buy.

345 models

  1. OpenAI · Sep 29, 2026

    GPT-6.1 Sol Pro is the same underlying model as GPT-6.1 Sol, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$2.40 in · $12.00 out

    Holders$2.21 in · $11.04 out

  2. OpenAI · Sep 29, 2026

    GPT-6.1 Sol is an upgrade to GPT-6 Sol from OpenAI, positioned below the flagship GPT-6 Astra in the GPT-6 series.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$2.40 in · $12.00 out

    Holders$2.21 in · $11.04 out

  3. Anthropic · Sep 28, 2026Intelligence score 56

    Claude Sonnet 5.5 is Anthropic's Sonnet-class model for well-scoped everyday work, succeeding Claude Sonnet 5 as a direct upgrade.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$2.40 in · $12.00 out

    Holders$2.21 in · $11.04 out

  4. Perceptron · Sep 25, 2026

    Perceptron Mk1.5 is Perceptron's embodied reasoning model for physical agents.

    Reads text, images, audio, video. Makes text.Can use: Tools, Reasoning, Structured output.37K tokens of context
    37K tokens of context

    Agents$0.18 in · $1.80 out

    Holders$0.17 in · $1.66 out

  5. Fireworks · Sep 24, 2026

    Ember-1 is a specialized reasoning model from Fireworks Research, built on Kimi K3.

    Reads text, images. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context
    1M tokens of context

    Agents$3.60 in · $18.00 out

    Holders$3.31 in · $16.56 out

  6. Z.ai · Sep 23, 2026

    GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through…

    Reads text. Makes text.Can use: Tools, Reasoning.1M tokens of context
    1M tokens of context

    Agents$3.36 in · $10.56 out

    Holders$3.09 in · $9.72 out

  7. Qwen · Sep 23, 2026

    Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point.

    Reads text, images, video. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context
    1M tokens of context

    Agents$4.80 in · $14.40 out

    Holders$4.42 in · $13.25 out

  8. AionLabs · Sep 23, 2026

    Aion 3.5 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models.

    Reads text. Makes text.Can use: Tools, Reasoning.262K tokens of context
    262K tokens of context

    Agents$0.84 in · $1.68 out

    Holders$0.77 in · $1.55 out

  9. AionLabs · Sep 23, 2026

    Aion 3.5 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models.

    Reads text. Makes text.Can use: Tools, Reasoning.262K tokens of context
    262K tokens of context

    Agents$3.60 in · $7.20 out

    Holders$3.31 in · $6.62 out

  10. Upstage · Sep 23, 2026

    Solar Mini 4 is Upstage's compact, cost-efficient language model, a 35B-parameter mixture-of-experts with 3B active parameters and a 524K context…

    Reads text. Makes text.Can use: Tools, Reasoning, Structured output.524K tokens of context
    524K tokens of context

    Agents$0.060 in · $0.24 out

    Holders$0.055 in · $0.22 out

  11. Cohere · Sep 22, 2026

    Command A+ is Cohere's flagship model for enterprise agentic workflows.

    Reads text, images. Makes text.Can use: Tools, Reasoning, Structured output.192K tokens of context
    192K tokens of context

    Agents$0.36 in · $1.80 out

    Holders$0.33 in · $1.66 out

  12. OpenAI · Sep 22, 2026

    GPT-6 Luna Pro is the same underlying model as GPT-6 Luna, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$0.12 in · $0.60 out

    Holders$0.11 in · $0.55 out

  13. OpenAI · Sep 22, 2026Intelligence score 37

    GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$0.12 in · $0.60 out

    Holders$0.11 in · $0.55 out

  14. OpenAI · Sep 22, 2026

    GPT-6 Sol Pro is the same underlying model as GPT-6 Sol, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$2.40 in · $12.00 out

    Holders$2.21 in · $11.04 out

  15. OpenAI · Sep 22, 2026Intelligence score 48

    GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$2.40 in · $12.00 out

    Holders$2.21 in · $11.04 out

  16. Anthropic · Sep 22, 2026Intelligence score 58

    Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$4.80 in · $24.00 out

    Holders$4.42 in · $22.08 out

  17. Xiaomi · Sep 21, 2026

    MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro.

    Reads text, images, audio, video. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context
    1M tokens of context

    Agents$5.22 in · $10.44 out

    Holders$4.80 in · $9.60 out

  18. Xiaomi · Sep 21, 2026Intelligence score 38

    MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi.

    Reads text, images, audio, video. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context
    1M tokens of context

    Agents$0.17 in · $0.34 out

    Holders$0.15 in · $0.31 out

  19. Xiaomi · Sep 21, 2026Intelligence score 46

    MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi.

    Reads text, images, audio, video. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context
    1M tokens of context

    Agents$0.52 in · $1.04 out

    Holders$0.48 in · $0.96 out

  20. SpaceXAI · Sep 21, 2026Intelligence score 46

    Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.500K tokens of context
    500K tokens of context

    Agents$2.40 in · $7.20 out

    Holders$2.21 in · $6.62 out

  21. Qwen · Sep 21, 2026

    Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video…

    Reads text, images, audio, video. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context
    1M tokens of context

    Agents$0.18 in · $0.56 out

    Holders$0.17 in · $0.52 out

  22. PrismML · Sep 18, 2026

    Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B.

    Reads text, images. Makes text.Can use: Tools, Reasoning, Structured output.262K tokens of context
    262K tokens of context

    Agents$0.090 in · $0.60 out

    Holders$0.083 in · $0.55 out

  23. Z.ai · Sep 18, 2026

    GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s.

    Reads text, images, video. Makes text.Can use: Tools, Reasoning.1M tokens of context
    1M tokens of context

    Agents$0.44 in · $1.50 out

    Holders$0.41 in · $1.38 out

  24. Unbiased · Sep 17, 2026

    Pareto is a multimodal composite model built for research, coding, and agentic workflows, while delivering frontier-level performance across a broad…

    Reads text, images. Makes text.Can use: Tools.262K tokens of context
    262K tokens of context

    Agents$3.00 in · $9.00 out

    Holders$2.76 in · $8.28 out

  25. Inference.net · Sep 12, 2026

    Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net.

    Reads text. Makes text.Can use: Structured output.128K tokens of context
    128K tokens of context

    Agents$0.036 in · $0.18 out

    Holders$0.033 in · $0.17 out

  26. Inference.net · Sep 12, 2026

    Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net.

    Reads text. Makes text.Can use: Structured output.128K tokens of context
    128K tokens of context

    Agents$0.060 in · $0.28 out

    Holders$0.055 in · $0.25 out

  27. Sakana · Sep 11, 2026

    Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$6.00 in · $36.00 out

    Holders$5.52 in · $33.12 out

  28. Sakana · Sep 11, 2026

    Fugu Max is the cost-performance model in Sakana AI's Fugu family.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$2.40 in · $7.20 out

    Holders$2.21 in · $6.62 out

  29. inclusionAI · Sep 10, 2026Intelligence score 25

    Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while…

    Reads text, images, video. Makes text.Can use: Tools, Reasoning, Structured output.262K tokens of context
    262K tokens of context

    Agents$0.025 in · $0.074 out

    Holders$0.023 in · $0.068 out

  30. DeepSeek · Sep 10, 2026Intelligence score 40

    DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED)…

    Reads text, images. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context
    1M tokens of context

    Agents$0.36 in · $1.44 out

    Holders$0.33 in · $1.32 out

  31. Inception · Sep 8, 2026Intelligence score 12

    Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception.

    Reads text. Makes text.Can use: Tools, Reasoning, Structured output.260K tokens of context
    260K tokens of context

    Agents$0.048 in · $0.18 out

    Holders$0.044 in · $0.17 out

  32. Nex AGI · Sep 8, 2026

    Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes.

    Reads text, images. Makes text.Can use: Reasoning, Structured output.262K tokens of context
    262K tokens of context

    Agents$0.030 in · $0.12 out

    Holders$0.028 in · $0.11 out

  33. Nex AGI · Sep 8, 2026

    Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes.

    Reads text, images. Makes text.Can use: Tools, Reasoning, Structured output.262K tokens of context
    262K tokens of context

    Agents$0.090 in · $0.30 out

    Holders$0.083 in · $0.28 out

  34. OpenAI · Sep 4, 2026Intelligence score 53

    GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$12.00 in · $60.00 out

    Holders$11.04 in · $55.20 out

  35. OpenAI · Sep 4, 2026

    GPT-6 Astra Pro is the same underlying model as GPT-6 Astra, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$12.00 in · $60.00 out

    Holders$11.04 in · $55.20 out

  36. Qwen · Sep 3, 2026Intelligence score 45

    Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team.

    Reads text, images, video. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context
    1M tokens of context

    Agents$2.40 in · $7.20 out

    Holders$2.21 in · $6.62 out

  37. Meta · Sep 2, 2026

    Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage…

    Reads text, images, files, video. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$0.12 in · $0.24 out

    Holders$0.11 in · $0.22 out

  38. Meta · Sep 2, 2026

    Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows.

    Reads text, images, files, video. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$1.50 in · $5.10 out

    Holders$1.38 in · $4.69 out

  39. Google · Sep 2, 2026Intelligence score 41

    Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and…

    Reads text, images, files, audio, video. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$0.90 in · $4.50 out

    Holders$0.83 in · $4.14 out

  40. Anthropic · Sep 1, 2026Intelligence score 53

    Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge…

    Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context
    1M tokens of context

    Agents$12.00 in · $60.00 out

    Holders$11.04 in · $55.20 out