Every model, one key
345 models from 51 makers. Agents pay the prices shown from their AI budget. $BINF holders pay up to 8% less for the credit they buy.
OpenAI · Sep 29, 2026
GPT-6.1 Sol Pro is the same underlying model as GPT-6.1 Sol, served with reasoning.mode set to pro for higher-quality responses on complex tasks.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$2.40 in · $12.00 out
Holders$2.21 in · $11.04 out
- GPT-6.1 SolNew
OpenAI · Sep 29, 2026
GPT-6.1 Sol is an upgrade to GPT-6 Sol from OpenAI, positioned below the flagship GPT-6 Astra in the GPT-6 series.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$2.40 in · $12.00 out
Holders$2.21 in · $11.04 out
Anthropic · Sep 28, 2026Intelligence score 56
Claude Sonnet 5.5 is Anthropic's Sonnet-class model for well-scoped everyday work, succeeding Claude Sonnet 5 as a direct upgrade.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$2.40 in · $12.00 out
Holders$2.21 in · $11.04 out
Perceptron · Sep 25, 2026
Perceptron Mk1.5 is Perceptron's embodied reasoning model for physical agents.
Reads text, images, audio, video. Makes text.Can use: Tools, Reasoning, Structured output.37K tokens of context37K tokens of contextAgents$0.18 in · $1.80 out
Holders$0.17 in · $1.66 out
- Ember-1New
Fireworks · Sep 24, 2026
Ember-1 is a specialized reasoning model from Fireworks Research, built on Kimi K3.
Reads text, images. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context1M tokens of contextAgents$3.60 in · $18.00 out
Holders$3.31 in · $16.56 out
Z.ai · Sep 23, 2026
GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through…
Reads text. Makes text.Can use: Tools, Reasoning.1M tokens of context1M tokens of contextAgents$3.36 in · $10.56 out
Holders$3.09 in · $9.72 out
Qwen · Sep 23, 2026
Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point.
Reads text, images, video. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context1M tokens of contextAgents$4.80 in · $14.40 out
Holders$4.42 in · $13.25 out
AionLabs · Sep 23, 2026
Aion 3.5 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models.
Reads text. Makes text.Can use: Tools, Reasoning.262K tokens of context262K tokens of contextAgents$0.84 in · $1.68 out
Holders$0.77 in · $1.55 out
- Aion 3.5New
AionLabs · Sep 23, 2026
Aion 3.5 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models.
Reads text. Makes text.Can use: Tools, Reasoning.262K tokens of context262K tokens of contextAgents$3.60 in · $7.20 out
Holders$3.31 in · $6.62 out
- Solar Mini 4New
Upstage · Sep 23, 2026
Solar Mini 4 is Upstage's compact, cost-efficient language model, a 35B-parameter mixture-of-experts with 3B active parameters and a 524K context…
Reads text. Makes text.Can use: Tools, Reasoning, Structured output.524K tokens of context524K tokens of contextAgents$0.060 in · $0.24 out
Holders$0.055 in · $0.22 out
- Command A+New
Cohere · Sep 22, 2026
Command A+ is Cohere's flagship model for enterprise agentic workflows.
Reads text, images. Makes text.Can use: Tools, Reasoning, Structured output.192K tokens of context192K tokens of contextAgents$0.36 in · $1.80 out
Holders$0.33 in · $1.66 out
OpenAI · Sep 22, 2026
GPT-6 Luna Pro is the same underlying model as GPT-6 Luna, served with reasoning.mode set to pro for higher-quality responses on complex tasks.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$0.12 in · $0.60 out
Holders$0.11 in · $0.55 out
- GPT-6 LunaNew
OpenAI · Sep 22, 2026Intelligence score 37
GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$0.12 in · $0.60 out
Holders$0.11 in · $0.55 out
OpenAI · Sep 22, 2026
GPT-6 Sol Pro is the same underlying model as GPT-6 Sol, served with reasoning.mode set to pro for higher-quality responses on complex tasks.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$2.40 in · $12.00 out
Holders$2.21 in · $11.04 out
- GPT-6 SolNew
OpenAI · Sep 22, 2026Intelligence score 48
GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$2.40 in · $12.00 out
Holders$2.21 in · $11.04 out
Anthropic · Sep 22, 2026Intelligence score 58
Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$4.80 in · $24.00 out
Holders$4.42 in · $22.08 out
Xiaomi · Sep 21, 2026
MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro.
Reads text, images, audio, video. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context1M tokens of contextAgents$5.22 in · $10.44 out
Holders$4.80 in · $9.60 out
Xiaomi · Sep 21, 2026Intelligence score 38
MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi.
Reads text, images, audio, video. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context1M tokens of contextAgents$0.17 in · $0.34 out
Holders$0.15 in · $0.31 out
Xiaomi · Sep 21, 2026Intelligence score 46
MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi.
Reads text, images, audio, video. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context1M tokens of contextAgents$0.52 in · $1.04 out
Holders$0.48 in · $0.96 out
- Grok 4.7New
SpaceXAI · Sep 21, 2026Intelligence score 46
Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.500K tokens of context500K tokens of contextAgents$2.40 in · $7.20 out
Holders$2.21 in · $6.62 out
Qwen · Sep 21, 2026
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video…
Reads text, images, audio, video. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context1M tokens of contextAgents$0.18 in · $0.56 out
Holders$0.17 in · $0.52 out
PrismML · Sep 18, 2026
Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B.
Reads text, images. Makes text.Can use: Tools, Reasoning, Structured output.262K tokens of context262K tokens of contextAgents$0.090 in · $0.60 out
Holders$0.083 in · $0.55 out
Z.ai · Sep 18, 2026
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s.
Reads text, images, video. Makes text.Can use: Tools, Reasoning.1M tokens of context1M tokens of contextAgents$0.44 in · $1.50 out
Holders$0.41 in · $1.38 out
- ParetoNew
Unbiased · Sep 17, 2026
Pareto is a multimodal composite model built for research, coding, and agentic workflows, while delivering frontier-level performance across a broad…
Reads text, images. Makes text.Can use: Tools.262K tokens of context262K tokens of contextAgents$3.00 in · $9.00 out
Holders$2.76 in · $8.28 out
Inference.net · Sep 12, 2026
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net.
Reads text. Makes text.Can use: Structured output.128K tokens of context128K tokens of contextAgents$0.036 in · $0.18 out
Holders$0.033 in · $0.17 out
Inference.net · Sep 12, 2026
Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net.
Reads text. Makes text.Can use: Structured output.128K tokens of context128K tokens of contextAgents$0.060 in · $0.28 out
Holders$0.055 in · $0.25 out
Sakana · Sep 11, 2026
Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$6.00 in · $36.00 out
Holders$5.52 in · $33.12 out
- Fugu MaxNew
Sakana · Sep 11, 2026
Fugu Max is the cost-performance model in Sakana AI's Fugu family.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$2.40 in · $7.20 out
Holders$2.21 in · $6.62 out
inclusionAI · Sep 10, 2026Intelligence score 25
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while…
Reads text, images, video. Makes text.Can use: Tools, Reasoning, Structured output.262K tokens of context262K tokens of contextAgents$0.025 in · $0.074 out
Holders$0.023 in · $0.068 out
DeepSeek · Sep 10, 2026Intelligence score 40
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED)…
Reads text, images. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context1M tokens of contextAgents$0.36 in · $1.44 out
Holders$0.33 in · $1.32 out
- Mercury 2.5New
Inception · Sep 8, 2026Intelligence score 12
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception.
Reads text. Makes text.Can use: Tools, Reasoning, Structured output.260K tokens of context260K tokens of contextAgents$0.048 in · $0.18 out
Holders$0.044 in · $0.17 out
Nex AGI · Sep 8, 2026
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes.
Reads text, images. Makes text.Can use: Reasoning, Structured output.262K tokens of context262K tokens of contextAgents$0.030 in · $0.12 out
Holders$0.028 in · $0.11 out
- Nex-N2.5-ProNew
Nex AGI · Sep 8, 2026
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes.
Reads text, images. Makes text.Can use: Tools, Reasoning, Structured output.262K tokens of context262K tokens of contextAgents$0.090 in · $0.30 out
Holders$0.083 in · $0.28 out
- GPT-6 AstraNew
OpenAI · Sep 4, 2026Intelligence score 53
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$12.00 in · $60.00 out
Holders$11.04 in · $55.20 out
OpenAI · Sep 4, 2026
GPT-6 Astra Pro is the same underlying model as GPT-6 Astra, served with reasoning.mode set to pro for higher-quality responses on complex tasks.
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$12.00 in · $60.00 out
Holders$11.04 in · $55.20 out
Qwen · Sep 3, 2026Intelligence score 45
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team.
Reads text, images, video. Makes text.Can use: Tools, Reasoning, Structured output.1M tokens of context1M tokens of contextAgents$2.40 in · $7.20 out
Holders$2.21 in · $6.62 out
Meta · Sep 2, 2026
Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage…
Reads text, images, files, video. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$0.12 in · $0.24 out
Holders$0.11 in · $0.22 out
Meta · Sep 2, 2026
Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows.
Reads text, images, files, video. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$1.50 in · $5.10 out
Holders$1.38 in · $4.69 out
Google · Sep 2, 2026Intelligence score 41
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and…
Reads text, images, files, audio, video. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$0.90 in · $4.50 out
Holders$0.83 in · $4.14 out
Anthropic · Sep 1, 2026Intelligence score 53
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge…
Reads text, images, files. Makes text.Can use: Tools, Reasoning, Structured output, Web search.1M tokens of context1M tokens of contextAgents$12.00 in · $60.00 out
Holders$11.04 in · $55.20 out