Model pricing

Complete Model Catalog

Full access to all offered AI models and detailed pricing.

Available Models
41+
Models Available
Claude · ChatGPT · DeepSeek · Gemini · MiniMax · Step · Xiaomi MiMo · Doubao Seed · Qwen · GLM · Kimi · Grok (xAI)
Status
Live
Filter: Model weightages explained in FAQ Token-pack pricing is up to 70% cheaper than official rates.
GLM 5.3 Flash
71% Off
Zhipu AI
Model IDglm-5.3-flash
Fast, cost-efficient model for high-volume chat and coding.
Official input
$0.15/M tokens
Official output
$0.50/M tokens
Context128K
VisionNo
Weightage1x
Savings70%
AccessToken packs
GPT-5.6 Sol
New
OpenAI
Model IDgpt-5.6-sol
Frontier reasoning for advanced coding and agent workflows.
Official input
$5.00/M tokens
Official output
$30.00/M tokens
Context1.05M
Weightage10x
AccessToken packs
GPT-5.6 Terra
New
OpenAI
Model IDgpt-5.6-terra
Balanced general reasoning for production workloads.
Official input
$2.50/M tokens
Official output
$15.00/M tokens
Context1.05M
Weightage5x
AccessToken packs
GPT-5.6 Luna
New
OpenAI
Model IDgpt-5.6-luna
Fast, lower-cost GPT-5.6 option for high-volume work.
Official input
$1.00/M tokens
Official output
$6.00/M tokens
Context1.05M
Weightage1x
AccessToken packs
Kimi K3
New
Moonshot AI
Model IDkimi-k3
2.8T parameter flagship for long-horizon coding and knowledge work.
Official input
$3.00/M tokens
Official output
$15.00/M tokens
Context1M
VisionYes
Weightage15x
AccessToken packs
Qwen 3.8 Max
Preview
Alibaba Cloud
Model IDqwen3.8-max
2.4T parameter flagship via Alibaba Cloud Token Plan.
Official input
From $6/mo
Official output
Token Plan
Context1M
VisionYes
Weightage20x
AccessToken packs
Claude Opus
77% Off
Anthropic
Model IDclaude-opus-4-7
Updated Opus. Best usage and pricing synced from the current workbook.
Input
$5.00 $1.14/M tokens
Output
$25.00 $5.71/M tokens
Context 1M
Vision Yes
Weightage 20x
Savings 77%
Claude Opus
77% Off
Anthropic
Model IDclaude-opus-4-6
Top-tier reasoning. Best usage and pricing synced from the current workbook.
Input
$5.00 $1.14/M tokens
Output
$25.00 $5.71/M tokens
Context 1M
Vision Yes
Weightage 20x
Savings 77%
Claude Sonnet
91% Off
Anthropic
Model IDclaude-sonnet-4-6
Best balance. Best usage and pricing synced from the current workbook.
Input
$3.00 $0.23/M tokens
Output
$15.00 $1.14/M tokens
Context 1M
Vision Yes
Weightage 4x
Savings 92%
Gemini 3.1 Pro
86% Off
Google
Model IDgemini-3.1-pro
Flagship Google. Best usage and pricing synced from the current workbook.
Input
$4.00 $0.57/M tokens
Output
$18.00 $2.86/M tokens
Context 1M
Vision Yes
Weightage 10x
Savings 86%
Grok 4.5
86% Off
xAI
Model IDgrok-4.5
xAI flagship. Best usage and pricing synced from the current workbook.
Official input
$0.34/M tokens
Official output
$1.03/M tokens
Context1M
VisionYes
Weightage6x
Savings57%
AccessToken packs
Grok 4.6
83% Off
xAI
Model IDgrok-4.6
Latest xAI model. Best usage and pricing synced from the current workbook.
Official input
$0.57/M tokens
Official output
$1.71/M tokens
Context1M
VisionYes
Weightage10x
Savings29%
AccessToken packs
Claude Opus 5
71% Off
Anthropic
Model IDclaude-opus-5
Latest Opus. Best usage and pricing synced from the current workbook.
Official input
$5.00/M tokens
Official output
$25.00/M tokens
Context1M
VisionYes
Weightage20x
Savings77%
AccessToken packs
Step-3.7 Flash
79% Off
Stepfun
Model IDstep-3.7-flash
Cheapest base model. Best usage and pricing synced from the current workbook.
Input
$0.19 $0.06/M tokens
Output
$1.16 $0.29/M tokens
Context 256K
Vision Yes
Weightage 1x
Savings 70%
Xiaomi MiMo V2.5 Pro
70% Off
Xiaomi
Model IDmimo-v2.5-pro
Xiaomi pro. Best usage and pricing synced from the current workbook.
Input
$0.43 $0.06/M tokens
Output
$0.86 $0.29/M tokens
Context 1M
Vision No
Weightage 4x
Savings 87%
GPT-5.4 Mini
80% Off
OpenAI
Model IDgpt-5.4-mini
Small context. Best usage and pricing synced from the current workbook.
Input
$0.75 $0.09/M tokens
Output
$4.50 $0.52/M tokens
Context 128K
Vision Yes
Weightage 1.2x
Savings 88%
Claude Haiku
91% Off
Anthropic
Model IDclaude-haiku-4-5-20251001
Fast & cheap. Best usage and pricing synced from the current workbook.
Input
$1.00 $0.11/M tokens
Output
$5.00 $0.57/M tokens
Context 1M
Vision Yes
Weightage 2x
Savings 89%
MiniMax M3
89% Off
MiniMax
Model IDMiniMax-M3
Latest MiniMax. Best usage and pricing synced from the current workbook.
Input
$0.60 $0.06/M tokens
Output
$2.40 $0.29/M tokens
Context 1M
Vision Yes
Weightage 1x
Savings 90%
MiniMax M2.7
90% Off
MiniMax
Model IDMiniMax-M2.7
Previous gen. Best usage and pricing synced from the current workbook.
Input
$4.00 $0.06/M tokens
Output
$15.00 $0.29/M tokens
Context 1M
Vision Yes
Weightage 1x
Savings 90%
Doubao Seed 2.0 Pro
90% Off
ByteDance
Model IDdoubao-seed-2.0-pro
ByteDance flagship. Best usage and pricing synced from the current workbook.
Input
$0.69 $0.11/M tokens
Output
$3.46 $0.57/M tokens
Context 128K
Vision Yes
Weightage 2x
Savings 83%
Qwen 3.7 Plus
83% Off
Qwen (Alibaba)
Model IDqwen3.7-plus
Updated Qwen. Best usage and pricing synced from the current workbook.
Input
$0.29 $0.23/M tokens
Output
$1.14 $1.14/M tokens
Context 1M
Vision Yes
Weightage 4x
Savings 20%
Qwen 3.7 Max
20% Off
Qwen (Alibaba)
Model IDqwen3.7-max
Top Qwen. Best usage and pricing synced from the current workbook.
Input
$1.71 $0.46/M tokens
Output
$5.14 $2.29/M tokens
Context 1M
Vision No
Weightage 12x
Savings 73%
GLM 5.1
73% Off
Zhipu
Model IDglm-5.1
Latest Zhipu. Best usage and pricing synced from the current workbook.
Input
$0.86 $0.29/M tokens
Output
$3.43 $1.43/M tokens
Context 256K
Vision Yes
Weightage 6x
Savings 67%
GLM 5.2
67% Off
Zhipu
Model IDglm-5.2
Flagship Zhipu. Best usage and pricing synced from the current workbook.
Input
$0.86 $0.46/M tokens
Output
$3.43 $2.29/M tokens
Context 1M
Vision Yes
Weightage 10x
Savings 47%
Kimi K2.6
47% Off
Kimi (Moonshot)
Model IDkimi-k2.6
Latest Kimi. Best usage and pricing synced from the current workbook.
Input
$0.93 $0.11/M tokens
Output
$3.86 $0.57/M tokens
Context 256K
Vision Yes
Weightage 2x
Savings 88%
Kimi K2.7
88% Off
Kimi (Moonshot)
Model IDkimi-k2.7
Premium Kimi. Best usage and pricing synced from the current workbook.
Input
$0.93 $0.23/M tokens
Output
$3.86 $1.14/M tokens
Context 256K
Vision Yes
Weightage 4x
Savings 75%
MiniMax M3 Highspeed
75% Off
MiniMax
Model IDMiniMax-M3-highspeed
Fast variant. Best usage and pricing synced from the current workbook.
Input
$0.90 $0.11/M tokens
Output
$3.60 $0.57/M tokens
Context 1M
Vision Yes
Weightage 2x
Savings 87%
MiniMax M2.7 Highspeed
87% Off
MiniMax
Model IDMiniMax-M2.7-highspeed
Fast variant. Best usage and pricing synced from the current workbook.
Input
$5.50 $0.11/M tokens
Output
$22.00 $0.57/M tokens
Context 1M
Vision Yes
Weightage 2x
Savings 87%
Doubao Seed 2.1 Turbo
New
ByteDance
Model IDdoubao-seed-2.1-turbo
Latest Doubao. Best usage and pricing synced from the current workbook.
Official input
$0.34/M tokens
Official output
$1.71/M tokens
Context256K
VisionYes
Weightage6x
AccessToken packs
GPT-5.3 Codex Spark
20% Off
OpenAI
Model IDgpt-5.3-codex-spark
Code specialist. Best usage and pricing synced from the current workbook.
Input
$1.75 $0.11/M tokens
Output
$14.00 $0.69/M tokens
Context 128K
Vision No
Weightage 2x
Savings 93%
Kimi for Coding
93% Off
Kimi (Moonshot)
Model IDkimi-for-coding
Coding variant. Best usage and pricing synced from the current workbook.
Input
$5.00 $0.17/M tokens
Output
$20.00 $0.86/M tokens
Context 256K
Vision Yes
Weightage 3x
Savings 80%
Doubao Seed 2.0 Code
80% Off
ByteDance
Model IDdoubao-seed-2.0-code
Coding specialist. Best usage and pricing synced from the current workbook.
Input
$0.69 $0.11/M tokens
Output
$3.46 $0.57/M tokens
Context 200K
Vision Yes
Weightage 2x
Savings 83%
DeepSeek V4 Flash
4x
DeepSeek
Model IDdeepseek-v4-flash
Newer than v3.2. Higher-cost option with premium capability or routing.
Input
$0.14 $0.23/M tokens
Output
$0.43 $0.46/M tokens
Context 1M
Vision No
Weightage 6x
Savings More expensive
DeepSeek V4 Pro
12x
DeepSeek
Model IDdeepseek-v4-pro
Premium DeepSeek. Higher-cost option with premium capability or routing.
Input
$0.43 $0.69/M tokens
Output
$0.86 $1.37/M tokens
Context 1M
Vision No
Weightage 18x
Savings More expensive
Gemini 3.1 Flash
20% Off
Google
Model IDgemini-3.1-flash
Fast & cheap. Best usage and pricing synced from the current workbook.
Input
$0.25 $0.17/M tokens
Output
$1.50 $0.86/M tokens
Context 1M
Vision Yes
Weightage 3x
Savings 31%
Gemini 3.5 Flash
31% Off
Google
Model IDgemini-3.5-flash
Newer flash. Best usage and pricing synced from the current workbook.
Input
$1.50 $0.34/M tokens
Output
$9.00 $1.71/M tokens
Context 1M
Vision Yes
Weightage 6x
Savings 77%
Python
# Use a model from the catalog with your API key from openai import OpenAI client = OpenAI( base_url="https://api.apitokendeal.com/v1", api_key="your_apitokendeal_key" ) models = ["gpt-5.5"] for model in models: response = client.chat.completions.create( model=model, messages=[{"role": "user", "content": "Hello"}] ) print(model, response.choices[0].message.content)