Model pricing
Complete Model Catalog
Full access to all offered AI models and detailed pricing.
Available Models
33+
Models Available
Claude · ChatGPT · DeepSeek · Gemini · MiniMax · Step · Xiaomi MiMo · Doubao Seed · Qwen · GLM · Kimi
Status
Live
Filter:
Model weightages explained in FAQ
Premium Models
Claude Sonnet 5
New
Anthropic
Model ID
claude-sonnet-5High-performance coding and agent model.
Official input
$3.00/M tokens
Official output
$15.00/M tokens
Context1M
Weightage6x
AccessToken packs
GPT-5.6 Sol
New
OpenAI
Model ID
gpt-5.6-solFrontier reasoning for advanced coding and agent workflows.
Official input
$5.00/M tokens
Official output
$30.00/M tokens
Context1.05M
Weightage10x
AccessToken packs
GPT-5.6 Terra
New
OpenAI
Model ID
gpt-5.6-terraBalanced general reasoning for production workloads.
Official input
$2.50/M tokens
Official output
$15.00/M tokens
Context1.05M
Weightage5x
AccessToken packs
GPT-5.6 Luna
New
OpenAI
Model ID
gpt-5.6-lunaFast, lower-cost GPT-5.6 option for high-volume work.
Official input
$1.00/M tokens
Official output
$6.00/M tokens
Context1.05M
Weightage2x
AccessToken packs
Claude Fable 5
New
Anthropic
Model ID
claude-fable-5Frontier intelligence for the hardest reasoning and coding tasks.
Official input
$10.00/M tokens
Official output
$50.00/M tokens
Context1M
VisionYes
Weightage50x
AccessToken packs
Kimi K3
New
Moonshot AI
Model ID
kimi-k32.8T parameter flagship for long-horizon coding and knowledge work.
Official input
$3.00/M tokens
Official output
$15.00/M tokens
Context1M
VisionYes
Weightage12x
AccessToken packs
Qwen 3.8 Max
Preview
Alibaba Cloud
Model ID
qwen3.8-max2.4T parameter flagship via Alibaba Cloud Token Plan.
Official input
From $6/mo
Official output
Token Plan
Context1M
VisionYes
Weightage12x
AccessToken packs
Claude Opus
77% Off
Anthropic
Model ID
claude-opus-4-8Latest Opus. Best usage and pricing synced from the current workbook.
Input
$5.00
$1.14/M tokens
Output
$25.00
$5.71/M tokens
Context
1M
Vision
Yes
Weightage
20x
Savings
77%
Claude Opus
77% Off
Anthropic
Model ID
claude-opus-4-7Updated Opus. Best usage and pricing synced from the current workbook.
Input
$5.00
$1.14/M tokens
Output
$25.00
$5.71/M tokens
Context
1M
Vision
Yes
Weightage
20x
Savings
77%
Claude Opus
77% Off
Anthropic
Model ID
claude-opus-4-6Top-tier reasoning. Best usage and pricing synced from the current workbook.
Input
$5.00
$1.14/M tokens
Output
$25.00
$5.71/M tokens
Context
1M
Vision
Yes
Weightage
20x
Savings
77%
GPT-5.5
91% Off
OpenAI
Model ID
gpt-5.5Flagship model. Best usage and pricing synced from the current workbook.
Input
$5.00
$0.46/M tokens
Output
$30.00
$2.29/M tokens
Context
256K
Vision
Yes
Weightage
8x
Savings
91%
Claude Sonnet
92% Off
Anthropic
Model ID
claude-sonnet-4-6Best balance. Best usage and pricing synced from the current workbook.
Input
$3.00
$0.23/M tokens
Output
$15.00
$1.14/M tokens
Context
1M
Vision
Yes
Weightage
4x
Savings
92%
GPT-5.4
91% Off
OpenAI
Model ID
gpt-5.4Best value flagship. Best usage and pricing synced from the current workbook.
Input
$2.50
$0.23/M tokens
Output
$15.00
$1.14/M tokens
Context
1M
Vision
Yes
Weightage
4x
Savings
91%
Gemini 3.1 Pro
86% Off
Google
Model ID
gemini-3.1-proFlagship Google. Best usage and pricing synced from the current workbook.
Input
$4.00
$0.57/M tokens
Output
$18.00
$2.86/M tokens
Context
1M
Vision
Yes
Weightage
10x
Savings
86%
Budget-Friendly Models
Xiaomi MiMo V2.5
80% Off
Xiaomi
Model ID
mimo-v2.5Cheapest overall. Best usage and pricing synced from the current workbook.
Input
$0.14
$0.03/M tokens
Output
$0.29
$0.14/M tokens
Context
1M
Vision
Yes
Weightage
0.5x
Savings
80%
Step-3.7 Flash
70% Off
Stepfun
Model ID
step-3.7-flashCheapest base model. Best usage and pricing synced from the current workbook.
Input
$0.19
$0.06/M tokens
Output
$1.16
$0.29/M tokens
Context
256K
Vision
Yes
Weightage
1x
Savings
70%
Xiaomi MiMo V2.5 Pro
87% Off
Xiaomi
Model ID
mimo-v2.5-proXiaomi pro. Best usage and pricing synced from the current workbook.
Input
$0.43
$0.06/M tokens
Output
$0.86
$0.29/M tokens
Context
1M
Vision
No
Weightage
1x
Savings
87%
GPT-5.4 Mini
88% Off
OpenAI
Model ID
gpt-5.4-miniSmall context. Best usage and pricing synced from the current workbook.
Input
$0.75
$0.09/M tokens
Output
$4.50
$0.52/M tokens
Context
128K
Vision
Yes
Weightage
2x
Savings
88%
Claude Haiku
89% Off
Anthropic
Model ID
claude-haiku-4-5-20251001Fast & cheap. Best usage and pricing synced from the current workbook.
Input
$1.00
$0.11/M tokens
Output
$5.00
$0.57/M tokens
Context
1M
Vision
Yes
Weightage
2x
Savings
89%
MiniMax M3
90% Off
MiniMax
Model ID
MiniMax-M3Latest MiniMax. Best usage and pricing synced from the current workbook.
Input
$0.60
$0.06/M tokens
Output
$2.40
$0.29/M tokens
Context
1M
Vision
Yes
Weightage
1x
Savings
90%
MiniMax M2.7
90% Off
MiniMax
Model ID
MiniMax-M2.7Previous gen. Best usage and pricing synced from the current workbook.
Input
$4.00
$0.06/M tokens
Output
$15.00
$0.29/M tokens
Context
1M
Vision
Yes
Weightage
1x
Savings
90%
Chinese & Multilingual Models
Doubao Seed 2.0 Pro
83% Off
ByteDance
Model ID
doubao-seed-2.0-proByteDance flagship. Best usage and pricing synced from the current workbook.
Input
$0.69
$0.11/M tokens
Output
$3.46
$0.57/M tokens
Context
128K
Vision
Yes
Weightage
2x
Savings
83%
Qwen 3.7 Plus
20% Off
Qwen (Alibaba)
Model ID
qwen3.7-plusUpdated Qwen. Best usage and pricing synced from the current workbook.
Input
$0.29
$0.23/M tokens
Output
$1.14
$1.14/M tokens
Context
1M
Vision
Yes
Weightage
4x
Savings
20%
Qwen 3.7 Max
73% Off
Qwen (Alibaba)
Model ID
qwen3.7-maxTop Qwen. Best usage and pricing synced from the current workbook.
Input
$1.71
$0.46/M tokens
Output
$5.14
$2.29/M tokens
Context
1M
Vision
No
Weightage
8x
Savings
73%
GLM 5.1
67% Off
Zhipu
Model ID
glm-5.1Latest Zhipu. Best usage and pricing synced from the current workbook.
Input
$0.86
$0.29/M tokens
Output
$3.43
$1.43/M tokens
Context
256K
Vision
Yes
Weightage
5x
Savings
67%
GLM 5.2
47% Off
Zhipu
Model ID
glm-5.2Flagship Zhipu. Best usage and pricing synced from the current workbook.
Input
$0.86
$0.46/M tokens
Output
$3.43
$2.29/M tokens
Context
1M
Vision
Yes
Weightage
8x
Savings
47%
Kimi K2.6
88% Off
Kimi (Moonshot)
Model ID
kimi-k2.6Latest Kimi. Best usage and pricing synced from the current workbook.
Input
$0.93
$0.11/M tokens
Output
$3.86
$0.57/M tokens
Context
256K
Vision
Yes
Weightage
2x
Savings
88%
Kimi K2.7
75% Off
Kimi (Moonshot)
Model ID
kimi-k2.7Premium Kimi. Best usage and pricing synced from the current workbook.
Input
$0.93
$0.23/M tokens
Output
$3.86
$1.14/M tokens
Context
256K
Vision
Yes
Weightage
4x
Savings
75%
MiniMax M3 Highspeed
87% Off
MiniMax
Model ID
MiniMax-M3-highspeedFast variant. Best usage and pricing synced from the current workbook.
Input
$0.90
$0.11/M tokens
Output
$3.60
$0.57/M tokens
Context
1M
Vision
Yes
Weightage
2x
Savings
87%
MiniMax M2.7 Highspeed
87% Off
MiniMax
Model ID
MiniMax-M2.7-highspeedFast variant. Best usage and pricing synced from the current workbook.
Input
$5.50
$0.11/M tokens
Output
$22.00
$0.57/M tokens
Context
1M
Vision
Yes
Weightage
2x
Savings
87%
Coding & Dev Models
GPT-5.3 Codex Spark
93% Off
OpenAI
Model ID
gpt-5.3-codex-sparkCode specialist. Best usage and pricing synced from the current workbook.
Input
$1.75
$0.11/M tokens
Output
$14.00
$0.69/M tokens
Context
128K
Vision
No
Weightage
2x
Savings
93%
Kimi for Coding
80% Off
Kimi (Moonshot)
Model ID
kimi-for-codingCoding variant. Best usage and pricing synced from the current workbook.
Input
$5.00
$0.17/M tokens
Output
$20.00
$0.86/M tokens
Context
256K
Vision
Yes
Weightage
3x
Savings
80%
Doubao Seed 2.0 Code
83% Off
ByteDance
Model ID
doubao-seed-2.0-codeCoding specialist. Best usage and pricing synced from the current workbook.
Input
$0.69
$0.11/M tokens
Output
$3.46
$0.57/M tokens
Context
200K
Vision
Yes
Weightage
2x
Savings
83%
Fast & Specialized Models
DeepSeek V4 Flash
4x
DeepSeek
Model ID
deepseek-v4-flashNewer than v3.2. Higher-cost option with premium capability or routing.
Input
$0.14
$0.23/M tokens
Output
$0.43
$0.46/M tokens
Context
1M
Vision
No
Weightage
4x
Savings
More expensive
DeepSeek V4 Pro
12x
DeepSeek
Model ID
deepseek-v4-proPremium DeepSeek. Higher-cost option with premium capability or routing.
Input
$0.43
$0.69/M tokens
Output
$0.86
$1.37/M tokens
Context
1M
Vision
No
Weightage
12x
Savings
More expensive
Gemini 3.1 Flash
31% Off
Google
Model ID
gemini-3.1-flashFast & cheap. Best usage and pricing synced from the current workbook.
Input
$0.25
$0.17/M tokens
Output
$1.50
$0.86/M tokens
Context
1M
Vision
Yes
Weightage
3x
Savings
31%
Gemini 3.5 Flash
77% Off
Google
Model ID
gemini-3.5-flashNewer flash. Best usage and pricing synced from the current workbook.
Input
$1.50
$0.34/M tokens
Output
$9.00
$1.71/M tokens
Context
1M
Vision
Yes
Weightage
6x
Savings
77%
Use any model with one API key
# Use a model from the catalog with your API key
from openai import OpenAI
client = OpenAI(
base_url="https://api.apitokendeal.com/v1",
api_key="your_apitokendeal_key"
)
models = ["gpt-5.5"]
for model in models:
response = client.chat.completions.create(
model=model,
messages=[{"role": "user", "content": "Hello"}]
)
print(model, response.choices[0].message.content)