Model pricing

Complete Model Catalog

Full access to all offered AI models and detailed pricing.

Available Models
33+
Models Available
Claude · ChatGPT · DeepSeek · Gemini · MiniMax · Step · Xiaomi MiMo · Doubao Seed · Qwen · GLM · Kimi
Status
Live
Filter: Model weightages explained in FAQ
GPT-5.6 Sol
New
OpenAI
Model IDgpt-5.6-sol
Frontier reasoning for advanced coding and agent workflows.
Official input
$5.00/M tokens
Official output
$30.00/M tokens
Context1.05M
Weightage10x
AccessToken packs
GPT-5.6 Terra
New
OpenAI
Model IDgpt-5.6-terra
Balanced general reasoning for production workloads.
Official input
$2.50/M tokens
Official output
$15.00/M tokens
Context1.05M
Weightage5x
AccessToken packs
GPT-5.6 Luna
New
OpenAI
Model IDgpt-5.6-luna
Fast, lower-cost GPT-5.6 option for high-volume work.
Official input
$1.00/M tokens
Official output
$6.00/M tokens
Context1.05M
Weightage2x
AccessToken packs
Kimi K3
New
Moonshot AI
Model IDkimi-k3
2.8T parameter flagship for long-horizon coding and knowledge work.
Official input
$3.00/M tokens
Official output
$15.00/M tokens
Context1M
VisionYes
Weightage12x
AccessToken packs
Qwen 3.8 Max
Preview
Alibaba Cloud
Model IDqwen3.8-max
2.4T parameter flagship via Alibaba Cloud Token Plan.
Official input
From $6/mo
Official output
Token Plan
Context1M
VisionYes
Weightage12x
AccessToken packs
Claude Opus
77% Off
Anthropic
Model IDclaude-opus-4-7
Updated Opus. Best usage and pricing synced from the current workbook.
Input
$5.00 $1.14/M tokens
Output
$25.00 $5.71/M tokens
Context 1M
Vision Yes
Weightage 20x
Savings 77%
Claude Opus
77% Off
Anthropic
Model IDclaude-opus-4-6
Top-tier reasoning. Best usage and pricing synced from the current workbook.
Input
$5.00 $1.14/M tokens
Output
$25.00 $5.71/M tokens
Context 1M
Vision Yes
Weightage 20x
Savings 77%
Claude Sonnet
92% Off
Anthropic
Model IDclaude-sonnet-4-6
Best balance. Best usage and pricing synced from the current workbook.
Input
$3.00 $0.23/M tokens
Output
$15.00 $1.14/M tokens
Context 1M
Vision Yes
Weightage 4x
Savings 92%
Gemini 3.1 Pro
86% Off
Google
Model IDgemini-3.1-pro
Flagship Google. Best usage and pricing synced from the current workbook.
Input
$4.00 $0.57/M tokens
Output
$18.00 $2.86/M tokens
Context 1M
Vision Yes
Weightage 10x
Savings 86%
Step-3.7 Flash
70% Off
Stepfun
Model IDstep-3.7-flash
Cheapest base model. Best usage and pricing synced from the current workbook.
Input
$0.19 $0.06/M tokens
Output
$1.16 $0.29/M tokens
Context 256K
Vision Yes
Weightage 1x
Savings 70%
Xiaomi MiMo V2.5 Pro
87% Off
Xiaomi
Model IDmimo-v2.5-pro
Xiaomi pro. Best usage and pricing synced from the current workbook.
Input
$0.43 $0.06/M tokens
Output
$0.86 $0.29/M tokens
Context 1M
Vision No
Weightage 1x
Savings 87%
GPT-5.4 Mini
88% Off
OpenAI
Model IDgpt-5.4-mini
Small context. Best usage and pricing synced from the current workbook.
Input
$0.75 $0.09/M tokens
Output
$4.50 $0.52/M tokens
Context 128K
Vision Yes
Weightage 2x
Savings 88%
Claude Haiku
89% Off
Anthropic
Model IDclaude-haiku-4-5-20251001
Fast & cheap. Best usage and pricing synced from the current workbook.
Input
$1.00 $0.11/M tokens
Output
$5.00 $0.57/M tokens
Context 1M
Vision Yes
Weightage 2x
Savings 89%
MiniMax M3
90% Off
MiniMax
Model IDMiniMax-M3
Latest MiniMax. Best usage and pricing synced from the current workbook.
Input
$0.60 $0.06/M tokens
Output
$2.40 $0.29/M tokens
Context 1M
Vision Yes
Weightage 1x
Savings 90%
MiniMax M2.7
90% Off
MiniMax
Model IDMiniMax-M2.7
Previous gen. Best usage and pricing synced from the current workbook.
Input
$4.00 $0.06/M tokens
Output
$15.00 $0.29/M tokens
Context 1M
Vision Yes
Weightage 1x
Savings 90%
Doubao Seed 2.0 Pro
83% Off
ByteDance
Model IDdoubao-seed-2.0-pro
ByteDance flagship. Best usage and pricing synced from the current workbook.
Input
$0.69 $0.11/M tokens
Output
$3.46 $0.57/M tokens
Context 128K
Vision Yes
Weightage 2x
Savings 83%
Qwen 3.7 Plus
20% Off
Qwen (Alibaba)
Model IDqwen3.7-plus
Updated Qwen. Best usage and pricing synced from the current workbook.
Input
$0.29 $0.23/M tokens
Output
$1.14 $1.14/M tokens
Context 1M
Vision Yes
Weightage 4x
Savings 20%
Qwen 3.7 Max
73% Off
Qwen (Alibaba)
Model IDqwen3.7-max
Top Qwen. Best usage and pricing synced from the current workbook.
Input
$1.71 $0.46/M tokens
Output
$5.14 $2.29/M tokens
Context 1M
Vision No
Weightage 8x
Savings 73%
GLM 5.1
67% Off
Zhipu
Model IDglm-5.1
Latest Zhipu. Best usage and pricing synced from the current workbook.
Input
$0.86 $0.29/M tokens
Output
$3.43 $1.43/M tokens
Context 256K
Vision Yes
Weightage 5x
Savings 67%
GLM 5.2
47% Off
Zhipu
Model IDglm-5.2
Flagship Zhipu. Best usage and pricing synced from the current workbook.
Input
$0.86 $0.46/M tokens
Output
$3.43 $2.29/M tokens
Context 1M
Vision Yes
Weightage 8x
Savings 47%
Kimi K2.6
88% Off
Kimi (Moonshot)
Model IDkimi-k2.6
Latest Kimi. Best usage and pricing synced from the current workbook.
Input
$0.93 $0.11/M tokens
Output
$3.86 $0.57/M tokens
Context 256K
Vision Yes
Weightage 2x
Savings 88%
Kimi K2.7
75% Off
Kimi (Moonshot)
Model IDkimi-k2.7
Premium Kimi. Best usage and pricing synced from the current workbook.
Input
$0.93 $0.23/M tokens
Output
$3.86 $1.14/M tokens
Context 256K
Vision Yes
Weightage 4x
Savings 75%
MiniMax M3 Highspeed
87% Off
MiniMax
Model IDMiniMax-M3-highspeed
Fast variant. Best usage and pricing synced from the current workbook.
Input
$0.90 $0.11/M tokens
Output
$3.60 $0.57/M tokens
Context 1M
Vision Yes
Weightage 2x
Savings 87%
MiniMax M2.7 Highspeed
87% Off
MiniMax
Model IDMiniMax-M2.7-highspeed
Fast variant. Best usage and pricing synced from the current workbook.
Input
$5.50 $0.11/M tokens
Output
$22.00 $0.57/M tokens
Context 1M
Vision Yes
Weightage 2x
Savings 87%
GPT-5.3 Codex Spark
93% Off
OpenAI
Model IDgpt-5.3-codex-spark
Code specialist. Best usage and pricing synced from the current workbook.
Input
$1.75 $0.11/M tokens
Output
$14.00 $0.69/M tokens
Context 128K
Vision No
Weightage 2x
Savings 93%
Kimi for Coding
80% Off
Kimi (Moonshot)
Model IDkimi-for-coding
Coding variant. Best usage and pricing synced from the current workbook.
Input
$5.00 $0.17/M tokens
Output
$20.00 $0.86/M tokens
Context 256K
Vision Yes
Weightage 3x
Savings 80%
Doubao Seed 2.0 Code
83% Off
ByteDance
Model IDdoubao-seed-2.0-code
Coding specialist. Best usage and pricing synced from the current workbook.
Input
$0.69 $0.11/M tokens
Output
$3.46 $0.57/M tokens
Context 200K
Vision Yes
Weightage 2x
Savings 83%
DeepSeek V4 Flash
4x
DeepSeek
Model IDdeepseek-v4-flash
Newer than v3.2. Higher-cost option with premium capability or routing.
Input
$0.14 $0.23/M tokens
Output
$0.43 $0.46/M tokens
Context 1M
Vision No
Weightage 4x
Savings More expensive
DeepSeek V4 Pro
12x
DeepSeek
Model IDdeepseek-v4-pro
Premium DeepSeek. Higher-cost option with premium capability or routing.
Input
$0.43 $0.69/M tokens
Output
$0.86 $1.37/M tokens
Context 1M
Vision No
Weightage 12x
Savings More expensive
Gemini 3.1 Flash
31% Off
Google
Model IDgemini-3.1-flash
Fast & cheap. Best usage and pricing synced from the current workbook.
Input
$0.25 $0.17/M tokens
Output
$1.50 $0.86/M tokens
Context 1M
Vision Yes
Weightage 3x
Savings 31%
Gemini 3.5 Flash
77% Off
Google
Model IDgemini-3.5-flash
Newer flash. Best usage and pricing synced from the current workbook.
Input
$1.50 $0.34/M tokens
Output
$9.00 $1.71/M tokens
Context 1M
Vision Yes
Weightage 6x
Savings 77%
Python
# Use a model from the catalog with your API key from openai import OpenAI client = OpenAI( base_url="https://api.apitokendeal.com/v1", api_key="your_apitokendeal_key" ) models = ["gpt-5.5"] for model in models: response = client.chat.completions.create( model=model, messages=[{"role": "user", "content": "Hello"}] ) print(model, response.choices[0].message.content)