Token Plans
Plan team AI usage before buying credits
Token Plan is useful for AI coding and team workflows. API quota is better for product traffic. We separate those paths so customers understand what they are paying for.
Checked price references
Key Model Studio price references before a written quote
The full catalog now lives on the Models page. This pricing page keeps the buyer-oriented reference rows readable, while every model detail page shows its copied official detail-price summary where the console exposed one.
Catalog lanes
9
Flagship, Cost-optimized, Visual, Wan, Audio, Multimodal, Embeddings, Third-party, and Older.
Detail pages
96
90 detail-price summaries were cleanly matched; alias-style rows stay marked for console confirmation.
Full model catalog
Review every model detail
Open model catalogqwen3.7-max
FamilyQwen3.7 flagship
UseText-only flagship reasoning, coding, office, and autonomous agent work
Input$2.5 / 1M tokens
Output$7.5 / 1M tokens
NoteSingapore international original price. Temporary promotional pricing may apply separately.
qwen3.7-plus
FamilyQwen3.7 multimodal
UseText, image, video input, GUI perception, coding, and tool-use workflows
Input$0.4 / 1M tokens
Output$1.6 / 1M tokens
NoteInternational console original price for the <=256K input tier.
qwen3.6-plus
FamilyQwen3.6 balanced
UseVision-language production, coding, OCR-like extraction, and business automation
Input$0.5 / 1M tokens
Output$3 / 1M tokens
NoteInternational console original price for the <=256K input tier.
qwen3.6-flash
FamilyQwen3.6 cost-optimized
UseFast multimodal production, support, extraction, code, and math workloads
Input$0.25 / 1M tokens
Output$1.5 / 1M tokens
NoteInternational console original price for the <=256K input tier.
qwen3.6-max-preview
FamilyQwen3.6 preview
UseText-only preview lane for high-end coding and agent evaluation
Input$1.3 / 1M tokens
Output$7.8 / 1M tokens
NoteInternational console original price for the <=128K input tier.
qwen3.6-27b
FamilyQwen3.6 open source
UseDense open-source vision-language lane for coding, STEM, and visual tasks
Input$0.6 / 1M tokens
Output$3.6 / 1M tokens
NoteInternational console listed price.
deepseek-v4-pro
FamilyDeepSeek
UseStrong reasoning, code, math, research, and long-form technical analysis
Input$1.65 / 1M tokens
Output$3.301 / 1M tokens
NoteChinese Mainland console detail price captured for this model. International support and final quote must be confirmed.
deepseek-v4-flash
FamilyDeepSeek
UseLightweight, high-concurrency reasoning and batch text processing
Input$0.138 / 1M tokens
Output$0.275 / 1M tokens
NoteChinese Mainland console detail price captured for this model. International support and final quote must be confirmed.
deepseek-v3.2
FamilyDeepSeek
UseStable DeepSeek lane with sparse attention and reasoning/tool-use support
Input$0.287 / 1M tokens
OutputConfirm output row
NoteChinese Mainland console detail price captured for input/cache rows. Confirm exact output tier in the official console before quoting.
kimi-k2.7-code
FamilyKimi
UseLong-context coding, visual input, instruction following, conversation, and agent tasks
Input$0.8939 / 1M tokens
Output$3.7131 / 1M tokens
NoteChinese Mainland official detail price.
glm-5.2
FamilyGLM
UseLong-horizon reasoning, 1M-context comprehension, code generation, and enterprise assistants
Input$1.1 / 1M tokens
Output$3.851 / 1M tokens
NoteChinese Mainland official detail price.
mimo-v2.5-pro
FamilyMiMo
UseCoding, agent, reasoning, and long-context software engineering workflows
InputCNY 7 / 1M tokens
OutputCNY 21 / 1M tokens
NoteChinese Mainland original price for 0<Token<=256K. For 256K<Token<=1M, input is CNY 14 and output is CNY 42 per 1M tokens.
MiniMax-M2.5
FamilyMiniMax
UseAgent-style coding, tool invocation, search, productivity, and office work
Input$0.304 / 1M tokens
Output$1.213 / 1M tokens
NoteChinese Mainland console detail price captured for this model.
Image generation / editing
qwen-image-2.0
$0.035 / image
Input
Text / image
Output
Image
High-quality image generation / editing
qwen-image-2.0-pro
$0.075 / image
Input
Text / image
Output
Image
Wan image generation / editing
wan2.7-image
$0.03 / image
Input
Text / image
Output
Image
Wan image generation / editing
wan2.7-image-pro
$0.075 / image
Input
Text / image
Output
Image
Text to video
happyhorse-1.1-t2v
From $0.14/s
Input
Text
Output
Video
Image to video
happyhorse-1.1-i2v
From $0.14/s
Input
Text / image
Output
Video
Reference to video
happyhorse-1.1-r2v
From $0.14/s
Input
Text / image
Output
Video
Text to video
happyhorse-1.0-t2v
$0.14/s 720P · $0.24/s 1080P
Input
Text
Output
Video
Image to video
happyhorse-1.0-i2v
$0.14/s 720P · $0.24/s 1080P
Input
Text / image
Output
Video
Reference to video
happyhorse-1.0-r2v
$0.14/s 720P · $0.24/s 1080P
Input
Text / image
Output
Video
Video editing
happyhorse-1.0-video-edit
$0.14/s 720P · $0.24/s 1080P
Input
Image / video
Output
Video
Embedding
text-embedding-v4
$0.07 / 1M tokens
Input
Text
Output
Vector
Reranking
qwen3-rerank
$0.1 / 1M tokens
Input
Text
Output
Ranked text
This section is intentionally a reference index, not a complete quote table. Token, cache, media, audio, video-second, and tool-call rows use different billing units, so final procurement review should start from the exact model detail page and end with a written quote.
Standard
USD 30 / user / month
25,000 Credits
Light daily AI usage
Pro
USD 100 / user / month
100,000 Credits
Frequent AI coding and content work
Max
USD 200 / user / month
250,000 Credits
Core users who rely on AI throughout the day
Shared quota pack
USD 700 / pack
1,000,000 Credits
Elastic overage pool for teams
Token Plan pricing references the Alibaba Cloud Model Studio Token Plan Team Edition documentation and is separate from model invocation pricing above. Final quotes must verify official plan availability, billing currency, taxes, exchange rate, payment method, account eligibility, and service fees.Official source
Billing structure
Separate official usage costs from procurement service fees
This is the most important trust point. Buyers should see the platform cost, service fee, payment cost, and delivery scope before payment.
New-user free quota
Check whether the buyer has unused trial allowance before purchasing credits.
Model invocation pricing
Confirm input and output token rates for the selected model and deployment region.
Training and deployment
Separate fine-tuning, deployment, and hosting costs from normal API usage.
Savings plans
Estimate whether volume commitments or prepaid packages make sense.
Bills and cost management
Review monthly usage, alerts, payment route, and replenishment timing.
Service scope
What ModelSmarter charges for
The customer should not confuse our service fee with official model usage cost. This section makes the service scope explicit.
Service
Model Review
$149
Service
Procurement Setup
From $790
Service
Managed Team
Custom
Quote process
Four checks before payment
A written quote should be the conversion point, not a generic checkout button.
01
Choose the model lane
Match workload, modality, latency needs, region, and context length before quoting.
02
Confirm billing route
Separate official model usage, Token Plan seats, shared quota, payment fees, and service scope.
03
Set up access
Coordinate account route, API key, base URL, endpoint region, and team handoff.
04
Monitor usage
Track consumption, rate limits, free quota, cost controls, and monthly replenishment needs.