Token Plans

Plan team AI usage before buying credits

Token Plan is useful for AI coding and team workflows. API quota is better for product traffic. We separate those paths so customers understand what they are paying for.

Checked price references

Key Model Studio price references before a written quote

The full catalog now lives on the Models page. This pricing page keeps the buyer-oriented reference rows readable, while every model detail page shows its copied official detail-price summary where the console exposed one.

Current highlighted models and prices were rechecked on July 21, 2026 against the official Alibaba Cloud Model Studio model and pricing pages. The broader catalog retains earlier detail evidence where noted and now covers 9 lanes, 99 listed entries, and 96 unique model detail pages. This is not an Alibaba Cloud official quote or a final ModelSmarter quote. Region, deployment mode, account route, quota, taxes, promotions, and availability still need final confirmation.Official source

Catalog lanes

9

Flagship, Cost-optimized, Visual, Wan, Audio, Multimodal, Embeddings, Third-party, and Older.

Detail pages

96

90 detail-price summaries were cleanly matched; alias-style rows stay marked for console confirmation.

Full model catalog

Review every model detail

Open model catalog

qwen3.7-max

FamilyQwen3.7 flagship

UseText-only flagship reasoning, coding, office, and autonomous agent work

Input$2.5 / 1M tokens

Output$7.5 / 1M tokens

NoteSingapore international original price. Temporary promotional pricing may apply separately.

qwen3.7-plus

FamilyQwen3.7 multimodal

UseText, image, video input, GUI perception, coding, and tool-use workflows

Input$0.4 / 1M tokens

Output$1.6 / 1M tokens

NoteInternational console original price for the <=256K input tier.

qwen3.6-plus

FamilyQwen3.6 balanced

UseVision-language production, coding, OCR-like extraction, and business automation

Input$0.5 / 1M tokens

Output$3 / 1M tokens

NoteInternational console original price for the <=256K input tier.

qwen3.6-flash

FamilyQwen3.6 cost-optimized

UseFast multimodal production, support, extraction, code, and math workloads

Input$0.25 / 1M tokens

Output$1.5 / 1M tokens

NoteInternational console original price for the <=256K input tier.

qwen3.6-max-preview

FamilyQwen3.6 preview

UseText-only preview lane for high-end coding and agent evaluation

Input$1.3 / 1M tokens

Output$7.8 / 1M tokens

NoteInternational console original price for the <=128K input tier.

qwen3.6-27b

FamilyQwen3.6 open source

UseDense open-source vision-language lane for coding, STEM, and visual tasks

Input$0.6 / 1M tokens

Output$3.6 / 1M tokens

NoteInternational console listed price.

deepseek-v4-pro

FamilyDeepSeek

UseStrong reasoning, code, math, research, and long-form technical analysis

Input$1.65 / 1M tokens

Output$3.301 / 1M tokens

NoteChinese Mainland console detail price captured for this model. International support and final quote must be confirmed.

deepseek-v4-flash

FamilyDeepSeek

UseLightweight, high-concurrency reasoning and batch text processing

Input$0.138 / 1M tokens

Output$0.275 / 1M tokens

NoteChinese Mainland console detail price captured for this model. International support and final quote must be confirmed.

deepseek-v3.2

FamilyDeepSeek

UseStable DeepSeek lane with sparse attention and reasoning/tool-use support

Input$0.287 / 1M tokens

OutputConfirm output row

NoteChinese Mainland console detail price captured for input/cache rows. Confirm exact output tier in the official console before quoting.

kimi-k2.7-code

FamilyKimi

UseLong-context coding, visual input, instruction following, conversation, and agent tasks

Input$0.8939 / 1M tokens

Output$3.7131 / 1M tokens

NoteChinese Mainland official detail price.

glm-5.2

FamilyGLM

UseLong-horizon reasoning, 1M-context comprehension, code generation, and enterprise assistants

Input$1.1 / 1M tokens

Output$3.851 / 1M tokens

NoteChinese Mainland official detail price.

mimo-v2.5-pro

FamilyMiMo

UseCoding, agent, reasoning, and long-context software engineering workflows

InputCNY 7 / 1M tokens

OutputCNY 21 / 1M tokens

NoteChinese Mainland original price for 0<Token<=256K. For 256K<Token<=1M, input is CNY 14 and output is CNY 42 per 1M tokens.

MiniMax-M2.5

FamilyMiniMax

UseAgent-style coding, tool invocation, search, productivity, and office work

Input$0.304 / 1M tokens

Output$1.213 / 1M tokens

NoteChinese Mainland console detail price captured for this model.

Image generation / editing

qwen-image-2.0

$0.035 / image

Input

Text / image

Output

Image

High-quality image generation / editing

qwen-image-2.0-pro

$0.075 / image

Input

Text / image

Output

Image

Wan image generation / editing

wan2.7-image

$0.03 / image

Input

Text / image

Output

Image

Wan image generation / editing

wan2.7-image-pro

$0.075 / image

Input

Text / image

Output

Image

Text to video

happyhorse-1.1-t2v

From $0.14/s

Input

Text

Output

Video

Image to video

happyhorse-1.1-i2v

From $0.14/s

Input

Text / image

Output

Video

Reference to video

happyhorse-1.1-r2v

From $0.14/s

Input

Text / image

Output

Video

Text to video

happyhorse-1.0-t2v

$0.14/s 720P · $0.24/s 1080P

Input

Text

Output

Video

Image to video

happyhorse-1.0-i2v

$0.14/s 720P · $0.24/s 1080P

Input

Text / image

Output

Video

Reference to video

happyhorse-1.0-r2v

$0.14/s 720P · $0.24/s 1080P

Input

Text / image

Output

Video

Video editing

happyhorse-1.0-video-edit

$0.14/s 720P · $0.24/s 1080P

Input

Image / video

Output

Video

Embedding

text-embedding-v4

$0.07 / 1M tokens

Input

Text

Output

Vector

Reranking

qwen3-rerank

$0.1 / 1M tokens

Input

Text

Output

Ranked text

This section is intentionally a reference index, not a complete quote table. Token, cache, media, audio, video-second, and tool-call rows use different billing units, so final procurement review should start from the exact model detail page and end with a written quote.

Standard

USD 30 / user / month

25,000 Credits

Light daily AI usage

Pro

USD 100 / user / month

100,000 Credits

Frequent AI coding and content work

Max

USD 200 / user / month

250,000 Credits

Core users who rely on AI throughout the day

Shared quota pack

USD 700 / pack

1,000,000 Credits

Elastic overage pool for teams

Token Plan pricing references the Alibaba Cloud Model Studio Token Plan Team Edition documentation and is separate from model invocation pricing above. Final quotes must verify official plan availability, billing currency, taxes, exchange rate, payment method, account eligibility, and service fees.Official source

Billing structure

Separate official usage costs from procurement service fees

This is the most important trust point. Buyers should see the platform cost, service fee, payment cost, and delivery scope before payment.

New-user free quota

Check whether the buyer has unused trial allowance before purchasing credits.

Model invocation pricing

Confirm input and output token rates for the selected model and deployment region.

Training and deployment

Separate fine-tuning, deployment, and hosting costs from normal API usage.

Savings plans

Estimate whether volume commitments or prepaid packages make sense.

Bills and cost management

Review monthly usage, alerts, payment route, and replenishment timing.

Service scope

What ModelSmarter charges for

The customer should not confuse our service fee with official model usage cost. This section makes the service scope explicit.

Service

Model Review

$149

Use-case review
Model shortlisting
Region notes
Buying checklist
Recommended

Service

Procurement Setup

From $790

Token Plan or API quota planning
Payment coordination
Account-route support
Delivery handoff

Service

Managed Team

Custom

Monthly planning
Usage review
Shared quota strategy
Priority support path

Quote process

Four checks before payment

A written quote should be the conversion point, not a generic checkout button.

01

Choose the model lane

Match workload, modality, latency needs, region, and context length before quoting.

02

Confirm billing route

Separate official model usage, Token Plan seats, shared quota, payment fees, and service scope.

03

Set up access

Coordinate account route, API key, base URL, endpoint region, and team handoff.

04

Monitor usage

Track consumption, rate limits, free quota, cost controls, and monthly replenishment needs.