This topic describes the billing rules and pricing for model training and model deployment on Alibaba Cloud Model Studio.
Training billing
Text generation models – Qwen
For the training workflow, see Introduction to model fine-tuning. After training completes, deploy the new model before evaluating or calling it.
|
Method |
Billed by training tokens |
|
Formula |
Model training fee = (Total tokens in training data + Total tokens in mixed training data) × Number of epochs × Training unit price (Minimum billing unit: 1 token) View the estimated training fee at the bottom of the model training console, and click Computing Details to view the total number of training tokens, number of epochs, and training unit price. |
Qwen
Model service | Model code | Price |
Qwen3.6-Flash-2026-04-16 | qwen3.6-flash-2026-04-16 | CNY 0.05/1,000 tokens |
Qwen3.5-27B | qwen3.5-27b | CNY 0.05/1,000 tokens |
Qwen3.5-9B | qwen3.5-9b | CNY 0.02/1,000 tokens |
Qwen3.5-Flash-2026-02-23 | qwen3.5-flash-2026-02-23 | CNY 0.05/1,000 tokens |
Qwen3-32B | qwen3-32b | CNY 0.04/1,000 tokens |
Qwen3-30B-A3B-Instruct-2507 | qwen3-30b-a3b-instruct-2507 | CNY 0.03/1,000 tokens |
Qwen3-14B | qwen3-14b | CNY 0.03/1,000 tokens |
Qwen3-8B | qwen3-8b | CNY 0.006/1,000 tokens |
Qwen3-4B-Instruct-2507 | qwen3-4b-instruct-2507 | CNY 0.006/1,000 tokens |
Qwen3-1.7B | qwen3-1.7b | CNY 0.0045/1,000 tokens |
Qwen3-0.6B | qwen3-0.6b | CNY 0.003/1,000 tokens |
Qwen2.5-72B-Instruct | qwen2.5-72b-instruct | CNY 0.15/1,000 tokens |
Qwen2.5-32B-Instruct | qwen2.5-32b-instruct | CNY 0.03/1,000 tokens |
Qwen2.5-14B-Instruct | qwen2.5-14b-instruct | CNY 0.03/1,000 tokens |
Qwen2.5-7B-Instruct | qwen2.5-7b-instruct | CNY 0.006/1,000 tokens |
Qwen-Plus-Character-2025-11-06 | qwen-plus-character-2025-11-06 | CNY 0.15/1,000 tokens |
Qwen-VL
Model service | Model code | Price |
Qwen3-VL-8B-Instruct | qwen3-vl-8b-instruct | CNY 0.012/1,000 tokens |
Qwen3-VL-8B-Thinking | qwen3-vl-8b-thinking | CNY 0.012/1,000 tokens |
Qwen3-VL-4B-Instruct | qwen3-vl-4b-instruct | CNY 0.006/1,000 tokens |
Qwen2.5-VL-72B-Instruct | qwen2.5-vl-72b-instruct | CNY 0.05/1,000 tokens |
Qwen2.5-VL-32B-Instruct | qwen2.5-vl-32b-instruct | CNY 0.02/1,000 tokens |
Qwen2.5-VL-7B-Instruct | qwen2.5-vl-7b-instruct | CNY 0.01/1,000 tokens |
Image generation models – Wan
For the training workflow, see Fine-tune image generation models. After training completes, deploy the new model before calling it.
|
Method |
Billed by training tokens |
|
Formula |
Model training fee = Total training tokens × Training unit price (Billing unit: per 1,000 tokens) |
|
Model |
Code |
Training price (per 1K tokens) |
|
Wan image generation |
wan2.7-image-pro |
CNY 0.08 |
|
Wan image generation |
wan2.7-image |
CNY 0.08 |
Video generation models – Wan
For the training workflow, see Fine-tuning video generation models. After training completes, deploy the new model before calling it.
|
Method |
Billed by training tokens |
|
Formula |
Model training fee = Total training tokens × Training unit price (Billing unit: per 1,000 tokens) |
|
Model |
Code |
Training price (per 1K tokens) |
|
Wan image-to-video (first frame-based) |
wan2.7-i2v |
CNY 2 |
|
wan2.2-i2v-flash |
CNY 0.06 |
|
|
wan2.5-i2v-preview |
CNY 0.32 |
|
|
Image-to-video (first and last frame-based) |
wan2.2-kf2v-flash |
CNY 0.06 |
Deployment billing
Text generation models: Qwen
Time-based billing (Provisioned Throughput)
Cost = Usage Duration × (Input TPM Unit Price × Input TPM + Output TPM Unit Price × Output TPM)
For the pay-as-you-go method, usage is billed hourly, and the unit price is based on the hourly rates in the table below. For the subscription method, usage is billed daily, and the unit price is based on the daily rates in the table below.
-
Subscription orders take effect immediately after payment. An N-day subscription is valid until 23:59 on the Nth day. If an order is placed after 22:00, the expiration date is automatically extended by one day.
-
After a subscription order expires, the service is stopped after a 2-hour grace period. After the service is stopped, the resources are retained for 14 hours and then released.
-
Subscription orders cannot be terminated early.
-
For the pay-as-you-go method, if your account has an overdue payment, the deployed resources are retained and continue to be billed for 24 hours, during which the service remains available. After 24 hours, the system stops billing, and the model deployment enters an overdue state. The underlying resources are deleted, but the model deployment task is retained. After you pay the overdue amount, the system reallocates resources, restores the service, and resumes billing. To stop incurring charges, you must delete the model deployment task. Billing stops after the task is successfully deleted.
If the model input exceeds the maximum input tokens or the purchased TPM, the call automatically switches to the pay-as-you-go mode for the current model. In this case, inference performance may decrease and will be subject to the public traffic control of the current snapshot model in the workspace. Costs are charged based on the model invocation (pay-as-you-go) standard.
-
In this case, the API call returns a header that contains
x-dashscope-ptu-overflow:true. -
To view TPM statistics, go to Model Monitoring (Beijing).
For the specific refund rules for scale-in scenarios (downgrades), see Refund rules for downgrades.
Qwen
|
Model name |
Model code |
Max input tokens |
Pay-as-you-go input Per 10k TPM/hour |
Pay-as-you-go output Per 1k TPM/hour |
Subscription input Per 10k TPM/day |
Subscription output Per 1k TPM/day |
|
Qwen3.7-Max-2026-05-20 |
qwen3.7-max-2026-05-20 |
256K |
CNY 28.8 |
CNY 8.64 |
CNY 345.6 |
CNY 103.68 |
|
Qwen3.7-Plus-2026-05-26 |
qwen3.7-plus-2026-05-26 |
256K |
CNY 4.8 |
CNY 1.92 |
CNY 57.6 |
CNY 23.04 |
|
Qwen3.6-Plus-2026-04-02 |
qwen3.6-plus-2026-04-02 |
128K |
CNY 4.8 |
CNY 2.88 |
CNY 57.6 |
CNY 34.56 |
|
Qwen3.5-Plus-2026-04-20 |
qwen3.5-plus-2026-04-20 |
128K |
CNY 1.92 |
CNY 1.15 |
CNY 23.04 |
CNY 13.82 |
|
Qwen3-Max-2025-09-23 |
qwen3-max-2025-09-23 |
128K |
CNY 7.68 |
CNY 3.08 |
CNY 92.16 |
CNY 36.96 |
|
Qwen-Flash-2025-07-28 |
qwen-flash-2025-07-28 |
128K |
CNY 0.36 |
CNY 0.36 |
CNY 4.32 |
CNY 4.32 |
|
Qwen-Plus-2025-12-01 |
qwen-plus-2025-12-01 |
128K |
CNY 1.92 |
Non-thinking mode: CNY 0.48 Thinking mode: CNY 1.92 |
CNY 23.04 |
Non-thinking mode: CNY 5.76 Thinking mode: CNY 23.04 |
DeepSeek
|
Model name |
Model code |
Max input tokens |
Pay-as-you-go input Per 10k TPM/hour |
Pay-as-you-go output Per 1k TPM/hour |
Subscription input Per 10k TPM/day |
Subscription output Per 1k TPM/day |
|
DeepSeek-v4-Flash |
deepseek-v4-flash |
256K |
CNY 3.6 |
CNY 0.72 |
CNY 43.2 |
CNY 8.64 |
|
DeepSeek-v4-Pro |
deepseek-v4-pro |
256K |
CNY 43.2 |
CNY 8.64 |
CNY 518.4 |
CNY 103.68 |
|
DeepSeek-v3.2 |
deepseek-v3.2 |
64K |
CNY 7.2 |
CNY 1.08 |
CNY 86.4 |
CNY 12.96 |
|
DeepSeek-v3 |
deepseek-v3 |
64K |
CNY 7.2 |
CNY 2.88 |
CNY 86.4 |
CNY 34.56 |
Qwen-VL
|
Model name |
Model code |
Max input tokens |
Pay-as-you-go input Per 10k TPM/hour |
Pay-as-you-go output Per 1k TPM/hour |
Subscription input Per 10k TPM/day |
Subscription output Per 1k TPM/day |
|
Qwen3-VL-Plus-2025-09-23 |
qwen3-vl-plus-2025-09-23 |
128K |
CNY 2.4 |
CNY 2.4 |
CNY 28.8 |
CNY 28.8 |
More models
|
Model name |
Model code |
Max input tokens |
Pay-as-you-go input Per 10k TPM/hour |
Pay-as-you-go output Per 1k TPM/hour |
Subscription input Per 10k TPM/day |
Subscription output Per 1k TPM/day |
|
GLM-5.1 |
glm-5.1 |
64K |
CNY 21.6 |
CNY 8.64 |
CNY 259.2 |
CNY 103.68 |
Time-based billing (Model Unit)
Cost = Usage Duration (hours) × Number of Model Units × Model Unit Price
For the pay-as-you-go method, the "Model Unit Price" is the "Hourly Price" from the table below. For the monthly subscription method, the formula is: Number of Months × Number of Model Units × Monthly Price.
-
For subscriptions, if you unsubscribe within the first month, the daily unit price (≈ monthly unit price / 30) is charged at 1.2 times the standard rate. Usage for less than a day is billed as a full day.
For the Model Unit pay-as-you-go method, computing power resources are allocated on a first-come, first-served basis. A full refund is issued if the purchase is unsuccessful.
Text generation
Qwen
|
Model name |
Model code |
Model unit specification |
Hourly price (CNY) Minimum billing unit: minute |
Monthly price (CNY) Minimum billing unit: day |
|
Qwen3.7-Plus-2026-05-26 |
qwen3.7-plus-2026-05-26 |
MU3 x 8 |
CNY 1,096 |
CNY 527,752 |
|
Qwen3.6-35B-A3B |
qwen3.6-35b-a3b |
MU8 x 1 |
CNY 47 |
CNY 22,400 |
|
MU9 x 1 |
CNY 51 |
CNY 24,600 |
||
|
Qwen3.6-27B |
qwen3.6-27b |
MU9 x 1 |
CNY 51 |
CNY 24,600 |
|
Qwen3.6-Flash-2026-04-16 |
qwen3.6-flash-2026-04-16 |
MU1 x 2 |
CNY 108 |
CNY 52,236 |
|
Qwen3.6-Plus-2026-04-02 |
qwen3.6-plus-2026-04-02 |
MU1 x 8 MU1 x 16 (PD separation mode) |
CNY 432 PD separation mode: CNY 864 |
CNY 208,944 PD separation mode: CNY 417,888 |
|
Qwen3.5-397B-A17B |
qwen3.5-397b-a17b |
MU2 x 8 |
CNY 504 |
CNY 240,288 |
|
MU3 x 8 MU3 x 16 (PD separation mode) |
CNY 1,096 PD separation mode: CNY 2,192 |
CNY 527,752 PD separation mode: CNY 1,055,504 |
||
|
MU6 x 16 |
CNY 400 |
CNY 193,424 |
||
|
Qwen3.5-122B-A10B |
qwen3.5-122b-a10b |
MU1 x 4 |
CNY 216 |
CNY 104,472 |
|
MU2 x 8 |
CNY 504 |
CNY 240,288 |
||
|
MU6 x 16 |
CNY 400 |
CNY 193,424 |
||
|
MU9 x 2 |
CNY 102 |
CNY 49,200 |
||
|
Qwen3.5-35B-A3B |
qwen3.5-35b-a3b |
MU1 x 2 |
CNY 108 |
CNY 52,236 |
|
MU2 x 8 |
CNY 504 |
CNY 240,288 |
||
|
MU8 x 1 |
CNY 47 |
CNY 22,400 |
||
|
MU9 x 1 |
CNY 51 |
CNY 24,600 |
||
|
Qwen3.5-27B |
qwen3.5-27b |
MU9 x 1 |
CNY 51 |
CNY 24,600 |
|
Qwen3.5-9B |
qwen3.5-9b |
MU8 x 1 |
CNY 47 |
CNY 22,400 |
|
MU9 x 1 |
CNY 51 |
CNY 24,600 |
||
|
Qwen3.5-Flash-2026-02-23 |
qwen3.5-flash-2026-02-23 |
MU1 x 2 |
CNY 108 |
CNY 52,236 |
|
Qwen3.5-Plus-2026-02-15 |
qwen3.5-plus-2026-02-15 |
MU1 x 16 (PD separation mode) |
PD separation mode: CNY 864 |
PD separation mode: CNY 417,888 |
|
MU3 x 8 MU3 x 16 (PD separation mode) |
CNY 1,096 PD separation mode: CNY 2,192 |
CNY 527,752 PD separation mode: CNY 1,055,504 |
||
|
Qwen3-235B-A22B-Instruct-2507 |
qwen3-235b-a22b-instruct-2507 |
MU1 x 4 |
CNY 216 |
CNY 104,472 |
|
MU2 x 8 |
CNY 504 |
CNY 240,288 |
||
|
Qwen3-Next-80B-A3B-Instruct |
qwen3-next-80b-a3b-instruct |
MU1 x 2 |
CNY 108 |
CNY 52,236 |
|
Qwen3-32B |
qwen3-32b |
MU1 x 4 |
CNY 216 |
CNY 104,472 |
|
MU6 x 4 |
CNY 100 |
CNY 48,356 |
||
|
Qwen3-30B-A3B |
qwen3-30b-a3b |
MU9 x 2 |
CNY 102 |
CNY 49,200 |
|
Qwen3-30B-A3B-Instruct-2507 |
qwen3-30b-a3b-instruct-2507 |
MU1 x 4 |
CNY 216 |
CNY 104,472 |
|
MU2 x 8 |
CNY 504 |
CNY 240,288 |
||
|
Qwen3-8B |
qwen3-8b |
MU1 x 2 |
CNY 108 |
CNY 52,236 |
|
MU2 x 2 |
CNY 126 |
CNY 60,072 |
||
|
MU5 x 1 |
CNY 21 |
CNY 10,139 |
||
|
Qwen3-4B |
qwen3-4b |
MU1 x 2 |
CNY 108 |
CNY 52,236 |
|
MU5 x 1 |
CNY 21 |
CNY 10,139 |
||
|
Qwen3-1.7B |
qwen3-1.7b |
MU1 x 2 |
CNY 108 |
CNY 52,236 |
|
MU5 x 1 |
CNY 21 |
CNY 10,139 |
||
|
Qwen3-Embedding-0.6B |
qwen3-embedding-0.6b |
MU5 x 1 |
CNY 21 |
CNY 10,139 |
|
MU6 x 1 |
CNY 25 |
CNY 12,089 |
||
|
Qwen3-MoE-Rerank-0.6B |
qwen3-moe-rerank-0.6b |
MU5 x 1 |
CNY 21 |
CNY 10,139 |
|
Qwen3-Rerank-0.6B |
qwen3-rerank-0.6b |
MU5 x 1 |
CNY 21 |
CNY 10,139 |
|
MU6 x 1 |
CNY 25 |
CNY 12,089 |
||
|
Qwen3-Max-2025-09-23 |
qwen3-max-2025-09-23 |
MU2 x 8 |
CNY 504 |
CNY 240,288 |
|
MU3 x 8 |
CNY 1,096 |
CNY 527,752 |
||
|
Qwen3-Rerank |
qwen3-rerank |
MU5 x 1 |
CNY 21 |
CNY 10,139 |
|
Qwen2.5-72B |
qwen2.5-72b-instruct |
MU1 x 4 |
CNY 216 |
CNY 104,472 |
|
Qwen2.5-Open-Source-32B |
qwen2.5-32b-instruct |
MU1 x 4 |
CNY 216 |
CNY 104,472 |
|
Qwen2.5-open-source-14B |
qwen2.5-14b-instruct |
MU1 x 2 |
CNY 108 |
CNY 52,236 |
|
Qwen2.5-7B |
qwen2.5-7b-instruct |
MU1 x 2 |
CNY 108 |
CNY 52,236 |
|
MU5 x 1 |
CNY 21 |
CNY 10,139 |
||
|
Qwen2.5-3B |
qwen2.5-3b-instruct |
MU5 x 1 |
CNY 21 |
CNY 10,139 |
|
Qwen-Flash-2025-07-28 |
qwen-flash-2025-07-28 |
MU1 x 4 |
CNY 216 |
CNY 104,472 |
|
Qwen-Plus-2025-07-28 |
qwen-plus-2025-07-28 |
MU1 x 4 MU1 x 16 (PD separation mode) |
CNY 216 PD separation mode: CNY 864 |
CNY 104,472 PD separation mode: CNY 417,888 |
|
Qwen-Plus-2025-12-01 |
qwen-plus-2025-12-01 |
MU1 x 4 |
CNY 216 |
CNY 104,472 |
GLM
|
Model name |
Model code |
Model unit specification |
Hourly price (CNY) Minimum billing unit: minute |
Monthly price (CNY) Minimum billing unit: day |
|
GLM-5.1 |
glm-5.1 |
MU2 x 8 MU2 x 16 (PD separation mode) |
CNY 504 PD separation mode: CNY 1,008 |
CNY 240,288 PD separation mode: CNY 480,576 |
|
MU3 x 16 (PD separation mode) |
PD separation mode: CNY 2,192 |
PD separation mode: CNY 1,055,504 |
||
|
MU6 x 16 |
CNY 400 |
CNY 193,424 |
||
|
GLM-5 |
glm-5 |
MU3 x 16 (PD separation mode) |
PD separation mode: CNY 2,192 |
PD separation mode: CNY 1,055,504 |
|
GLM-4.7 |
glm-4.7 |
MU6 x 32 (PD separation mode) |
PD separation mode: CNY 800 |
PD separation mode: CNY 386,848 |
DeepSeek
|
Model name |
Model code |
Model unit specification |
Hourly price (CNY) Minimum billing unit: minute |
Monthly price (CNY) Minimum billing unit: day |
|
DeepSeek-v4-Flash |
deepseek-v4-flash |
MU1 x 8 |
CNY 432 |
CNY 208,944 |
|
DeepSeek-v3.2 |
deepseek-v3.2 |
MU2 x 16 (PD separation mode) |
PD separation mode: CNY 1,008 |
PD separation mode: CNY 480,576 |
More models
|
Model name |
Model code |
Model unit specification |
Hourly price (CNY) Minimum billing unit: minute |
Monthly price (CNY) Minimum billing unit: day |
|
MiniMax-M2.5 |
MiniMax-M2.5 |
MU1 x 16 (PD separation mode) |
PD separation mode: CNY 864 |
PD separation mode: CNY 417,888 |
|
Kimi-K2.5 |
kimi-k2.5 |
MU2 x 8 |
CNY 504 |
CNY 240,288 |
Model types:
-
Instruct - The deployed model performs inference in non-thinking mode.
-
Thinking - The deployed model performs inference in thinking mode.
Model deployment types:
-
PD separation mode - Reduces first-token latency and improves throughput.
In this deployment mode, the model inference process splits the first-token calculation (Prefill) and subsequent token calculation (Decode) stages to run on different compute nodes.
Multimodal
Qwen-VL
|
Model name |
Model code |
Model unit specification |
Hourly price (CNY) Minimum billing unit: minute |
Monthly price (CNY) Minimum billing unit: day |
|
Qwen3-VL-235B-A22B-Instruct |
qwen3-vl-235b-a22b-instruct |
MU1 x 4 |
CNY 216 |
CNY 104,472 |
|
Qwen3-VL-235B-A22B-Thinking |
qwen3-vl-235b-a22b-thinking |
MU1 x 4 |
CNY 216 |
CNY 104,472 |
|
Qwen3-VL-32B-Instruct |
qwen3-vl-32b-instruct |
MU2 x 8 |
CNY 504 |
CNY 240,288 |
|
Qwen3-VL-8B-Instruct |
qwen3-vl-8b-instruct |
MU1 x 2 |
CNY 108 |
CNY 52,236 |
|
Qwen3-VL-4B-Instruct |
qwen3-vl-4b-instruct |
MU1 x 2 |
CNY 108 |
CNY 52,236 |
|
Qwen3-VL-2B-Instruct |
qwen3-vl-2b-instruct |
MU5 x 1 |
CNY 21 |
CNY 10,139 |
|
Qwen3-VL-Embedding-2B |
qwen3-vl-embedding-2b |
MU5 x 1 |
CNY 21 |
CNY 10,139 |
|
Qwen3-VL-Flash-2025-10-15 |
qwen3-vl-flash-2025-10-15 |
MU1 x 4 |
CNY 216 |
CNY 104,472 |
|
Qwen3-VL-Plus-2025-09-23 |
qwen3-vl-plus-2025-09-23 |
MU1 x 4 |
CNY 216 |
CNY 104,472 |
|
Qwen-VL-Max-2025-08-13 |
qwen-vl-max-2025-08-13 |
MU6 x 4 |
CNY 100 |
CNY 48,356 |
|
Qwen-VL-OCR-2025-11-20 |
qwen-vl-ocr-2025-11-20 |
MU6 x 4 |
CNY 100 |
CNY 48,356 |
Qwen Omni
|
Model name |
Model code |
Model unit specification |
Hourly price (CNY) Minimum billing unit: minute |
Monthly price (CNY) Minimum billing unit: day |
|
Qwen3.5-Omni-Flash |
qwen3.5-omni-flash |
MU8 x 1 |
CNY 47 |
CNY 22,400 |
|
MU9 x 1 |
CNY 51 |
CNY 24,600 |
||
|
Qwen3.5-Omni-Plus |
qwen3.5-omni-plus |
MU9 x 8 |
CNY 408 |
CNY 196,800 |
Model types:
-
Instruct - The deployed model performs inference in non-thinking mode.
-
Thinking - The deployed model performs inference in thinking mode.
-
Instruct/Thinking - You can choose whether to enable thinking mode when deploying the model.
Speech synthesis
CosyVoice
|
Model name |
Model code |
Model unit specification |
Hourly price (CNY) |
Monthly price (CNY) |
|
cosyvoice-v3-flash |
cosyvoice-v3-flash |
MU5 |
CNY 21 |
CNY 10,139 |
By model token usage
Cost = Number of Input Tokens × Input Unit Price + Number of Output Tokens × Output Unit Price (Minimum billing unit: 1 token)
-
Billing by model token usage is only supported after you have completed Supervised Fine-Tuning (SFT) for the following foundation models and you have obtained a custom model.
Qwen
|
Foundation model |
Model code |
Input CNY/1k tokens |
Output CNY/1k tokens |
|
Qwen3-32B |
qwen3-32b |
CNY 0.002 |
Non-thinking mode: CNY 0.008 Thinking mode: CNY 0.02 |
|
Qwen3-14B |
qwen3-14b |
CNY 0.001 |
Non-thinking mode: CNY 0.004 Thinking mode: CNY 0.01 |
|
Qwen3-8B |
qwen3-8b |
CNY 0.0005 |
Non-thinking mode: CNY 0.002 Thinking mode: CNY 0.005 |
|
Qwen2.5-72B |
qwen2.5-72b-instruct |
CNY 0.004 |
CNY 0.012 |
|
Qwen2.5-32B |
qwen2.5-32b-instruct |
CNY 0.002 |
CNY 0.006 |
|
Qwen2.5-Open-Source-14B |
qwen2.5-14b-instruct |
¥0.001 |
CNY 0.003 |
|
Qwen2.5-Open-Source-7B |
qwen2.5-7b-instruct |
CNY 0.0005 |
CNY 0.001 |
Qwen-VL
|
Foundation model |
Model code |
Input CNY/1k tokens |
Output CNY/1k tokens |
|
Qwen3-VL-8B-Instruct |
qwen3-vl-8b-instruct |
CNY 0.0005 |
CNY 0.002 |
|
Qwen2.5-VL-72B |
qwen2.5-vl-72b-instruct |
CNY 0.016 |
CNY 0.048 |
|
Qwen2.5-VL-32B |
qwen2.5-vl-32b-instruct |
CNY 0.008 |
CNY 0.024 |
|
Qwen2.5-VL-7B |
qwen2.5-vl-7b-instruct |
CNY 0.002 |
CNY 0.005 |
Image generation models – Wan
Deployment is free. Invocations are billed at the standard rate of the fine-tuned base model. For the training workflow, see Fine-tune image generation models.
|
Model ID |
LoRA Deployment & Invocation Price |
|
wan2.7-image-pro |
CNY 0.50/image |
|
wan2.7-image |
CNY 0.20/image |
FAQ
Q: When does billing for model deployment start?
A: Billing starts when the model status changes to Running. No charges apply during Deploying, Overdue Payment, or Deployment Failed.
For monthly subscriptions, the billing period starts when the status changes to Running.
Q: Am I charged if I cancel a training job?
A: Yes. If you cancel training manually, you are charged for all tokens processed before cancellation. Training jobs interrupted by system errors or other non-user causes are not charged.
Q: How do I view invocation statistics for a deployed model?
A: Visit the Model Monitoring (Beijing), Model Monitoring (Virginia), or Model Monitoring (Singapore) page.
