The enhanced version of the Qwen QwQ reasoning model, trained on the Qwen2.5 model, has significantly improved its reasoning capabilities through reinforcement learning. The model's core metrics in mathematics and coding (e.g., AIME 24/25, LiveCodeBench) as well as some general metrics (e.g., IFEval, LiveBench) have reached the level of the full version of DeepSeek-R1.
Inference Service Provider
The inference service provider for qwq-plus is Alibaba Cloud Model Studio.
Model Capabilities
China (Beijing)
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality |
Text |
Output Modality |
Text |
Model Experience |
Function Calling |
||
Structured Outputs |
Web Search |
||
Prefix Completion |
Context Caching |
||
Batch Inference |
Fine-tuning |
Singapore
Scope: International
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality |
Text |
Output Modality |
Text |
Model Experience |
Function Calling |
||
Structured Outputs |
Web Search |
||
Prefix Completion |
Context Caching |
||
Batch Inference |
Fine-tuning |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length |
98304 |
Max Output Length |
8192 |
Context Window |
131072 |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
China (Beijing)
| Billing Item | Price (CNY) | Unit |
|---|---|---|
Input |
1.6 |
Per 1M tokens |
Output |
4 |
Per 1M tokens |
Input(Batch File) |
0.8 |
Per 1M tokens |
Output(Batch File) |
2 |
Per 1M tokens |
Singapore
Scope: International
| Billing Item | Price (CNY) | Unit |
|---|---|---|
Input |
5.871 |
Per 1M tokens |
Output |
17.614 |
Per 1M tokens |
Rate Limits
China (Beijing)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
600 |
TPM (Tokens Per Minute) |
1,000,000 |
Singapore
Scope: International
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
60 |
TPM (Tokens Per Minute) |
100,000 |