qwen3.8-27b

Updated at:

The Qwen3.8 27B native vision-language dense model builds upon the 3.6-27B version, with key improvements in coding and office productivity capabilities across both text and visual modalities. It enables more reliable end-to-end completion of complex tasks, delivering consistently trustworthy results.

Inference Service Provider

The inference service provider for qwen3.8-27b is Alibaba Cloud Model Studio.

Model Capabilities

Capability

Support

Capability

Support

Input Modality

Image Text Video

Output Modality

Text

Model Experience

Supported

Function Calling

Supported

Structured Outputs

Supported

Web Search

Supported

Prefix Completion

Supported

Context Caching

Supported

Batch Inference

Unsupported

Fine-tuning

Unsupported

Context Limits

Parameter

Value

Parameter

Value

Max Input Length

991808

Max Output Length

131072

Max Input Length (Thinking Mode)

983616

Max Output Length (Thinking Mode)

131072

Context Window

1000000

Max Chain-of-Thought Length

262144

The supported model length may vary depending on different combinations of API input parameters.

Pricing

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.

China (Beijing)

Billing Item

Price (CNY)

Unit

Input

3

Per 1M tokens

Output

12

Per 1M tokens

Input(Implicit Cache)

0.6

Per 1M tokens

Explicit Cache Creation

3.75

Per 1M tokens

Explicit Cache Read

0.3

Per 1M tokens

Singapore

Scope: International

Billing Item

Price (CNY)

Unit

Input

3.646

Per 1M tokens

Output

21.875

Per 1M tokens

Input(Implicit Cache)

0.729

Per 1M tokens

Explicit Cache Creation

4.557

Per 1M tokens

Explicit Cache Read

0.365

Per 1M tokens

Rate Limits

China (Beijing)

Parameter

Value

RPM (Requests Per Minute)

5,000

TPM (Tokens Per Minute)

5,000,000

Singapore

Scope: International

Parameter

Value

RPM (Requests Per Minute)

5,000

TPM (Tokens Per Minute)

5,000,000