qwen-flash-character

更新时间:
复制 MD 格式

The Qwen Role-Playing Model Series is specifically optimized for muti-language anthropomorphic interaction scenarios. It demonstrates advanced capabilities in character consistency maintenance, context-aware dialogue progression, and empathetic engagement, enabling precise personalized character embodiment. This version significantly enhances Japanese linguistic localization (including dialects and honorifics), human-like role-playing authenticity, narrative coherence control, and scenario-based cognitive intelligence.

Inference Service Provider

The inference service provider for qwen-flash-character is Alibaba Cloud Model Studio.

Model Capabilities

China (Beijing)

Capability Support Capability Support

Input Modality

Text

Output Modality

Text

Model Experience

Supported

Function Calling

Unsupported

Structured Outputs

Unsupported

Web Search

Supported

Prefix Completion

Unsupported

Context Caching

Supported

Batch Inference

Unsupported

Fine-tuning

Unsupported

Singapore

Scope: International

Capability Support Capability Support

Input Modality

Text

Output Modality

Text

Model Experience

Supported

Function Calling

Unsupported

Structured Outputs

Unsupported

Web Search

Unsupported

Prefix Completion

Unsupported

Context Caching

Supported

Batch Inference

Unsupported

Fine-tuning

Unsupported

Context Limits

Parameter Value Parameter Value

Max Input Length

8000

Max Output Length

4096

Context Window

8192

Pricing

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.

China (Beijing)

Billing Item Price (CNY) Unit

Input

0.25

Per 1M tokens

Output

1.5

Per 1M tokens

Input(Implicit Cache)

0.05

Per 1M tokens

Singapore

Scope: International

Billing Item Price (CNY) Unit

Input

0.375

Per 1M tokens

Output

2.998

Per 1M tokens

Input(Implicit Cache)

0.075

Per 1M tokens

Rate Limits

China (Beijing)

Parameter Value

RPM (Requests Per Minute)

120

TPM (Tokens Per Minute)

500,000

Singapore

Scope: International

Parameter Value

RPM (Requests Per Minute)

120

TPM (Tokens Per Minute)

500,000

Snapshot Versions

qwen-flash-character-2026-02-26

, , , , , .2026226.

Inference Service Provider

The inference service provider for qwen-flash-character-2026-02-26 is Alibaba Cloud Model Studio.

Model Capabilities

Capability Support Capability Support

Input Modality

Text

Output Modality

Text

Model Experience

Supported

Function Calling

Unsupported

Structured Outputs

Unsupported

Web Search

Unsupported

Prefix Completion

Unsupported

Context Caching

Supported

Batch Inference

Unsupported

Fine-tuning

Unsupported

Context Limits

Parameter Value Parameter Value

Max Input Length

8192

Max Output Length

4096

Context Window

8192

Pricing

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.

China (Beijing)

Billing Item Price (CNY) Unit

Input

0.18

Per 1M tokens

Output

1.5

Per 1M tokens

Input(Implicit Cache)

0.036

Per 1M tokens

Rate Limits

China (Beijing)

Parameter Value

RPM (Requests Per Minute)

120

TPM (Tokens Per Minute)

500,000