Model Studio feature updates

Updated at:

Announcements

Feature updates

2026

August

Date

Feature module

Feature highlight

Description

August 4

Platform

Model upgrade notice

Model upgrade noticeLearn more

August 3

Platform

Alibaba Cloud Model Studio Personal Token Plan user benefits upgrade

Alibaba Cloud Model Studio Personal Token Plan user benefits upgradeLearn more

July

Date

Feature module

Feature highlight

Description

July 21

Platform

Memory commercialization notice

Memory commercialization noticeLearn more

July 16

Platform

Legacy Enterprise Knowledge Base sunset notice

Legacy Enterprise Knowledge Base sunset notice. Learn more

July 16

Platform

Managed Agent commercialization notice

Managed Agent commercialization notice. Learn more

July 14

Platform

GLM-5.2 Fast mode price reduction notice

GLM-5.2 Fast mode price reduction notice. Learn more

July 13

Platform

Gateway change notice

Gateway change notice. Learn more

July 10

Platform

Sunset notice for select legacy models

Sunset notice for select legacy models. Learn more

July 9

Platform

Delisting notice for Model Studio products: Xiyan GBI and Xiyan GBI Service Package

Delisting notice for Model Studio products: Xiyan GBI and Xiyan GBI Service Package. Learn more

July 6

Model service

Extended sunset schedule for select legacy models

The sunset dates for select legacy models have been extended. Learn more

June

Date

Feature module

Feature highlight

Description

June 28

Model service

qwen-turbo resource plan retirement

The qwen-turbo resource plan will stop accepting calls on July 13. Learn more

June 9

Model evaluation

Adds leaderboard and comprehensive evaluation

The model evaluation service now includes a leaderboard and comprehensive evaluation, supporting scoring methods such as BLEU_4. Learn More

June 5

Model import

Availability on the international site

The model import feature is now available on the international site, supporting the import of LoRA fine-tuned models from OSS. Learn More

June 3

Model import

Model import API launched

The API Reference now includes a new section for model import, covering APIs for import tasks and custom model objects. Learn More

June 3

API authentication

New documentation for temporary API keys

New documentation explains how to generate temporary API keys to avoid exposing permanent API keys in untrusted environments. Learn More

June 1

Responses API

Asynchronous calls

The Responses API now supports asynchronous calls. You can submit long-running tasks by setting background=true and then poll for the results. Learn More

June 1

Applications - Tongyi Tingwu Agent

Discounted ASR resource packs

The Tongyi Tingwu Agent now offers discounted ASR resource packs, designed for use cases such as automotive sales and customer profiling. Learn More

June 1

Spring AI for Alibaba Cloud

New documentation: Calling Model Studio applications

New documentation shows how to use the Spring AI for Alibaba Cloud framework to call agents and workflows in Model Studio. Learn More

May

Date

Feature module

Feature

Description

May 31

model tuning

New reinforcement learning training (invitation-only)

Introduces reinforcement learning (RL) training to optimize model strategies using reward signals. This feature is currently invitation-only. Learn more

May 28

model tuning

Support for image generation models

Model tuning supports custom training for image generation models, such as Wan / Wanx. Learn more

May 26

multimodal translation

Tongyi Multimodal Translation API reference now available

The API reference for Tongyi Multimodal Translation is now available, with APIs for text, image, document, and webpage translation.

May 25

model compression

New model compression module

The new model compression module uses quantization algorithms to convert fine-tuned, full-precision models into low-precision versions, reducing deployment costs. Learn more

May 21

Application components - data connection

New ListCategory API for data connection

The new ListCategory API for data connection lets you query the category list of the source application. Learn more

May 21

Tingwu Agent

WebSocket protocol support for industrial command transcription

The Tingwu Agent supports the WebSocket protocol for transcribing industrial commands. Documentation for this interaction protocol is now available. Learn more

May 15

Application - data connection (formerly application data)

New ChangeParseSetting API for data connection

The new ChangeParseSetting API for data connection lets you configure parsing methods by file type. Learn more

May 15

Model - API

Text generation API endpoint updated

The text generation API endpoint now includes categories for the OpenAI Responses and Anthropic Messages interfaces. Learn more

May 11

Application invocation API

New DashScope API for agent applications

The new DashScope API reference for agent applications is now available. It supports single-turn and multi-turn conversations, streaming responses, file-based Q&A, and visual understanding. Learn more

May 8

Token Plan

Team management for Token Plan Team Edition

The Token Plan Team Edition now includes team management features, supporting SSO/DingTalk logon, seat allocation, and Credits usage monitoring. Learn more

May 4

model tuning

Zero-code safety and compliance enhancement

The model tuning workflow for text generation models now includes a zero-code option to enhance the safety and compliance of large models. Learn more

April

Date

Feature module

Feature

Description

April 28

Application - multimodal interaction development kit

Launched the server-side Java SDK for the multimodal interaction development kit

The new server-side Java SDK includes documentation covering downloads, installation, key APIs, and code examples. Learn more

April 24

Knowledge base RAG

Added logging and monitoring to the knowledge base

All retrieval calls from the knowledge base are now delivered to Log Service, supporting auditing, troubleshooting, statistics, and alerts. Learn more

April 23

Model - API

Added support for EventBridge HTTP callbacks and RocketMQ for asynchronous tasks

Asynchronous tasks can now push completion events through EventBridge, eliminating the need for polling. Learn more

April 14

Application - multimodal interaction development kit

Launched the mobile Android SDK for the multimodal interaction development kit

The new mobile Android SDK includes documentation covering downloads, installation, key APIs, and code examples. Learn more

April 9

Multimodal interaction development kit

Added a license mode to the RTOS C SDK

The RTOS C SDK for the multimodal interaction development kit now supports a license mode, covering semi-managed and fully managed access on embedded hardware. Learn more

April 7

Application component - prompt engineering

Launched the prompt engineering API for application components

A new prompt engineering module and its API are now available as an application component, allowing you to manage prompt templates. Learn more

April 1

Application Publishing and Sharing - UI designer

Launched the UI designer

The Model Studio UI designer is now available. It integrates Mobi low-code capabilities, allowing you to build web applications with a drag-and-drop interface. Learn more

March

Date

Feature module

Feature

Description

March 20

Memory

Memory 2.0 is now available

Introduces long-term memory: automatically extracts memory segments and user profiles from conversations, and supports sharing a single Memory across multiple applications. Learn more

March 19

Qwen Web Search Agent

Web Search Agent now available

Available in the Marketplace with a product introduction, API reference, and usage instructions. Learn more

March 19

Quanmiao - PPT generation

PPT generation API for Quanmiao now available

The Quanmiao solution now includes a complete API set for PPT generation, with interfaces to manage templates, documents, and generated works. Learn more

February

Date

Feature module

Feature

Description

February 28

Multimodal Interaction Development Kit

New Linux C++ SDK

A new Linux C++ SDK is now available for the Multimodal Interaction Development Kit. The documentation provides download and installation instructions, and key interface examples. Learn more

February 22

Client Development Tools

New programming tools, including Kilo CLI

The client development tools now include Kilo CLI, accessible through Token Plan, Coding Plan, and pay-as-you-go. Learn more

February 6

Application - Tongyi Multimodal Translation

New Web Translation JSSDK

The newly available web translation JSSDK lets you make your website multilingual simply by inserting a script. Learn more

February 6

Plugins

Image Generation plugin now available on the international site

The official plugin list on the international site now includes text_to_image (Image Generation). Learn more

February 6

MCP

New official MCP service

The official MCP service is now live, supporting both in-platform integration and access by third-party services such as Amap Maps. Learn more

February 6

Multimodal Interaction Development Kit

New Android/iOS Lite SDKs

The Multimodal Interaction Development Kit now offers new Android and iOS Lite SDKs and a voice list. Learn more

February 6

Voice Cloning API

New documentation for CosyVoice Voice Cloning API

Detailed documentation for the CosyVoice Voice Cloning API is now available, covering how to create and query voice cloning tasks. Learn more

February 5

Application - Knowledge base

New ranking model and instruction intervention mode for knowledge base retrieval

The Retrieve interface now includes a ranking model option and an instruction intervention mode, enhancing the platform's retrieval capabilities. Learn more

January

Date

Feature module

Feature

Description

January 23

Model deployment

Deploy new pre-trained models via the API with time-based billing per Model Unit (MU)

The model deployment API now supports deploying pre-trained models such as qwen-flash and qwen-plus, and introduces time-based billing per Model Unit (MU). Learn More

January 22

Model fine-tuning

Support for visual understanding (VL) models

Model fine-tuning now supports the visual understanding (VL) model type, enabling you to perform customized training on multimodal models. Learn More

January 21

Application - Multimodal Interaction Development Kit

You can now configure the Multimodal Interaction Development Kit

The application configuration page now lets you configure four modules: voice interaction, understanding and generation, skills, and agents. Learn More

January 21

Model fine-tuning

Support for video generation models

Model fine-tuning now supports the video generation model type, enabling you to perform customized training on video models such as the Wan series. Learn More

2025

December

Date

Feature module

Feature

Description

December 22

model usage

New dashboard for free tier and usage statistics

Provides a centralized view of each model's free tier, including used/remaining quotas and call volume statistics. model usage

December 15

model evaluation

New leaderboard feature

You can now select multiple models at once and generate a comparison leaderboard based on evaluation metrics like BLEU and ROUGE for easy side-by-side performance comparison. model evaluation

December 15

model evaluation

New evaluator types

Supports comprehensive model performance assessment across mainstream evaluation scenarios, including string matching, text similarity calculation, model-based and manual classification, and model-based scoring. model evaluation

November

Date

Feature module

Feature

Description

November 6

model monitoring

Added support for viewing model inference logs

You can now view historical conversations between users and models, which helps with prompt optimization and troubleshooting. model monitoring

October

Date

Feature module

Feature

Description

October 24

model deployment

New model unit deployment method (time-based billing)

Key advantages ofmodel unitdeployment include:

  • Flexible performance adjustment

  • High service stability

  • Predictable fixed costs

October 21

Model fine-tuning and deployment

New:SFT model fine-tuning and deployment for Qwen3-VL models

Qwen3-VL-8B-Instruct and Qwen3-VL-8B-Thinking now support SFT model fine-tuning and deployment of fine-tuned models.

SFT (Supervised Fine-Tuning) supports both full-parameter fine-tuning and efficient fine-tuning (LoRA) methods.

September

Date

Feature module

Feature

Description

September 12

Model fine-tuning and deployment

DPO preference training support for Qwen2.5 and Qwen3 series

The Qwen3-32B, 14B, 8B models and the Qwen2.5-72B, 32B, 14B, 8B models now support DPO preference training.

DPO preference training uses negative feedback to reduce hallucination and better align model outputs with human preferences.

August

Date

Feature module

Feature

Description

August 14

Model fine-tuning and deployment

Qwen2.5-VL models now support fine-tuning and deployment

The Qwen2.5-VL-72B, 32B, and 7B models now support model fine-tuning and deployment.

July

Date

Feature module

Feature

Description

July 29

free tier

New "stop when free tier is depleted" feature

The stop when free tier is depleted feature, when enabled, prevents users from making calls after their free tier is exhausted. It returns an error with the code AllocationQuota.FreeTierOnly and prevents additional charges.

June

Date

Feature module

Feature

Description

June 5

model monitoring

New alerts and notifications

  • The system notifies you or your operations team when a configured metric, such as call statistics or performance indicators, becomes abnormal. model monitoring

April

Date

Feature module

Feature

Description

April 18

model monitoring

New advanced monitoring mode

  • Advanced monitoring provides minute-level data refresh with low latency and records failed model call details, including 4xx and 5xx error counts. model monitoring.

February

Date

Feature module

Feature

Description

February 8

Billing

Billing adjustment for DeepSeek series models

February 7

Billing

Price reduction for the qwen-max model

  • The input price for the qwen-max model has been reduced by 88%, and the output price has been reduced by 84%:

    • For real-time calls, the input price is reduced from CNY 0.02 per 1,000 tokens to CNY 0.0024 per 1,000 tokens, and the output price is reduced from CNY 0.06 per 1,000 tokens to CNY 0.0096 per 1,000 tokens.

    • For batch calls, the input price is reduced from CNY 0.01 per 1,000 tokens to CNY 0.0012 per 1,000 tokens, and the output price is reduced from CNY 0.03 per 1,000 tokens to CNY 0.0048 per 1,000 tokens.

  • For the latest billing details, see the Model Studio console.

January

Date

Feature module

Feature

Description

January 22

Billing

Price adjustment for some Qwen series models

  • The prices for the qwen2.5-14b-instruct and qwen2.5-7b-instruct models are reduced by 50%:

    • The input price for the qwen2.5-14b-instruct model is reduced to CNY 0.001 per 1,000 tokens, and the output price is reduced to CNY 0.003 per 1,000 tokens.

    • The input price for the qwen2.5-7b-instruct model is reduced to CNY 0.0005 per 1,000 tokens, and the output price is reduced to CNY 0.001 per 1,000 tokens.

  • The qwen2.5-3b-instruct and qwen2-vl-72b-instruct models have transitioned from a free trial to a paid service:

    • The input price for the qwen2.5-3b-instruct model is CNY 0.0003 per 1,000 tokens, and the output price is CNY 0.0009 per 1,000 tokens.

    • The input price for the qwen2-vl-72b-instruct model is CNY 0.016 per 1,000 tokens, and the output price is CNY 0.048 per 1,000 tokens.

  • For the latest billing details, see the Model Studio console.

January 21

model monitoring

New model monitoring capability

  • Model monitoring tracks model usage and performance, helping you identify issues and optimize performance. model monitoring.

2024

December
DateModuleFeatureDescription

December 31

billing

Price reduction for Qwen-VL models

December 27

model fine-tuning

SFT fine-tuning for qwen2.5-7b-instruct

  • The qwen2.5-7b-instruct model now supports both full-parameter and efficient SFT fine-tuning.

December 27

model deployment

Model deployment now supports pay-per-invocation billing

  • The pay-per-invocation billing option now supports the deployment of fine-tuned models, including qwen2.5-7B, 14B, 32B, 72B, and qwen2-7B.

December 24

ModelScope

New context cache feature

  • Context cache reduces redundant inference computations, improving response speed and lowering costs without affecting quality. context cache.

Currently, only the qwen-plus model is supported.

December 20

OpenAI API compatibility

Batch jobs now support task notifications.

Three more models are now supported for batch jobs.

December 20

ModelScope

New search_options parameter

  • You can now use the search_options parameter to configure settings for web search, such as search sources and the number of results. This parameter is available for the qwen-max, qwen-plus, and qwen-turbo models. For usage instructions, see the Qwen API Documentation - DashScope.
November

Date

Feature module

Feature

Description

November 7

data center

Data processing now supports canvas orchestration

  • With canvas orchestration, you can flexibly combine data cleaning nodes and data augmentation nodes to easily create complex data processing workflows.Data cleaning or augmentation

October

Date

Module

Feature

Description

October 17

Model Square

model sunsetting

  • The qwen-max-longcontext model will be sunset on October 17, 2024. We recommend migrating to qwen-max.

September

Date

Feature module

Feature

Description

September 20

billing

Prepaid support for model inference

  • Savings plans (prepaid) are now available to offset inference fees beyond the free tier for Qwen, Tongyi Farui, Baichuan-Open Source, ChatGLM, and OpenNLU models. Prepaid (savings plan).

September 19

billing

Free tier validity period extended

Price reduction for select Qwen series models

Model Square

New model types added

  • New models have been added, including Qwen Coder, Qwen Audio, and the open-source version of Qwen 2.5. You can find these new models in the model list.

throttling

Support for custom quota increase requests

September 18

Homepage

Model Studio console redesign

  • The console has been redesigned for a clearer and more user-friendly interface. Explore the new capabilities in the model playground.

September 10

billing

Model training billing adjustment

  • Model training billing updated: hybrid training is now a billable item. Formula: (Total training tokens + Total hybrid training tokens) × epoch × Unit price. billable item.

September 6

Model Tools

New models supported for model training

  • qwen-turbo-0624 and qwen-plus-0723 now support model training. For billing rules, see the Model Studio console.

September 6

Model Square

Billing adjustment for large models

  • The qwen-vl-max-0809, qwen-vl-max-0201, sensevoice-v1, and paraformer-realtime-v2 models are now billable. For billing rules, see the Model Studio console.

August
DateFeature moduleFeatureDescription

August 29

Model data

Enhanced model data import

  • Improved error details for failed and partially successful imports.
  • Status polling is now available for model data publishing and imports.

model data.

August 27

OpenAI API compatibility

New Vision invocation mode

  • The new Vision invocation mode lets you use the Qwen vision model by simply adjusting the API-KEY, BASE_URL, and model parameters in your existing framework. This enables seamless integration with OpenAI interfaces and tools. To get started, see OpenAI Compatibility - Vision.

August 23

OpenAI API compatibility

New batch invocation mode

This mode is ideal for high-volume, asynchronous inference tasks, such as data processing and cleansing with large models, large-scale data analysis, and automated extraction.

August 16

Workspace

New trusted workspace

August 8

Model tools

New model evaluation method

August 7

Model tools

New models for dedicated deployment

  • The qwen2-7b-instruct model now supports custom training and dedicated deployment.
  • The qwen2-72b-instruct model now supports custom training and dedicated deployment. To get started, see Fine-tune models in the console.
July

Date

Feature module

Feature

Description

July 30

data processing

New operator types for data processing

  • Added the detoxification operator, which automatically detects, analyzes, and removes sensitive or non-compliant content from your data. data cleaning or augmentation

July 23

system management

Updated query capability for call statistics

  • Sub-workspaces can now view only the details of their own business data. model monitoring

July 19

model square

Added best practice documentation

  • The console now includes best practice code samples.

July 19

model square

Upgraded Qwen series models

  • Performance for the qwen-plus and qwen-turbo models has been improved. text generation.

July 2

model tools

Added support for model fine-tuning for qwen-vl-plus

June

Date

Feature module

Feature

Description

June 13

data center

Data augmentation enhancements

  • Data augmentation now supports parameter and prompt configuration for more effective data processing.

May

Date

Feature module

Feature

Description

May 31

data center

data management supports data filtering

  • data management now supports batch deletion and data filtering by file name or file status.

May 21

commercialization

Price reduction for select models and separate billing for input and output

  • Input and output are now billed separately for select Qwen models, with an across-the-board price reduction.

commercialization

Free tier policy update for new users

  • New users receive a total of 36 million free tokens for mainstream Tongyi Qwen large models. (This promotion is no longer available.)

May 3

Home

A redesigned Home page that improves the onboarding experience

  • You can make up to 100 API calls to the max, turbo, and plus models before activating the service.

  • The streamlined Home page offers a guided onboarding experience. Product introduction .

Model Square

A completely redesigned, more intuitive, and clearer interface

  • The qwen-max model has been fully upgraded.

  • The redesigned interface supports filtering models by series and provides a global search filter.

  • View API call examples with a single click to quickly invoke models.

  • A one-stop workflow now covers the entire model lifecycle: invocation, model fine-tuning, billing, authorization, and service activation.

Model Playground

New users can try out select models directly online

  • Try models before activating the service.

  • Optimized the model selection mechanism to support selection by series and custom models.

  • View API call examples with one click.

  • The VL and Wan models are now fully available for online testing. Introduction to Model Playground.

model fine-tuning

New model types available for training

  • SFT and deployment are now available for the qwen-plus and qwen1.5-72B models.

  • SFT and deployment are now available for the hundred-billion-parameter open-source model, qwen1.5-110B. Fine-tune Models in the Console.

data processing

New automated data processing capabilities and enhanced data security

  • Automated data processing now includes operators for data deduplication, sensitive information removal, and more.

  • One-click processing for SFT data is now available to improve data quality and address issues such as privacy leaks and data duplication in training sets.

model data

You can now create datasets of different types, such as training and evaluation sets, and manage multiple data versions. Published datasets can be used for model fine-tuning or evaluation.

  • New multi-version management for model data enables a more efficient workflow.

  • Optimized the data display and configuration logic.

April

Date

Feature module

Feature

Description

April 25

Service agreement

Service agreement update

April 16

model experience center

New models available for testing

  • The Qianwen-VL model is now available for testing in the model experience center. Overview .

  • The Wanx model is now available for testing in the model experience center.

  • Improved the auto-scroll-to-bottom behavior in the model experience center.

April 3

model square

Model deprecation and upgrade

March

Date

Functional module

Feature

Description

March 15

large model service

Module updates for Alibaba Cloud Model Studio

March 15

Model Hub

New large models and billing adjustments

February

Date

Feature module

Feature

Description

February 8

model experience

Added the model experience center

February 5

Model Square

Added large model types

  • Official and third-party large models are now available.

February 4

model deployment

Added the Qwen-Plus model deployment type

  • Supports independent deployment for the Qwen-Plus large model.

January

Date

Feature module

Feature

Description

January 18

model training

New option for mixed training

  • Mixed training is now supported. You can specify a data ratio to automatically mix data for training.

January 11

model deployment

New offline status

  • You can now delete a dedicated model instance that is offline.

January 11

model training

New models supported for training

  • You can now train the Qwen-Plus and Qwen-Turbo models.