API reference

Updated at:

CosyVoice long-text speech synthesis converts text submitted in a single input into speech and streams the audio as binary data. With a voice that supports Speech Synthesis Markup Language (SSML), you can add background audio, set pauses, and correct pronunciation. This topic describes endpoints, the interaction process, request parameters, response events, and supported voices.

Billing and concurrency limits

  • CosyVoice long-text speech synthesis is available only in the commercial version and does not offer a free trial. Activate the commercial version before use. For activation instructions and billing methods, see Billing methods.

  • For billable items, see Billable items.

  • For concurrency limits, see Concurrency and QPS.

Features and limits

  • Supports PCM, WAV, MP3, and OPUS audio formats.

  • Accepts one text input per session. The text must use UTF-8 encoding and contain no more than 10,000 characters. Streaming text input is not supported, but synthesized audio is returned as a stream.

  • Supports voice selection and adjustments to speech rate, pitch, and volume.

  • Supports SSML and word-level timestamps, depending on the voice. For details, see the voice list.

Endpoints and authentication

Select an endpoint for the network environment and set the SDK URL to that endpoint.

Access type

Description

URL

Internet access

Available to servers with Internet connectivity.

wss://nls-gateway-cn-beijing.aliyuncs.com/ws/v1

Internal access from ECS

Available to Elastic Compute Service (ECS) instances in a virtual private cloud (VPC) in the China (Beijing) region. The classic network is not supported. Internal access does not incur Internet traffic fees for the ECS instance.

ws://nls-gateway-cn-beijing-internal.aliyuncs.com:80/ws/v1

The client authenticates with an NLS token when establishing a WebSocket connection. For instructions, see Obtain a token. The Appkey in the request identifies an Intelligent Speech Interaction project. Obtain it from the Intelligent Speech Interaction console.

Interaction process

image
  1. Authenticate: Establish a WebSocket connection using an NLS token.

  2. Configure parameters: Set the Appkey, voice, audio format, sample rate, and other parameters, and initiate synthesis.

  3. Send text: Send all text to synthesize in a single input. With an SDK, the text can be passed through the text parameter of startTts.

  4. Receive data: The server streams binary audio data. The client receives the data and saves or plays it. If timestamps are enabled, also handle subtitle events.

  5. Complete synthesis: The server sends a SynthesisCompleted event when all audio has been returned. This event does not indicate that the client has finished playing the audio.

Request parameters

The following tables describe the parameters. Configuration methods, parameter types, and defaults may differ across SDK languages. Use the interface for the SDK language in use. For example, the Java SDK uses setter methods on the StreamInputTts object and enums, whereas the Python SDK accepts synthesis parameters through startTts.

Synthesis parameters

Parameter

Type

Required

Description

appkey

String

Yes

The Appkey of the Intelligent Speech Interaction project.

voice

String

Yes

The voice name. Use a voice parameter value from the voice list.

format

Enum/String

No

The audio format: pcm, wav, mp3, or opus. Default: pcm. The Java SDK uses the OutputFormatEnum enum.

sample_rate

Enum/Integer

No

The audio sample rate in Hz: 8000, 16000, 24000, or 48000. Use a rate supported by the selected voice. The server default is 16000. An SDK may send a different default; for example, Python SDK startTts sends 24000 by default. The Java SDK uses SAMPLE_RATE_8K, SAMPLE_RATE_16K, SAMPLE_RATE_24K, and SAMPLE_RATE_48K in SampleRateEnum.

volume

Integer

No

The volume, from 0 to 100. Default: 50.

speech_rate

Integer

No

The speech rate, from -500 to 500. Default: 0.

pitch_rate

Integer

No

The pitch, from -500 to 500. Default: 0.

bit_rate

Integer

No

The audio bitrate in kbps. Applies only to OPUS. Valid range: 6 to 510. Default: 32.

enable_subtitle

Boolean

No

Set to true to enable word-level timestamps. Requires a voice that supports timestamps. Word-level timestamps are not returned unless enabled. With the Python SDK, pass {"enable_subtitle": True} through the ex parameter of startTts.

enable_aigc_tag

Boolean

No

Whether to embed an implicit AIGC identifier in the generated audio. Default: false. When set to true, the identifier is embedded in WAV, MP3, or OPUS audio.

aigc_propagator

String

No

The ContentPropagator field in the implicit AIGC identifier, which identifies the content distributor. Defaults to the Alibaba Cloud UID. Applies only when enable_aigc_tag is true.

aigc_propagate_id

String

No

The PropagateID field in the implicit AIGC identifier, which uniquely identifies a distribution activity. Defaults to the task ID of the synthesis request. Applies only when enable_aigc_tag is true.

Text to synthesize

Parameter

Type

Required

Description

text

String

Yes

The text to synthesize, in UTF-8 encoding, with no more than 10,000 characters. Separate English words with spaces. Send text only once per session.

With a voice that supports SSML, use SSML to control sentence and word segmentation, pronunciation, speed, pauses, pitch, and volume, or to add background music. For details, see SSML.

Responses

Synthesized audio is returned in binary WebSocket messages. Synthesis status and timestamps are returned as JSON events. The header.namespace field is FlowingSpeechSynthesizer.

Synthesis completion event

When the client receives SynthesisCompleted, all audio for the synthesis request has been returned. The following example shows the main fields, with redacted message_id and task_id values.

{
  "header": {
    "message_id": "05450bf69c53413f8d88aed1ee60****",
    "task_id": "640bc797bb684bd6960185651307****",
    "namespace": "FlowingSpeechSynthesizer",
    "name": "SynthesisCompleted",
    "status": 20000000,
    "status_text": "GATEWAY|SUCCESS|Success."
  },
  "payload": {
    "measureType": "TextLengthHD",
    "measureLength": 49
  }
}

Field

Type

Description

header.status

Integer

The status code. A value of 20000000 indicates success.

header.status_text

String

The status description.

payload.measureType

String

The metering type, such as TextLengthHD.

payload.measureLength

Integer

The metered character count for this request. Each Chinese character counts as two characters. Each English letter, punctuation mark, or space within a sentence counts as one character.

Word-level timestamps

Use a voice that supports timestamps and set enable_subtitle to true. For example, longxiaochun_v2 and longwan_v2 support timestamps, but longxiaochun does not. For support by voice, see the voice list.

Subtitles are returned in SentenceSynthesis and SentenceEnd events during synthesis, before the full synthesis completes. A SentenceBegin event marks the start of a sentence, with the sentence index in its payload. The main fields in subtitle events are described below.

Field

Type

Description

payload.index

Integer

The sentence index.

payload.subtitles

Array

The list of subtitles and their timestamps. The list may be empty if timestamps are not enabled or the voice does not support them. Even with timestamps enabled, some SentenceSynthesis events may contain an empty subtitle list.

payload.subtitles[].text

String

The subtitle text.

payload.subtitles[].begin_time

Integer

The subtitle start time, in milliseconds.

payload.subtitles[].end_time

Integer

The subtitle end time, in milliseconds.

Audio duration and text synchronization

Synthesized audio duration depends on the text, voice, speech rate, and other factors. Do not estimate it from the duration of another audio recording. To synchronize text with audio playback, use subtitle timestamps and the player's actual playback position, rather than distributing text progress based only on total audio duration.

The API does not return a precomputed total audio duration. When playing audio as it arrives, the complete audio is not yet available at the start. Determine the total duration from the audio data after receiving all audio. The last subtitle's end_time may differ from the total audio duration and cannot be used as a substitute.

Voice list

CosyVoice-V2

Name

Voice name

(voice parameter value)

Type

Scenario

Supported languages

Supported sample rates

Timestamp support

SSML support

Longcheng

longcheng_v2

Sunny male voice

Intelligent customer service, news broadcasting, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longhua

longhua_v2

Lively girl

Intelligent customer service, news broadcasting, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longshu

longshu_v2

Male news anchor voice

News broadcasting, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Bella2.0

loongbella_v2

Female news anchor voice

Intelligent customer service, news broadcasting, chat, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longwan

longwan_v2

Mandarin female voice

Intelligent customer service, news broadcasting, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longxiaochun

longxiaochun_v2

Gentle Sister

Intelligent customer service, news broadcasting, chat, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz/48 kHz

Yes

Yes

Longxiaoxia

longxiaoxia_v2

Gentle female voice

Intelligent customer service, news broadcasting, chat, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longhuhu

longhuhu

An innocent and carefree girl

Child's voice

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longanpei

longanpei

Young female teacher

Consumer electronics - Education and training

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longdaiyu

longdaiyu

Delicate and talented female voice

Short video dubbing

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longgaoseng

longgaoseng

Enlightened monk voice

Short video dubbing

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyingmu

longyingmu

An elegant and sophisticated woman

Customer service

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyingxun

longyingxun

A young, inexperienced man

Customer service

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyingcui

longyingcui

Stern male collection agent voice

Customer service

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyingda

longyingda

Cheerful high-pitched female voice

Customer service

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyingjing

longyingjing

An unassuming and composed woman

Customer service

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyingyan

longyingyan

A stern and righteous woman

Customer service

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyingtian

longyingtian

Gentle and sweet female voice

Customer service

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyingbing

longyingbing

A sharp and commanding woman

Customer service

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyingtao

longyingtao

A woman with a gentle and calm demeanor

Customer service

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyingling

longyingling

Mild and empathetic female voice

Customer service

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

YUMI

longyumi_v2

Proper young woman

Voice assistant

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longanran

longanran

Lively and textured female voice

Livestreaming e-commerce

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longanxuan

longanxuan

Female-Hosted Live Stream

Livestreaming e-commerce

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longanchong

longanchong

Enthusiastic Salesperson

Livestreaming e-commerce

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longanping

longanping

High-pitched female livestreamer voice

Livestreaming e-commerce

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longbaizhi

longbaizhi

Wise female narrator voice

Audiobook

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longsanshu

longsanshu

Steady and textured male voice

Audiobook

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longxiu

longxiu_v2

Erudite male storyteller voice

Audiobook

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longmiao

longmiao_v2

Cadenced female voice

Audiobook

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyue

longyue_v2

Warm and magnetic female voice

Audiobook

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longnan

longnan_v2

Wise young man

Audiobook

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyuan

longyuan_v2

Warm and healing female voice

Audiobook

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longanrou

longanrou

Gentle Best Friend

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longqiang

longqiang_v2

Romantic and Charming Woman

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longhan

longhan_v2

A caring and dedicated man

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longxing

longxing_v2

Simple and friendly

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longfeifei

longfeifei_v2

Sweet and pampered girl

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longxiaocheng

longxiaocheng_v2

Magnetic low-pitched male voice

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longzhe

longzhe_v2

The Endearingly Awkward Man

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longyan

longyan_v2

Warm Spring Breeze Woman

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longtian

longtian_v2

Magnetic-Rational Man

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longze

longze_v2

Warm and energetic male voice

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longshao

longshao_v2

A motivated and ambitious man

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longhao

longhao_v2

Passionate and Melancholy Man

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longshen

kabuleshen_v2

Talented male singer voice

Social companion

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longjielidou

longjielidou_v2

Cheerful and Mischievous Boy

Child's voice

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longling

longling_v2

A girl with a naive and expressionless face

Child's voice

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longke

longke_v2

Sweet and innocent girl

Child's voice

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longxian

longxian_v2

Bold and Cute Girl

Child's voice

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longlaotie

longlaotie_v2

A Straight Shooter

Northeastern dialect

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longjiayi

longjiayi_v2

Intellectual Cantonese female voice

Cantonese

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longtao

longtao_v2

Positive Cantonese female voice

Cantonese

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longfei

longfei_v2

Passionate and magnetic male voice

Poetry recitation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Libai

libai_v2

Ancient Immortal Poet

Poetry recitation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longjin

longjin_v2

Elegant and gentle male voice

Poetry recitation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longshuo

longshuo_v2

A skilled professional

News broadcasting

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longxiaobai

longxiaobai_v2

Steady female announcer voice

News broadcasting

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

Longjing

longjing_v2

Typical female announcer voice

News broadcasting

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

loongstella

loongstella_v2

An astute and capable woman

News broadcasting

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

Yes

Yes

loongeva

loongeva_v2

Intellectual English Woman

Overseas marketing

British English

8 kHz/16 kHz/24 kHz

No

No

loongbrian

loongbrian_v2

Steady English male voice

Overseas marketing

British English

8 kHz/16 kHz/24 kHz

No

No

loongluna

loongluna_v2

British English female voice

Overseas marketing

British English

8 kHz/16 kHz/24 kHz

No

No

loongluca

loongluca_v2

British English male voice

Overseas marketing

British English

8 kHz/16 kHz/24 kHz

No

No

loongemily

loongemily_v2

British English female voice

Overseas marketing

British English

8 kHz/16 kHz/24 kHz

No

No

loongeric

loongeric_v2

British English male voice

Overseas marketing

British English

8 kHz/16 kHz/24 kHz

No

No

loongabby

loongabby_v2

American English female voice

Overseas marketing

American English

8 kHz/16 kHz/24 kHz

No

No

loongannie

loongannie_v2

American English female voice

Overseas marketing

American English

8 kHz/16 kHz/24 kHz

No

No

loongandy

loongandy_v2

American English male voice

Overseas marketing

American English

8 kHz/16 kHz/24 kHz

No

No

loongava

loongava_v2

American English female voice

Overseas marketing

American English

8 kHz/16 kHz/24 kHz

No

No

loongbeth

loongbeth_v2

American English female voice

Overseas marketing

American English

8 kHz/16 kHz/24 kHz

No

No

loongbetty

loongbetty_v2

American English female voice

Overseas marketing

American English

8 kHz/16 kHz/24 kHz

No

No

loongcindy

loongcindy_v2

American English female voice

Overseas marketing

American English

8 kHz/16 kHz/24 kHz

No

No

loongcally

loongcally_v2

American English female voice

Overseas marketing

American English

8 kHz/16 kHz/24 kHz

No

No

loongdavid

loongdavid_v2

American English male voice

Overseas marketing

American English

8 kHz/16 kHz/24 kHz

No

No

loongdonna

loongdonna_v2

American English female voice

Overseas marketing

American English

8 kHz/16 kHz/24 kHz

No

No

loongkyong

loongkyong_v2

Korean female voice

Overseas marketing

Korean

8 kHz/16 kHz/24 kHz

No

No

loongtomoka

loongtomoka_v2

Japanese female voice

Overseas marketing

Japanese

8 kHz/16 kHz/24 kHz

No

No

loongtomoya

loongtomoya_v2

Japanese male voice

Overseas marketing

Japanese

8 kHz/16 kHz/24 kHz

No

No

CosyVoice-V1

Name

Voice name

(voice parameter value)

Type

Scenario

Supported languages

Supported sample rates

Timestamp support

SSML support

Longchen

longchen

Dubbed film male voice

Dubbed film

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longxiong

longxiong

Dubbed film male voice

Dubbed film

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longyu

longyu

Mature older sister voice

Chat, news broadcasting, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longjiao

longjiao

Mature older sister voice

Chat, news broadcasting, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longmei

longmei

Gentle female voice

Chat, news broadcasting, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longgui

longgui

Gentle female voice

Chat, news broadcasting, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longping

longping

Male sports commentator voice

Sports commentary

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longpang

longpang

Male sports commentator voice

Sports commentary

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longwu

longwu

Nonsensical male voice

Chat, news broadcasting, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longqi

longqi

Lively child's voice

Chat, news broadcasting, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longxian

longxian_normal

Sunny female voice

Chat, news broadcasting, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longfei

longfeifei

Mature female voice

Chat, news broadcasting, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longxiu

longxiu

Young male voice

Chat, news broadcasting, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longdachui

longdachui

Humorous male voice

Chat, news broadcasting, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longjiajia

longjiajia

Approachable female voice

Chat, news broadcasting, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longjiayi

longjiayi

Cantonese female voice

Chat, news broadcasting, audiobooks, in-car navigation

Cantonese and Cantonese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longtao

longtao

Cantonese female voice

Chat, news broadcasting, audiobooks, in-car navigation

Cantonese and Cantonese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longjiaxin

longjiaxin

Cantonese female voice

Chat, news broadcasting, audiobooks, in-car navigation

Cantonese and Cantonese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longcheng

longcheng

Sunny male voice

Intelligent customer service, news broadcasting, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longzhe

longzhe

Mature male voice

Chat, news broadcasting, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longnan

longnan

Young male voice

News broadcasting, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longyan

longyan

Friendly female voice

Intelligent customer service, chat, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longqiang

longqiang

Languid female voice

Chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longhua

longhua

Lively girl

Intelligent customer service, news broadcasting, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longxing

longxing

Heartwarming female voice

Intelligent customer service, chat

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longjin

longjin

Young male voice

Intelligent customer service, news broadcasting, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longhan

longhan

Young male voice

Intelligent customer service, news broadcasting, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longtian

longtian

Overbearing CEO male voice

Chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longshuo

longshuo

Steady male voice

Intelligent customer service, news broadcasting, audiobooks,

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Stella2.0

loongstella

Brisk female voice

Intelligent customer service, news broadcasting, chat, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longxiaocheng

longxiaocheng

Dapper gentleman

Intelligent customer service, news broadcasting, chat, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longxiaoxia

longxiaoxia

Gentle female voice

Intelligent customer service, news broadcasting, chat, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longxiaochun

longxiaochun

Gentle Sister

Intelligent customer service, news broadcasting, chat, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longxiaobai

longxiaobai

Female chat voice

News broadcasting, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longlaotie

longlaotie

Northeastern male voice

Chat, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longyue

longyue

Female storyteller voice

Intelligent customer service, news broadcasting, audiobooks,

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Bella2.0

loongbella

Female news anchor voice

Intelligent customer service, news broadcasting, chat, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longshu

longshu

Male news anchor voice

News broadcasting, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longjing

longjing

Serious female voice

News broadcasting, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longmiao

longmiao

Classy female voice

Intelligent customer service, news broadcasting, chat, audiobooks, in-car navigation

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longlaoli

libai

Mandarin male voice

Poetry recitation, prose, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longwan

longwan

Mandarin female voice

Intelligent customer service, news broadcasting, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longke

longke

Lively girl

Intelligent customer service, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longling

longling

Lively girl

Intelligent customer service, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longshao

longshao

Energetic male voice

Intelligent customer service, news broadcasting, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longze

longze

Sunny male voice

Intelligent customer service, news broadcasting, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No

Longhao

longhao

Warm male voice

Intelligent customer service, news broadcasting, chat, audiobooks

Chinese and Chinese-English mixed

8 kHz/16 kHz/24 kHz

No

No