API reference

Updated at:

Streaming text-to-speech (TTS) accepts text in multiple segments within a session and returns synthesized audio as a binary stream. It is suitable for reading aloud text generated word by word or in segments by a large language model. This topic describes the limits, endpoints, interaction flow, request parameters, and supported voices.

Features and limits

  • Supports audio output in PCM, WAV, MP3, and OPUS formats.

  • Supports voices for different scenarios and styles, as well as speech rate, pitch, and volume settings. These parameters can be configured only before the input stream starts.

  • Text can be sent in multiple segments within a session while audio is received with low latency. For real-time playback, use a player that supports the audio encoding format and streaming input.

  • Input text must be UTF-8 encoded. Each text submission can contain up to 10,000 characters, and the total for a session cannot exceed 100,000 characters. A Chinese character, English letter, punctuation mark, or space within a sentence counts as one character.

  • Streaming text input does not support SSML tags. To use SSML, select a voice that supports SSML and submit the complete text in a single request through the CosyVoice long-text speech synthesis API. This method still returns audio as a stream, but does not accept multiple text submissions in the same session.

Billing and concurrency limits

  • Streaming TTS is available only in the commercial version, with no free trial. Activate the commercial version before use. For activation instructions and billing methods, see Billing methods.

  • For the billable items, see Billable items.

  • For concurrency limits, see Concurrency and QPS limits.

Endpoints

Access type

Description

URL

Public network

Available to all servers. The SDK uses a public endpoint by default.

wss://nls-gateway-cn-beijing.aliyuncs.com/ws/v1

ECS internal network

ECS instances in the China (Beijing) region can access the service through a virtual private cloud (VPC). Instances in a classic network cannot access the speech service through AnyTunnel over the internal network. Using the internal endpoint incurs no public data transfer fees for the ECS instance.

ws://nls-gateway-cn-beijing-internal.aliyuncs.com:80/ws/v1

Access the nearest region

Streaming TTS supports access through nls-gateway.aliyuncs.com. We recommend this domain for end users. The service resolves it to a server in the nearest region based on the client's location. For example, a request from Beijing is routed to a server in the China (Beijing) region, with the same effect as specifying nls-gateway-cn-beijing.aliyuncs.com.

Interaction flow

image
  1. Authenticate: The client uses an NLS token for authentication when establishing a WebSocket connection. For instructions, see Obtain a token.

  2. Start synthesis: Send StartSynthesis to configure parameters such as the Appkey, voice, and audio format. Text can be sent after the server returns SynthesisStarted.

  3. Send text and receive audio: Send text segments through RunSynthesis. The server returns synthesized audio as binary data. The client saves or plays the audio in the order received. When timestamps are enabled, the client can also receive subtitle events.

  4. Finish input and wait for completion: After sending all text, send StopSynthesis. Continue receiving the remaining audio until the server returns SynthesisCompleted.

Request parameters

Use the SDK's FlowingSpeechSynthesizer object to configure synthesis parameters. The format and sample_rate values in the following table are protocol values. The Java SDK also supports the corresponding OutputFormatEnum and SampleRateEnum enumerations.

Parameter

Type

Required

Description

appkey

String

Yes

The Appkey of a project created in the console.

voice

String

Yes

The voice to use. For valid values and the sample rates, timestamps, and other capabilities supported by each voice, see the voice list.

format

String

No

The audio encoding format. Valid values: pcm, wav, mp3, and opus. Default value: pcm.

sample_rate

Integer

No

The audio sample rate in Hz. Valid values: 8000, 16000, 24000, and 48000. Select a rate supported by the voice. Default value: 16000. The corresponding Java SDK enum values are SAMPLE_RATE_8K, SAMPLE_RATE_16K, SAMPLE_RATE_24K, and SAMPLE_RATE_48K.

volume

Integer

No

The volume, from 0 to 100. Default value: 50.

speech_rate

Integer

No

The speech rate, from -500 to 500. Default value: 0.

pitch_rate

Integer

No

The pitch, from -500 to 500. Default value: 0.

enable_subtitle

Boolean

No

Specifies whether to enable word-level timestamps. Select a voice that supports timestamps. For usage instructions, see Timestamps.

enable_aigc_tag

Boolean

No

Specifies whether to add an implicit AIGC label to the generated audio. When set to true, the label is embedded in WAV, MP3, or OPUS audio. Default value: false.

aigc_propagator

String

No

The ContentPropagator field in the implicit AIGC label, which identifies the content distributor. This parameter takes effect only when enable_aigc_tag is true. The default value is the Alibaba Cloud UID.

aigc_propagate_id

String

No

The PropagateID field in the implicit AIGC label, which uniquely identifies a distribution event. This parameter takes effect only when enable_aigc_tag is true. The default value is the Task ID of the current speech synthesis request.

Responses

The server returns audio in binary WebSocket messages and synthesis status in JSON messages. When synthesis is complete, the server returns a SynthesisCompleted event. The following example shows only the main fields:

{
    "header": {
        "message_id": "05450bf69c53413f8d88aed1ee60****",
        "task_id": "640bc797bb684bd6960185651307****",
        "namespace": "FlowingSpeechSynthesizer",
        "name": "SynthesisCompleted",
        "status": 20000000,
        "status_text": "Gateway:SUCCESS:Success."
    },
    "payload": {
        "measureType": "TextLength",
        "measureLength": 12
    }
}

Field

Type

Description

header.name

String

The event name. SynthesisCompleted indicates that synthesis is complete.

header.status

Integer

The status code. 20000000 indicates success.

header.status_text

String

The status description.

payload.measureType

String

The metering type.

payload.measureLength

Integer

The length of the text synthesized in the request. This is metering information in the response, not a request parameter.

Voice list

Select a voice based on requirements such as the language, sample rate, and timestamp support, and use its voice parameter value.

Name

voice parameter value

Type

Scenario

Supported languages

Supported sample rates (Hz)

Word/sentence-level timestamp support

Erhua support

Voice quality

Abin

abin

Cantonese-accented Mandarin

Conversational digital human

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz/48 kHz

No

No

Standard Edition

Zhixiaobai

zhixiaobai

Mandarin female voice

Conversational digital human

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz/48 kHz

No

Yes

Standard Edition

Zhixiaoxia

zhixiaoxia

Mandarin female voice

Conversational digital human

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz/48 kHz

No

Yes

Standard Edition

Zhixiaomei

zhixiaomei

Mandarin female voice

Live streaming digital human

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz

Yes

Yes

Standard Edition

Zhigui

zhigui

Mandarin female voice

Live streaming digital human

Supports Chinese and mixed Chinese-English scenarios

8K/16K

Yes

Yes

Standard Edition

Zhishuo

zhishuo

Mandarin male voice

Customer service digital human

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz

Yes

Yes

Standard Edition

Aixia

aixia

Mandarin female voice

Customer service digital human

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz

Yes

Yes

Standard Edition

Cally

cally

American English female voice

Spoken English conversational digital human

Supports only pure English scenarios

8K or 16K

Yes

Yes

Standard Edition

Zhifeng_emo

zhifeng_emo

Multi-emotional male voice

General

Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz

Yes

Yes

Standard Edition

Zhibing_emo

zhibing_emo

Multi-emotional male voice

Common scenarios

Pure Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

Yes

Standard Edition

Zhimiao_emo

zhimiao_emo

Multi-emotional female voice

Chinese-English

Chinese and English scenarios

8K or 16K

Yes

Yes

Standard Edition

Zhimi_emo

zhimi_emo

Multi-emotional female voice

General

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

No

Standard Edition

Zhiyan_emo

zhiyan_emo

Multi-emotional female voice

General scenarios

Chinese and mixed Chinese-English scenarios

8 K/16 K

Yes

No

Standard Edition

Zhibei_emo

zhibei_emo

Multi-emotional child's voice

General

Chinese and mixed Chinese-English scenarios

8K/16K

Yes

No

Standard Edition

Zhitian_emo

zhitian_emo

Multi-emotional female voice

Common Scenarios

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

No

Standard Edition

Xiaoyun

xiaoyun

Standard female voice

General

Chinese and mixed Chinese-English scenarios

8K or 16K

No

No

Lite Edition

Xiaogang

xiaogang

Standard male voice

General

Chinese and mixed Chinese-English scenarios

8K or 16K

No

No

Lite Edition

Ruoxi

ruoxi

Gentle female voice

Common scenarios

Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz

No

No

Standard Edition

Siqi

siqi

Gentle female voice

Common scenarios

Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

Sijia

sijia

Standard female voice

General

Chinese and mixed Chinese-English scenarios

8K, 16K, or 24K

No

No

Standard Edition

Sicheng

sicheng

Standard male voice

Common scenarios

Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

Aiqi

aiqi

Gentle female voice

General

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

No

Standard Edition

Aijia

aijia

Standard female voice

General

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

No

Standard Edition

Aicheng

aicheng

Standard male voice

General

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

No

Standard Edition

Aida

aida

Standard male voice

General

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

No

Standard Edition

Ninger

ninger

Standard female voice

Common scenarios

Pure Chinese scenarios

8 kHz/16 kHz/24 kHz

No

No

Standard Edition

Ruilin

ruilin

Standard female voice

General scenarios

Pure Chinese scenarios

8 kHz/16 kHz/24 kHz

No

No

Standard Edition

Siyue

siyue

Gentle female voice

Customer service

Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

Aiya

aiya

Strict female voice

Customer service

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

No

Standard Edition

Aimei

aimei

Sweet female voice

Customer service

Chinese and mixed Chinese-English scenarios

8K/16K

Yes

No

Standard Edition

Aiyu

aiyu

Natural female voice

Customer service

Chinese and mixed Chinese-English scenarios

8 K/16 K

Yes

No

Standard Edition

Aiyue

aiyue

Gentle female voice

Customer service

Chinese and mixed Chinese-English scenarios

8 K/16 K

Yes

No

Standard Edition

Aijing

aijing

Strict female voice

Customer service

Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz

Yes

No

Standard Edition

Xiaomei

xiaomei

Sweet female voice

Customer service

Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz

No

No

Standard Edition

Aina

aina

Zhejiang-accented Mandarin female voice

Customer service

Pure Chinese scenarios

8K or 16K

Yes

No

Standard Edition

Yina

yina

Zhejiang-accented Mandarin female voice

Customer service

Pure Chinese scenarios

8 kHz/16 kHz/24 kHz

No

No

Standard Edition

Sijing

sijing

Strict female voice

Customer service

Pure Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

Sitong

sitong

Child's voice

Child's voice

Pure Chinese scenarios

8 kHz/16 kHz/24 kHz

No

No

Standard Edition

Xiaobei

xiaobei

Young female voice

Child's voice

Pure Chinese scenarios

8K, 16K, or 24K

Yes

No

Standard Edition

Aitong

aitong

Child's voice

Child's voice

Pure Chinese scenarios

8K or 16K

Yes

No

Standard Edition

Aiwei

aiwei

Young female voice

Child's voice

Pure Chinese scenarios

8 K/16 K

Yes

No

Standard Edition

Aibao

aibao

Young female voice

Child's voice

Pure Chinese scenarios

8 kHz/16 kHz

Yes

No

Standard Edition

Harry

harry

British English male voice

English Scenario

English scenario

8 K/16 K

No

No

Standard Edition

Abby

abby

American English female voice

English-language scenario

English

8K or 16K

Yes

No

Standard Edition

Andy

andy

American English male voice

English scenario

English-language scenario

8 K/16 K

Yes

No

Standard Edition

Eric

eric

British English male voice

Scenario in English

English Scenario

8K or 16K

Yes

No

Standard Edition

Emily

emily

British English female voice

Scenario: English

English Scenario

8 K/16 K

Yes

No

Standard Edition

Luna

luna

British English female voice

English Scenario

English Scenario

8 K/16 K

Yes

No

Standard Edition

Luca

luca

British English male voice

English-language scenario

English scenario

8K or 16K

Yes

No

Standard Edition

Wendy

wendy

British English female voice

English

English

8 kHz/16 kHz/24 kHz

No

No

Standard Edition

William

william

British English male voice

English Scenario

English Scenario

8 kHz/16 kHz/24 kHz

No

No

Standard Edition

Olivia

olivia

British English female voice

English-language scenario

English scenario

8 kHz/16 kHz/24 kHz

No

No

Standard Edition

Shanshan

shanshan

Cantonese female voice

Dialect

Standard Written Cantonese (Simplified) and mixed Cantonese-English scenarios

8 kHz/16 kHz/24 kHz

No

No

Standard Edition

Aiyuan

aiyuan

Warm, caring female voice

Literary

Chinese and mixed Chinese-English scenarios

8K/16K

Yes

Yes

Premium Edition

Aiying

aiying

Soft and cute child's voice

Literary scenario

Chinese and mixed Chinese-English scenarios

8K/16K

Yes

Yes

Premium Edition

Aixiang

aixiang

Magnetic male voice

Literary

Chinese and mixed Chinese-English scenarios

8K / 16K

Yes

Yes

Premium Edition

Aimo

aimo

Emotional male voice

Literature-based scenario

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Premium Edition

Aiye

aiye

Young male voice

Literary

Chinese and mixed Chinese-English scenarios

8 K/16 K

Yes

Yes

Premium Edition

Aiting

aiting

Radio female voice

Literature scenario

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Premium Edition

Aifan

aifan

Emotional female voice

Literature-based scenario

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Premium Edition

Lydia

lydia

Bilingual English-Chinese female voice

English

English and mixed English-Chinese scenarios

8K or 16K

Yes

No

Standard Edition

Xiaoyue

chuangirl

Sichuanese female voice

Dialect

Chinese and mixed Chinese-English scenarios

8K or 16K

No

No

Standard Edition

Aishuo

aishuo

Natural male voice

Customer service

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

No

Standard Edition

Qingqing

qingqing

Female voice with a Taiwan (China) accent

Dialect

Pure Chinese scenarios

8 K/16 K

No

No

Standard Edition

Cuijie

cuijie

Northeastern Mandarin female voice

Dialect

Pure Chinese scenarios

8K or 16K

Yes

Yes

Standard Edition

Xiaoze

xiaoze

Hunanese heavy-accented male voice

Dialect

Pure Chinese scenarios

8K or 16K

No

No

Standard Edition

Ainan

ainan

Advertisement male voice

Literature scenario

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Premium Edition

Aihao

aihao

News male voice

Literary

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Premium Edition

Aiming

aiming

Humorous male voice

Literary

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Premium Edition

Aixiao

aixiao

News female voice

Literary

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Premium Edition

Aichu

aichu

Gourmet-style male voice

Literary scenario

Chinese and mixed Chinese-English scenarios

8 K/16 K

Yes

Yes

Premium Edition

Aiqian

aiqian

News female voice

Literature scenario

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Premium Edition

Zhi Xiang

tomoka

Japanese female voice

Multilingual

Pure Japanese scenarios

8K/16K

Yes

No

Standard Edition

Tomoya

tomoya

Japanese male voice

Multilingual

Pure Japanese scenarios

8K or 16K

Yes

No

Standard Edition

Annie

annie

American English female voice

English Scenario

Pure English scenarios

8K/16K

Yes

No

Standard Edition

Aishu

aishu

News male voice

Literary

Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Premium Edition

Airu

airu

Newscast female voice

Literature scenario

Chinese and mixed Chinese-English scenarios

8K/16K

Yes

Yes

Premium Edition

Jiajia

jiajia

Cantonese female voice

Dialect scenarios

Standard Written Cantonese (Simplified) and mixed Cantonese-English scenarios

8K or 16K

Yes

No

Standard Edition

Indah

indah

Indonesian female voice

Multilingual

Pure Indonesian scenarios

8 K/16 K

No

No

Standard Edition

Peach

taozi

Cantonese female voice

Dialect

Supports Standard Written Cantonese (Simplified) and mixed Cantonese-English scenarios

8 kHz/16 kHz

Yes

No

Standard Edition

Sales Associate

guijie

Friendly female voice

General

Supports Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Standard Edition

Stella

stella

Intellectual female voice

Common scenarios

Supports Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Standard Edition

Stanley

stanley

Calm male voice

General

Supports Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Standard Edition

Kenny

kenny

Calm male voice

General scenarios

Supports Chinese and mixed Chinese-English scenarios

8 K/16 K

Yes

Yes

Standard Edition

Rosa

rosa

Natural female voice

General

Supports Chinese and mixed Chinese-English scenarios

8 K/16 K

Yes

Yes

Standard Edition

Farah

farah

Malay female voice

Multilingual

Supports only pure Malay scenarios

8K/16K

No

No

Standard Edition

Mashu

mashu

Children's drama male voice

General

General

8K/16K

Yes

No

Standard Edition

Zhiqi

zhiqi

Gentle female voice

Ultra-high definition

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz/48 kHz

Yes

No

Premium Edition

Zhichu

zhichu

Gourmet-style male voice

Ultra-high definition

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz/48 kHz

Yes

Yes

Premium Edition

Xiaoxian

xiaoxian

Friendly female voice

Live streaming

Supports Chinese and mixed Chinese-English scenarios

8K/16K

Yes

Yes

Standard Edition

Yuer

yuer

Children's drama female voice

Common scenarios

Supports only pure Chinese scenarios

8K or 16K

Yes

No

Standard Edition

Maoxiaomei

maoxiaomei

Energetic female voice

Live streaming

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz

Yes

Yes

Standard Edition

Zhixiang

zhixiang

Magnetic male voice

Ultra-high definition

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz/48 kHz

Yes

No

Premium Edition

Zhijia

zhijia

Standard female voice

Ultra-high definition

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz/48 kHz

Yes

No

Premium Edition

Zhinan

zhinan

Advertisement male voice

Ultra-high definition

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz/48 kHz

Yes

No

Premium Edition

Zhiqian

zhiqian

News female voice

Ultra-high definition

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz/48 kHz

Yes

No

Premium Edition

Zhiru

zhiru

Newscast female voice

Ultra-high definition

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz/48 kHz

Yes

No

Premium Edition

Zhide

zhide

Newscast male voice

Ultra-high-definition scenarios

Supports Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz/24 kHz/48 kHz

Yes

No

Premium Edition

Zhifei

zhifei

Passionate commentary

Ultra-high-definition scenarios

Supports Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

No

Premium Edition

Aifei

aifei

Passionate commentary

Live streaming

Supports Chinese and mixed Chinese-English scenarios

8 K/16 K

Yes

Yes

Standard Edition

Yaqun

yaqun

Store broadcast

Live streaming

Supports Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Standard Edition

Qiaowei

qiaowei

Store broadcast

Live streaming

Supports Chinese and mixed Chinese-English scenarios

8K/16K

Yes

Yes

Standard Edition

Dahu

dahu

Northeastern Mandarin male voice

Scenarios with dialects

Supports Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Standard Edition

ava

ava

American English (Female)

English scenario

Supports only pure English scenarios

8K or 16K

Yes

No

Standard Edition

Zhilun

zhilun

Suspense narration

Ultra-high definition

Supports Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

No

Premium Edition

Ailun

ailun

Suspense narration

Live streaming

Supports Chinese and mixed Chinese-English scenarios

8K or 16K

Yes

Yes

Standard Edition

Jielidou

jielidou

Soothing child's voice

Child's voice

Supports only pure Chinese scenarios

8 kHz/16 kHz

Yes

Yes

Standard Edition

Zhiwei

zhiwei

Young female voice

Ultra-high definition

Supports only pure Chinese scenarios

8 kHz/16 kHz/24 kHz/48 kHz

Yes

No

Premium Edition

Laotie

laotie

Northeastern buddy

Live streaming

Supports only pure Chinese scenarios

8K/16K

Yes

Yes

Standard Edition

Laomei

laomei

Hawking female voice

Live streaming

Supports only pure Chinese scenarios

8K/16K

Yes

Yes

Standard Edition

Aikan

aikan

Tianjin dialect male voice

Scenarios for dialects

Supports only pure Chinese scenarios

8K or 16K

Yes

Yes

Standard Edition

Tala

tala

Filipino female voice

Multilingual

Supports only Filipino scenarios

8K or 16K

No

No

Standard Edition

Zhitian

zhitian

Sweet female voice

Common scenarios

Supports Chinese and mixed Chinese-English scenarios

8 K/16 K

Yes

No

Premium Edition

Zhiqing

zhiqing

A girl speaking a dialect from Taiwan (China)

Dialect

Supports only pure Chinese scenarios

8K or 16K

Yes

No

Premium Edition

Tien

tien

Vietnamese female voice

Multilingual

Supports only Vietnamese scenarios

8 K / 16 K

No

No

Standard Edition

Becca

becca

American English customer service female voice

American English

Supports only pure English scenarios

8K or 16K

No

No

Standard Edition

Kyong

Kyong

Korean female voice

Korean

Korean

8K/16K

No

No

Standard Edition

masha

masha

Russian female voice

Russian

Russian

8K or 16K

No

No

Standard Edition

camila

camila

Spanish female voice

Spanish

Spanish

8 kHz/16 kHz

No

No

Standard Edition

perla

perla

Italian female voice

Italian scenario

Italian

8 kHz/16 kHz

No

No

Standard Edition

Zhimao

zhimao

Mandarin female voice

Live streaming

Chinese

8 kHz/16 kHz

Yes

No

Standard Edition

Zhiyuan

zhiyuan

Mandarin female voice

General

Chinese

8 kHz/16 kHz

Yes

No

Standard Edition

Zhiya

zhiya

Mandarin female voice

Customer service

Chinese

8 kHz/16 kHz

Yes

No

Standard Edition

Zhiyue

zhiyue

Mandarin female voice

Common scenarios

Chinese

8 kHz/16 kHz

Yes

No

Standard Edition

Zhida

zhida

Mandarin male voice

Common scenarios

Chinese and mixed Chinese-English scenarios

8 kHz/16 kHz

Yes

No

Standard Edition

Zhistella

zhistella

Mandarin female voice

Common scenarios

Chinese

8 kHz/16 kHz

Yes

No

Standard Edition

Kelly

kelly

Hong Kong Cantonese female voice

Dialect

Hong Kong Cantonese

8 kHz/16 kHz

Yes

No

Standard Edition

clara

clara

French female voice

General

French

8 kHz/16 kHz

No

No

Standard Edition

hanna

hanna

German female voice

General

German

8 kHz/16 kHz

No

No

Standard Edition

waan

waan

Thai female voice

General scenarios

Thai

8 kHz/16 kHz

No

No

Standard Edition

betty

betty

American English female voice

General scenarios

American English

8 kHz/16 kHz

Yes

No

Standard Edition

beth

beth

American English female voice

Common scenarios

American English

8 kHz/16 kHz

Yes

No

Standard Edition

cindy

cindy

American English female voice

Common scenarios

American English

8 kHz/16 kHz

Yes

No

Standard Edition

donna

donna

American English female voice

Common Scenarios

American English

8 kHz/16 kHz

Yes

No

Standard Edition

eva

eva

American English female voice

General

American English

8 kHz/16 kHz

Yes

No

Standard Edition

brian

brian

American English male voice

General

American English

8 kHz/16 kHz

Yes

No

Standard Edition

david

david

American English male voice

Common Scenarios

American English

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

abby_ecmix

abby_ecmix

American English female voice

Common scenarios

English and mixed English-Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

annie_ecmix

annie_ecmix

American English female voice

Common scenarios

English and mixed English-Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

andy_ecmix

andy_ecmix

American English male voice

Common scenarios

English and mixed English-Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

ava_ecmix

ava_ecmix

American English female voice

Common scenarios

English and mixed English-Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

betty_ecmix

betty_ecmix

American English female voice

Common scenarios

English and mixed English-Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

beth_ecmix

beth_ecmix

American English female voice

General

English and mixed English-Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

brian_ecmix

brian_ecmix

American English male voice

Common scenarios

English and mixed English-Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

cindy_ecmix

cindy_ecmix

American English female voice

Common scenarios

English and mixed English-Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

cally_ecmix

cally_ecmix

American English female voice

Common scenarios

English and mixed English-Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

donna_ecmix

donna_ecmix

American English female voice

General

English and mixed English-Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

david_ecmix

david_ecmix

American English male voice

General

English and mixed English-Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition

eva_ecmix

eva_ecmix

American English female voice

General scenarios

English and mixed English-Chinese scenarios

8 kHz/16 kHz/24 kHz

Yes

No

Standard Edition