Type
Speech synthesis (TTS)
Start free — $2 credit
StepFun API model
Step TTS 2 — expressive speech synthesis with official and cloned voices, speed and volume control, pronunciation mapping, and multiple output formats.
Specs
Transparent USD pricing. No mainland-China account or phone required. The displayed media rate may include a service margin covering provider input/output billing, payment processing, chargeback exposure, and operations. Any margin is included in the displayed rate and is not added separately.
Speech synthesis (TTS)
Text input, audio output
Audio, Text-to-Speech, Voice Cloning, Emotion, Pronunciation Control
$0.3889 per 10k characters
/v1/audio/speech
Step TTS 2 — expressive speech synthesis with official and cloned voices, speed and volume control, pronunciation mapping, and multiple output formats.
Pricing
$0.3889 per 10k characters. Paid in USD. The displayed media rate may include a service margin covering provider input/output billing, payment processing, chargeback exposure, and operations. Any margin is included in the displayed rate and is not added separately. View live pricing
Quickstart
curl -X POST https://api.chinaapi-ru.com/v1/audio/speech \
-H "Authorization: Bearer sk-..." \
-H "Content-Type: application/json" \
-d '{
"model": "step-tts-2",
"voice": "cixingnansheng",
"input": "(轻声)Welcome to ChinaAPI. 今天我们来测试自然语音。",
"response_format": "mp3"
}' \
--output speech.mp3
Internal links
Use step-tts-2 through ChinaAPI with an API key and the OpenAI-compatible endpoint shown above.
step-tts-2 is $0.3889 per 10k characters, paid in USD. The live pricing page is authoritative. The displayed media rate may include a service margin covering provider input/output billing, payment processing, chargeback exposure, and operations. Any margin is included in the displayed rate and is not added separately.
No. ChinaAPI provides access without a mainland-China account or phone number and includes a $2 free trial. Check live pricing for the displayed USD rate.