Audio
Kling TTS
Generate speech from text prompts and different voices using the Kling TTS model, which leverages advanced AI techniques to create high-quality text-to-speech.
- Text to speech
Sign in to create. Review the credit cost in the studio before running.
Catalog preview for Kling TTS. Review your own result before publishing.
A starting point.
Write the words or musical direction and upload a reference if this operation requires one.
Before you publish.
Listen to the complete audio and check pronunciation, lyrics, timing and sound quality.
Review the quote before running. A failed generation releases its reserved credits. Download the results you want to keep; stored media is retained for 30 days.
Inputs and settings
Use the current input form for this endpoint, then review the credit quote before generating.
Supported inputs
These fields come from the current studio form for kling-video/v1/tts. Other modes may require different inputs.
| Input | Type | Requirement | Options and limits |
|---|---|---|---|
text | Text | Required | Max. characters: 500 |
voice_id | Text | Optional | — |
voice_speed | Number | Optional | ≥ 0.8 · ≤ 2 |
Check the full model schema for all options and cross-field rules. Request a quote for your selected settings before generating.
Use Kling TTS through the API.
Start with this model ID for the endpoint featured on this page:
kling-video/v1/tts
Input requirements belong to this model ID. Check a different mode’s schema separately when switching modes.
Check the current model schema for supported inputs and options. Other tasks may use different model IDs.
From quote to result.
Request a credit quote for your input and settings before submitting a generation. A sample output or the price of another model is not a quote for your request.
Follow the generation workflow, then poll the job for its status and result. The API and studio use the same balance.