Skip to content
Explore models

Audio

Kling TTS

Generate speech from text prompts and different voices using the Kling TTS model, which leverages advanced AI techniques to create high-quality text-to-speech.

  • Text to speech
Create with Kling TTS
Preview for Kling TTS

Catalog preview for Kling TTS. Review your own result before publishing.

A starting point.

Write the words or musical direction and upload a reference if this operation requires one.

Before you publish.

Listen to the complete audio and check pronunciation, lyrics, timing and sound quality.

Review the quote before running. A failed generation releases its reserved credits. Download the results you want to keep; stored media is retained for 30 days.

How credits work · Use the API · Content guidelines

Inputs and settings

Use the current input form for this endpoint, then review the credit quote before generating.

Supported inputs

These fields come from the current studio form for kling-video/v1/tts. Other modes may require different inputs.

Input types, requirements and limits for kling-video/v1/tts
InputTypeRequirementOptions and limits
textTextRequiredMax. characters: 500
voice_idTextOptional—
voice_speedNumberOptional≥ 0.8 · ≤ 2

Check the full model schema for all options and cross-field rules. Request a quote for your selected settings before generating.

Use Kling TTS through the API.

Start with this model ID for the endpoint featured on this page:

kling-video/v1/tts

Input requirements belong to this model ID. Check a different mode’s schema separately when switching modes.

Check the current model schema for supported inputs and options. Other tasks may use different model IDs.

From quote to result.

Request a credit quote for your input and settings before submitting a generation. A sample output or the price of another model is not a quote for your request.

Follow the generation workflow, then poll the job for its status and result. The API and studio use the same balance.