# Qwen 3 TTS

> Audio AI Model by Qwen 3 Tts — available on Layer

Multilingual TTS with zero-shot voice cloning and prompt-based voice design.

Qwen 3 TTS (1.7B) is a multilingual text-to-speech model with zero-shot voice cloning. Clone a voice from a single audio sample, or design a new voice from a text prompt; the resulting speaker embedding can then narrate any text. Supports English, Chinese, Spanish, French, German, Italian, Japanese, Korean, Portuguese and Russian.

## Specifications

| Property | Value |
|----------|-------|
| Provider | Qwen 3 Tts |
| Category | Audio |
| Price | 0.03 / character count per character_count |

## Capabilities

- Text to Speech
- Auto Duration

---

[Try Qwen 3 TTS](https://next.app.layer.ai/me/session/new?baseModelId=QWEN_3_TTS) | [All Audio Models](https://layer.ai/models/audio) | [All Models](https://layer.ai/models)
