$ AICostSaverbeta

Qwen3 TTS

Input, lowest

—

Output, lowest

—

Providers

1

Providers & prices

ProviderInputOutputContextLast sync
D DigitalOcean — — 32,768

About this model

A 1.7B-parameter text-to-speech model covering 10 languages Hugging Face with remarkably low latency. Use this when you need production-ready voice synthesis with instant customization — it supports 3-second voice cloning, natural language instruction control over tone and emotion, and streaming generation