zenlm/zen3-tts-0.6B

🤗 Hugging Face sourcetext-to-speechapache-2.0915M params1.8 GBsafetensorsHF checksums availableupdated today
No torrent yet

Zen3 TTS 0.6B

Compact ~0.6B Zen3 text-to-speech model at 12 Hz, sized for edge and latency-sensitive synthesis.

Derived by fine-tuning Qwen/Qwen3-TTS-12Hz-0.6B-Base (Alibaba Cloud, Apache-2.0).

Weights

This repository contains the model weights: model.safetensors (talker) plus a speech_tokenizer/ module (12 Hz codec), config and tokenizer files.

The model uses the qwen3_tts architecture and loads with transformers (>= 4.57). It is API-compatible with the upstream base — follow the inference recipe on the base model card Qwen/Qwen3-TTS-12Hz-0.6B-Base.

Provenance

Fine-tuned from Qwen/Qwen3-TTS-12Hz-0.6B-Base (Apache-2.0). See NOTICE for full attribution.