aufklarer/Chatterbox-Multilingual-MLX-fp16

🤗 Hugging Face sourcetext-to-speechmit643M params1.3 GBsafetensors✓ 3 checksumsupdated today
Submit in one command

Run it next to your model folder. It makes the torrent, checks your files against Hugging Face, and submits it. You just start seeding and paste your key from your account. It only reads your files and never changes them. Read the script first if you like.

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo aufklarer/Chatterbox-Multilingual-MLX-fp16 ./model-folder
Needs a seeder →

aufklarer/Chatterbox-Multilingual-MLX-fp16

Multilingual Chatterbox (Resemble AI, MIT) converted to MLX in genuine fp16 for Apple-silicon inference. Unlike mlx-community/chatterbox-fp16 (stored as F32, ~2.6 GB), this bundle stores all float tensors as fp16 (~1.3 GB) with an identical key layout.

Zero-shot voice cloning from a short reference clip across 23 languages (incl. Arabic and Hindi). Component prefixes: ve.* (voice encoder), t3.* (text→speech-token T3), s3gen.* (token→waveform S3Gen).

Note: requires the S3Tokenizer weights from mlx-community/S3TokenizerV2, downloaded automatically at runtime.

Use with mlx-audio

pip install -U mlx-audio
mlx_audio.tts.generate --model aufklarer/Chatterbox-Multilingual-MLX-fp16 --text "[ar] مرحبا" --ref_audio reference.wav

Converted with models/chatterbox/export/convert.py (speech-models).