phasefield-audio/Irodori-TTS-v4.1-Anime

🤗 Hugging Face 来源text-to-speechmit766M 参数3.1 GBsafetensors✓ 6 个校验和今天更新
磁力链接🌱 1✓ 与 Hugging Face 一致

Irodori-TTS-v4.1-Anime

A Japanese text-to-speech model fine-tuned from Aratako/Irodori-TTS-v4.1-Small using anime-style speech data.

The base model's annotation pipeline is not publicly documented, so the fine-tuning data was annotated independently. Consequently, caption conditioning and emoji controls may behave differently from the base model.

Checkpoints

The full-precision checkpoint is available at the repository root.

Quantized variants are provided in the following subdirectories:

  • int8-weight-only
  • int8-dynamic
  • int4-weight-only
  • float8-weight-only
  • float8-dynamic

For inference and installation instructions, see the original Irodori-TTS repository.

License

This model follows the same MIT License and ethical restrictions as the base model.